olmo-2-1b-instruct / README.md
zakkeown's picture
Publish olmo-2-1b-instruct@1 (aether publish)
e8706cd verified
|
Raw History Blame Contribute Delete
2.91 kB
---
license: apache-2.0
base_model: allenai/OLMo-2-0425-1B-Instruct
pipeline_tag: text-generation
tags:
- aether
- core-ai
- apple
- macos
- ios
---
# olmo-2-1b-instruct
A Core AI bundle of [allenai/OLMo-2-0425-1B-Instruct](https://huggingface.co/allenai/OLMo-2-0425-1B-Instruct) for the Aether SDK (iOS and macOS 27+).
- **Source:** `allenai/OLMo-2-0425-1B-Instruct` at revision `48d788eca847d4d7548f375ad03d3c9312f6139e`, licence Apache-2.0.
The licence is included as `LICENSE`. `allenai/OLMo-2-0425-1B-Instruct` declares apache-2.0 in its model card but ships no licence file; `LICENSE` is the canonical text from apache.org.
- **Changes from the source:** converted from PyTorch to Core AI (`.aimodel`) by Aether forge (recipe
`olmo-2-1b-instruct@1`). Weights are int8-linear-perchannel (8-bit weights). The tokenizer files are the source's own.
## Variants
| Variant | Platform | Arch | Compute | Compiled | Assets | Download |
|---|---|---|---|---|---|---|
| `macos-any-gpu` | macos | any | gpu | no (specialized on first load) | `olmo_2_1b_instruct.aimodel` 1.69 GB | 1.7 GB |
| `ios-any-gpu` | ios | any | gpu | no (specialized on first load) | `olmo_2_1b_instruct.aimodel` 1.69 GB | 1.7 GB |
| `ios-h18p-gpu` | ios | h18p | gpu | yes | `olmo_2_1b_instruct.h18p.aimodelc` 1.69 GB | 1.7 GB |
## Verification
Every row is a record in `verification/` about exactly these bytes (matched by bundle digest). Reference rows
are strict T2 passes of the unquantized export on the same fixture, in `verification/reference/`.
| Variant | Tier | Result | Detail | Device | OS build | Compute | Record |
|---|---|---|---|---|---|---|---|
| `ios-any-gpu` | T0 | pass | | iPhone18,2 | 24A446 | target | `eb2601c9` |
| `ios-any-gpu` | T2 | pass | 19/19 strict; profile quantized-8bit; fixture `7b97ff11db95049d` | iPhone18,2 | 24A446 | target | `6022bca4` |
| `ios-h18p-gpu` | T0 | pass | | iPhone18,2 | 24A446 | target | `f9346e21` |
| `ios-h18p-gpu` | T2 | pass | 19/19 strict; profile quantized-8bit; fixture `7b97ff11db95049d` | iPhone18,2 | 24A446 | target | `d80ed5ef` |
| `macos-any-gpu` | T0 | pass | | Mac17,6 | 26A434 | target | `cb07238b` |
| `macos-any-gpu` | T1 | pass | | Mac17,6 | 26A434 | target | `c9a063cd` |
| `macos-any-gpu` | T2 | pass | 19/19 strict; 19/19 strict; profile quantized-8bit; fixture `7b97ff11db95049d` | Mac17,6 | 26A434 | target | `8e88f60c` |
| unquantized reference (not published) | T2 | pass | 19/19 strict; 19/19 strict; profile strict; fixture `7b97ff11db95049d` | Mac17,6 | 26A434 | target | `87dd4aec` |
The quantized profile also requires: T2 strict on the unquantized reference export (met by the reference row).
## Use
```sh
aether run olmo-2-1b-instruct --prompt "Hello"
```
```swift
import Aether
let aether = try Aether()
let chat = try await aether.chat("olmo-2-1b-instruct")
let reply = try await chat.respond(to: "Hello")
print(reply.text)
```