erratum: the eco-chain row published .867 as its solo score; in the ECO ledger .867 is pre_train_chain_read, the LOADED arm read before this stage (a different weight file). This arm alone reads .833 (stop member masked off) and .853 in the pair. Corrected against the ledger in the index, the card and the Space catalog; .867 kept as the named comparator. Card also notes every packaged arm is the wide adapter shape and the narrow certified twins are not packaged.
Browse files- README.md +7 -1
- arms/index.json +9 -5
README.md
CHANGED
|
@@ -63,7 +63,7 @@ result is.
|
|
| 63 |
| `rules-minted-seed1` | the second seed of the ruled dial; the quietest arm in the package on general text | arm | minted words .647 · unseen .593 · general-text cost .0066 bpb | DIAL | settled (2 seeds, 2026-09-22) |
|
| 64 |
| `chain` | reads five if-then rules and writes out every step: 'So Wren is …' five times | arm | chain .867 · general-text cost .003 bpb | E-I (trained with an abstention term, so it composes ungated) | settled (2 seeds) |
|
| 65 |
| `chain-plain` | the same step-by-step chain writing, trained without the abstention term: the best chain number in the library, but it writes on off-domain text too | arm | chain .887 | C1 | settled (2 seeds) |
|
| 66 |
-
| `eco-chain` | a loaded chain arm carried through a second training stage beside a fresh turn-end arm, and still detachable | arm |
|
| 67 |
| `stop` | ends its turn cleanly with a blank line instead of rambling on | arm | clean stop .83 (second seed 1.00) | E-I | settled (2 seeds) |
|
| 68 |
| `stop-seed1` | the second seed: every probe ends its turn | arm | clean stop 1.00 | E-I | settled (2 seeds) |
|
| 69 |
| `eco-stop` | the fresh turn-end arm trained in that second stage, with its own abstention term | arm | turn end .953 in the pair · chat 1.00 | ECO / L4 | settled (2 seeds) |
|
|
@@ -199,6 +199,12 @@ state dict is unchanged by any number of mount/detach cycles.
|
|
| 199 |
- The rule-chain arms are a controlled capability on made-up words, not
|
| 200 |
general reasoning: they read five if-then rules and either write the
|
| 201 |
steps out or answer with the final word.
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 202 |
- `rules-minted` was trained on the lexicon minted for the **next** craft's
|
| 203 |
curriculum; it is packaged here because it is the newest settled arm and
|
| 204 |
the same recipe, not because this core saw those words in pretraining.
|
|
|
|
| 63 |
| `rules-minted-seed1` | the second seed of the ruled dial; the quietest arm in the package on general text | arm | minted words .647 · unseen .593 · general-text cost .0066 bpb | DIAL | settled (2 seeds, 2026-09-22) |
|
| 64 |
| `chain` | reads five if-then rules and writes out every step: 'So Wren is …' five times | arm | chain .867 · general-text cost .003 bpb | E-I (trained with an abstention term, so it composes ungated) | settled (2 seeds) |
|
| 65 |
| `chain-plain` | the same step-by-step chain writing, trained without the abstention term: the best chain number in the library, but it writes on off-domain text too | arm | chain .887 | C1 | settled (2 seeds) |
|
| 66 |
+
| `eco-chain` | a loaded chain arm carried through a second training stage beside a fresh turn-end arm, and still detachable | arm | alone .833 · in the pair .853 · quiet .0054 bpb (the loaded arm read .867 before this stage) | ECO / L4 | settled (2 seeds) |
|
| 67 |
| `stop` | ends its turn cleanly with a blank line instead of rambling on | arm | clean stop .83 (second seed 1.00) | E-I | settled (2 seeds) |
|
| 68 |
| `stop-seed1` | the second seed: every probe ends its turn | arm | clean stop 1.00 | E-I | settled (2 seeds) |
|
| 69 |
| `eco-stop` | the fresh turn-end arm trained in that second stage, with its own abstention term | arm | turn end .953 in the pair · chat 1.00 | ECO / L4 | settled (2 seeds) |
|
|
|
|
| 199 |
- The rule-chain arms are a controlled capability on made-up words, not
|
| 200 |
general reasoning: they read five if-then rules and either write the
|
| 201 |
steps out or answer with the final word.
|
| 202 |
+
- Every packaged arm is the **wide** adapter shape: 32 slots of dimension
|
| 203 |
+
4 read through a 64-atom address, 8.57M parameters. The certified twins
|
| 204 |
+
of these same recipes at the narrower 16-atom, dimension-8 shape (a
|
| 205 |
+
near-identical 8.56M parameters) live on the training repo and are
|
| 206 |
+
**not** packaged here, so nothing in this table is a comparison
|
| 207 |
+
between the two shapes.
|
| 208 |
- `rules-minted` was trained on the lexicon minted for the **next** craft's
|
| 209 |
curriculum; it is packaged here because it is the newest settled arm and
|
| 210 |
the same recipe, not because this core saw those words in pretraining.
|
arms/index.json
CHANGED
|
@@ -382,17 +382,18 @@
|
|
| 382 |
"family": "chain",
|
| 383 |
"status": "settled (2 seeds)",
|
| 384 |
"rank": 12,
|
| 385 |
-
"score": "
|
| 386 |
"measured": {
|
| 387 |
-
"
|
| 388 |
-
"masked_chain": 0.8333,
|
| 389 |
"pair_chain": 0.8533,
|
|
|
|
| 390 |
"general_text_cost_bpb": 0.0054
|
| 391 |
},
|
| 392 |
"cell": "ECO / L4",
|
| 393 |
"examples": [
|
| 394 |
"If someone is sook, then they are quen. If someone is quen, then they are harl. If someone is harl, then they are torv. If someone is torv, then they are prin. If someone is prin, then they are mund. Wren is sook. What follows?"
|
| 395 |
-
]
|
|
|
|
| 396 |
},
|
| 397 |
{
|
| 398 |
"id": "stop",
|
|
@@ -749,6 +750,9 @@
|
|
| 749 |
"detach": "the core's state dict is unchanged by mount/detach, and the restored logits are bit-identical to the pre-mount fingerprint",
|
| 750 |
"when": "2026-09-22",
|
| 751 |
"device": "RTX 4090, fp32, greedy",
|
| 752 |
-
"why": "bit-identical logits under a greedy decode give bit-identical bytes, so the campaign's scores are the package's scores; re-scoring here could only reproduce them"
|
|
|
|
|
|
|
|
|
|
| 753 |
}
|
| 754 |
}
|
|
|
|
| 382 |
"family": "chain",
|
| 383 |
"status": "settled (2 seeds)",
|
| 384 |
"rank": 12,
|
| 385 |
+
"score": "alone .833 · in the pair .853 · quiet .0054 bpb (the loaded arm read .867 before this stage)",
|
| 386 |
"measured": {
|
| 387 |
+
"alone_chain": 0.8333,
|
|
|
|
| 388 |
"pair_chain": 0.8533,
|
| 389 |
+
"pre_stage_solo_chain": 0.867,
|
| 390 |
"general_text_cost_bpb": 0.0054
|
| 391 |
},
|
| 392 |
"cell": "ECO / L4",
|
| 393 |
"examples": [
|
| 394 |
"If someone is sook, then they are quen. If someone is quen, then they are harl. If someone is harl, then they are torv. If someone is torv, then they are prin. If someone is prin, then they are mund. Wren is sook. What follows?"
|
| 395 |
+
],
|
| 396 |
+
"erratum": "2026-09-22: first published as 'chain .867 solo'; .867 is the ledger's pre_train_chain_read — the loaded arm BEFORE this stage, a different weight file. This arm alone reads .833."
|
| 397 |
},
|
| 398 |
{
|
| 399 |
"id": "stop",
|
|
|
|
| 750 |
"detach": "the core's state dict is unchanged by mount/detach, and the restored logits are bit-identical to the pre-mount fingerprint",
|
| 751 |
"when": "2026-09-22",
|
| 752 |
"device": "RTX 4090, fp32, greedy",
|
| 753 |
+
"why": "bit-identical logits under a greedy decode give bit-identical bytes, so the campaign's scores are the package's scores; re-scoring here could only reproduce them",
|
| 754 |
+
"errata": [
|
| 755 |
+
"2026-09-22: first published as 'chain .867 solo'; .867 is the ledger's pre_train_chain_read — the loaded arm BEFORE this stage, a different weight file. This arm alone reads .833."
|
| 756 |
+
]
|
| 757 |
}
|
| 758 |
}
|