Built at a2f3d0b967c85a203269a1bdb4cbaf9e80045c22
Browse files- README.md +12 -12
- config.json +4 -3
- model.safetensors +2 -2
README.md
CHANGED
|
@@ -27,7 +27,7 @@ are fixed operations of the engine; the network only chooses between them.
|
|
| 27 |
What it knows is in the tree, where it can be read and changed, and not in the weights. The
|
| 28 |
network learns only how to shape and read that memory. That is why it is small: it is not a
|
| 29 |
transformer, it was not trained on a large corpus, and it does not write free text. These files
|
| 30 |
-
hold 5,
|
| 31 |
permanent state the project's seeds make. It runs on a processor, and in a browser through
|
| 32 |
WebAssembly, where the Datamine Network chat runs it with nothing sent to a server.
|
| 33 |
|
|
@@ -40,7 +40,7 @@ which DAM1 is not; they run in float16 on an NVIDIA RTX 3080 Ti, DAM1 on the pro
|
|
| 40 |
|
| 41 |
| Model | Parameters | Weights | Memory | Answered | Exact | One answer |
|
| 42 |
|---|---:|---:|---:|---:|---:|---:|
|
| 43 |
-
| **DAM1** | **5,
|
| 44 |
| Qwen2.5-0.5B-Instruct | 494,032,768 | 988 MB | 1.01 GB | 63.8% | 39.6% | 109 ms |
|
| 45 |
| Qwen3-0.6B | 596,049,920 | 1.50 GB | 1.22 GB | 52.7% | 41.4% | 80 ms |
|
| 46 |
| LFM2-350M | 354,483,968 | 709 MB | 727 MB | 52.2% | 44.1% | 53 ms |
|
|
@@ -166,7 +166,7 @@ it is a Hugging Face dataset, so the card lists no `datasets`.
|
|
| 166 |
- `model/data/train`: 274 lesson files in 16 folders, from basics and Grade 1 to Grade 12 and
|
| 167 |
university, with the world lessons and the debug cases agreed while the design was made. Each
|
| 168 |
line is text with the shape its tree should take and the answers expected. A line is either
|
| 169 |
-
learned or held out of training: 2,
|
| 170 |
cases are graded and never trained on.
|
| 171 |
- `model/data/facts` and `model/data/seeds`: plain fact sentences and the 60 seeds files built
|
| 172 |
from them, told into the tree before anything is read. They are not used to train the network.
|
|
@@ -297,23 +297,23 @@ expected answer.
|
|
| 297 |
|---|---:|---:|
|
| 298 |
| Basics | 446 of 446 (100.0%) | 142 of 142 (100.0%) |
|
| 299 |
| Grade 1 | 73 of 73 (100.0%) | 24 of 24 (100.0%) |
|
| 300 |
-
| Grade 2 |
|
| 301 |
| Grade 3 | 246 of 246 (100.0%) | 97 of 97 (100.0%) |
|
| 302 |
-
| Grade 4 |
|
| 303 |
-
| Grade 5 |
|
| 304 |
-
| Grade 6 |
|
| 305 |
| Grade 7 | 83 of 83 (100.0%) | 35 of 35 (100.0%) |
|
| 306 |
| Grade 8 | 79 of 79 (100.0%) | 33 of 33 (100.0%) |
|
| 307 |
-
| Grade 9 | 149 of 149 (100.0%) |
|
| 308 |
| Grade 10 | 26 of 26 (100.0%) | 9 of 9 (100.0%) |
|
| 309 |
| Grade 11 | 82 of 82 (100.0%) | 23 of 23 (100.0%) |
|
| 310 |
-
| Grade 12 |
|
| 311 |
| University | 47 of 47 (100.0%) | 13 of 13 (100.0%) |
|
| 312 |
| World and debug cases | 36 of 36 (100.0%) | 32 of 32 (100.0%) |
|
| 313 |
-
| **All** | **
|
| 314 |
|
| 315 |
-
It reads every learned line of the curriculum. The
|
| 316 |
-
|
| 317 |
|
| 318 |
The held-out lines use the same sentence shapes and many of the same words as the learned lines.
|
| 319 |
They measure new words and numbers in taught shapes. They are not an independent benchmark. No
|
|
|
|
| 27 |
What it knows is in the tree, where it can be read and changed, and not in the weights. The
|
| 28 |
network learns only how to shape and read that memory. That is why it is small: it is not a
|
| 29 |
transformer, it was not trained on a large corpus, and it does not write free text. These files
|
| 30 |
+
hold 5,596,839 numbers, three networks of 1,865,613 numbers each that read by vote, and the
|
| 31 |
permanent state the project's seeds make. It runs on a processor, and in a browser through
|
| 32 |
WebAssembly, where the Datamine Network chat runs it with nothing sent to a server.
|
| 33 |
|
|
|
|
| 40 |
|
| 41 |
| Model | Parameters | Weights | Memory | Answered | Exact | One answer |
|
| 42 |
|---|---:|---:|---:|---:|---:|---:|
|
| 43 |
+
| **DAM1** | **5,596,839** | **22 MB** | **43 MB** | **100.0%** | **100.0%** | **3 ms** |
|
| 44 |
| Qwen2.5-0.5B-Instruct | 494,032,768 | 988 MB | 1.01 GB | 63.8% | 39.6% | 109 ms |
|
| 45 |
| Qwen3-0.6B | 596,049,920 | 1.50 GB | 1.22 GB | 52.7% | 41.4% | 80 ms |
|
| 46 |
| LFM2-350M | 354,483,968 | 709 MB | 727 MB | 52.2% | 44.1% | 53 ms |
|
|
|
|
| 166 |
- `model/data/train`: 274 lesson files in 16 folders, from basics and Grade 1 to Grade 12 and
|
| 167 |
university, with the world lessons and the debug cases agreed while the design was made. Each
|
| 168 |
line is text with the shape its tree should take and the answers expected. A line is either
|
| 169 |
+
learned or held out of training: 2,575 learned lines and 942 held-out lines. The 12 debug
|
| 170 |
cases are graded and never trained on.
|
| 171 |
- `model/data/facts` and `model/data/seeds`: plain fact sentences and the 60 seeds files built
|
| 172 |
from them, told into the tree before anything is read. They are not used to train the network.
|
|
|
|
| 297 |
|---|---:|---:|
|
| 298 |
| Basics | 446 of 446 (100.0%) | 142 of 142 (100.0%) |
|
| 299 |
| Grade 1 | 73 of 73 (100.0%) | 24 of 24 (100.0%) |
|
| 300 |
+
| Grade 2 | 676 of 676 (100.0%) | 239 of 239 (100.0%) |
|
| 301 |
| Grade 3 | 246 of 246 (100.0%) | 97 of 97 (100.0%) |
|
| 302 |
+
| Grade 4 | 265 of 265 (100.0%) | 88 of 88 (100.0%) |
|
| 303 |
+
| Grade 5 | 172 of 172 (100.0%) | 72 of 72 (100.0%) |
|
| 304 |
+
| Grade 6 | 156 of 156 (100.0%) | 58 of 58 (100.0%) |
|
| 305 |
| Grade 7 | 83 of 83 (100.0%) | 35 of 35 (100.0%) |
|
| 306 |
| Grade 8 | 79 of 79 (100.0%) | 33 of 33 (100.0%) |
|
| 307 |
+
| Grade 9 | 149 of 149 (100.0%) | 65 of 66 (98.5%) |
|
| 308 |
| Grade 10 | 26 of 26 (100.0%) | 9 of 9 (100.0%) |
|
| 309 |
| Grade 11 | 82 of 82 (100.0%) | 23 of 23 (100.0%) |
|
| 310 |
+
| Grade 12 | 39 of 39 (100.0%) | 11 of 11 (100.0%) |
|
| 311 |
| University | 47 of 47 (100.0%) | 13 of 13 (100.0%) |
|
| 312 |
| World and debug cases | 36 of 36 (100.0%) | 32 of 32 (100.0%) |
|
| 313 |
+
| **All** | **2575 of 2575 (100.0%)** | **941 of 942 (99.9%)** |
|
| 314 |
|
| 315 |
+
It reads every learned line of the curriculum. The one held-out line it misses is in the lesson
|
| 316 |
+
on the parts a kind of thing has, which asks a part of a thing that is not a creature.
|
| 317 |
|
| 318 |
The held-out lines use the same sentence shapes and many of the same words as the learned lines.
|
| 319 |
They measure new words and numbers in taught shapes. They are not an independent benchmark. No
|
config.json
CHANGED
|
@@ -9,7 +9,7 @@
|
|
| 9 |
"slots": 200,
|
| 10 |
"item": 16,
|
| 11 |
"hidden": 512,
|
| 12 |
-
"items":
|
| 13 |
"pointer": true,
|
| 14 |
"voters": 3,
|
| 15 |
"classes": [
|
|
@@ -40,6 +40,7 @@
|
|
| 40 |
"{setProperty speed: @}",
|
| 41 |
"{setProperty temperature: @}",
|
| 42 |
"{check @}",
|
|
|
|
| 43 |
"{step newest}",
|
| 44 |
"{step top}",
|
| 45 |
"{setProperty feeling: @}",
|
|
@@ -49,7 +50,6 @@
|
|
| 49 |
"{get count: @}",
|
| 50 |
"{step user: @}",
|
| 51 |
"{compute @}",
|
| 52 |
-
"{children add: @ type: relation}",
|
| 53 |
"{get amount}",
|
| 54 |
"{number set: @}",
|
| 55 |
"{number add: @}",
|
|
@@ -60,6 +60,7 @@
|
|
| 60 |
"{release}",
|
| 61 |
"{get past}",
|
| 62 |
"{get relation: @}",
|
|
|
|
| 63 |
"{join}",
|
| 64 |
"{activity}",
|
| 65 |
"{children add: @ type: deed}",
|
|
@@ -108,7 +109,7 @@
|
|
| 108 |
]
|
| 109 |
},
|
| 110 |
"build": {
|
| 111 |
-
"commit": "
|
| 112 |
"network": "data/network/release.bin",
|
| 113 |
"state": "data/seeds: all"
|
| 114 |
}
|
|
|
|
| 9 |
"slots": 200,
|
| 10 |
"item": 16,
|
| 11 |
"hidden": 512,
|
| 12 |
+
"items": 4275,
|
| 13 |
"pointer": true,
|
| 14 |
"voters": 3,
|
| 15 |
"classes": [
|
|
|
|
| 40 |
"{setProperty speed: @}",
|
| 41 |
"{setProperty temperature: @}",
|
| 42 |
"{check @}",
|
| 43 |
+
"{children add: @ type: relation}",
|
| 44 |
"{step newest}",
|
| 45 |
"{step top}",
|
| 46 |
"{setProperty feeling: @}",
|
|
|
|
| 50 |
"{get count: @}",
|
| 51 |
"{step user: @}",
|
| 52 |
"{compute @}",
|
|
|
|
| 53 |
"{get amount}",
|
| 54 |
"{number set: @}",
|
| 55 |
"{number add: @}",
|
|
|
|
| 60 |
"{release}",
|
| 61 |
"{get past}",
|
| 62 |
"{get relation: @}",
|
| 63 |
+
"{children add: @ type: relation time: past}",
|
| 64 |
"{join}",
|
| 65 |
"{activity}",
|
| 66 |
"{children add: @ type: deed}",
|
|
|
|
| 109 |
]
|
| 110 |
},
|
| 111 |
"build": {
|
| 112 |
+
"commit": "a2f3d0b967c85a203269a1bdb4cbaf9e80045c22",
|
| 113 |
"network": "data/network/release.bin",
|
| 114 |
"state": "data/seeds: all"
|
| 115 |
}
|
model.safetensors
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:94792089c4b7f17da5f4b16a10c58e2ccbef26fc18af1d2e5868237f89a4038c
|
| 3 |
+
size 22492156
|