Optitransfer commited on
Commit
cf56563
路
0 Parent(s):

Optitransfer: the converge programme

Browse files
Files changed (3) hide show
  1. .gitattributes +35 -0
  2. README.md +120 -0
  3. index.html +67 -0
.gitattributes ADDED
@@ -0,0 +1,35 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ *.7z filter=lfs diff=lfs merge=lfs -text
2
+ *.arrow filter=lfs diff=lfs merge=lfs -text
3
+ *.bin filter=lfs diff=lfs merge=lfs -text
4
+ *.bz2 filter=lfs diff=lfs merge=lfs -text
5
+ *.ckpt filter=lfs diff=lfs merge=lfs -text
6
+ *.ftz filter=lfs diff=lfs merge=lfs -text
7
+ *.gz filter=lfs diff=lfs merge=lfs -text
8
+ *.h5 filter=lfs diff=lfs merge=lfs -text
9
+ *.joblib filter=lfs diff=lfs merge=lfs -text
10
+ *.lfs.* filter=lfs diff=lfs merge=lfs -text
11
+ *.mlmodel filter=lfs diff=lfs merge=lfs -text
12
+ *.model filter=lfs diff=lfs merge=lfs -text
13
+ *.msgpack filter=lfs diff=lfs merge=lfs -text
14
+ *.npy filter=lfs diff=lfs merge=lfs -text
15
+ *.npz filter=lfs diff=lfs merge=lfs -text
16
+ *.onnx filter=lfs diff=lfs merge=lfs -text
17
+ *.ot filter=lfs diff=lfs merge=lfs -text
18
+ *.parquet filter=lfs diff=lfs merge=lfs -text
19
+ *.pb filter=lfs diff=lfs merge=lfs -text
20
+ *.pickle filter=lfs diff=lfs merge=lfs -text
21
+ *.pkl filter=lfs diff=lfs merge=lfs -text
22
+ *.pt filter=lfs diff=lfs merge=lfs -text
23
+ *.pth filter=lfs diff=lfs merge=lfs -text
24
+ *.rar filter=lfs diff=lfs merge=lfs -text
25
+ *.safetensors filter=lfs diff=lfs merge=lfs -text
26
+ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
27
+ *.tar.* filter=lfs diff=lfs merge=lfs -text
28
+ *.tar filter=lfs diff=lfs merge=lfs -text
29
+ *.tflite filter=lfs diff=lfs merge=lfs -text
30
+ *.tgz filter=lfs diff=lfs merge=lfs -text
31
+ *.wasm filter=lfs diff=lfs merge=lfs -text
32
+ *.xz filter=lfs diff=lfs merge=lfs -text
33
+ *.zip filter=lfs diff=lfs merge=lfs -text
34
+ *.zst filter=lfs diff=lfs merge=lfs -text
35
+ *tfevents* filter=lfs diff=lfs merge=lfs -text
README.md ADDED
@@ -0,0 +1,120 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ title: Optitransfer
3
+ emoji: 馃敩
4
+ colorFrom: indigo
5
+ colorTo: blue
6
+ sdk: static
7
+ pinned: true
8
+ tags:
9
+ - model-merging
10
+ - open-weights
11
+ - llm
12
+ - qwen2.5
13
+ - 7b
14
+ - text-generation
15
+ - reproducible-evaluation
16
+ - benchmarks
17
+ - leaderboards
18
+ ---
19
+
20
+ <div align="center">
21
+
22
+ # Optitransfer
23
+
24
+ **Converge: building frontier-level open models from the work the open-source community has already published.**
25
+
26
+ *One repeatable process at every model class. Every result is measured against its base model and published with per-item evidence.*
27
+
28
+ [Models](https://huggingface.co/Optitransfer) 路 [crdt-merge](https://github.com/mgillr/crdt-merge) 路 [acfa-rs](https://github.com/mgillr/acfa-rs) 路 [Paper](https://arxiv.org/abs/2605.19373)
29
+
30
+ </div>
31
+
32
+ ---
33
+
34
+ ## Purpose
35
+
36
+ Frontier AI is concentrated in a handful of laboratories, because each generation of pre-training costs more than
37
+ almost anyone else can spend. The open ecosystem has already paid that cost many times over. Thousands of fine-tuned
38
+ models sit on public repositories, each one somebody's GPU-hours and each adding a skill to a shared base: mathematics,
39
+ code, instruction following, a domain, a language.
40
+
41
+ **Converge** is built to turn that work into frontier-level open models. It does not out-train the laboratories. It brings what
42
+ the community has built into a single model for each model class, measures every step, keeps every gain and
43
+ publishes every regression. The model is never finished: it improves release after release.
44
+
45
+ ## The goal
46
+
47
+ **Reach frontier-level performance in every model class.** The programme starts with 7B, then moves to each larger
48
+ size class and to other model families. At every class the model must improve on every axis that matters (reasoning,
49
+ mathematics, code, instruction following and knowledge) without a material regression against its base. Each class
50
+ is complete only when no public fine-tune improves any axis further.
51
+
52
+ ## The process
53
+
54
+ The process is defined once and run the same way at every class, so every release can be rebuilt and re-measured.
55
+
56
+ 1. **Start** from a strong open base model in the class.
57
+ 2. **Discover** the compatible fine-tunes the community has published for it.
58
+ 3. **Assimilate** them into one model.
59
+ 4. **Measure** every axis against the base model on identical items, in the same session, on public benchmark protocols.
60
+ 5. **Release** a model only when it holds against its base and the previous release on the whole board, and publish every regression next to every gain, with per-item evidence. This is the release rule from the corrected releases on; v1 and v2 predate it, and their cards list their regressions.
61
+ 6. **Repeat** until nothing more improves, then carry the process to the next class.
62
+
63
+ The construction method is proprietary. The evaluation is fully public: protocols, per-item outputs and the tools
64
+ that recompute every number.
65
+
66
+ ## Where the programme stands
67
+
68
+ | Release | Status |
69
+ |---|---|
70
+ | [converge v1](https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v1) | Public. The first 7B release, on Qwen2.5-7B-Instruct. It carries a merge defect that has since been corrected; see the notice on its card. |
71
+ | [converge v2](https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v2) | Public. The current release, with a full validation board. It is built on v1, so it inherits the defect. |
72
+ | corrected v1 and v2 | Built with the same recipes and the defect removed. Being benchmarked across the full board now. |
73
+ | converge v3 | Next. Released only once it holds against its base model and the previous release on the whole board. |
74
+
75
+ ## How we report
76
+
77
+ - **Paired, never selected.** Every score is the model minus its base on the same items, with a 95% confidence
78
+ interval and an exact significance test.
79
+ - **Regressions are first-class.** Each card lists what went down as prominently as what went up.
80
+ - **Reproducible only.** Public numbers come from reproducible runs on public protocols: HELM-protocol GSM8K and
81
+ MMLU, MMLU-Pro (TIGER-Lab), ZeroEval, EvalPlus HumanEval+ and MBPP+, MATH, AIME, ARC and IFEval.
82
+ - **Corrected in public.** When a defect is found, the affected cards say so first.
83
+
84
+ ## Who it is for
85
+
86
+ - **The open community.** A stronger open model at each class, with the evidence to check it.
87
+ - **Healthcare.** A hospital network fine-tunes models on patient data that cannot leave its walls. Converge brings
88
+ those internal fine-tunes together inside its boundary, with a record of every change.
89
+ - **Finance.** A regulated bank must show auditors how each model update was made, what changed and that nothing
90
+ regressed. Converge's per-item record is that audit trail.
91
+ - **Sovereign AI.** A national programme needs frontier capability without depending on a few foreign vendors.
92
+ Converge grows by contribution, not by procurement.
93
+
94
+ ## How it is offered
95
+
96
+ **A free model, with paid guarantees.** The public model stays open, and the free tier is never limited to sell a
97
+ paid one. Around it, the first commercial product is auditable private aggregation: Optitransfer brings an
98
+ organisation's own fine-tunes together inside its boundary, with the same measurement and evidence. Compliance
99
+ evidence, sovereign deployments and private federations follow.
100
+
101
+ ## Open components
102
+
103
+ - **[crdt-merge](https://github.com/mgillr/crdt-merge):** conflict-free (CRDT) merging for data, JSON, models and
104
+ agents, with the E4 trust-scoring engine ([paper](https://arxiv.org/abs/2605.19373)).
105
+ - **[acfa-rs](https://github.com/mgillr/acfa-rs):** accountable federated aggregation with signed, offline-verifiable
106
+ receipts (Apache-2.0, [arXiv:2607.10305](https://arxiv.org/abs/2607.10305)).
107
+
108
+ ## Team
109
+
110
+ **Ryan Gillespie**, founder, architect and operator 路 [@Optitransfer](https://huggingface.co/Optitransfer)
111
+
112
+ - Fifteen years of verification and validation on safety-critical systems: railway signalling (Crossrail, ETCS),
113
+ medical devices taken to CE mark (Roche), and aviation (Lilium). Converge's evidence-first design is that
114
+ discipline applied to model weights.
115
+ - Built the full stack, from the merge protocol and its CRDT substrate to the evaluation and evidence system.
116
+ - Ran the research programme behind every published measurement.
117
+ - Author of the paper on CRDT-based model merging ([arXiv:2605.19373](https://arxiv.org/abs/2605.19373)).
118
+
119
+ <sub>Current releases are licensed for non-commercial research, a condition inherited from an upstream source model.
120
+ Each model card states its licence in full.</sub>
index.html ADDED
@@ -0,0 +1,67 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ <!doctype html>
2
+ <html lang="en">
3
+ <head>
4
+ <meta charset="utf-8">
5
+ <meta name="viewport" content="width=device-width, initial-scale=1">
6
+ <title>Optitransfer 路 converge</title>
7
+ <meta name="description" content="Converge: building frontier-level open models from the work the open-source community has already published. One repeatable process at every model class, with every result published with per-item evidence.">
8
+ <style>
9
+ :root { --bg: #ffffff; --fg: #1a1d24; --muted: #5b6270; --line: #e6e8ec; --accent: #3b4fd8; }
10
+ @media (prefers-color-scheme: dark) { :root { --bg: #0f1115; --fg: #e8eaf0; --muted: #9aa1ae; --line: #252a33; --accent: #8ea0ff; } }
11
+ * { box-sizing: border-box; }
12
+ body { margin: 0; background: var(--bg); color: var(--fg); font: 16px/1.6 system-ui, -apple-system, "Segoe UI", sans-serif; }
13
+ main { max-width: 760px; margin: 0 auto; padding: 56px 16px 72px; }
14
+ h1 { font-size: 2.2rem; margin: 0 0 .25rem; letter-spacing: -.01em; }
15
+ .lede { font-size: 1.15rem; margin: 0 0 .25rem; }
16
+ .sub { color: var(--muted); margin: 0 0 2rem; }
17
+ h2 { font-size: 1.05rem; text-transform: uppercase; letter-spacing: .06em; color: var(--muted); margin: 2.25rem 0 .5rem; }
18
+ a { color: var(--accent); text-decoration: none; } a:hover { text-decoration: underline; }
19
+ ul, ol { padding-left: 1.2rem; } li { margin: .3rem 0; }
20
+ .rel { border-top: 1px solid var(--line); }
21
+ .rel div { display: flex; gap: 1rem; padding: .6rem 0; border-bottom: 1px solid var(--line); }
22
+ .rel b { min-width: 9.5rem; flex-shrink: 0; }
23
+ footer { margin-top: 3rem; color: var(--muted); font-size: .9rem; }
24
+ @media (max-width: 520px) { .rel div { flex-direction: column; gap: .1rem; } }
25
+ </style>
26
+ </head>
27
+ <body>
28
+ <main>
29
+ <h1>Optitransfer</h1>
30
+ <p class="lede"><strong>Converge:</strong> building frontier-level open models from the work the open-source community has already published.</p>
31
+ <p class="sub">One repeatable process at every model class. Every result is measured against its base model and published with per-item evidence.</p>
32
+
33
+ <h2>Purpose</h2>
34
+ <p>The community has published thousands of fine-tuned models, each adding some skill to a shared base. Converge brings that work together into a single open model for each class, without pre-training a new one. It measures every change, keeps every gain and publishes every regression.</p>
35
+
36
+ <h2>The goal</h2>
37
+ <p>Reach frontier-level performance in every model class: 7B first, then each larger size class and other model families. The model must improve on every axis without a material regression against its base, until no public fine-tune improves any axis further.</p>
38
+
39
+ <h2>The process</h2>
40
+ <ol>
41
+ <li><strong>Start</strong> from a strong open base model in the class.</li>
42
+ <li><strong>Discover</strong> the compatible fine-tunes the community has published.</li>
43
+ <li><strong>Assimilate</strong> them into one model.</li>
44
+ <li><strong>Measure</strong> every axis against the base on identical items, on public protocols.</li>
45
+ <li><strong>Release</strong> a model only when it holds on the whole board, and publish every regression. This is the release rule from the corrected releases on.</li>
46
+ <li><strong>Repeat</strong>, then carry the process to the next class.</li>
47
+ </ol>
48
+ <p>The construction method is proprietary. The evaluation is fully public.</p>
49
+
50
+ <h2>Releases</h2>
51
+ <div class="rel">
52
+ <div><b><a href="https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v1">converge v1</a></b><span>Public. The first 7B release. It carries a corrected merge defect; see its card.</span></div>
53
+ <div><b><a href="https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v2">converge v2</a></b><span>Public. The current release. It is built on v1, so it inherits the defect.</span></div>
54
+ <div><b>corrected v1 and v2</b><span>Same recipes, defect removed. Being benchmarked across the full board now.</span></div>
55
+ <div><b>converge v3</b><span>Next. Released only once it holds against its base and the previous release.</span></div>
56
+ </div>
57
+
58
+ <h2>Open components</h2>
59
+ <ul>
60
+ <li><a href="https://github.com/mgillr/crdt-merge">crdt-merge</a>: conflict-free merging for data, models and agents, with the E4 trust engine.</li>
61
+ <li><a href="https://github.com/mgillr/acfa-rs">acfa-rs</a>: accountable federated aggregation with offline-verifiable receipts.</li>
62
+ </ul>
63
+
64
+ <footer>Ryan Gillespie, founder 路 <a href="https://huggingface.co/Optitransfer">@Optitransfer</a></footer>
65
+ </main>
66
+ </body>
67
+ </html>