Spaces:
Running
Running
Commit 路
cf56563
0
Parent(s):
Optitransfer: the converge programme
Browse files- .gitattributes +35 -0
- README.md +120 -0
- index.html +67 -0
.gitattributes
ADDED
|
@@ -0,0 +1,35 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
*.7z filter=lfs diff=lfs merge=lfs -text
|
| 2 |
+
*.arrow filter=lfs diff=lfs merge=lfs -text
|
| 3 |
+
*.bin filter=lfs diff=lfs merge=lfs -text
|
| 4 |
+
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
| 5 |
+
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
| 6 |
+
*.ftz filter=lfs diff=lfs merge=lfs -text
|
| 7 |
+
*.gz filter=lfs diff=lfs merge=lfs -text
|
| 8 |
+
*.h5 filter=lfs diff=lfs merge=lfs -text
|
| 9 |
+
*.joblib filter=lfs diff=lfs merge=lfs -text
|
| 10 |
+
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
| 11 |
+
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
| 12 |
+
*.model filter=lfs diff=lfs merge=lfs -text
|
| 13 |
+
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
| 14 |
+
*.npy filter=lfs diff=lfs merge=lfs -text
|
| 15 |
+
*.npz filter=lfs diff=lfs merge=lfs -text
|
| 16 |
+
*.onnx filter=lfs diff=lfs merge=lfs -text
|
| 17 |
+
*.ot filter=lfs diff=lfs merge=lfs -text
|
| 18 |
+
*.parquet filter=lfs diff=lfs merge=lfs -text
|
| 19 |
+
*.pb filter=lfs diff=lfs merge=lfs -text
|
| 20 |
+
*.pickle filter=lfs diff=lfs merge=lfs -text
|
| 21 |
+
*.pkl filter=lfs diff=lfs merge=lfs -text
|
| 22 |
+
*.pt filter=lfs diff=lfs merge=lfs -text
|
| 23 |
+
*.pth filter=lfs diff=lfs merge=lfs -text
|
| 24 |
+
*.rar filter=lfs diff=lfs merge=lfs -text
|
| 25 |
+
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
| 26 |
+
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
| 27 |
+
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
| 28 |
+
*.tar filter=lfs diff=lfs merge=lfs -text
|
| 29 |
+
*.tflite filter=lfs diff=lfs merge=lfs -text
|
| 30 |
+
*.tgz filter=lfs diff=lfs merge=lfs -text
|
| 31 |
+
*.wasm filter=lfs diff=lfs merge=lfs -text
|
| 32 |
+
*.xz filter=lfs diff=lfs merge=lfs -text
|
| 33 |
+
*.zip filter=lfs diff=lfs merge=lfs -text
|
| 34 |
+
*.zst filter=lfs diff=lfs merge=lfs -text
|
| 35 |
+
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
README.md
ADDED
|
@@ -0,0 +1,120 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
---
|
| 2 |
+
title: Optitransfer
|
| 3 |
+
emoji: 馃敩
|
| 4 |
+
colorFrom: indigo
|
| 5 |
+
colorTo: blue
|
| 6 |
+
sdk: static
|
| 7 |
+
pinned: true
|
| 8 |
+
tags:
|
| 9 |
+
- model-merging
|
| 10 |
+
- open-weights
|
| 11 |
+
- llm
|
| 12 |
+
- qwen2.5
|
| 13 |
+
- 7b
|
| 14 |
+
- text-generation
|
| 15 |
+
- reproducible-evaluation
|
| 16 |
+
- benchmarks
|
| 17 |
+
- leaderboards
|
| 18 |
+
---
|
| 19 |
+
|
| 20 |
+
<div align="center">
|
| 21 |
+
|
| 22 |
+
# Optitransfer
|
| 23 |
+
|
| 24 |
+
**Converge: building frontier-level open models from the work the open-source community has already published.**
|
| 25 |
+
|
| 26 |
+
*One repeatable process at every model class. Every result is measured against its base model and published with per-item evidence.*
|
| 27 |
+
|
| 28 |
+
[Models](https://huggingface.co/Optitransfer) 路 [crdt-merge](https://github.com/mgillr/crdt-merge) 路 [acfa-rs](https://github.com/mgillr/acfa-rs) 路 [Paper](https://arxiv.org/abs/2605.19373)
|
| 29 |
+
|
| 30 |
+
</div>
|
| 31 |
+
|
| 32 |
+
---
|
| 33 |
+
|
| 34 |
+
## Purpose
|
| 35 |
+
|
| 36 |
+
Frontier AI is concentrated in a handful of laboratories, because each generation of pre-training costs more than
|
| 37 |
+
almost anyone else can spend. The open ecosystem has already paid that cost many times over. Thousands of fine-tuned
|
| 38 |
+
models sit on public repositories, each one somebody's GPU-hours and each adding a skill to a shared base: mathematics,
|
| 39 |
+
code, instruction following, a domain, a language.
|
| 40 |
+
|
| 41 |
+
**Converge** is built to turn that work into frontier-level open models. It does not out-train the laboratories. It brings what
|
| 42 |
+
the community has built into a single model for each model class, measures every step, keeps every gain and
|
| 43 |
+
publishes every regression. The model is never finished: it improves release after release.
|
| 44 |
+
|
| 45 |
+
## The goal
|
| 46 |
+
|
| 47 |
+
**Reach frontier-level performance in every model class.** The programme starts with 7B, then moves to each larger
|
| 48 |
+
size class and to other model families. At every class the model must improve on every axis that matters (reasoning,
|
| 49 |
+
mathematics, code, instruction following and knowledge) without a material regression against its base. Each class
|
| 50 |
+
is complete only when no public fine-tune improves any axis further.
|
| 51 |
+
|
| 52 |
+
## The process
|
| 53 |
+
|
| 54 |
+
The process is defined once and run the same way at every class, so every release can be rebuilt and re-measured.
|
| 55 |
+
|
| 56 |
+
1. **Start** from a strong open base model in the class.
|
| 57 |
+
2. **Discover** the compatible fine-tunes the community has published for it.
|
| 58 |
+
3. **Assimilate** them into one model.
|
| 59 |
+
4. **Measure** every axis against the base model on identical items, in the same session, on public benchmark protocols.
|
| 60 |
+
5. **Release** a model only when it holds against its base and the previous release on the whole board, and publish every regression next to every gain, with per-item evidence. This is the release rule from the corrected releases on; v1 and v2 predate it, and their cards list their regressions.
|
| 61 |
+
6. **Repeat** until nothing more improves, then carry the process to the next class.
|
| 62 |
+
|
| 63 |
+
The construction method is proprietary. The evaluation is fully public: protocols, per-item outputs and the tools
|
| 64 |
+
that recompute every number.
|
| 65 |
+
|
| 66 |
+
## Where the programme stands
|
| 67 |
+
|
| 68 |
+
| Release | Status |
|
| 69 |
+
|---|---|
|
| 70 |
+
| [converge v1](https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v1) | Public. The first 7B release, on Qwen2.5-7B-Instruct. It carries a merge defect that has since been corrected; see the notice on its card. |
|
| 71 |
+
| [converge v2](https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v2) | Public. The current release, with a full validation board. It is built on v1, so it inherits the defect. |
|
| 72 |
+
| corrected v1 and v2 | Built with the same recipes and the defect removed. Being benchmarked across the full board now. |
|
| 73 |
+
| converge v3 | Next. Released only once it holds against its base model and the previous release on the whole board. |
|
| 74 |
+
|
| 75 |
+
## How we report
|
| 76 |
+
|
| 77 |
+
- **Paired, never selected.** Every score is the model minus its base on the same items, with a 95% confidence
|
| 78 |
+
interval and an exact significance test.
|
| 79 |
+
- **Regressions are first-class.** Each card lists what went down as prominently as what went up.
|
| 80 |
+
- **Reproducible only.** Public numbers come from reproducible runs on public protocols: HELM-protocol GSM8K and
|
| 81 |
+
MMLU, MMLU-Pro (TIGER-Lab), ZeroEval, EvalPlus HumanEval+ and MBPP+, MATH, AIME, ARC and IFEval.
|
| 82 |
+
- **Corrected in public.** When a defect is found, the affected cards say so first.
|
| 83 |
+
|
| 84 |
+
## Who it is for
|
| 85 |
+
|
| 86 |
+
- **The open community.** A stronger open model at each class, with the evidence to check it.
|
| 87 |
+
- **Healthcare.** A hospital network fine-tunes models on patient data that cannot leave its walls. Converge brings
|
| 88 |
+
those internal fine-tunes together inside its boundary, with a record of every change.
|
| 89 |
+
- **Finance.** A regulated bank must show auditors how each model update was made, what changed and that nothing
|
| 90 |
+
regressed. Converge's per-item record is that audit trail.
|
| 91 |
+
- **Sovereign AI.** A national programme needs frontier capability without depending on a few foreign vendors.
|
| 92 |
+
Converge grows by contribution, not by procurement.
|
| 93 |
+
|
| 94 |
+
## How it is offered
|
| 95 |
+
|
| 96 |
+
**A free model, with paid guarantees.** The public model stays open, and the free tier is never limited to sell a
|
| 97 |
+
paid one. Around it, the first commercial product is auditable private aggregation: Optitransfer brings an
|
| 98 |
+
organisation's own fine-tunes together inside its boundary, with the same measurement and evidence. Compliance
|
| 99 |
+
evidence, sovereign deployments and private federations follow.
|
| 100 |
+
|
| 101 |
+
## Open components
|
| 102 |
+
|
| 103 |
+
- **[crdt-merge](https://github.com/mgillr/crdt-merge):** conflict-free (CRDT) merging for data, JSON, models and
|
| 104 |
+
agents, with the E4 trust-scoring engine ([paper](https://arxiv.org/abs/2605.19373)).
|
| 105 |
+
- **[acfa-rs](https://github.com/mgillr/acfa-rs):** accountable federated aggregation with signed, offline-verifiable
|
| 106 |
+
receipts (Apache-2.0, [arXiv:2607.10305](https://arxiv.org/abs/2607.10305)).
|
| 107 |
+
|
| 108 |
+
## Team
|
| 109 |
+
|
| 110 |
+
**Ryan Gillespie**, founder, architect and operator 路 [@Optitransfer](https://huggingface.co/Optitransfer)
|
| 111 |
+
|
| 112 |
+
- Fifteen years of verification and validation on safety-critical systems: railway signalling (Crossrail, ETCS),
|
| 113 |
+
medical devices taken to CE mark (Roche), and aviation (Lilium). Converge's evidence-first design is that
|
| 114 |
+
discipline applied to model weights.
|
| 115 |
+
- Built the full stack, from the merge protocol and its CRDT substrate to the evaluation and evidence system.
|
| 116 |
+
- Ran the research programme behind every published measurement.
|
| 117 |
+
- Author of the paper on CRDT-based model merging ([arXiv:2605.19373](https://arxiv.org/abs/2605.19373)).
|
| 118 |
+
|
| 119 |
+
<sub>Current releases are licensed for non-commercial research, a condition inherited from an upstream source model.
|
| 120 |
+
Each model card states its licence in full.</sub>
|
index.html
ADDED
|
@@ -0,0 +1,67 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
<!doctype html>
|
| 2 |
+
<html lang="en">
|
| 3 |
+
<head>
|
| 4 |
+
<meta charset="utf-8">
|
| 5 |
+
<meta name="viewport" content="width=device-width, initial-scale=1">
|
| 6 |
+
<title>Optitransfer 路 converge</title>
|
| 7 |
+
<meta name="description" content="Converge: building frontier-level open models from the work the open-source community has already published. One repeatable process at every model class, with every result published with per-item evidence.">
|
| 8 |
+
<style>
|
| 9 |
+
:root { --bg: #ffffff; --fg: #1a1d24; --muted: #5b6270; --line: #e6e8ec; --accent: #3b4fd8; }
|
| 10 |
+
@media (prefers-color-scheme: dark) { :root { --bg: #0f1115; --fg: #e8eaf0; --muted: #9aa1ae; --line: #252a33; --accent: #8ea0ff; } }
|
| 11 |
+
* { box-sizing: border-box; }
|
| 12 |
+
body { margin: 0; background: var(--bg); color: var(--fg); font: 16px/1.6 system-ui, -apple-system, "Segoe UI", sans-serif; }
|
| 13 |
+
main { max-width: 760px; margin: 0 auto; padding: 56px 16px 72px; }
|
| 14 |
+
h1 { font-size: 2.2rem; margin: 0 0 .25rem; letter-spacing: -.01em; }
|
| 15 |
+
.lede { font-size: 1.15rem; margin: 0 0 .25rem; }
|
| 16 |
+
.sub { color: var(--muted); margin: 0 0 2rem; }
|
| 17 |
+
h2 { font-size: 1.05rem; text-transform: uppercase; letter-spacing: .06em; color: var(--muted); margin: 2.25rem 0 .5rem; }
|
| 18 |
+
a { color: var(--accent); text-decoration: none; } a:hover { text-decoration: underline; }
|
| 19 |
+
ul, ol { padding-left: 1.2rem; } li { margin: .3rem 0; }
|
| 20 |
+
.rel { border-top: 1px solid var(--line); }
|
| 21 |
+
.rel div { display: flex; gap: 1rem; padding: .6rem 0; border-bottom: 1px solid var(--line); }
|
| 22 |
+
.rel b { min-width: 9.5rem; flex-shrink: 0; }
|
| 23 |
+
footer { margin-top: 3rem; color: var(--muted); font-size: .9rem; }
|
| 24 |
+
@media (max-width: 520px) { .rel div { flex-direction: column; gap: .1rem; } }
|
| 25 |
+
</style>
|
| 26 |
+
</head>
|
| 27 |
+
<body>
|
| 28 |
+
<main>
|
| 29 |
+
<h1>Optitransfer</h1>
|
| 30 |
+
<p class="lede"><strong>Converge:</strong> building frontier-level open models from the work the open-source community has already published.</p>
|
| 31 |
+
<p class="sub">One repeatable process at every model class. Every result is measured against its base model and published with per-item evidence.</p>
|
| 32 |
+
|
| 33 |
+
<h2>Purpose</h2>
|
| 34 |
+
<p>The community has published thousands of fine-tuned models, each adding some skill to a shared base. Converge brings that work together into a single open model for each class, without pre-training a new one. It measures every change, keeps every gain and publishes every regression.</p>
|
| 35 |
+
|
| 36 |
+
<h2>The goal</h2>
|
| 37 |
+
<p>Reach frontier-level performance in every model class: 7B first, then each larger size class and other model families. The model must improve on every axis without a material regression against its base, until no public fine-tune improves any axis further.</p>
|
| 38 |
+
|
| 39 |
+
<h2>The process</h2>
|
| 40 |
+
<ol>
|
| 41 |
+
<li><strong>Start</strong> from a strong open base model in the class.</li>
|
| 42 |
+
<li><strong>Discover</strong> the compatible fine-tunes the community has published.</li>
|
| 43 |
+
<li><strong>Assimilate</strong> them into one model.</li>
|
| 44 |
+
<li><strong>Measure</strong> every axis against the base on identical items, on public protocols.</li>
|
| 45 |
+
<li><strong>Release</strong> a model only when it holds on the whole board, and publish every regression. This is the release rule from the corrected releases on.</li>
|
| 46 |
+
<li><strong>Repeat</strong>, then carry the process to the next class.</li>
|
| 47 |
+
</ol>
|
| 48 |
+
<p>The construction method is proprietary. The evaluation is fully public.</p>
|
| 49 |
+
|
| 50 |
+
<h2>Releases</h2>
|
| 51 |
+
<div class="rel">
|
| 52 |
+
<div><b><a href="https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v1">converge v1</a></b><span>Public. The first 7B release. It carries a corrected merge defect; see its card.</span></div>
|
| 53 |
+
<div><b><a href="https://huggingface.co/Optitransfer/Qwen2.5-7B-Instruct-converge-collective-v2">converge v2</a></b><span>Public. The current release. It is built on v1, so it inherits the defect.</span></div>
|
| 54 |
+
<div><b>corrected v1 and v2</b><span>Same recipes, defect removed. Being benchmarked across the full board now.</span></div>
|
| 55 |
+
<div><b>converge v3</b><span>Next. Released only once it holds against its base and the previous release.</span></div>
|
| 56 |
+
</div>
|
| 57 |
+
|
| 58 |
+
<h2>Open components</h2>
|
| 59 |
+
<ul>
|
| 60 |
+
<li><a href="https://github.com/mgillr/crdt-merge">crdt-merge</a>: conflict-free merging for data, models and agents, with the E4 trust engine.</li>
|
| 61 |
+
<li><a href="https://github.com/mgillr/acfa-rs">acfa-rs</a>: accountable federated aggregation with offline-verifiable receipts.</li>
|
| 62 |
+
</ul>
|
| 63 |
+
|
| 64 |
+
<footer>Ryan Gillespie, founder 路 <a href="https://huggingface.co/Optitransfer">@Optitransfer</a></footer>
|
| 65 |
+
</main>
|
| 66 |
+
</body>
|
| 67 |
+
</html>
|