Neural Solver Synthesis
Models, datasets, certified evaluations, and final evidence. Interactive reports: https://wandb.ai/neural-solver-synthesis/neural-solver-synthesis
Viewer • Updated • 10k • 28Note SDS training dataset (seed 101).
IDEALLab/OpenR1-SDS-10k-seed202
Viewer • Updated • 10k • 36Note SDS training dataset (seed 202).
IDEALLab/OpenR1-SDS-10k-seed303
Viewer • Updated • 10k • 24Note SDS training dataset (seed 303).
IDEALLab/ShinkaEvolve-SDS-1000-v2-seed101
Viewer • Updated • 1k • 13Note Refreshed neutral-prompt Shinka comparison dataset (seed 101).
IDEALLab/ShinkaEvolve-SDS-1000-v2-seed202
Viewer • Updated • 1k • 21Note Refreshed neutral-prompt Shinka comparison dataset (seed 202).
IDEALLab/ShinkaEvolve-SDS-1000-v2-seed303
Viewer • Updated • 1k • 10Note Refreshed neutral-prompt Shinka comparison dataset (seed 303).
IDEALLab/OpenR1-SDS-Base-Generations-seed101
Viewer • Updated • 64k • 4Note Raw Base model SDS generations (seed 101).
IDEALLab/OpenR1-SDS-Base-Generations-seed202
Viewer • Updated • 64k • 5Note Raw Base model SDS generations (seed 202).
IDEALLab/OpenR1-SDS-Base-Generations-seed303
Viewer • Updated • 64k • 7Note Raw Base model SDS generations (seed 303).
IDEALLab/OpenR1-SDS-Universal-Search-seed101
Preview • Updated • 11Note Universal search artifact bundle over Base-generated SDS solvers (seed 101).
IDEALLab/OpenR1-SDS-Universal-Search-seed202
Preview • Updated • 12Note Universal search artifact bundle over Base-generated SDS solvers (seed 202).
IDEALLab/OpenR1-SDS-Universal-Search-seed303
Preview • Updated • 6Note Universal search artifact bundle over Base-generated SDS solvers (seed 303).
IDEALLab/OpenR1-JSSP-ContractV4-10k-seed101
Viewer • Updated • 10k • 24Note JSSP ContractV4 training dataset (seed 101).
IDEALLab/OpenR1-JSSP-ContractV4-10k-seed202
Viewer • Updated • 10k • 9Note JSSP ContractV4 training dataset (seed 202).
IDEALLab/OpenR1-JSSP-ContractV4-10k-seed303
Viewer • Updated • 10k • 13Note JSSP ContractV4 training dataset (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed101
Reinforcement Learning • 15B • Updated • 8Note Main SDS Hero checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed202
Reinforcement Learning • 15B • Updated • 17Note Main SDS Hero checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303
Reinforcement Learning • 15B • Updated • 4Note Main SDS Hero checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Oracle-seed101
Reinforcement Learning • 15B • Updated • 11Note Original SDS oracle ablation checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Oracle-seed202
Reinforcement Learning • 15B • Updated • 4Note Original SDS oracle ablation checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Oracle-seed303
Reinforcement Learning • 15B • Updated • 10Note Original SDS oracle ablation checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Diversity-seed101
Reinforcement Learning • 15B • Updated • 10Note Original SDS diversity ablation checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Diversity-seed202
Reinforcement Learning • 15B • Updated • 6Note Original SDS diversity ablation checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Diversity-seed303
Reinforcement Learning • 15B • Updated • 10Note Original SDS diversity ablation checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Prompt-seed101
Reinforcement Learning • 15B • Updated • 11Note Original SDS prompt ablation checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Prompt-seed202
Reinforcement Learning • 15B • Updated • 10Note Original SDS prompt ablation checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Prompt-seed303
Reinforcement Learning • 15B • Updated • 4Note Original SDS prompt ablation checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Minimalist-seed101
Reinforcement Learning • 15B • Updated • 6Note Original SDS minimalist checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Minimalist-seed202
Reinforcement Learning • 15B • Updated • 2Note Original SDS minimalist checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Minimalist-seed303
Reinforcement Learning • 15B • Updated • 3Note Original SDS minimalist checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-SoftGate-seed101
Text Generation • 15B • Updated • 12Note SDS soft-gate ablation checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-SoftGate-seed202
Text Generation • 15B • Updated • 11Note SDS soft-gate ablation checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-SoftGate-seed303
Text Generation • 15B • Updated • 8Note SDS soft-gate ablation checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-RewardNormalization-seed101
Text Generation • 15B • Updated • 16Note SDS reward-normalization ablation checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-RewardNormalization-seed202
Text Generation • 15B • Updated • 4Note SDS reward-normalization ablation checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-RewardNormalization-seed303
Text Generation • 15B • Updated • 10Note SDS reward-normalization ablation checkpoint (seed 303).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-JSSP-Hero-V4-seed101
Text Generation • 15B • Updated • 12Note JSSP Hero V4 transfer checkpoint (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-JSSP-Hero-V4-seed202
Text Generation • 15B • Updated • 10Note JSSP Hero V4 transfer checkpoint (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-JSSP-Hero-V4-seed303
Text Generation • 15B • Updated • 4Note JSSP Hero V4 transfer checkpoint (seed 303).
IDEALLab/OpenR1-SDS-SoftGate-Eval-v1-seed101
Viewer • Updated • 3 • 13Note SDS soft-gate evaluation bundle (seed 101).
IDEALLab/OpenR1-SDS-SoftGate-Eval-v1-seed202
Viewer • Updated • 3 • 36Note SDS soft-gate evaluation bundle (seed 202).
IDEALLab/OpenR1-SDS-SoftGate-Eval-v1-seed303
Viewer • Updated • 3 • 23Note SDS soft-gate evaluation bundle (seed 303).
IDEALLab/OpenR1-SDS-RewardNormalization-Eval-v1-seed101
Viewer • Updated • 3 • 15Note SDS reward-normalization evaluation bundle (seed 101).
IDEALLab/OpenR1-SDS-RewardNormalization-Eval-v1-seed202
Viewer • Updated • 3 • 45Note SDS reward-normalization evaluation bundle (seed 202).
IDEALLab/OpenR1-SDS-RewardNormalization-Eval-v1-seed303
Viewer • Updated • 3 • 24Note SDS reward-normalization evaluation bundle (seed 303).
IDEALLab/OpenR1-SDS-BaselineEvidence-Eval-v1
Viewer • Updated • 26 • 28Note Frozen-solver, hand-written SA, timing, and baseline-evidence bundle.
IDEALLab/OpenR1-SDS-FeasibilitySparsity-Logs-v1
Preview • Updated • 97Note Raw feasibility-sparsity and generation-trace logs.
IDEALLab/OpenR1-JSSP-V4-Rebuttal-Eval-v1
Updated • 12Note Three-seed JSSP V4 evaluation and aggregate results.
IDEALLab/Neural-Solver-Synthesis-Final-Evidence-v1
Preview • Updated • 106Note Checksum-verified final evidence, claim index, and provenance pointers.
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed101
Reinforcement Learning • 15B • Updated • 27Note TSP RL policy checkpoint (seed 101); boundary experiment.
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed202
Reinforcement Learning • 15B • Updated • 61Note TSP RL policy checkpoint (seed 202); boundary experiment.
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-TSP-Hero-seed303
Reinforcement Learning • 15B • Updated • 22Note TSP RL policy checkpoint (seed 303); boundary experiment.
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed101
Reinforcement Learning • 15B • Updated • 42Note SDS training-time prompt-ablation checkpoint at step 90 (seed 101).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed202
Reinforcement Learning • 15B • Updated • 24Note SDS training-time prompt-ablation checkpoint at step 90 (seed 202).
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-NoHypothesize-step90-seed303
Reinforcement Learning • 15B • Updated • 30Note SDS training-time prompt-ablation checkpoint at step 90 (seed 303).