Intern-S2-Preview-OPD

Intern-S2-Preview-OPD is a post-trained version of Intern-S2-Preview, an efficient 35B scientific multimodal foundation model.

The model is post-trained using the method introduced in SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation of Long-Context Reasoning. This post-training process substantially improves the model's reasoning performance, particularly on challenging proof and mathematical reasoning benchmarks.

Evaluation Results

Model ProofBench AnswerBench AIME25 AMOBench
Intern-S2-Preview 21.70 76.03 88.33 58.00
Intern-S2-Preview-OPD 44.50 (+22.80) 80.10 (+4.07) 95.00 (+6.67) 59.50 (+1.50)
SU-01 (Teacher Model) 45.00 77.50 94.60 61.75

The values in parentheses indicate absolute improvements over the base model.

Usage

Intern-S2-Preview-OPD uses the same model architecture and inference interface as Intern-S2-Preview. Please refer to the Intern-S2-Preview model card for deployment instructions and recommended inference settings.

License

This model is released under the Apache License 2.0.

Downloads last month
17
Safetensors
Model size
36B params
Tensor type
F32
·
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for bingyang-lei/Intern-S2-Preview-SimpleOPD

Finetuned
(1)
this model

Collection including bingyang-lei/Intern-S2-Preview-SimpleOPD