EOPSA-DeepSeek-R1-7B

EOPSA (Efficient On-Policy Self-Distilled Safety Alignment) checkpoint of DeepSeek-R1-Distill-Qwen-7B.

Load

from transformers import AutoModelForCausalLM, AutoTokenizer

model_id = "neuqrui/EOPSA-DeepSeek-R1-7B"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
    model_id,
    torch_dtype="auto",
    device_map="auto",
)
Downloads last month
142
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for neuqrui/EOPSA-DeepSeek-R1-7B

Finetuned
(307)
this model
Quantizations
1 model

Collection including neuqrui/EOPSA-DeepSeek-R1-7B