How to use from
vLLM
Install from pip and serve model
# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "allenai/Dense_1b_130B"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "allenai/Dense_1b_130B",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'
Use Docker
docker model run hf.co/allenai/Dense_1b_130B
Quick Links

Dense_1b_130B

An active-parameter-matched dense baseline released alongside EMO: Pretraining Mixture of Experts for Emergent Modularity — referred to as "Dense @ 8" in Figure 1 of the paper. Not midtrained.

1B parameter dense decoder-only Transformer (no MoE) pretrained from scratch on 130B tokens of the OLMoE pretraining mix. Provides an active-parameter-matched comparison point against 8-expert subsets carved out of the larger 1B/14B EMO models.

Usage

from transformers import AutoModelForCausalLM, AutoTokenizer

model_id = "allenai/Dense_1b_130B"
model = AutoModelForCausalLM.from_pretrained(model_id, trust_remote_code=True)
tokenizer = AutoTokenizer.from_pretrained(model_id, trust_remote_code=True)

inputs = tokenizer(["Language modeling is "], return_tensors="pt", return_token_type_ids=False)
out = model.generate(**inputs, max_new_tokens=100, do_sample=True, temperature=1.0, top_p=0.7)
print(tokenizer.batch_decode(out, skip_special_tokens=True)[0])

Citation

@article{wang2026emo,
  title  = {EMO: Pretraining Mixture of Experts for Emergent Modularity},
  author = {Wang, Ryan and Bhagia, Akshita and Min, Sewon},
  year   = {2026},
  url    = {https://arxiv.org/abs/2605.06663}
}

License

This model is licensed under Apache 2.0. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines

Links

Downloads last month
290
Safetensors
Model size
1B params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Dataset used to train allenai/Dense_1b_130B

Collection including allenai/Dense_1b_130B

Paper for allenai/Dense_1b_130B

Article mentioning allenai/Dense_1b_130B