🏛️ Zen Alta 1-3B (Initial Baseline / Legacy)

This repository contains the initial unpruned 28-layer baseline (3.21B parameters) from the early stages of the Zen Alta project, including raw FP16 safetensors shards and an unpruned Q4_K_M GGUF.

Active Production Releases: For the faster, 24-layer pruned & DPO-aligned production model and its speculative draft counterpart, please see:


📦 Repository Files

  • model-00001-of-00002.safetensors / model-00002-of-00002.safetensors: Full 28-layer base weights
  • llama-3.2-3b-instruct.Q4_K_M.gguf: Unpruned baseline GGUF (1.92 GB)
  • Modelfile: Ollama configuration
Downloads last month
137
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for ZenithLLM/zen-alta-1-3b

Quantized
(538)
this model