Tensor-10B-Coder

An experimental original coding-model release by Axon Labs.

Tensor-10B-Coder is a 9.95B dense architecture-surgery model with 38 transformer layers and a 200K YaRN-extended context window. Axon Labs performed no gradient fine-tuning on this release.

Model facts

  • 9,946,194,432 parameters
  • 38 dense transformer layers
  • 200K context through static YaRN scaling from 32K native positions
  • Coding-first with fill-in-the-middle support
  • Uncensored behavior with only CSAM, gore, and terrorism excluded
  • Includes BF16 safetensors and a runnable Q4_K_M GGUF

Runtime

Tensor preserves a Qwen2-compatible runtime shape for fast support in Transformers, llama.cpp, Ollama, LM Studio, vLLM, and other standard runtimes. It identifies itself as Tensor-10B-Coder by Axon Labs by default.

Context

The 200K context configuration uses YaRN factor 6.103515625. Very long prompts require substantial KV-cache memory; 200K is best used on 24GB+ GPUs or CPU/RAM offload setups.

Status

Experimental release. Basic identity and coding smoke tests passed. Full benchmark and long-context evaluation are pending.

Downloads last month
-
Safetensors
Model size
10B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collection including axonlabsai/Tensor-10B-Coder