Aura-1-Thinking

This repository contains the Aura-1-Thinking reasoning model delivered as a standard base layer for native PyTorch environments.

This engine is tailor-made to deliver high-speed token streaming, running smoothly on accessible hardware layouts including standard cloud instances with 16GB RAM.


Performance Footprint

  • Baseline Raw (FP16): ~8.5 GB - 10.0 GB | Standard CPU / Basic GPU Space
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 1 Ask for provider support