README
π
Atlas Inference is a pure Rust & ROCm/CUDA LLM inference eng
We built Atlas Inference, a pure Rust LLM inference engine with custom CUDA and ROCm kernels for the NVIDIA DGX Spark GB10 and AMD Strix Halo. Super tiny builds, no external dependencies.