We built Atlas Inference, a pure Rust LLM inference engine with custom CUDA and ROCm kernels for the NVIDIA DGX Spark GB10 and AMD Strix Halo. Super tiny builds, no external dependencies.