arxiv:2609.02652
Malandrino PRO
Pier-Jean
AI & ML interests
Document parsing
Quantization
Recent Activity
liked a Space 2 days ago
convaiinnovations/laya-demo authored a paper 21 days ago
Unfolding the Leech Lattice: Fused Multi-Shell Decoding and VRAM Layouts for 2-Bit LLM Weights published an article 26 days ago
I put a 24 dimensional lattice in a CUDA kernel to run Qwen3-4B in 2.6 GB