Prefill-Free Cross-Family KV Cache Transfer for Heterogeneous Multi-Agent LLMs Paper • 2609.32259 • Published 9 days ago • 95
Tailoring the Quantization Space for 1-Bit KV Cache Compression Paper • 2610.03027 • Published 6 days ago • 3
EXAONE 4.5 Collection LG's First Open-Weight Vision-Language Model for Industrial Intelligence • 5 items • Updated Apr 22 • 47
NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache Paper • 2505.18231 • Published May 23, 2025 • 3
Gaussian Weight Sampling for Scalable, Efficient and Stable Pseudo-Quantization Training Paper • 2505.11170 • Published May 16, 2025 • 3