SWAMP: Sparse Weight Averaging with Multiple Particles for Iterative Magnitude Pruning Paper • 2305.14852 • Published May 24, 2023 • 3
Harness-Aware Distillation for Small Language Model Agents Paper • 2610.02858 • Published 7 days ago • 8
TAPE: Tool-Guided Adaptive Planning and Constrained Execution in Language Model Agents Paper • 2602.19633 • Published Feb 23 • 10
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens Paper • 2412.10208 • Published Dec 13, 2024 • 19