Layer-wise Curriculum Learning for Efficient LLM Compression Paper • 2609.19213 • Published 20 days ago
Train Overcomplete, Deploy Compact: Scaling Recovery Capacity for Structured LLM Pruning Paper • 2609.06974 • Published 29 days ago
Re-calibrated Contrastive Loss for Transformation-Aware Prompt Conditioning in Vision-Language Models Paper • 2609.06967 • Published 29 days ago
meta-llama/Llama-4-Scout-17B-16E-Instruct Image-Text-to-Text • 109B • Updated May 22, 2025 • 173k • • 1.36k