amd/Instella-MoE-16B-A3B-SFT-w4a16-llmcompressor Text Generation • 15B • Updated 4 days ago • 115 • 1
amd/Instella-MoE-16B-A3B-SFT-w4a16-llmcompressor Text Generation • 15B • Updated 4 days ago • 115 • 1
Instella: Fully Open Language Models with Stellar Performance Paper • 2511.10628 • Published Nov 13, 2025 • 7
PARD: Accelerating LLM Inference with Low-Cost PARallel Draft Model Adaptation Paper • 2504.18583 • Published Apr 23, 2025
SAND-Math: Using LLMs to Generate Novel, Difficult and Useful Mathematics Questions and Answers Paper • 2507.20527 • Published Jul 28, 2025 • 7
DL-QAT: Weight-Decomposed Low-Rank Quantization-Aware Training for Large Language Models Paper • 2504.09223 • Published Apr 12, 2025
AMD-Hummingbird: Towards an Efficient Text-to-Video Model Paper • 2503.18559 • Published Mar 24, 2025 • 5
Agent Laboratory: Using LLM Agents as Research Assistants Paper • 2501.04227 • Published Jan 8, 2025 • 95
The Nature of Mathematical Modeling and Probabilistic Optimization Engineering in Generative AI Paper • 2410.18441 • Published Oct 24, 2024 • 6
TalkMosaic: Interactive PhotoMosaic with Multi-modal LLM Q&A Interactions Paper • 2409.13941 • Published Sep 20, 2024
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks Paper • 2409.09240 • Published Sep 14, 2024