Efficient Solar Open Family Collection Solar Open models officially optimized by Nota AI for Korea’s government-led Sovereign AI initiative as part of the Upstage consortium. • 7 items • Updated 8 days ago • 25
Quantize the Target, Quantize the Drafter: Efficient Inference with Qwen3.5-4B Paper • 2607.04244 • Published 26 days ago • 1
Efficient MoE-based LLM Collection Mixture-of-Experts Large Language Models with Advanced Quantization • 5 items • Updated Mar 11 • 25
Efficient Large Vision-Language Model Collection ERGO: LVLM trained with RL on efficiency objectives; https://github.com/nota-github/ERGO • 3 items • Updated Feb 22 • 27
Llama 3.1 Collection This collection hosts the transformers and original repos of the Llama 3.1, Llama Guard 3 and Prompt Guard models • 11 items • Updated Dec 6, 2024 • 715
Shortened LLaMA: A Simple Depth Pruning for Large Language Models Paper • 2402.02834 • Published Feb 5, 2024 • 17
Efficient Large Language Model Collection Shortened LLMs from Depth Pruning; https://github.com/Nota-NetsPresso/shortened-llm • 14 items • Updated Apr 2, 2025 • 5
Efficient Stable Diffusion Collection Block-removed Knowledge-distilled SD models; https://github.com/Nota-NetsPresso/BK-SDM • 9 items • Updated Jul 1, 2024 • 6