Make AOT compilation conditional for models >= 2B parameters to optimize free tier usage 4500f92 luigi-liu-tw commited on Oct 12, 2025
Add AOT compilation optimization for ZeroGPU acceleration a7866ff luigi-liu-tw commited on Oct 12, 2025
add 4 20b+ models after enabling dynamic gpu duration fea2910 verified Luigi commited on Oct 12, 2025
Add dynamic duration calculation for ZeroGPU acceleration 6073cc2 luigi-liu-tw commited on Oct 12, 2025
disable two models that cannot run or too run too slowly on hf spaces with zerogpu 3dc7ced luigi-liu-tw commited on Oct 11, 2025
feat(models): add Granite-4.0-Micro and Qwen3-4B-Instruct-2507 to MODELS registry c30a7f7 verified Luigi commited on Oct 9, 2025
add parser_model_ner_gemma_v0 based on gemma 3 370m it bc1bd75 verified Luigi commited on Aug 29, 2025
remove prevously added breeze models (as it didn't work), add smollm 135m taiwan b3fd72e luigi-liu-tw commited on Aug 4, 2025
add Qwen2.5-Taiwan-3B-Reason-GRPO & Llama-3.2-Taiwan-1B f82b9e0 luigi-liu-tw commited on Jul 31, 2025