feat(zerogpu): clean CPU-to-CUDA model movement and teardown in @spaces.GPU 2ff0e02 verified devinblaze commited on 1 day ago
fix(scores): parse score criteria into list of levels cda313d verified devinblaze commited on 1 day ago
fix(vram): enforce single model in VRAM and flexible criteria parsing bf6cb06 verified devinblaze commited on 1 day ago
fix: order 3b model detection before 7b fallback in resolve_model_id 8d59ccf AI Bot commited on 1 day ago
feat: use verified Qwen2.5-Coder-3B and Qwen2.5-7B-abliterated models 082f79a AI Bot commited on 1 day ago
fix: make chat_completions synchronous and return traceback on error db3fa9b AI Bot commited on 1 day ago
fix: deploy 3-model hybrid engine with ZeroGPU startup handshake e8ee6f4 AI Bot commited on 1 day ago
fix: parse ChatInterface messages and launch uvicorn for root FastAPI endpoints f18f3d3 AI Bot commited on 1 day ago
fix: import spaces before torch and initialize ZeroGPU with demo.launch f684d52 AI Bot commited on 1 day ago
fix: decorate Gradio handlers with @spaces.GPU and remove pypi spaces from requirements 3f78a4c AI Bot commited on 1 day ago
feat: deploy Laya System 1 + 3B abliterated + 7B abliterated super-stack 685ca6a AI Bot commited on 1 day ago