Spaces:
Running
Running
|
Download config/codemap.md from Deign86/mathpulse-api-v3test: direct link, hf CLI and curl.
- Browser
- Download file 1.38 kB
-
https://huggingface.co/spaces/Deign86/mathpulse-api-v3test/resolve/main/config/codemap.md
- Command line
-
hf download hf://spaces/Deign86/mathpulse-api-v3test/config/codemap.md
-
curl -L -o codemap.md https://huggingface.co/spaces/Deign86/mathpulse-api-v3test/resolve/main/config/codemap.md
1.38 kB
backend/config/
Responsibility
Holds model routing and AI API cost configuration: models.yaml defines generation/embedding models and task policies; ai_pricing.py exposes DeepSeek pricing tiers.
Design
models.yamlis declarative:models,model_capabilities,routing.task_model_map,task_fallback_model_map, andtask_provider_map.deepseek-reasoneris sequential-only and enabled for reasoning tasks; fallback routes usedeepseek-chat. Embeddings useBAAI/bge-small-en-v1.5and are configured separately from generation.DEEPSEEK_PRICINGcontains per-million-token cache-hit, cache-miss, and output rates; promotional expiry is compared with UTC at call time.get_active_pricing(model_id)andget_full_pricing(model_id)return pricing dictionaries and raiseValueErrorfor unknown IDs.
Flow
Model-routing consumers resolve a task to provider/model and fallback using the YAML mappings; pricing consumers pass model IDs to the pricing helpers, which select active promotion or full rates.
Integration
backend/main.py and inference/model-routing services consume model configuration; RAG uses the embedding model for curriculum retrieval. Cost/reporting code imports config.ai_pricing.get_active_pricing or get_full_pricing; YAML embedding configuration is explicitly separate from the admin-swappable generation pipeline.