{ "intent_classifier_ms": 0.309, "anomaly_detector_ms": 17.703, "kb_retrieval_ms": 0.446, "note": "LLM generation latency depends on the external Inference API call and is measured live in the app, not benchmarked here." }