{ "intent_classifier_ms": 0.633, "anomaly_detector_ms": 14.676, "kb_retrieval_ms": 0.862, "note": "LLM generation latency depends on the external Inference API call and is measured live in the app, not benchmarked here." }