A cost-aware decision layer for routing text requests: a cheap model first, an LLM only when confidence is low, a person when both are unsure. FastAPI service with a React dashboard, measured for accuracy, cost and latency.
react python docker typescript text-classification scikit-learn routing mlops fastapi llm model-cascade
-
Updated
Oct 7, 2026 - TypeScript