Fast EU LLMs for intent detection: 19 deployments compared through a production request path
We compare 1,140 model-scenario runs across short and long conversational contexts. The results show why model size alone is a poor proxy for the accuracy, response time and cost balance a live voice or support agent needs.
Read the full benchmarkIntent30/30
Short p95338 ms
Long p95401 ms
Rows19
Fastest p95 response among perfect-accuracy deployments in each context suite.