
Teaching a language model to stop thinking out loud
In non-reasoning mode, GLM-5.2 was leaking its reasoning into the AI assistant’s response in about one conversation in a hundred.

In non-reasoning mode, GLM-5.2 was leaking its reasoning into the AI assistant’s response in about one conversation in a hundred.
We reduced voice agent hallucination rates from 4-5% to less than 1% in production without adding latency. LLMs generate text faster than humans can speak it. That speed gap is where we run detection.
Giga's AI agents handle complex workflows at scale, from live delivery issues to compliance decisions, while maintaining over 90% resolution accuracy in production.