New API Rates AI Model Hallucination Scores

TL;DR. A new API offers per-token hallucination scores for frontier language models by using a white-box proxy model to infer internal confusion. - The API works by routing black-box LLM queries through an observable open-weight model. - This proxy model is trained against human-judged GPT-5.4 Nano outputs to identify hallucinated claims. - Scores indicate the likelihood of the proxy model being confused, agreeing with human judges over 95% of the time.

Sources

Back to QLANKR News