AI accuracy & hallucination report
Every Hanah region is graded by the same harness, the same ground-truth corpus, and the same published methodology: Word Error Rate against verbatim transcripts, plus an LLM jury panel grading coverage and hallucination. Each region runs its own speech and language stack, so each is graded — and published — separately. The headline numbers below render from the exact data each regional report publishes; open a region for the per-recording transcripts, generated documents, and per-juror verdicts behind them.
Headline numbers are rendered client-side from each region's published harness aggregate — the same file that drives the full report, so this page can never disagree with it. A hallucination is only counted when at least two independent jurors agree.
Looking for security controls, subprocessors and data residency instead? Those are documented per region in the Trust Center.
