LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes and What Recovers It
- View PDF HTML (experimental) Abstract:Ambient AI scribes draft clinical notes, and published audits find their dominant error is omission: information the encounter established that the note fails to record.
- The standard check is an LLM judge: a second model reads the note against the transcript and flags problems.
Unverified
- View PDF HTML (experimental) Abstract:Ambient AI scribes draft clinical notes, and published audits find their dominant error is omission: information the encounter established that the note fails to record.
- The standard check is an LLM judge: a second model reads the note against the transcript and flags problems.
Sources: Arxiv