OossaAI is evolving fast. We explain it simply.
Newsletter

Note · 1 min read

LLM filter trims errors in AI transcriptions of noisy police audio

Researchers introduced an LLM‑based filter to improve pseudo‑labeling for speech‑recognition on noisy broadcast police communications, lowering transcription errors.

Oossa · About Oossa

The team led by Kaavya Chaparala evaluated two foundation speech‑recognition models—OpenAI’s Whisper and Qwen3‑ASR—on noisy broadcast police communications from Baltimore and Chicago. They found the models’ built‑in confidence scores could not separate good from bad machine‑generated transcripts. By using a large language model as an external judge to discard context‑implausible pseudo‑labels, they cut the word‑error rate of the training data. The paper, accepted to the SLT 2026 conference, also proposes swapping pseudo‑labels between models as a future direction.

Why it matters

The method makes it cheaper and more accurate to turn noisy police recordings into readable text without extensive human labeling.

Was this article useful?

Sources & references

#SourceOutletDateKey takeaway
1Pretrained ASR Pseudo-labeling for Noisy Police Audio ↗arXivSep 28, 2026arXiv:2609.30469v1 Announce Type: new Abstract: Pretrained ASR systems perform poorly on noisy Broadcast Police Communication (BPC), hinderi

1 sources

Last updated:

Oossallms.txt.md

Share

Oossa · Newsletter

The week in AI, explained

Every Monday: the stories worth knowing, in plain language. Free, no spam.

LLM filter trims errors in AI transcriptions of noisy police audio – Oossa