# Qwen2.5-Omni-3B adapters boost entity recall in accented conversational ASR

Digest AI · Research · published 2026-09-21T04:00:00Z

Canonical: https://digestai.news/story/qwen2-5-omni-3b-adapters-boost-entity-recall-in-accented-conversationa

## Summary

A new three‑stage pipeline targets accented English speech from India, Indonesia and Latin America, where standard ASR systems optimized for word error rate often miss named entities and filler words. First, heuristic SQL filters curate training data that is 2.8 × richer in entities than random samples. Second, regional LoRA adapters fine‑tuned on the Qwen2.5‑Omni‑3B model generate both verbatim and corrected transcripts in a single forward pass. Third, a six‑category error taxonomy is evaluated by an LLM‑based judge, achieving 83.8 % agreement on 210 human‑labelled samples.

On a 6 k‑utterance test set the pipeline reaches 80‑85 % entity recall (up from 53‑55 %) and 76‑86 % filler recall (up from under 5 %), while keeping word error rate between 6 % and 10 %. It outperforms Whisper and a commercial ASR on entity recall and matches a zero‑shot 30 B‑parameter model with ten‑fold fewer parameters. Paired bootstrap tests attribute 2.8‑4.2 percentage‑point gains in entity recall to data curation alone (p < 0.0001).

## Key points

- Heuristic SQL filters produce training data with 2.8× higher entity density than random sampling.
- LoRA adapters fine‑tuned on Qwen2.5‑Omni‑3B raise entity recall to 80‑85% and filler recall to 76‑86% on accented speech.
- Pipeline matches a zero‑shot 30B model’s performance with only 3B parameters and 10× fewer resources.

## Why it matters

Improving entity and filler detection in accented ASR helps language‑learning tools provide more accurate feedback, expanding accessibility for non‑native speakers.

## Sources

1. [Beyond WER: Entity and Disfluency Recall in Accented Conversational ASR](https://arxiv.org/abs/2609.20828) (arXiv cs.CL, 2026-09-21, primary source)

Part of the developing story: [AI Whisper Advances Multilingual Speech Recognition](https://digestai.news/thread/token-merging-boosts-whisper-efficiency-across-16-languages-with-minimal) (2 stories)

## Cite

Digest AI, "Qwen2.5-Omni-3B adapters boost entity recall in accented conversational ASR", 21 September 2026, https://digestai.news/story/qwen2-5-omni-3b-adapters-boost-entity-recall-in-accented-conversationa

---

Written by Digest AI's editorial model from the linked sources; the sources are the record. Headlines, digests and key points are written by Digest AI and may be quoted with a link to the story page. Linked articles belong to their publishers. Terms: https://digestai.news/terms#reuse
JSON: https://digestai.news/story/qwen2-5-omni-3b-adapters-boost-entity-recall-in-accented-conversationa.json
