Pith. sign in

REVIEW 3 cited by

Beyond Transcription: Mechanistic Interpretability in ASR

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2508.15882 v1 pith:ICAEJSIF submitted 2025-08-21 cs.SD cs.CLcs.LGeess.AS

Beyond Transcription: Mechanistic Interpretability in ASR

classification cs.SD cs.CLcs.LGeess.AS
keywords interpretabilityacoustichallucinationsinsightsmethodsmodelrecognitionrepresentations
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Interpretability methods have recently gained significant attention, particularly in the context of large language models, enabling insights into linguistic representations, error detection, and model behaviors such as hallucinations and repetitions. However, these techniques remain underexplored in automatic speech recognition (ASR), despite their potential to advance both the performance and interpretability of ASR systems. In this work, we adapt and systematically apply established interpretability methods such as logit lens, linear probing, and activation patching, to examine how acoustic and semantic information evolves across layers in ASR systems. Our experiments reveal previously unknown internal dynamics, including specific encoder-decoder interactions responsible for repetition hallucinations and semantic biases encoded deep within acoustic representations. These insights demonstrate the benefits of extending and applying interpretability techniques to speech recognition, opening promising directions for future research on improving model transparency and robustness.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. HALAS: A Human-Annotated Dataset of Hallucinations of Modern ASR Systems

    cs.SD 2026-06 unverdicted novelty 7.0

    HALAS is a human-annotated dataset of ASR hallucinations on unprocessed real audio that shows simple metrics outperform current detection methods at 81% ROC-AUC versus 53.1% F1.

  2. Interleaved Speech Language Models Latently Work In Text

    cs.CL 2026-06 unverdicted novelty 7.0

    Interleaved SLMs implicitly transcribe spoken words to text tokens in middle layers (top candidate for 77% of data) before predicting in text space and returning to speech.

  3. From Text Metrics to Model Internals: A Study of Whisper ASR Hallucination Detection

    cs.SD 2026-06 unverdicted novelty 5.0

    Internal decoder probing of Whisper yields strongest hallucination detection without references, with late fusion of text and internal features performing best overall.