Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-03T15:17:52.966144Z
Paper Citation Record · LEDGER
As of 23 July 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2607.01733.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-03T15:17:52.966144Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-20T06:30:07.809122+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f57687bf-c0dd-4370-85cd-f6b593a5903f · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c14b926b-2300-44fb-a8c9-dc6991f28493 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e0a4d62e-4afe-47d2-8805-8271eb57c0d5 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving DeepSeek-V3 Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation eebeee84-2976-4e6f-a23b-273115040c83 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Gemma 3 Technical Report
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 6b030ef0-c96b-4223-b2b0-11a820079f2e · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Qwen3 Technical Report
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e31fa4ac-17ff-4e38-9ed5-fc0c8c32a2c2 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving On decoder-only architecture for speech-to-text and large language model integration,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c6928632-7829-4ac6-914f-87605b590ba5 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Moshi: a speech-text foundation model for real-time dialogue
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c29d5d96-2aab-4c0c-ab5e-b27a92a4c594 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving SALMONN: Towards generic hearing abilities for large language models,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation f9023051-45ca-4da1-873d-621a9a42150e · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation be537fd5-64fa-4663-ab30-6a2d14204140 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Kimi-Audio Technical Report
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 0eabac8e-4d8b-4533-9cd6-a8a2bd50c0c7 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Step-Audio 2 Technical Report
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 0607b9e9-9c57-4a93-8989-ad7adb3987e7 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Qwen3-Omni Technical Report
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 37054825-977e-43ca-9461-dfec0a768d26 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving SLM-S2ST: A multimodal language model for direct speech- to-speech translation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation dc79e153-5fc8-4d81-a68f-7c8ee5c5e2e1 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Fun-audio-chat technical report
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation acc687e0-3e9e-45ef-bc4b-ed5ef2f99f97 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e18ee8f2-5b9b-449a-9983-987079b5e7ee · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation baac7d1e-82d0-4cc3-8c7a-5cbd29965572 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Fun-ASR technical report
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 97134e9f-f059-4a2d-b161-c1ef3e4f075f · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Index-asr technical report
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation b43b9f70-3b2c-4534-9adb-97d9ca5a04ec · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Speech recognition meets large language model: Benchmarking, models, and exploration,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 2fe4713a-4cc0-4c1f-b756-3dd2136996bf · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Efficient Scaling for LLM-based ASR
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation dc0ac3f7-8d2d-4672-aea4-6602281cda39 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Qwen3-ASR Technical Report
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation ebfd6284-540c-4606-99ae-a97fd732d2ee · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Transducer-Llama: Integrating LLMs into streamable transducer-based speech recognition,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 98c3d0c3-1d63-4b9d-a12f-d15df855b221 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 7d9bec23-5f67-4efb-aea1-38cc4007313e · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Train short, infer long: Speech-llm enables zero-shot streamable joint asr and di- arization on long audio
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 6873758e-e06e-483d-99fa-95bc7cb828e4 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Rlbr: Reinforcement learning with biasing rewards for contextual speech large language models,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 0094fa6d-1b31-4dbc-a700-7eaf59c40f09 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Wav2Prompt: End-to-end speech prompt learning and task-based fine-tuning for text-based LLMs,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 77991ea3-a50e-4d17-bde4-23c4292930af · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Alignformer: Modality matching can achieve better zero-shot instruction-following speech-LLM,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 9bc29d3b-2f9f-481c-bfbf-79b0afc3686f · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Qwen2.5-Omni Technical Report
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation fe7dcb65-5b71-4ae1-8988-2dca377b95f5 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Voxtral
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation dc8365fa-c3df-49c8-88a1-862a96b22064 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving SpiRit- LM: Interleaved spoken and written language model,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e1fd989b-9ad9-4711-90a9-02e36a14790e · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Available: https://aclanthology.org/2025.tacl-1.2/
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation ed1c8b03-cc89-4e2c-9120-56f69f53ecad · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 8b3f074c-f435-42a4-9c1f-2dae65772be2 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Injecting text in self-supervised speech pretraining,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 21415aeb-b8a2-45dd-b510-86247b70f52e · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 8a97ec94-c0d9-4817-b836-b62020c01b01 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving SpeechT5: Unified-modal encoder- decoder pre-training for spoken language processing,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 3a4a95b5-2717-4677-8651-6b50b3fe1c73 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving SLAM: A Unified Encoder for Speech and Language Modeling via Speech-Text Joint Pre-Training
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation cfda2c70-b20e-4817-a848-9eb39bbaa4f0 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving JOIST: A joint speech and text streaming model for ASR,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation a186f019-b909-431d-9b18-7cd3a0f7e4bd · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Joint unsupervised and supervised training for multilingual ASR,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c86a68ee-96a5-42c8-88e1-b5c91b4d4447 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving FastInject: Injecting unpaired text data into CTC-based ASR training,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c141dfe8-06de-4a95-80b0-92621b7dc88b · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Multitask training with text data for end-to-end speech recognition,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 73544a1e-b578-4357-9f5b-ad08ee41205c · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving An attention-based joint acoustic and text on-device end-to-end model,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 4c4564b0-70a4-49b4-b86a-137e4b02e2ae · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Maestro: Matched speech text representations through modality matching,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation da805f65-5027-46ca-9a29-cb9a8717769f · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Improving joint speech-text repre- sentations without alignment,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 32ec8a36-d24d-4a32-b485-a94b63aa6d2a · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Measuring Massive Multitask Language Understanding
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 4cd0231e-454d-43fa-a999-c319ff7b00a3 · outbound
Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Ultraeval-audio: A unified framework for comprehensive evaluation of audio foundation models,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
No inbound Pith citation observations are available.