Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

When Text Misleads: Inconsistent-Aware Reasoning for Audio-Grounded Dialogue

A scalable framework that identifies text-biased surface interpretations and converts disagreement regions into conflict QA examples and develops an agentic-style reasoning framework that converts speech into an Audio Twin, a text-readable representation of localized acoustic cues that exposes acoustic evidence to the reasoning model.

Yen-Ju Lu, Yu-Zhe Wang, Yao-Han Guan et al. · 1 citation

Reconstruct! Don't Encode: Self-Supervised Representation Reconstruction Loss for High-Intelligibility and Low-Latency Streaming Neural Audio Codec

It is demonstrated that self-supervised representation reconstruction (SSRR) loss fundamentally improves codec training and performance, and enhances intelligibility by reconstructing distilled self-supervised representations from codec outputs.

Junhyeok Lee, Xiluo He, Jihwan Lee et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.