Skip to content
VuFind
  • Login
    • English
    • Deutsch
    • Español
    • Français
    • Italiano
Advanced
  • Cite this
  • Text this
  • Email this
  • Print
  • Export Record
    • Export to RefWorks
    • Export to EndNoteWeb
    • Export to EndNote
  • Save to List
  • Permanent link
Cover Image

Saved in:
Bibliographic Details
Main Authors: Zhang, Yuxin, Zhang, Xiangyu Tony, Liu, Daijiao, Tian, Fei, Deng, Yayue, Chen, Jun, Lin, Qingjian, Zhang, Haoyang, Li, Yuxin, Gong, Jinglan, Huang, Yechang, Zhao, Liang, Yao, Chengyuan, Liu, Hexin, Chng, Eng Siong, Yang, Xuerui, Yu, Gang, Zhang, Xiangyu, Jiang, Daxin
Format: Preprint
Published: 2026
Subjects:
Audio and Speech Processing
Online Access:https://arxiv.org/abs/2604.25719
Tags: Add Tag
No Tags, Be the first to tag this record!
  • Holdings
  • Description
  • Table of Contents
  • Comments
  • Similar Items
  • Staff View

Internet

https://arxiv.org/abs/2604.25719

Similar Items

  • Step-Audio-R1 Technical Report
    by: Tian, Fei, et al.
    Published: (2025)
  • DuplexSLA: A Full-Duplex Spoken Language Model with Synchronized Speech, Language, and Action
    by: Zhang, Haoyang, et al.
    Published: (2026)
  • MULTI-Bench: A Multi-Turn Interactive Benchmark for Assessing Emotional Intelligence ability of Spoken Dialogue Models
    by: Deng, Yayue, et al.
    Published: (2025)
  • Speaking in Wavelet Domain: A Simple and Efficient Approach to Speed up Speech Diffusion Model
    by: Zhang, Xiangyu, et al.
    Published: (2024)
  • Code-switching Speech Recognition Under the Lens: Model- and Data-Centric Perspectives
    by: Liu, Hexin, et al.
    Published: (2025)

Search Options

  • Search History
  • Advanced Search

Find More

  • Browse the Catalog
  • Browse Alphabetically
  • Explore Channels
  • Course Reserves
  • New Items

Need Help?

  • Search Tips
  • Ask a Librarian
  • FAQs