Saved in:
| Main Authors: | Wu, Haibin, Do, Bach Viet, Suda, Naveen, Chan, Julian, R, Madhavan C, Yang, Gene-Ping, Wu, Yi-Chiao, Kanda, Naoyuki, Adi, Yossef, Lei, Xin, Liu, Yue, Metze, Florian, Liu, Yuzong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.20094 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning
by: Ouyang, Zhicheng, et al.
Published: (2026)
by: Ouyang, Zhicheng, et al.
Published: (2026)
MimiTalk: Revolutionizing Qualitative Research with Dual-Agent AI
by: Liu, Fengming, et al.
Published: (2025)
by: Liu, Fengming, et al.
Published: (2025)
Juego de ladrones / Director, Mimi Leder
Published: (2009)
Published: (2009)
Llama-Mimi: Exploring the Limits of Flattened Speech Language Modeling
by: Sugiura, Issa, et al.
Published: (2025)
by: Sugiura, Issa, et al.
Published: (2025)
Crevasse Detection using MimiNet and Sentinel-1 SAR on Greenland
by: Khan, Naureen, et al.
Published: (2025)
by: Khan, Naureen, et al.
Published: (2025)
Reseña de "Consuming the Caribbean: From Arawaks to Zombies" de Mimi Sheller
by: Linden Lewis
Published: (2005)
by: Linden Lewis
Published: (2005)
MimiC: Combating Client Dropouts in Federated Learning by Mimicking Central Updates
by: Sun, Yuchang, et al.
Published: (2023)
by: Sun, Yuchang, et al.
Published: (2023)
MimiCAT: Mimic with Correspondence-Aware Cascade-Transformer for Category-Free 3D Pose Transfer
by: Chai, Zenghao, et al.
Published: (2025)
by: Chai, Zenghao, et al.
Published: (2025)
TS3-Codec: Transformer-Based Simple Streaming Single Codec
by: Wu, Haibin, et al.
Published: (2024)
by: Wu, Haibin, et al.
Published: (2024)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
by: Choi, Kanghyun, et al.
Published: (2024)
by: Choi, Kanghyun, et al.
Published: (2024)
JASTIN: Aligning LLMs for Zero-Shot Audio and Speech Evaluation via Natural Language Instructions
by: Zhang, Leying, et al.
Published: (2026)
by: Zhang, Leying, et al.
Published: (2026)
A Simple HMM with Self-Supervised Representations for Phone Segmentation
by: Yang, Gene-Ping, et al.
Published: (2024)
by: Yang, Gene-Ping, et al.
Published: (2024)
Equipping LLM with Directional Multi-Talker Speech Understanding Capabilities
by: Lin, Ju, et al.
Published: (2026)
by: Lin, Ju, et al.
Published: (2026)
Reseña de "Toward a Global PhD? Forces & Forms in Doctoral Education Worldwide" de NERAD, Maresi; HEGGELUND, Mimi
by:
Published: (2011)
by:
Published: (2011)
A filtered life: Social media on a college campus By Nicole Taylor and Mimi Nichter. New York: Routledge, 2022. 210 pp.
by: Patricia G. Lange
Published: (2024)
by: Patricia G. Lange
Published: (2024)
Thelma & Louise [Película] : un final inesperado = Thelma & Louis / director y productor, Ridley Scott ; productor, Mimi Polk ; guionista, Callie Khouri
by: Scott, Ridley
Published: (1991)
by: Scott, Ridley
Published: (1991)
Arqueología del nuevo mundo [Documental] : Canibalismo = Cannibals and ice age crossings / Mimi Edmunds, directora ; Tom Naughton y Nicolas Valcour, productores
by: Edmunds, Mimi
Published: (1993)
by: Edmunds, Mimi
Published: (1993)
E2 TTS: Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS
by: Eskimez, Sefik Emre, et al.
Published: (2024)
by: Eskimez, Sefik Emre, et al.
Published: (2024)
Forecasting Automotive Supply Chain Shortfalls with Heterogeneous Time Series
by: Do, Bach Viet, et al.
Published: (2024)
by: Do, Bach Viet, et al.
Published: (2024)
An Investigation of Noise Robustness for Flow-Matching-Based Zero-Shot TTS
by: Wang, Xiaofei, et al.
Published: (2024)
by: Wang, Xiaofei, et al.
Published: (2024)
SLM-TTA: A Framework for Test-Time Adaptation of Generative Spoken Language Models
by: Wu, Yuan-Kuei, et al.
Published: (2025)
by: Wu, Yuan-Kuei, et al.
Published: (2025)
Enhancing TTS Stability in Hebrew using Discrete Semantic Units
by: Zeldes, Ella, et al.
Published: (2024)
by: Zeldes, Ella, et al.
Published: (2024)
Spontaneous Perovskite Passivators Effectively Combined with PTAA Hole‐Transport Materials in Perovskite Solar Cells
by: Naoyuki Nishimura, et al.
Published: (2025)
by: Naoyuki Nishimura, et al.
Published: (2025)
A Language Modeling Approach to Diacritic-Free Hebrew TTS
by: Roth, Amit, et al.
Published: (2024)
by: Roth, Amit, et al.
Published: (2024)
On the Optimality of Decode and Forward for Some Cooperative Broadcast Channels
by: Gouic, Nicolas Le, et al.
Published: (2026)
by: Gouic, Nicolas Le, et al.
Published: (2026)
Profile-Error-Tolerant Target-Speaker Voice Activity Detection
by: Wang, Dongmei, et al.
Published: (2023)
by: Wang, Dongmei, et al.
Published: (2023)
Uncovering Heterogeneity of Solar Flare Mechanism With Mixture Models
by: Do, Bach Viet, et al.
Published: (2024)
by: Do, Bach Viet, et al.
Published: (2024)
MASV: Speaker Verification with Global and Local Context Mamba
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Chapter Final Phases of Medieval Hebraism
by: Schwartz, Yossef
Published: (2019)
by: Schwartz, Yossef
Published: (2019)
Relay Channels with Unreliable Helpers
by: Steinberg, Yossef
Published: (2024)
by: Steinberg, Yossef
Published: (2024)
Degradedness Under Cooperation
by: Steinberg, Yossef
Published: (2025)
by: Steinberg, Yossef
Published: (2025)
Hydrophilic or Hydrophobic? Spontaneous Chemical Capping with Bis(trifluoromethanesulfonyl)imide‐Based Additive for Photoabsorbers in Perovskite Solar Cells
by: Naoyuki Nishimura, et al.
Published: (2025)
by: Naoyuki Nishimura, et al.
Published: (2025)
CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
PhoneWorld: Scaling Phone-Use Agent Environments
by: Tang, Zhengyang, et al.
Published: (2026)
by: Tang, Zhengyang, et al.
Published: (2026)
Kratos: An FPGA Benchmark for Unrolled DNNs with Fine-Grained Sparsity and Mixed Precision
by: Dai, Xilai, et al.
Published: (2024)
by: Dai, Xilai, et al.
Published: (2024)
On the Effectiveness of Acoustic BPE in Decoder-Only TTS
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Laugh Now Cry Later: Controlling Time-Varying Emotional States of Flow-Matching-Based Zero-Shot Text-to-Speech
by: Wu, Haibin, et al.
Published: (2024)
by: Wu, Haibin, et al.
Published: (2024)
HD-PPT: Hierarchical Decoding of Content- and Prompt-Preference Tokens for Instruction-based TTS
by: Nie, Sihang, et al.
Published: (2025)
by: Nie, Sihang, et al.
Published: (2025)
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch
by: Song, Xingchen, et al.
Published: (2024)
by: Song, Xingchen, et al.
Published: (2024)
Influence of Pond System on Surface Temperature and Urban Cooling
by: Quang‐Viet Nguyen, et al.
Published: (2025)
by: Quang‐Viet Nguyen, et al.
Published: (2025)
Similar Items
-
Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning
by: Ouyang, Zhicheng, et al.
Published: (2026) -
MimiTalk: Revolutionizing Qualitative Research with Dual-Agent AI
by: Liu, Fengming, et al.
Published: (2025) -
Juego de ladrones / Director, Mimi Leder
Published: (2009) -
Llama-Mimi: Exploring the Limits of Flattened Speech Language Modeling
by: Sugiura, Issa, et al.
Published: (2025) -
Crevasse Detection using MimiNet and Sentinel-1 SAR on Greenland
by: Khan, Naureen, et al.
Published: (2025)