TextBFGS: A Case-Based Reasoning Approach to Code Optimization via Error-Operator Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Zizheng, Liao, Yuyang, Chen, Chen, He, Jian, Wu, Dun, Yu, Qianjin, Gao, Yanqin, Yang, Jin, Zhang, Kailai, Chng, Eng Siong, Zhong, Xionghu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bridging Speech and Text: Enhancing ASR with Pinyin-to-Character Pre-training in LLMs
di: Yuhang, Yang, et al.
Pubblicazione: (2024)
di: Yuhang, Yang, et al.
Pubblicazione: (2024)
UniArray: Unified Spectral-Spatial Modeling for Array-Geometry-Agnostic Speech Separation
di: Chen, Weiguang, et al.
Pubblicazione: (2025)
di: Chen, Weiguang, et al.
Pubblicazione: (2025)
Noise-Aware Speech Separation with Contrastive Learning
di: Zhang, Zizheng, et al.
Pubblicazione: (2023)
di: Zhang, Zizheng, et al.
Pubblicazione: (2023)
Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization
di: Hu, Yuchen, et al.
Pubblicazione: (2024)
di: Hu, Yuchen, et al.
Pubblicazione: (2024)
Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback
di: Chen, Chen, et al.
Pubblicazione: (2024)
di: Chen, Chen, et al.
Pubblicazione: (2024)
Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
Wav2code: Restore Clean Speech Representations via Codebook Lookup for Noise-Robust ASR
di: Hu, Yuchen, et al.
Pubblicazione: (2023)
di: Hu, Yuchen, et al.
Pubblicazione: (2023)
Zero-shot Context Biasing with Trie-based Decoding using Synthetic Multi-Pronunciation
di: Liu, Changsong, et al.
Pubblicazione: (2025)
di: Liu, Changsong, et al.
Pubblicazione: (2025)
Hierarchical Self-Supervised Representation Learning for Depression Detection from Speech
di: Li, Yuxin, et al.
Pubblicazione: (2025)
di: Li, Yuxin, et al.
Pubblicazione: (2025)
Bi-directional Context-Enhanced Speech Large Language Models for Multilingual Conversational ASR
di: Peng, Yizhou, et al.
Pubblicazione: (2025)
di: Peng, Yizhou, et al.
Pubblicazione: (2025)
Text-based Talking Video Editing with Cascaded Conditional Diffusion
di: Han, Bo, et al.
Pubblicazione: (2024)
di: Han, Bo, et al.
Pubblicazione: (2024)
Proactive for Uncertainty: Cause-Aware Error Diagnosis and Interactive Clarification for Spoken Dialogue Systems
di: Peng, Yizhou, et al.
Pubblicazione: (2026)
di: Peng, Yizhou, et al.
Pubblicazione: (2026)
Prosodic Boundary-Aware Streaming Generation for LLM-Based TTS with Streaming Text Input
di: Liu, Changsong, et al.
Pubblicazione: (2026)
di: Liu, Changsong, et al.
Pubblicazione: (2026)
CS-Sum: A Benchmark for Code-Switching Dialogue Summarization and the Limits of Large Language Models
di: Suresh, Sathya Krishnan, et al.
Pubblicazione: (2025)
di: Suresh, Sathya Krishnan, et al.
Pubblicazione: (2025)
Code-switching Speech Recognition Under the Lens: Model- and Data-Centric Perspectives
di: Liu, Hexin, et al.
Pubblicazione: (2025)
di: Liu, Hexin, et al.
Pubblicazione: (2025)
StreamVoiceAnon+: Emotion-Preserving Streaming Speaker Anonymization via Frame-Level Acoustic Distillation
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin
di: Rao, Abhinav, et al.
Pubblicazione: (2022)
di: Rao, Abhinav, et al.
Pubblicazione: (2022)
Noise-aware Speech Enhancement using Diffusion Probabilistic Model
di: Hu, Yuchen, et al.
Pubblicazione: (2023)
di: Hu, Yuchen, et al.
Pubblicazione: (2023)
VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models
di: Chen, Yukun, et al.
Pubblicazione: (2026)
di: Chen, Yukun, et al.
Pubblicazione: (2026)
EASY: Emotion-aware Speaker Anonymization via Factorized Distillation
di: Yao, Jixun, et al.
Pubblicazione: (2025)
di: Yao, Jixun, et al.
Pubblicazione: (2025)
Continual Learning Optimizations for Auto-regressive Decoder of Multilingual ASR systems
di: Kwok, Chin Yuen, et al.
Pubblicazione: (2024)
di: Kwok, Chin Yuen, et al.
Pubblicazione: (2024)
Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission
di: Thakur, Nirmalya Mallick, et al.
Pubblicazione: (2025)
di: Thakur, Nirmalya Mallick, et al.
Pubblicazione: (2025)
Next-Frame Feature Prediction for Multimodal Deepfake Detection and Temporal Localization
di: Anshul, Ashutosh, et al.
Pubblicazione: (2025)
di: Anshul, Ashutosh, et al.
Pubblicazione: (2025)
Improving Synthetic Data Training for Contextual Biasing Models with a Keyword-Aware Cost Function
di: Kwok, Chin Yuen, et al.
Pubblicazione: (2025)
di: Kwok, Chin Yuen, et al.
Pubblicazione: (2025)
Continual Learning with Embedding Layer Surgery and Task-wise Beam Search using Whisper
di: Kwok, Chin Yuen, et al.
Pubblicazione: (2025)
di: Kwok, Chin Yuen, et al.
Pubblicazione: (2025)
Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models
di: Wu, Donghang, et al.
Pubblicazione: (2025)
di: Wu, Donghang, et al.
Pubblicazione: (2025)
Multi-band Frequency Reconstruction for Neural Psychoacoustic Coding
di: Ng, Dianwen, et al.
Pubblicazione: (2025)
di: Ng, Dianwen, et al.
Pubblicazione: (2025)
Improving Code-Switching Speech Recognition with TTS Data Augmentation
di: Yeo, Yue Heng, et al.
Pubblicazione: (2026)
di: Yeo, Yue Heng, et al.
Pubblicazione: (2026)
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
di: Yao, Jixun, et al.
Pubblicazione: (2025)
di: Yao, Jixun, et al.
Pubblicazione: (2025)
Cross-Modality and Within-Modality Regularization for Audio-Visual DeepFake Detection
di: Zou, Heqing, et al.
Pubblicazione: (2024)
di: Zou, Heqing, et al.
Pubblicazione: (2024)
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
di: Hu, Yuchen, et al.
Pubblicazione: (2024)
di: Hu, Yuchen, et al.
Pubblicazione: (2024)
Aligning Speech to Languages to Enhance Code-switching Speech Recognition
di: Liu, Hexin, et al.
Pubblicazione: (2024)
di: Liu, Hexin, et al.
Pubblicazione: (2024)
DiaSynth: Synthetic Dialogue Generation Framework for Low Resource Dialogue Applications
di: Suresh, Sathya Krishnan, et al.
Pubblicazione: (2024)
di: Suresh, Sathya Krishnan, et al.
Pubblicazione: (2024)
Stream-Voice-Anon: Enhancing Utility of Real-Time Speaker Anonymization via Neural Audio Codec and Language Models
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
di: Kuzmin, Nikita, et al.
Pubblicazione: (2026)
Speech Enhancement Using Continuous Embeddings of Neural Audio Codec
di: Li, Haoyang, et al.
Pubblicazione: (2025)
di: Li, Haoyang, et al.
Pubblicazione: (2025)
From KAN to GR-KAN: Advancing Speech Enhancement with KAN-Based Methodology
di: Li, Haoyang, et al.
Pubblicazione: (2024)
di: Li, Haoyang, et al.
Pubblicazione: (2024)
GenTranslate: Large Language Models are Generative Multilingual Speech and Machine Translators
di: Hu, Yuchen, et al.
Pubblicazione: (2024)
di: Hu, Yuchen, et al.
Pubblicazione: (2024)
LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation
di: Luong, Hieu-Thi, et al.
Pubblicazione: (2024)
di: Luong, Hieu-Thi, et al.
Pubblicazione: (2024)
Large Language Models Meet Contrastive Learning: Zero-Shot Emotion Recognition Across Languages
di: Zou, Heqing, et al.
Pubblicazione: (2025)
di: Zou, Heqing, et al.
Pubblicazione: (2025)
Speech Separation using Neural Audio Codecs with Embedding Loss
di: Yip, Jia Qi, et al.
Pubblicazione: (2024)
di: Yip, Jia Qi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Bridging Speech and Text: Enhancing ASR with Pinyin-to-Character Pre-training in LLMs
di: Yuhang, Yang, et al.
Pubblicazione: (2024) -
UniArray: Unified Spectral-Spatial Modeling for Array-Geometry-Agnostic Speech Separation
di: Chen, Weiguang, et al.
Pubblicazione: (2025) -
Noise-Aware Speech Separation with Contrastive Learning
di: Zhang, Zizheng, et al.
Pubblicazione: (2023) -
Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization
di: Hu, Yuchen, et al.
Pubblicazione: (2024) -
Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback
di: Chen, Chen, et al.
Pubblicazione: (2024)