DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Yu, An, Sohyun, Deng, Haikang, Yin, Da, Peng, Clark, Hsieh, Cho-Jui, Chang, Kai-Wei, Peng, Nanyun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Decoupling Task-Solving and Output Formatting in LLM Generation
por: Deng, Haikang, et al.
Publicado: (2025)
por: Deng, Haikang, et al.
Publicado: (2025)
On the Loss of Context-awareness in General Instruction Fine-tuning
por: Wang, Yihan, et al.
Publicado: (2024)
por: Wang, Yihan, et al.
Publicado: (2024)
FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback
por: Wu, Xueqing, et al.
Publicado: (2025)
por: Wu, Xueqing, et al.
Publicado: (2025)
Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
por: Hong, Yunqi, et al.
Publicado: (2025)
por: Hong, Yunqi, et al.
Publicado: (2025)
GenEARL: A Training-Free Generative Framework for Multimodal Event Argument Role Labeling
por: Bansal, Hritik, et al.
Publicado: (2024)
por: Bansal, Hritik, et al.
Publicado: (2024)
A Dialectic Pipeline for Improving LLM Robustness
por: Candussio, Sara
Publicado: (2026)
por: Candussio, Sara
Publicado: (2026)
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
por: Altakrori, Malik H., et al.
Publicado: (2025)
por: Altakrori, Malik H., et al.
Publicado: (2025)
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning
por: Wu, Xueqing, et al.
Publicado: (2024)
por: Wu, Xueqing, et al.
Publicado: (2024)
Verbalized Representation Learning for Interpretable Few-Shot Generalization
por: Yang, Cheng-Fu, et al.
Publicado: (2024)
por: Yang, Cheng-Fu, et al.
Publicado: (2024)
DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to Models
por: Bafna, Niyati, et al.
Publicado: (2025)
por: Bafna, Niyati, et al.
Publicado: (2025)
Dolphin-CN-Dialect: Where Chinese Dialects Matter
por: Meng, Yangyang, et al.
Publicado: (2026)
por: Meng, Yangyang, et al.
Publicado: (2026)
Exploiting Dialect Identification in Automatic Dialectal Text Normalization
por: Alhafni, Bashar, et al.
Publicado: (2024)
por: Alhafni, Bashar, et al.
Publicado: (2024)
Phonotactic Complexity across Dialects
por: Shim, Ryan Soh-Eun, et al.
Publicado: (2024)
por: Shim, Ryan Soh-Eun, et al.
Publicado: (2024)
Reward-Augmented Decoding: Efficient Controlled Text Generation With a Unidirectional Reward Model
por: Deng, Haikang, et al.
Publicado: (2023)
por: Deng, Haikang, et al.
Publicado: (2023)
ProductWebGen: Benchmarking Multimodal Product Webpage Generation
por: Liu, Zhihong, et al.
Publicado: (2026)
por: Liu, Zhihong, et al.
Publicado: (2026)
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation
por: Blaschke, Verena, et al.
Publicado: (2025)
por: Blaschke, Verena, et al.
Publicado: (2025)
Algerian Dialect
por: Benmounah, Zakaria, et al.
Publicado: (2025)
por: Benmounah, Zakaria, et al.
Publicado: (2025)
Improving Dialectal Slot and Intent Detection with Auxiliary Tasks: A Multi-Dialectal Bavarian Case Study
por: Krückl, Xaver Maria, et al.
Publicado: (2025)
por: Krückl, Xaver Maria, et al.
Publicado: (2025)
The Hrunting of AI: Where and How to Improve English Dialectal Fairness
por: Li, Wei, et al.
Publicado: (2026)
por: Li, Wei, et al.
Publicado: (2026)
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles
por: Deng, Yihe, et al.
Publicado: (2025)
por: Deng, Yihe, et al.
Publicado: (2025)
Guiding Through Complexity: What Makes Good Supervision for Hard Math Reasoning Tasks?
por: He, Xuan, et al.
Publicado: (2024)
por: He, Xuan, et al.
Publicado: (2024)
SafeWorld: Geo-Diverse Safety Alignment
por: Yin, Da, et al.
Publicado: (2024)
por: Yin, Da, et al.
Publicado: (2024)
Towards Comprehensive Semantic Speech Embeddings for Chinese Dialects
por: Chang, Kalvin, et al.
Publicado: (2026)
por: Chang, Kalvin, et al.
Publicado: (2026)
Saudi-Dialect-ALLaM: LoRA Fine-Tuning for Dialectal Arabic Generation
por: Barmandah, Hassan
Publicado: (2025)
por: Barmandah, Hassan
Publicado: (2025)
Synchronous Faithfulness Monitoring for Trustworthy Retrieval-Augmented Generation
por: Wu, Di, et al.
Publicado: (2024)
por: Wu, Di, et al.
Publicado: (2024)
Analyzing Dialectical Biases in LLMs for Knowledge and Reasoning Benchmarks
por: Pan, Eileen, et al.
Publicado: (2025)
por: Pan, Eileen, et al.
Publicado: (2025)
OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks
por: Hu, Wenbo, et al.
Publicado: (2026)
por: Hu, Wenbo, et al.
Publicado: (2026)
A Catalog of Basque Dialectal Resources: Online Collections and Standard-to-Dialectal Adaptations
por: Bengoetxea, Jaione, et al.
Publicado: (2026)
por: Bengoetxea, Jaione, et al.
Publicado: (2026)
ArabicDialectHub: A Cross-Dialectal Arabic Learning Resource and Platform
por: Lahlou, Salem
Publicado: (2026)
por: Lahlou, Salem
Publicado: (2026)
A Novel Dialect-Aware Framework for the Classification of Arabic Dialects and Emotions
por: Alsadhan, Nasser A
Publicado: (2025)
por: Alsadhan, Nasser A
Publicado: (2025)
Computational Linguistics Meets Libyan Dialect: A Study on Dialect Identification
por: Essgaer, Mansour, et al.
Publicado: (2025)
por: Essgaer, Mansour, et al.
Publicado: (2025)
Extracting Lexical Features from Dialects via Interpretable Dialect Classifiers
por: Xie, Roy, et al.
Publicado: (2024)
por: Xie, Roy, et al.
Publicado: (2024)
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
por: Wadhawan, Rohan, et al.
Publicado: (2024)
por: Wadhawan, Rohan, et al.
Publicado: (2024)
On Asymmetric Optimization of Reasoning and Perception in Vision-Language Model Post-Training
por: Wu, Xueqing, et al.
Publicado: (2026)
por: Wu, Xueqing, et al.
Publicado: (2026)
DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English
por: Oh, Jio, et al.
Publicado: (2026)
por: Oh, Jio, et al.
Publicado: (2026)
DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects
por: Lucas, Jason, et al.
Publicado: (2026)
por: Lucas, Jason, et al.
Publicado: (2026)
CODET: A Benchmark for Contrastive Dialectal Evaluation of Machine Translation
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2023)
por: Alam, Md Mahfuz Ibn, et al.
Publicado: (2023)
MaiBaam: A Multi-Dialectal Bavarian Universal Dependency Treebank
por: Blaschke, Verena, et al.
Publicado: (2024)
por: Blaschke, Verena, et al.
Publicado: (2024)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
por: Abdullah, Badr M., et al.
Publicado: (2025)
por: Abdullah, Badr M., et al.
Publicado: (2025)
Evaluating Dialect Robustness of Language Models via Conversation Understanding
por: Srirag, Dipankar, et al.
Publicado: (2024)
por: Srirag, Dipankar, et al.
Publicado: (2024)
Ejemplares similares
-
Decoupling Task-Solving and Output Formatting in LLM Generation
por: Deng, Haikang, et al.
Publicado: (2025) -
On the Loss of Context-awareness in General Instruction Fine-tuning
por: Wang, Yihan, et al.
Publicado: (2024) -
FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback
por: Wu, Xueqing, et al.
Publicado: (2025) -
Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
por: Hong, Yunqi, et al.
Publicado: (2025) -
GenEARL: A Training-Free Generative Framework for Multimodal Event Argument Role Labeling
por: Bansal, Hritik, et al.
Publicado: (2024)