Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Xintong, Pan, Jingheng, Liu, Yixiao, Zhao, Xiaohu, Lyu, Chenyang, Wu, Minghao, Biemann, Chris, Wang, Longyue, Xu, Linlong, Luo, Weihua, Zhang, Kaifu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
por: Pan, Jingheng, et al.
Publicado: (2026)
por: Pan, Jingheng, et al.
Publicado: (2026)
Chinese Toxic Language Mitigation via Sentiment Polarity Consistent Rewrites
por: Wang, Xintong, et al.
Publicado: (2025)
por: Wang, Xintong, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
por: Wang, Xintong, et al.
Publicado: (2024)
por: Wang, Xintong, et al.
Publicado: (2024)
Challenging Multilingual LLMs: A New Taxonomy and Benchmark for Unraveling Hallucination in Translation
por: Wu, Xinwei, et al.
Publicado: (2025)
por: Wu, Xinwei, et al.
Publicado: (2025)
New Trends for Modern Machine Translation with Large Reasoning Models
por: Liu, Sinuo, et al.
Publicado: (2025)
por: Liu, Sinuo, et al.
Publicado: (2025)
The Bitter Lesson Learned from 2,000+ Multilingual Benchmarks
por: Wu, Minghao, et al.
Publicado: (2025)
por: Wu, Minghao, et al.
Publicado: (2025)
CogSteer: Cognition-Inspired Selective Layer Intervention for Efficiently Steering Large Language Models
por: Wang, Xintong, et al.
Publicado: (2024)
por: Wang, Xintong, et al.
Publicado: (2024)
Finding the Translation Switch: Discovering and Exploiting the Task-Initiation Features in LLMs
por: Wu, Xinwei, et al.
Publicado: (2026)
por: Wu, Xinwei, et al.
Publicado: (2026)
Beyond Single-Reward: Multi-Pair, Multi-Perspective Preference Optimization for Machine Translation
por: Wang, Hao, et al.
Publicado: (2025)
por: Wang, Hao, et al.
Publicado: (2025)
(Perhaps) Beyond Human Translation: Harnessing Multi-Agent Collaboration for Translating Ultra-Long Literary Texts
por: Wu, Minghao, et al.
Publicado: (2024)
por: Wu, Minghao, et al.
Publicado: (2024)
TransBench: Benchmarking Machine Translation for Industrial-Scale Applications
por: Li, Haijun, et al.
Publicado: (2025)
por: Li, Haijun, et al.
Publicado: (2025)
Towards Lightweight, Adaptive and Attribute-Aware Multi-Aspect Controllable Text Generation with Large Language Models
por: Zhu, Chenyu, et al.
Publicado: (2025)
por: Zhu, Chenyu, et al.
Publicado: (2025)
Marco-LLM: Bridging Languages via Massive Multilingual Training for Cross-Lingual Enhancement
por: Ming, Lingfeng, et al.
Publicado: (2024)
por: Ming, Lingfeng, et al.
Publicado: (2024)
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling
por: Jiang, Fan, et al.
Publicado: (2026)
por: Jiang, Fan, et al.
Publicado: (2026)
LayAlign: Enhancing Multilingual Reasoning in Large Language Models via Layer-Wise Adaptive Fusion and Alignment Strategy
por: Ruan, Zhiwen, et al.
Publicado: (2025)
por: Ruan, Zhiwen, et al.
Publicado: (2025)
Marco-Bench-MIF: On Multilingual Instruction-Following Capability of Large Language Models
por: Zeng, Bo, et al.
Publicado: (2025)
por: Zeng, Bo, et al.
Publicado: (2025)
Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions
por: Zhao, Yu, et al.
Publicado: (2024)
por: Zhao, Yu, et al.
Publicado: (2024)
Probing Large Language Models from A Human Behavioral Perspective
por: Wang, Xintong, et al.
Publicado: (2023)
por: Wang, Xintong, et al.
Publicado: (2023)
Incentivizing Parametric Knowledge via Reinforcement Learning with Verifiable Rewards for Cross-Cultural Entity Translation
por: Zhou, Jiang, et al.
Publicado: (2026)
por: Zhou, Jiang, et al.
Publicado: (2026)
MVL-SIB: A Massively Multilingual Vision-Language Benchmark for Cross-Modal Topical Matching
por: Schmidt, Fabian David, et al.
Publicado: (2025)
por: Schmidt, Fabian David, et al.
Publicado: (2025)
Marco-o1 v2: Towards Widening The Distillation Bottleneck for Reasoning Models
por: Yin, Huifeng, et al.
Publicado: (2025)
por: Yin, Huifeng, et al.
Publicado: (2025)
Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model
por: Geigle, Gregor, et al.
Publicado: (2025)
por: Geigle, Gregor, et al.
Publicado: (2025)
A Unified Agentic Framework for Evaluating Conditional Image Generation
por: Wang, Jifang, et al.
Publicado: (2025)
por: Wang, Jifang, et al.
Publicado: (2025)
A Paradigm Shift: The Future of Machine Translation Lies with Large Language Models
por: Lyu, Chenyang, et al.
Publicado: (2023)
por: Lyu, Chenyang, et al.
Publicado: (2023)
ComfyUI-R1: Exploring Reasoning Models for Workflow Generation
por: Xu, Zhenran, et al.
Publicado: (2025)
por: Xu, Zhenran, et al.
Publicado: (2025)
Dataset of Quotation Attribution in German News Articles
por: Petersen-Frey, Fynn, et al.
Publicado: (2024)
por: Petersen-Frey, Fynn, et al.
Publicado: (2024)
Beyond Probabilities: Unveiling the Misalignment in Evaluating Large Language Models
por: Lyu, Chenyang, et al.
Publicado: (2024)
por: Lyu, Chenyang, et al.
Publicado: (2024)
Marco-Voice Technical Report
por: Tian, Fengping, et al.
Publicado: (2025)
por: Tian, Fengping, et al.
Publicado: (2025)
The Affective Bridge: Preserving Speech Representations while Enhancing Deepfake Detection vian emotional Constraints
por: Li, Yupei, et al.
Publicado: (2025)
por: Li, Yupei, et al.
Publicado: (2025)
Retrieval-augmented Multi-modal Chain-of-Thoughts Reasoning for Large Language Models
por: Liu, Bingshuai, et al.
Publicado: (2023)
por: Liu, Bingshuai, et al.
Publicado: (2023)
MTabVQA: Evaluating Multi-Tabular Reasoning of Language Models in Visual Space
por: Singh, Anshul, et al.
Publicado: (2025)
por: Singh, Anshul, et al.
Publicado: (2025)
ComfyUI-Copilot: An Intelligent Assistant for Automated Workflow Development
por: Xu, Zhenran, et al.
Publicado: (2025)
por: Xu, Zhenran, et al.
Publicado: (2025)
Just Use XML: Revisiting Joint Translation and Label Projection
por: DK, Thennal, et al.
Publicado: (2026)
por: DK, Thennal, et al.
Publicado: (2026)
DeepWideSearch: Benchmarking Depth and Width in Agentic Information Seeking
por: Lan, Tian, et al.
Publicado: (2025)
por: Lan, Tian, et al.
Publicado: (2025)
Marco-ASR: A Principled and Metric-Driven Framework for Fine-Tuning Large-Scale ASR Models for Domain Adaptation
por: Ni, Xuanfan, et al.
Publicado: (2025)
por: Ni, Xuanfan, et al.
Publicado: (2025)
Perspectives - Interactive Document Clustering in the Discourse Analysis Tool Suite
por: Fischer, Tim, et al.
Publicado: (2026)
por: Fischer, Tim, et al.
Publicado: (2026)
Massively Multilingual Adaptation of Large Language Models Using Bilingual Translation Data
por: Ji, Shaoxiong, et al.
Publicado: (2025)
por: Ji, Shaoxiong, et al.
Publicado: (2025)
GPT4Video: A Unified Multimodal Large Language Model for lnstruction-Followed Understanding and Safety-Aware Generation
por: Wang, Zhanyu, et al.
Publicado: (2023)
por: Wang, Zhanyu, et al.
Publicado: (2023)
Findings of the WMT 2024 Shared Task on Discourse-Level Literary Translation
por: Wang, Longyue, et al.
Publicado: (2024)
por: Wang, Longyue, et al.
Publicado: (2024)
Vectra: A New Metric, Dataset, and Model for Visual Quality Assessment in E-Commerce In-Image Machine Translation
por: Wu, Qingyu, et al.
Publicado: (2026)
por: Wu, Qingyu, et al.
Publicado: (2026)
Ejemplares similares
-
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
por: Pan, Jingheng, et al.
Publicado: (2026) -
Chinese Toxic Language Mitigation via Sentiment Polarity Consistent Rewrites
por: Wang, Xintong, et al.
Publicado: (2025) -
Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding
por: Wang, Xintong, et al.
Publicado: (2024) -
Challenging Multilingual LLMs: A New Taxonomy and Benchmark for Unraveling Hallucination in Translation
por: Wu, Xinwei, et al.
Publicado: (2025) -
New Trends for Modern Machine Translation with Large Reasoning Models
por: Liu, Sinuo, et al.
Publicado: (2025)