Salvato in:
| Autori principali: | Bowen, Braeden, Vijayan, Vipin, Grigsby, Scott, Anderson, Timothy, Gwinnup, Jeremy |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2403.03075 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Case for Evaluating Multimodal Translation Models on Text Datasets
di: Vijayan, Vipin, et al.
Pubblicazione: (2024)
di: Vijayan, Vipin, et al.
Pubblicazione: (2024)
Adding Multimodal Capabilities to a Text-only Translation Model
di: Vijayan, Vipin, et al.
Pubblicazione: (2024)
di: Vijayan, Vipin, et al.
Pubblicazione: (2024)
Exploring the Necessity of Visual Modality in Multimodal Machine Translation using Authentic Datasets
di: Long, Zi, et al.
Pubblicazione: (2024)
di: Long, Zi, et al.
Pubblicazione: (2024)
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
di: Pan, Jingheng, et al.
Pubblicazione: (2026)
di: Pan, Jingheng, et al.
Pubblicazione: (2026)
Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation
di: Chen, Andong, et al.
Pubblicazione: (2024)
di: Chen, Andong, et al.
Pubblicazione: (2024)
Towards Zero-Shot Multimodal Machine Translation
di: Futeral, Matthieu, et al.
Pubblicazione: (2024)
di: Futeral, Matthieu, et al.
Pubblicazione: (2024)
Dual-branch Prompting for Multimodal Machine Translation
di: Wang, Jie, et al.
Pubblicazione: (2025)
di: Wang, Jie, et al.
Pubblicazione: (2025)
Registering Source Tokens to Target Language Spaces in Multilingual Neural Machine Translation
di: Qu, Zhi, et al.
Pubblicazione: (2025)
di: Qu, Zhi, et al.
Pubblicazione: (2025)
Extend Adversarial Policy Against Neural Machine Translation via Unknown Token
di: Zou, Wei, et al.
Pubblicazione: (2025)
di: Zou, Wei, et al.
Pubblicazione: (2025)
LLM Reasoning for Machine Translation: Synthetic Data Generation over Thinking Tokens
di: Zebaze, Armel, et al.
Pubblicazione: (2025)
di: Zebaze, Armel, et al.
Pubblicazione: (2025)
MatViX: Multimodal Information Extraction from Visually Rich Articles
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2024)
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2024)
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation
di: Dai, Huangyu, et al.
Pubblicazione: (2024)
di: Dai, Huangyu, et al.
Pubblicazione: (2024)
Translate, then Detect: Leveraging Machine Translation for Cross-Lingual Toxicity Classification
di: Bell, Samuel J., et al.
Pubblicazione: (2025)
di: Bell, Samuel J., et al.
Pubblicazione: (2025)
Scalable Multilingual Multimodal Machine Translation with Speech-Text Fusion
di: Du, Yexing, et al.
Pubblicazione: (2026)
di: Du, Yexing, et al.
Pubblicazione: (2026)
CaMMT: Benchmarking Culturally Aware Multimodal Machine Translation
di: Villa-Cueva, Emilio, et al.
Pubblicazione: (2025)
di: Villa-Cueva, Emilio, et al.
Pubblicazione: (2025)
BatchGEMBA: Token-Efficient Machine Translation Evaluation with Batched Prompting and Prompt Compression
di: Larionov, Daniil, et al.
Pubblicazione: (2025)
di: Larionov, Daniil, et al.
Pubblicazione: (2025)
Testing the Limits of Machine Translation from One Book
di: Shaw, Jonathan, et al.
Pubblicazione: (2025)
di: Shaw, Jonathan, et al.
Pubblicazione: (2025)
Why do Large Language Models Fail in Low-resource Translation? Unraveling the Token Dynamics of Large Language Models for Machine Translation
di: Qian, Shenbin, et al.
Pubblicazione: (2026)
di: Qian, Shenbin, et al.
Pubblicazione: (2026)
Efficient Pre-Training with Token Superposition
di: Peng, Bowen, et al.
Pubblicazione: (2026)
di: Peng, Bowen, et al.
Pubblicazione: (2026)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
di: Chung, Jiwan, et al.
Pubblicazione: (2025)
di: Chung, Jiwan, et al.
Pubblicazione: (2025)
Exploring Machine Learning and Language Models for Multimodal Depression Detection
di: Hong, Javier Si Zhao, et al.
Pubblicazione: (2025)
di: Hong, Javier Si Zhao, et al.
Pubblicazione: (2025)
M3PO: Multimodal-Model-Guided Preference Optimization for Visual Instruction Following
di: Gao, Ruirui, et al.
Pubblicazione: (2025)
di: Gao, Ruirui, et al.
Pubblicazione: (2025)
Context-Informed Machine Translation of Manga using Multimodal Large Language Models
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
EMMeTT: Efficient Multimodal Machine Translation Training
di: Żelasko, Piotr, et al.
Pubblicazione: (2024)
di: Żelasko, Piotr, et al.
Pubblicazione: (2024)
Automatic Machine Translation Detection Using a Surrogate Multilingual Translation Model
di: García-Romero, Cristian, et al.
Pubblicazione: (2025)
di: García-Romero, Cristian, et al.
Pubblicazione: (2025)
GIIFT: Graph-guided Inductive Image-free Multimodal Machine Translation
di: Xiong, Jiafeng, et al.
Pubblicazione: (2025)
di: Xiong, Jiafeng, et al.
Pubblicazione: (2025)
TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation
di: Lin, Haokun, et al.
Pubblicazione: (2025)
di: Lin, Haokun, et al.
Pubblicazione: (2025)
Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection
di: Wu, Chenwang, et al.
Pubblicazione: (2026)
di: Wu, Chenwang, et al.
Pubblicazione: (2026)
TopicVD: A Topic-Based Dataset of Video-Guided Multimodal Machine Translation for Documentaries
di: Lv, Jinze, et al.
Pubblicazione: (2025)
di: Lv, Jinze, et al.
Pubblicazione: (2025)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
di: Li, Yanhong, et al.
Pubblicazione: (2025)
di: Li, Yanhong, et al.
Pubblicazione: (2025)
Can LLMs Detect Intrinsic Hallucinations in Paraphrasing and Machine Translation?
di: Gogoulou, Evangelia, et al.
Pubblicazione: (2025)
di: Gogoulou, Evangelia, et al.
Pubblicazione: (2025)
Enhanced Hallucination Detection in Neural Machine Translation through Simple Detector Aggregation
di: Himmi, Anas, et al.
Pubblicazione: (2024)
di: Himmi, Anas, et al.
Pubblicazione: (2024)
Groma: Localized Visual Tokenization for Grounding Multimodal Large Language Models
di: Ma, Chuofan, et al.
Pubblicazione: (2024)
di: Ma, Chuofan, et al.
Pubblicazione: (2024)
Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation
di: Gigant, Théo, et al.
Pubblicazione: (2026)
di: Gigant, Théo, et al.
Pubblicazione: (2026)
Token Masking Improves Transformer-Based Text Classification
di: Xu, Xianglong, et al.
Pubblicazione: (2025)
di: Xu, Xianglong, et al.
Pubblicazione: (2025)
ShortV: Efficient Multimodal Large Language Models by Freezing Visual Tokens in Ineffective Layers
di: Yuan, Qianhao, et al.
Pubblicazione: (2025)
di: Yuan, Qianhao, et al.
Pubblicazione: (2025)
On the Hallucination in Simultaneous Machine Translation
di: Zhong, Meizhi, et al.
Pubblicazione: (2024)
di: Zhong, Meizhi, et al.
Pubblicazione: (2024)
Paraphrase-Aligned Machine Translation
di: Chang, Ke-Ching, et al.
Pubblicazione: (2024)
di: Chang, Ke-Ching, et al.
Pubblicazione: (2024)
Vision-Grounded Machine Interpreting: Improving the Translation Process through Visual Cues
di: Fantinuoli, Claudio
Pubblicazione: (2025)
di: Fantinuoli, Claudio
Pubblicazione: (2025)
Estimating Machine Translation Difficulty
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025)
di: Proietti, Lorenzo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The Case for Evaluating Multimodal Translation Models on Text Datasets
di: Vijayan, Vipin, et al.
Pubblicazione: (2024) -
Adding Multimodal Capabilities to a Text-only Translation Model
di: Vijayan, Vipin, et al.
Pubblicazione: (2024) -
Exploring the Necessity of Visual Modality in Multimodal Machine Translation using Authentic Datasets
di: Long, Zi, et al.
Pubblicazione: (2024) -
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
di: Pan, Jingheng, et al.
Pubblicazione: (2026) -
Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation
di: Chen, Andong, et al.
Pubblicazione: (2024)