Detecting Concrete Visual Tokens for Multimodal Machine Translation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bowen, Braeden, Vijayan, Vipin, Grigsby, Scott, Anderson, Timothy, Gwinnup, Jeremy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Case for Evaluating Multimodal Translation Models on Text Datasets
von: Vijayan, Vipin, et al.
Veröffentlicht: (2024)
von: Vijayan, Vipin, et al.
Veröffentlicht: (2024)
Adding Multimodal Capabilities to a Text-only Translation Model
von: Vijayan, Vipin, et al.
Veröffentlicht: (2024)
von: Vijayan, Vipin, et al.
Veröffentlicht: (2024)
Exploring the Necessity of Visual Modality in Multimodal Machine Translation using Authentic Datasets
von: Long, Zi, et al.
Veröffentlicht: (2024)
von: Long, Zi, et al.
Veröffentlicht: (2024)
Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation
von: Chen, Andong, et al.
Veröffentlicht: (2024)
von: Chen, Andong, et al.
Veröffentlicht: (2024)
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
von: Pan, Jingheng, et al.
Veröffentlicht: (2026)
von: Pan, Jingheng, et al.
Veröffentlicht: (2026)
Towards Zero-Shot Multimodal Machine Translation
von: Futeral, Matthieu, et al.
Veröffentlicht: (2024)
von: Futeral, Matthieu, et al.
Veröffentlicht: (2024)
Registering Source Tokens to Target Language Spaces in Multilingual Neural Machine Translation
von: Qu, Zhi, et al.
Veröffentlicht: (2025)
von: Qu, Zhi, et al.
Veröffentlicht: (2025)
Extend Adversarial Policy Against Neural Machine Translation via Unknown Token
von: Zou, Wei, et al.
Veröffentlicht: (2025)
von: Zou, Wei, et al.
Veröffentlicht: (2025)
LLM Reasoning for Machine Translation: Synthetic Data Generation over Thinking Tokens
von: Zebaze, Armel, et al.
Veröffentlicht: (2025)
von: Zebaze, Armel, et al.
Veröffentlicht: (2025)
Dual-branch Prompting for Multimodal Machine Translation
von: Wang, Jie, et al.
Veröffentlicht: (2025)
von: Wang, Jie, et al.
Veröffentlicht: (2025)
Translate, then Detect: Leveraging Machine Translation for Cross-Lingual Toxicity Classification
von: Bell, Samuel J., et al.
Veröffentlicht: (2025)
von: Bell, Samuel J., et al.
Veröffentlicht: (2025)
MatViX: Multimodal Information Extraction from Visually Rich Articles
von: Khalighinejad, Ghazal, et al.
Veröffentlicht: (2024)
von: Khalighinejad, Ghazal, et al.
Veröffentlicht: (2024)
Scalable Multilingual Multimodal Machine Translation with Speech-Text Fusion
von: Du, Yexing, et al.
Veröffentlicht: (2026)
von: Du, Yexing, et al.
Veröffentlicht: (2026)
CaMMT: Benchmarking Culturally Aware Multimodal Machine Translation
von: Villa-Cueva, Emilio, et al.
Veröffentlicht: (2025)
von: Villa-Cueva, Emilio, et al.
Veröffentlicht: (2025)
BatchGEMBA: Token-Efficient Machine Translation Evaluation with Batched Prompting and Prompt Compression
von: Larionov, Daniil, et al.
Veröffentlicht: (2025)
von: Larionov, Daniil, et al.
Veröffentlicht: (2025)
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation
von: Dai, Huangyu, et al.
Veröffentlicht: (2024)
von: Dai, Huangyu, et al.
Veröffentlicht: (2024)
Testing the Limits of Machine Translation from One Book
von: Shaw, Jonathan, et al.
Veröffentlicht: (2025)
von: Shaw, Jonathan, et al.
Veröffentlicht: (2025)
Why do Large Language Models Fail in Low-resource Translation? Unraveling the Token Dynamics of Large Language Models for Machine Translation
von: Qian, Shenbin, et al.
Veröffentlicht: (2026)
von: Qian, Shenbin, et al.
Veröffentlicht: (2026)
Context-Informed Machine Translation of Manga using Multimodal Large Language Models
von: Lippmann, Philip, et al.
Veröffentlicht: (2024)
von: Lippmann, Philip, et al.
Veröffentlicht: (2024)
Efficient Pre-Training with Token Superposition
von: Peng, Bowen, et al.
Veröffentlicht: (2026)
von: Peng, Bowen, et al.
Veröffentlicht: (2026)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
M3PO: Multimodal-Model-Guided Preference Optimization for Visual Instruction Following
von: Gao, Ruirui, et al.
Veröffentlicht: (2025)
von: Gao, Ruirui, et al.
Veröffentlicht: (2025)
Exploring Machine Learning and Language Models for Multimodal Depression Detection
von: Hong, Javier Si Zhao, et al.
Veröffentlicht: (2025)
von: Hong, Javier Si Zhao, et al.
Veröffentlicht: (2025)
Automatic Machine Translation Detection Using a Surrogate Multilingual Translation Model
von: García-Romero, Cristian, et al.
Veröffentlicht: (2025)
von: García-Romero, Cristian, et al.
Veröffentlicht: (2025)
Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection
von: Wu, Chenwang, et al.
Veröffentlicht: (2026)
von: Wu, Chenwang, et al.
Veröffentlicht: (2026)
EMMeTT: Efficient Multimodal Machine Translation Training
von: Żelasko, Piotr, et al.
Veröffentlicht: (2024)
von: Żelasko, Piotr, et al.
Veröffentlicht: (2024)
GIIFT: Graph-guided Inductive Image-free Multimodal Machine Translation
von: Xiong, Jiafeng, et al.
Veröffentlicht: (2025)
von: Xiong, Jiafeng, et al.
Veröffentlicht: (2025)
TopicVD: A Topic-Based Dataset of Video-Guided Multimodal Machine Translation for Documentaries
von: Lv, Jinze, et al.
Veröffentlicht: (2025)
von: Lv, Jinze, et al.
Veröffentlicht: (2025)
TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
Enhanced Hallucination Detection in Neural Machine Translation through Simple Detector Aggregation
von: Himmi, Anas, et al.
Veröffentlicht: (2024)
von: Himmi, Anas, et al.
Veröffentlicht: (2024)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
von: Li, Yanhong, et al.
Veröffentlicht: (2025)
von: Li, Yanhong, et al.
Veröffentlicht: (2025)
On the Hallucination in Simultaneous Machine Translation
von: Zhong, Meizhi, et al.
Veröffentlicht: (2024)
von: Zhong, Meizhi, et al.
Veröffentlicht: (2024)
Paraphrase-Aligned Machine Translation
von: Chang, Ke-Ching, et al.
Veröffentlicht: (2024)
von: Chang, Ke-Ching, et al.
Veröffentlicht: (2024)
Estimating Machine Translation Difficulty
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)
von: Proietti, Lorenzo, et al.
Veröffentlicht: (2025)
Open Machine Translation for Esperanto
von: de Gibert, Ona, et al.
Veröffentlicht: (2026)
von: de Gibert, Ona, et al.
Veröffentlicht: (2026)
Can LLMs Detect Intrinsic Hallucinations in Paraphrasing and Machine Translation?
von: Gogoulou, Evangelia, et al.
Veröffentlicht: (2025)
von: Gogoulou, Evangelia, et al.
Veröffentlicht: (2025)
ShortV: Efficient Multimodal Large Language Models by Freezing Visual Tokens in Ineffective Layers
von: Yuan, Qianhao, et al.
Veröffentlicht: (2025)
von: Yuan, Qianhao, et al.
Veröffentlicht: (2025)
Machine Translation Models are Zero-Shot Detectors of Translation Direction
von: Wastl, Michelle, et al.
Veröffentlicht: (2024)
von: Wastl, Michelle, et al.
Veröffentlicht: (2024)
The Impact of Syntactic and Semantic Proximity on Machine Translation with Back-Translation
von: Guerin, Nicolas, et al.
Veröffentlicht: (2024)
von: Guerin, Nicolas, et al.
Veröffentlicht: (2024)
Grounded Concreteness: Human-Like Concreteness Sensitivity in Vision-Language Models
von: Roy, Aryan, et al.
Veröffentlicht: (2026)
von: Roy, Aryan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Case for Evaluating Multimodal Translation Models on Text Datasets
von: Vijayan, Vipin, et al.
Veröffentlicht: (2024) -
Adding Multimodal Capabilities to a Text-only Translation Model
von: Vijayan, Vipin, et al.
Veröffentlicht: (2024) -
Exploring the Necessity of Visual Modality in Multimodal Machine Translation using Authentic Datasets
von: Long, Zi, et al.
Veröffentlicht: (2024) -
Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation
von: Chen, Andong, et al.
Veröffentlicht: (2024) -
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
von: Pan, Jingheng, et al.
Veröffentlicht: (2026)