Guardado en:
| Autores principales: | Vijayan, Vipin, Bowen, Braeden, Grigsby, Scott, Anderson, Timothy, Gwinnup, Jeremy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2403.03045 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Case for Evaluating Multimodal Translation Models on Text Datasets
por: Vijayan, Vipin, et al.
Publicado: (2024)
por: Vijayan, Vipin, et al.
Publicado: (2024)
Detecting Concrete Visual Tokens for Multimodal Machine Translation
por: Bowen, Braeden, et al.
Publicado: (2024)
por: Bowen, Braeden, et al.
Publicado: (2024)
Improving Language Transfer Capability of Decoder-only Architecture in Multilingual Neural Machine Translation
por: Qu, Zhi, et al.
Publicado: (2024)
por: Qu, Zhi, et al.
Publicado: (2024)
Exploring the Capabilities of Large Multimodal Models on Dense Text
por: Zhang, Shuo, et al.
Publicado: (2024)
por: Zhang, Shuo, et al.
Publicado: (2024)
From TOWER to SPIRE: Adding the Speech Modality to a Translation-Specialist LLM
por: Ambilduke, Kshitij, et al.
Publicado: (2025)
por: Ambilduke, Kshitij, et al.
Publicado: (2025)
Enhancing Code-switched Text-to-Speech Synthesis Capability in Large Language Models with only Monolingual Corpora
por: Xu, Jing, et al.
Publicado: (2024)
por: Xu, Jing, et al.
Publicado: (2024)
Decoder-only Streaming Transformer for Simultaneous Translation
por: Guo, Shoutao, et al.
Publicado: (2024)
por: Guo, Shoutao, et al.
Publicado: (2024)
Wings: Learning Multimodal LLMs without Text-only Forgetting
por: Zhang, Yi-Kai, et al.
Publicado: (2024)
por: Zhang, Yi-Kai, et al.
Publicado: (2024)
Adding Chocolate to Mint: Mitigating Metric Interference in Machine Translation
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
Adding Alignment Control to Language Models
por: Zhu, Wenhong, et al.
Publicado: (2025)
por: Zhu, Wenhong, et al.
Publicado: (2025)
Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation
por: Luo, Yingfeng, et al.
Publicado: (2025)
por: Luo, Yingfeng, et al.
Publicado: (2025)
Scalable Multilingual Multimodal Machine Translation with Speech-Text Fusion
por: Du, Yexing, et al.
Publicado: (2026)
por: Du, Yexing, et al.
Publicado: (2026)
Text-only Synthesis for Image Captioning
por: Zhou, Qing, et al.
Publicado: (2024)
por: Zhou, Qing, et al.
Publicado: (2024)
Investigating Decoder-only Large Language Models for Speech-to-text Translation
por: Huang, Chao-Wei, et al.
Publicado: (2024)
por: Huang, Chao-Wei, et al.
Publicado: (2024)
Decoding the Multimodal Mind: Generalizable Brain-to-Text Translation via Multimodal Alignment and Adaptive Routing
por: Ye, Chunyu, et al.
Publicado: (2025)
por: Ye, Chunyu, et al.
Publicado: (2025)
Looking Right is Sometimes Right: Investigating the Capabilities of Decoder-only LLMs for Sequence Labeling
por: Dukić, David, et al.
Publicado: (2024)
por: Dukić, David, et al.
Publicado: (2024)
Continually Adding New Languages to Multilingual Language Models
por: Owodunni, Abraham Toluwase, et al.
Publicado: (2025)
por: Owodunni, Abraham Toluwase, et al.
Publicado: (2025)
A Novel Paradigm Boosting Translation Capabilities of Large Language Models
por: Guo, Jiaxin, et al.
Publicado: (2024)
por: Guo, Jiaxin, et al.
Publicado: (2024)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
por: Rajaee, Sara, et al.
Publicado: (2026)
por: Rajaee, Sara, et al.
Publicado: (2026)
Proverbs Run in Pairs: Evaluating Proverb Translation Capability of Large Language Model
por: Wang, Minghan, et al.
Publicado: (2025)
por: Wang, Minghan, et al.
Publicado: (2025)
ControlMed: Adding Reasoning Control to Medical Language Model
por: Lee, Sung-Min, et al.
Publicado: (2025)
por: Lee, Sung-Min, et al.
Publicado: (2025)
Forecasting Frontier Language Model Agent Capabilities
por: Pimpale, Govind, et al.
Publicado: (2025)
por: Pimpale, Govind, et al.
Publicado: (2025)
Advancing Semantic Textual Similarity Modeling: A Regression Framework with Translated ReLU and Smooth K2 Loss
por: Zhang, Bowen, et al.
Publicado: (2024)
por: Zhang, Bowen, et al.
Publicado: (2024)
Generating Difficult-to-Translate Texts
por: Zouhar, Vilém, et al.
Publicado: (2025)
por: Zouhar, Vilém, et al.
Publicado: (2025)
Gemini: A Family of Highly Capable Multimodal Models
por: Gemini Team, et al.
Publicado: (2023)
por: Gemini Team, et al.
Publicado: (2023)
Merge to Learn: Efficiently Adding Skills to Language Models with Model Merging
por: Morrison, Jacob, et al.
Publicado: (2024)
por: Morrison, Jacob, et al.
Publicado: (2024)
Imp: Highly Capable Large Multimodal Models for Mobile Devices
por: Shao, Zhenwei, et al.
Publicado: (2024)
por: Shao, Zhenwei, et al.
Publicado: (2024)
HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models
por: Que, Haoran, et al.
Publicado: (2024)
por: Que, Haoran, et al.
Publicado: (2024)
PROST-LLM: Progressively Enhancing the Speech-to-Speech Translation Capability in LLMs
por: Xu, Jing, et al.
Publicado: (2026)
por: Xu, Jing, et al.
Publicado: (2026)
Emergent Explainability: Adding a causal chain to neural network inference
por: Perrett, Adam
Publicado: (2024)
por: Perrett, Adam
Publicado: (2024)
Translating Step-by-Step: Decomposing the Translation Process for Improved Translation Quality of Long-Form Texts
por: Briakou, Eleftheria, et al.
Publicado: (2024)
por: Briakou, Eleftheria, et al.
Publicado: (2024)
Self-Translate-Train: Enhancing Cross-Lingual Transfer of Large Language Models via Inherent Capability
por: Ri, Ryokan, et al.
Publicado: (2024)
por: Ri, Ryokan, et al.
Publicado: (2024)
Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations
por: Miller, Evan
Publicado: (2024)
por: Miller, Evan
Publicado: (2024)
Merge and Conquer: Instructing Multilingual Models by Adding Target Language Weights
por: Valero, Eneko, et al.
Publicado: (2026)
por: Valero, Eneko, et al.
Publicado: (2026)
CT2C-QA: Multimodal Question Answering over Chinese Text, Table and Chart
por: Zhao, Bowen, et al.
Publicado: (2024)
por: Zhao, Bowen, et al.
Publicado: (2024)
Speculative Thinking: Enhancing Small-Model Reasoning with Large Model Guidance at Inference Time
por: Yang, Wang, et al.
Publicado: (2025)
por: Yang, Wang, et al.
Publicado: (2025)
ACE-$M^3$: Automatic Capability Evaluator for Multimodal Medical Models
por: Zhang, Xiechi, et al.
Publicado: (2024)
por: Zhang, Xiechi, et al.
Publicado: (2024)
Can They Dixit? Yes they Can! Dixit as a Playground for Multimodal Language Model Capabilities
por: Balepur, Nishant, et al.
Publicado: (2025)
por: Balepur, Nishant, et al.
Publicado: (2025)
Chitchat as Interference: Adding User Backstories to Task-Oriented Dialogues
por: Stricker, Armand, et al.
Publicado: (2024)
por: Stricker, Armand, et al.
Publicado: (2024)
M3PO: Multimodal-Model-Guided Preference Optimization for Visual Instruction Following
por: Gao, Ruirui, et al.
Publicado: (2025)
por: Gao, Ruirui, et al.
Publicado: (2025)
Ejemplares similares
-
The Case for Evaluating Multimodal Translation Models on Text Datasets
por: Vijayan, Vipin, et al.
Publicado: (2024) -
Detecting Concrete Visual Tokens for Multimodal Machine Translation
por: Bowen, Braeden, et al.
Publicado: (2024) -
Improving Language Transfer Capability of Decoder-only Architecture in Multilingual Neural Machine Translation
por: Qu, Zhi, et al.
Publicado: (2024) -
Exploring the Capabilities of Large Multimodal Models on Dense Text
por: Zhang, Shuo, et al.
Publicado: (2024) -
From TOWER to SPIRE: Adding the Speech Modality to a Translation-Specialist LLM
por: Ambilduke, Kshitij, et al.
Publicado: (2025)