Multimodal Prompt Optimization: Why Not Leverage Multiple Modalities for MLLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Choi, Yumin, Kim, Dongki, Baek, Jinheon, Hwang, Sung Ju |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
System Prompt Optimization with Meta-Learning
di: Choi, Yumin, et al.
Pubblicazione: (2025)
di: Choi, Yumin, et al.
Pubblicazione: (2025)
Efficient Real-time Refinement of Language Model Text Generation
di: Ko, Joonho, et al.
Pubblicazione: (2025)
di: Ko, Joonho, et al.
Pubblicazione: (2025)
Sketch-of-Thought: Efficient LLM Reasoning with Adaptive Cognitive-Inspired Sketching
di: Aytes, Simon A., et al.
Pubblicazione: (2025)
di: Aytes, Simon A., et al.
Pubblicazione: (2025)
Retrieval-Augmented Data Augmentation for Low-Resource Domain Tasks
di: Seo, Minju, et al.
Pubblicazione: (2024)
di: Seo, Minju, et al.
Pubblicazione: (2024)
ResearchAgent: Iterative Research Idea Generation over Scientific Literature with Large Language Models
di: Baek, Jinheon, et al.
Pubblicazione: (2024)
di: Baek, Jinheon, et al.
Pubblicazione: (2024)
UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities
di: Yeo, Woongyeong, et al.
Pubblicazione: (2025)
di: Yeo, Woongyeong, et al.
Pubblicazione: (2025)
Chain of Retrieval: Multi-Aspect Iterative Search Expansion and Post-Order Search Aggregation for Full Paper Retrieval
di: Park, Sangwoo, et al.
Pubblicazione: (2025)
di: Park, Sangwoo, et al.
Pubblicazione: (2025)
VideoRAG: Retrieval-Augmented Generation over Video Corpus
di: Jeong, Soyeong, et al.
Pubblicazione: (2025)
di: Jeong, Soyeong, et al.
Pubblicazione: (2025)
PREPING: Building Agent Memory without Tasks
di: Choi, Yumin, et al.
Pubblicazione: (2026)
di: Choi, Yumin, et al.
Pubblicazione: (2026)
HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents
di: Yeo, Woongyeng, et al.
Pubblicazione: (2026)
di: Yeo, Woongyeng, et al.
Pubblicazione: (2026)
Rethinking Code Refinement: Learning to Judge Code Efficiency
di: Seo, Minju, et al.
Pubblicazione: (2024)
di: Seo, Minju, et al.
Pubblicazione: (2024)
OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources
di: Baek, Jinheon, et al.
Pubblicazione: (2026)
di: Baek, Jinheon, et al.
Pubblicazione: (2026)
Unified Multimodal Interleaved Document Representation for Retrieval
di: Lee, Jaewoo, et al.
Pubblicazione: (2024)
di: Lee, Jaewoo, et al.
Pubblicazione: (2024)
GFlowPO: Generative Flow Network as a Language Model Prompt Optimizer
di: Cho, Junmo, et al.
Pubblicazione: (2026)
di: Cho, Junmo, et al.
Pubblicazione: (2026)
ES-Merging: Biological MLLM Merging via Embedding Space Signals
di: Lee, Wonbin, et al.
Pubblicazione: (2026)
di: Lee, Wonbin, et al.
Pubblicazione: (2026)
Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents
di: Kim, Suji, et al.
Pubblicazione: (2026)
di: Kim, Suji, et al.
Pubblicazione: (2026)
Adaptive-RAG: Learning to Adapt Retrieval-Augmented Large Language Models through Question Complexity
di: Jeong, Soyeong, et al.
Pubblicazione: (2024)
di: Jeong, Soyeong, et al.
Pubblicazione: (2024)
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
di: Lee, Chanuk, et al.
Pubblicazione: (2026)
di: Lee, Chanuk, et al.
Pubblicazione: (2026)
It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs
di: Park, Sangwoo, et al.
Pubblicazione: (2026)
di: Park, Sangwoo, et al.
Pubblicazione: (2026)
One Prompt is not Enough: Automated Construction of a Mixture-of-Expert Prompts
di: Wang, Ruochen, et al.
Pubblicazione: (2024)
di: Wang, Ruochen, et al.
Pubblicazione: (2024)
Revisiting In-Context Learning with Long Context Language Models
di: Baek, Jinheon, et al.
Pubblicazione: (2024)
di: Baek, Jinheon, et al.
Pubblicazione: (2024)
Knowledge Base Construction for Knowledge-Augmented Text-to-SQL
di: Baek, Jinheon, et al.
Pubblicazione: (2025)
di: Baek, Jinheon, et al.
Pubblicazione: (2025)
WorldMM: Dynamic Multimodal Memory Agent for Long Video Reasoning
di: Yeo, Woongyeong, et al.
Pubblicazione: (2025)
di: Yeo, Woongyeong, et al.
Pubblicazione: (2025)
When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs
di: Jeong, Soyeong, et al.
Pubblicazione: (2025)
di: Jeong, Soyeong, et al.
Pubblicazione: (2025)
Database-Augmented Query Representation for Information Retrieval
di: Jeong, Soyeong, et al.
Pubblicazione: (2024)
di: Jeong, Soyeong, et al.
Pubblicazione: (2024)
Knowledge-Augmented Large Language Models for Personalized Contextual Query Suggestion
di: Baek, Jinheon, et al.
Pubblicazione: (2023)
di: Baek, Jinheon, et al.
Pubblicazione: (2023)
Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models?
di: Kang, Deokhyung, et al.
Pubblicazione: (2025)
di: Kang, Deokhyung, et al.
Pubblicazione: (2025)
From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities
di: Jiang, Shixin, et al.
Pubblicazione: (2024)
di: Jiang, Shixin, et al.
Pubblicazione: (2024)
AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML
di: Trirat, Patara, et al.
Pubblicazione: (2024)
di: Trirat, Patara, et al.
Pubblicazione: (2024)
Training-Free Exponential Context Extension via Cascading KV Cache
di: Willette, Jeffrey, et al.
Pubblicazione: (2024)
di: Willette, Jeffrey, et al.
Pubblicazione: (2024)
IPCGRL: Language-Instructed Reinforcement Learning for Procedural Level Generation
di: Baek, In-Chang, et al.
Pubblicazione: (2025)
di: Baek, In-Chang, et al.
Pubblicazione: (2025)
By My Eyes: Grounding Multimodal Large Language Models with Sensor Data via Visual Prompting
di: Yoon, Hyungjun, et al.
Pubblicazione: (2024)
di: Yoon, Hyungjun, et al.
Pubblicazione: (2024)
Prompt-SAW: Leveraging Relation-Aware Graphs for Textual Prompt Compression
di: Ali, Muhammad Asif, et al.
Pubblicazione: (2024)
di: Ali, Muhammad Asif, et al.
Pubblicazione: (2024)
Integrating Pre-trained Language Model into Neural Machine Translation
di: Hwang, Soon-Jae, et al.
Pubblicazione: (2023)
di: Hwang, Soon-Jae, et al.
Pubblicazione: (2023)
Automatic Prompt Optimization with Prompt Distillation
di: Dyagin, Ernest A., et al.
Pubblicazione: (2025)
di: Dyagin, Ernest A., et al.
Pubblicazione: (2025)
Explain in Your Own Words: Improving Reasoning via Token-Selective Dual Knowledge Distillation
di: Kim, Minsang, et al.
Pubblicazione: (2026)
di: Kim, Minsang, et al.
Pubblicazione: (2026)
Leveraging Zero-Shot Prompting for Efficient Language Model Distillation
di: Vöge, Lukas, et al.
Pubblicazione: (2024)
di: Vöge, Lukas, et al.
Pubblicazione: (2024)
Local Prompt Optimization
di: Jain, Yash, et al.
Pubblicazione: (2025)
di: Jain, Yash, et al.
Pubblicazione: (2025)
Hard Prompts Made Interpretable: Sparse Entropy Regularization for Prompt Tuning with RL
di: Choi, Yunseon, et al.
Pubblicazione: (2024)
di: Choi, Yunseon, et al.
Pubblicazione: (2024)
PromptWizard: Task-Aware Prompt Optimization Framework
di: Agarwal, Eshaan, et al.
Pubblicazione: (2024)
di: Agarwal, Eshaan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
System Prompt Optimization with Meta-Learning
di: Choi, Yumin, et al.
Pubblicazione: (2025) -
Efficient Real-time Refinement of Language Model Text Generation
di: Ko, Joonho, et al.
Pubblicazione: (2025) -
Sketch-of-Thought: Efficient LLM Reasoning with Adaptive Cognitive-Inspired Sketching
di: Aytes, Simon A., et al.
Pubblicazione: (2025) -
Retrieval-Augmented Data Augmentation for Low-Resource Domain Tasks
di: Seo, Minju, et al.
Pubblicazione: (2024) -
ResearchAgent: Iterative Research Idea Generation over Scientific Literature with Large Language Models
di: Baek, Jinheon, et al.
Pubblicazione: (2024)