Can Large Multimodal Models Uncover Deep Semantics Behind Images?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yixin, Li, Zheng, Dong, Qingxiu, Xia, Heming, Sui, Zhifang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Single Frames: Can LMMs Comprehend Temporal and Contextual Narratives in Image Sequences?
von: Wang, Xiaochen, et al.
Veröffentlicht: (2025)
von: Wang, Xiaochen, et al.
Veröffentlicht: (2025)
RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection
von: Yang, Yixin, et al.
Veröffentlicht: (2025)
von: Yang, Yixin, et al.
Veröffentlicht: (2025)
Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding
von: Xia, Heming, et al.
Veröffentlicht: (2024)
von: Xia, Heming, et al.
Veröffentlicht: (2024)
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Self-Boosting Large Language Models with Synthetic Preference Data
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)
Taking a Deep Breath: Enhancing Language Modeling of Large Language Models with Sentinel Tokens
von: Luo, Weiyao, et al.
Veröffentlicht: (2024)
von: Luo, Weiyao, et al.
Veröffentlicht: (2024)
Towards Harmonized Uncertainty Estimation for Large Language Models
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
Decoding in Geometry: Alleviating Embedding-Space Crowding for Complex Reasoning
von: Yang, Yixin, et al.
Veröffentlicht: (2026)
von: Yang, Yixin, et al.
Veröffentlicht: (2026)
Reinforcement Pre-Training
von: Dong, Qingxiu, et al.
Veröffentlicht: (2025)
von: Dong, Qingxiu, et al.
Veröffentlicht: (2025)
Can Large Language Models Always Solve Easy Problems if They Can Solve Harder Ones?
von: Yang, Zhe, et al.
Veröffentlicht: (2024)
von: Yang, Zhe, et al.
Veröffentlicht: (2024)
LLM-REVal: Can We Trust LLM Reviewers Yet?
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
A Survey on In-context Learning
von: Dong, Qingxiu, et al.
Veröffentlicht: (2022)
von: Dong, Qingxiu, et al.
Veröffentlicht: (2022)
HauntAttack: When Attack Follows Reasoning as a Shadow
von: Ma, Jingyuan, et al.
Veröffentlicht: (2025)
von: Ma, Jingyuan, et al.
Veröffentlicht: (2025)
Can MLLMs Understand the Deep Implication Behind Chinese Images?
von: Zhang, Chenhao, et al.
Veröffentlicht: (2024)
von: Zhang, Chenhao, et al.
Veröffentlicht: (2024)
PeriodicLoRA: Breaking the Low-Rank Bottleneck in LoRA Optimization
von: Meng, Xiangdi, et al.
Veröffentlicht: (2024)
von: Meng, Xiangdi, et al.
Veröffentlicht: (2024)
CoLT: Reasoning with Chain of Latent Tool Calls
von: Zhu, Fangwei, et al.
Veröffentlicht: (2026)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2026)
Language Models Encode the Value of Numbers Linearly
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
DySem: Uncovering Dynamic Semantic Components of Large Language Models for Calculating Semantic Textual Similarity
von: Zheng, Kaijie, et al.
Veröffentlicht: (2026)
von: Zheng, Kaijie, et al.
Veröffentlicht: (2026)
Towards Better RL Training Data Utilization via Second-Order Rollout
von: Yang, Zhe, et al.
Veröffentlicht: (2026)
von: Yang, Zhe, et al.
Veröffentlicht: (2026)
Exploring Activation Patterns of Parameters in Language Models
von: Wang, Yudong, et al.
Veröffentlicht: (2024)
von: Wang, Yudong, et al.
Veröffentlicht: (2024)
Enhancing Tool Retrieval with Iterative Feedback from Large Language Models
von: Xu, Qiancheng, et al.
Veröffentlicht: (2024)
von: Xu, Qiancheng, et al.
Veröffentlicht: (2024)
Plug-and-Play Training Framework for Preference Optimization
von: Ma, Jingyuan, et al.
Veröffentlicht: (2024)
von: Ma, Jingyuan, et al.
Veröffentlicht: (2024)
Multilingual Brain Surgeon: Large Language Models Can be Compressed Leaving No Language Behind
von: Zeng, Hongchuan, et al.
Veröffentlicht: (2024)
von: Zeng, Hongchuan, et al.
Veröffentlicht: (2024)
AdaMMS: Model Merging for Heterogeneous Multimodal Large Language Models with Unsupervised Coefficient Optimization
von: Du, Yiyang, et al.
Veröffentlicht: (2025)
von: Du, Yiyang, et al.
Veröffentlicht: (2025)
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
Not All Demonstration Examples are Equally Beneficial: Reweighting Demonstration Examples for In-Context Learning
von: Yang, Zhe, et al.
Veröffentlicht: (2023)
von: Yang, Zhe, et al.
Veröffentlicht: (2023)
Towards Optimal Learning of Language Models
von: Gu, Yuxian, et al.
Veröffentlicht: (2024)
von: Gu, Yuxian, et al.
Veröffentlicht: (2024)
Online Experiential Learning for Language Models
von: Ye, Tianzhu, et al.
Veröffentlicht: (2026)
von: Ye, Tianzhu, et al.
Veröffentlicht: (2026)
PEToolLLM: Towards Personalized Tool Learning in Large Language Models
von: Xu, Qiancheng, et al.
Veröffentlicht: (2025)
von: Xu, Qiancheng, et al.
Veröffentlicht: (2025)
Large AI Model Empowered Multimodal Semantic Communications
von: Jiang, Feibo, et al.
Veröffentlicht: (2023)
von: Jiang, Feibo, et al.
Veröffentlicht: (2023)
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis
von: Zhu, Wenhao, et al.
Veröffentlicht: (2023)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2023)
Social Good or Scientific Curiosity? Uncovering the Research Framing Behind NLP Artefacts
von: Chamoun, Eric, et al.
Veröffentlicht: (2025)
von: Chamoun, Eric, et al.
Veröffentlicht: (2025)
Reward Reasoning Model
von: Guo, Jiaxin, et al.
Veröffentlicht: (2025)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2025)
Chain-of-Thought Tokens are Computer Program Variables
von: Zhu, Fangwei, et al.
Veröffentlicht: (2025)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2025)
Think Only When You Need with Large Hybrid-Reasoning Models
von: Jiang, Lingjie, et al.
Veröffentlicht: (2025)
von: Jiang, Lingjie, et al.
Veröffentlicht: (2025)
Text as Images: Can Multimodal Large Language Models Follow Printed Instructions in Pixels?
von: Li, Xiujun, et al.
Veröffentlicht: (2023)
von: Li, Xiujun, et al.
Veröffentlicht: (2023)
Can Large Language Models Automatically Jailbreak GPT-4V?
von: Wu, Yuanwei, et al.
Veröffentlicht: (2024)
von: Wu, Yuanwei, et al.
Veröffentlicht: (2024)
Data Selection via Optimal Control for Language Models
von: Gu, Yuxian, et al.
Veröffentlicht: (2024)
von: Gu, Yuxian, et al.
Veröffentlicht: (2024)
Towards a Unified View of Preference Learning for Large Language Models: A Survey
von: Gao, Bofei, et al.
Veröffentlicht: (2024)
von: Gao, Bofei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Beyond Single Frames: Can LMMs Comprehend Temporal and Contextual Narratives in Image Sequences?
von: Wang, Xiaochen, et al.
Veröffentlicht: (2025) -
RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection
von: Yang, Yixin, et al.
Veröffentlicht: (2025) -
Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding
von: Xia, Heming, et al.
Veröffentlicht: (2024) -
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation
von: Li, Rui, et al.
Veröffentlicht: (2025) -
Self-Boosting Large Language Models with Synthetic Preference Data
von: Dong, Qingxiu, et al.
Veröffentlicht: (2024)