Cascaded Self-Evaluation Augmented Training for Lightweight Multimodal LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Lv, Zheqi, Wang, Wenkai, Wang, Jiawei, Zhang, Shengyu, Wu, Fei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Rolling Stone Gathers No Moss: Adaptive Policy Optimization for Stable Self-Evaluation in Large Multimodal Models
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation
di: Tang, Zihao, et al.
Pubblicazione: (2024)
di: Tang, Zihao, et al.
Pubblicazione: (2024)
MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference
di: Li, Kunxi, et al.
Pubblicazione: (2025)
di: Li, Kunxi, et al.
Pubblicazione: (2025)
Collaboration of Large Language Models and Small Recommendation Models for Device-Cloud Recommendation
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
A Lightweight Framework for Trigger-Guided LoRA-Based Self-Adaptation in LLMs
di: Wei, Jiacheng, et al.
Pubblicazione: (2025)
di: Wei, Jiacheng, et al.
Pubblicazione: (2025)
ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference
di: Ouyang, Haojie, et al.
Pubblicazione: (2025)
di: Ouyang, Haojie, et al.
Pubblicazione: (2025)
Node Importance Estimation Leveraging LLMs for Semantic Augmentation in Knowledge Graphs
di: Lin, Xinyu, et al.
Pubblicazione: (2024)
di: Lin, Xinyu, et al.
Pubblicazione: (2024)
Order Doesn't Matter, But Reasoning Does: Training LLMs with Order-Centric Augmentation
di: He, Qianxi, et al.
Pubblicazione: (2025)
di: He, Qianxi, et al.
Pubblicazione: (2025)
SCALM: Detecting Bad Practices in Smart Contracts Through LLMs
di: Li, Zongwei, et al.
Pubblicazione: (2025)
di: Li, Zongwei, et al.
Pubblicazione: (2025)
Reinforcement Learning Enhanced LLMs: A Survey
di: Wang, Shuhe, et al.
Pubblicazione: (2024)
di: Wang, Shuhe, et al.
Pubblicazione: (2024)
Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
di: Yang, Zhuolin, et al.
Pubblicazione: (2026)
di: Yang, Zhuolin, et al.
Pubblicazione: (2026)
InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models
di: Liu, Zeyu, et al.
Pubblicazione: (2025)
di: Liu, Zeyu, et al.
Pubblicazione: (2025)
BalanceRAG: Joint Risk Calibration for Cascaded Retrieval-Augmented Generation
di: Jia, Zijun, et al.
Pubblicazione: (2026)
di: Jia, Zijun, et al.
Pubblicazione: (2026)
Persona-Augmented Benchmarking: Evaluating LLMs Across Diverse Writing Styles
di: Truong, Kimberly Le, et al.
Pubblicazione: (2025)
di: Truong, Kimberly Le, et al.
Pubblicazione: (2025)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
di: Li, Yanhong, et al.
Pubblicazione: (2025)
di: Li, Yanhong, et al.
Pubblicazione: (2025)
LaRA: Benchmarking Retrieval-Augmented Generation and Long-Context LLMs -- No Silver Bullet for LC or RAG Routing
di: Li, Kuan, et al.
Pubblicazione: (2025)
di: Li, Kuan, et al.
Pubblicazione: (2025)
Understanding the Role of LLMs in Multimodal Evaluation Benchmarks
di: Jiang, Botian, et al.
Pubblicazione: (2024)
di: Jiang, Botian, et al.
Pubblicazione: (2024)
World-Model-Augmented Web Agents with Action Correction
di: Shen, Zhouzhou, et al.
Pubblicazione: (2026)
di: Shen, Zhouzhou, et al.
Pubblicazione: (2026)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
Detecting and Mitigating Bias in LLMs through Knowledge Graph-Augmented Training
di: Kumar, Rajeev, et al.
Pubblicazione: (2025)
di: Kumar, Rajeev, et al.
Pubblicazione: (2025)
Lightweight Clinical Decision Support System using QLoRA-Fine-Tuned LLMs and Retrieval-Augmented Generation
di: Ansari, Mohammad Shoaib, et al.
Pubblicazione: (2025)
di: Ansari, Mohammad Shoaib, et al.
Pubblicazione: (2025)
Evaluating Role-Consistency in LLMs for Counselor Training
di: Rudolph, Eric, et al.
Pubblicazione: (2026)
di: Rudolph, Eric, et al.
Pubblicazione: (2026)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
di: Chang, Kai-Po, et al.
Pubblicazione: (2025)
di: Chang, Kai-Po, et al.
Pubblicazione: (2025)
Memory-Augmented Agent Training for Business Document Understanding
di: Liu, Jiale, et al.
Pubblicazione: (2024)
di: Liu, Jiale, et al.
Pubblicazione: (2024)
PAVE: Premise-Aware Validation and Editing for Retrieval-Augmented LLMs
di: Huang, Tianyi, et al.
Pubblicazione: (2026)
di: Huang, Tianyi, et al.
Pubblicazione: (2026)
SelfAug: Mitigating Catastrophic Forgetting in Retrieval-Augmented Generation via Distribution Self-Alignment
di: Huang, Yuqing, et al.
Pubblicazione: (2025)
di: Huang, Yuqing, et al.
Pubblicazione: (2025)
ActionStudio: A Lightweight Framework for Data and Training of Large Action Models
di: Zhang, Jianguo, et al.
Pubblicazione: (2025)
di: Zhang, Jianguo, et al.
Pubblicazione: (2025)
Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs
di: Li, Wu, et al.
Pubblicazione: (2026)
di: Li, Wu, et al.
Pubblicazione: (2026)
Self-Evolved Reward Learning for LLMs
di: Huang, Chenghua, et al.
Pubblicazione: (2024)
di: Huang, Chenghua, et al.
Pubblicazione: (2024)
RPTS: Tree-Structured Reasoning Process Scoring for Faithful Multimodal Evaluation
di: Wang, Haofeng, et al.
Pubblicazione: (2025)
di: Wang, Haofeng, et al.
Pubblicazione: (2025)
A Novel Differential Feature Learning for Effective Hallucination Detection and Classification
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering
di: Cocchi, Federico, et al.
Pubblicazione: (2024)
di: Cocchi, Federico, et al.
Pubblicazione: (2024)
Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs
di: Wang, Ziliang, et al.
Pubblicazione: (2025)
di: Wang, Ziliang, et al.
Pubblicazione: (2025)
Finding and Reactivating Post-Trained LLMs' Hidden Safety Mechanisms
di: Li, Mingjie, et al.
Pubblicazione: (2026)
di: Li, Mingjie, et al.
Pubblicazione: (2026)
AskToAct: Enhancing LLMs Tool Use via Self-Correcting Clarification
di: Zhang, Xuan, et al.
Pubblicazione: (2025)
di: Zhang, Xuan, et al.
Pubblicazione: (2025)
Step-On-Feet Tuning: Scaling Self-Alignment of LLMs via Bootstrapping
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
Augmenting Black-box LLMs with Medical Textbooks for Biomedical Question Answering
di: Wang, Yubo, et al.
Pubblicazione: (2023)
di: Wang, Yubo, et al.
Pubblicazione: (2023)
Documenti analoghi
-
A Rolling Stone Gathers No Moss: Adaptive Policy Optimization for Stable Self-Evaluation in Large Multimodal Models
di: Wang, Wenkai, et al.
Pubblicazione: (2025) -
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
di: Lv, Zheqi, et al.
Pubblicazione: (2025) -
ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation
di: Tang, Zihao, et al.
Pubblicazione: (2024) -
MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference
di: Li, Kunxi, et al.
Pubblicazione: (2025) -
Collaboration of Large Language Models and Small Recommendation Models for Device-Cloud Recommendation
di: Lv, Zheqi, et al.
Pubblicazione: (2025)