Salvato in:
| Autori principali: | Lv, Zheqi, Wang, Wenkai, Wang, Jiawei, Zhang, Shengyu, Wu, Fei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2501.05662 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Rolling Stone Gathers No Moss: Adaptive Policy Optimization for Stable Self-Evaluation in Large Multimodal Models
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation
di: Tang, Zihao, et al.
Pubblicazione: (2024)
di: Tang, Zihao, et al.
Pubblicazione: (2024)
Collaboration of Large Language Models and Small Recommendation Models for Device-Cloud Recommendation
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
di: Lv, Zheqi, et al.
Pubblicazione: (2025)
MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference
di: Li, Kunxi, et al.
Pubblicazione: (2025)
di: Li, Kunxi, et al.
Pubblicazione: (2025)
A Lightweight Framework for Trigger-Guided LoRA-Based Self-Adaptation in LLMs
di: Wei, Jiacheng, et al.
Pubblicazione: (2025)
di: Wei, Jiacheng, et al.
Pubblicazione: (2025)
ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference
di: Ouyang, Haojie, et al.
Pubblicazione: (2025)
di: Ouyang, Haojie, et al.
Pubblicazione: (2025)
Reinforcement Learning Enhanced LLMs: A Survey
di: Wang, Shuhe, et al.
Pubblicazione: (2024)
di: Wang, Shuhe, et al.
Pubblicazione: (2024)
Node Importance Estimation Leveraging LLMs for Semantic Augmentation in Knowledge Graphs
di: Lin, Xinyu, et al.
Pubblicazione: (2024)
di: Lin, Xinyu, et al.
Pubblicazione: (2024)
SCALM: Detecting Bad Practices in Smart Contracts Through LLMs
di: Li, Zongwei, et al.
Pubblicazione: (2025)
di: Li, Zongwei, et al.
Pubblicazione: (2025)
Order Doesn't Matter, But Reasoning Does: Training LLMs with Order-Centric Augmentation
di: He, Qianxi, et al.
Pubblicazione: (2025)
di: He, Qianxi, et al.
Pubblicazione: (2025)
InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
di: Yang, Zhuolin, et al.
Pubblicazione: (2026)
di: Yang, Zhuolin, et al.
Pubblicazione: (2026)
Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models
di: Liu, Zeyu, et al.
Pubblicazione: (2025)
di: Liu, Zeyu, et al.
Pubblicazione: (2025)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
di: Wu, Mengsong, et al.
Pubblicazione: (2025)
BalanceRAG: Joint Risk Calibration for Cascaded Retrieval-Augmented Generation
di: Jia, Zijun, et al.
Pubblicazione: (2026)
di: Jia, Zijun, et al.
Pubblicazione: (2026)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
di: Li, Yanhong, et al.
Pubblicazione: (2025)
di: Li, Yanhong, et al.
Pubblicazione: (2025)
Persona-Augmented Benchmarking: Evaluating LLMs Across Diverse Writing Styles
di: Truong, Kimberly Le, et al.
Pubblicazione: (2025)
di: Truong, Kimberly Le, et al.
Pubblicazione: (2025)
World-Model-Augmented Web Agents with Action Correction
di: Shen, Zhouzhou, et al.
Pubblicazione: (2026)
di: Shen, Zhouzhou, et al.
Pubblicazione: (2026)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
di: Chang, Kai-Po, et al.
Pubblicazione: (2025)
di: Chang, Kai-Po, et al.
Pubblicazione: (2025)
Forward Once for All: Structural Parameterized Adaptation for Efficient Cloud-coordinated On-device Recommendation
di: Fu, Kairui, et al.
Pubblicazione: (2025)
di: Fu, Kairui, et al.
Pubblicazione: (2025)
LaRA: Benchmarking Retrieval-Augmented Generation and Long-Context LLMs -- No Silver Bullet for LC or RAG Routing
di: Li, Kuan, et al.
Pubblicazione: (2025)
di: Li, Kuan, et al.
Pubblicazione: (2025)
Understanding the Role of LLMs in Multimodal Evaluation Benchmarks
di: Jiang, Botian, et al.
Pubblicazione: (2024)
di: Jiang, Botian, et al.
Pubblicazione: (2024)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
Detecting and Mitigating Bias in LLMs through Knowledge Graph-Augmented Training
di: Kumar, Rajeev, et al.
Pubblicazione: (2025)
di: Kumar, Rajeev, et al.
Pubblicazione: (2025)
InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
Lightweight Clinical Decision Support System using QLoRA-Fine-Tuned LLMs and Retrieval-Augmented Generation
di: Ansari, Mohammad Shoaib, et al.
Pubblicazione: (2025)
di: Ansari, Mohammad Shoaib, et al.
Pubblicazione: (2025)
Memory-Augmented Agent Training for Business Document Understanding
di: Liu, Jiale, et al.
Pubblicazione: (2024)
di: Liu, Jiale, et al.
Pubblicazione: (2024)
Evaluating Role-Consistency in LLMs for Counselor Training
di: Rudolph, Eric, et al.
Pubblicazione: (2026)
di: Rudolph, Eric, et al.
Pubblicazione: (2026)
FlagEvalMM: A Flexible Framework for Comprehensive Multimodal Model Evaluation
di: He, Zheqi, et al.
Pubblicazione: (2025)
di: He, Zheqi, et al.
Pubblicazione: (2025)
PAVE: Premise-Aware Validation and Editing for Retrieval-Augmented LLMs
di: Huang, Tianyi, et al.
Pubblicazione: (2026)
di: Huang, Tianyi, et al.
Pubblicazione: (2026)
SelfAug: Mitigating Catastrophic Forgetting in Retrieval-Augmented Generation via Distribution Self-Alignment
di: Huang, Yuqing, et al.
Pubblicazione: (2025)
di: Huang, Yuqing, et al.
Pubblicazione: (2025)
Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering
di: Cocchi, Federico, et al.
Pubblicazione: (2024)
di: Cocchi, Federico, et al.
Pubblicazione: (2024)
A Novel Differential Feature Learning for Effective Hallucination Detection and Classification
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
di: Wang, Wenkai, et al.
Pubblicazione: (2025)
ActionStudio: A Lightweight Framework for Data and Training of Large Action Models
di: Zhang, Jianguo, et al.
Pubblicazione: (2025)
di: Zhang, Jianguo, et al.
Pubblicazione: (2025)
CtrlCoT: Dual-Granularity Chain-of-Thought Compression for Controllable Reasoning
di: Fan, Zhenxuan, et al.
Pubblicazione: (2026)
di: Fan, Zhenxuan, et al.
Pubblicazione: (2026)
RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction
di: Liu, Sihao, et al.
Pubblicazione: (2026)
di: Liu, Sihao, et al.
Pubblicazione: (2026)
Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs
di: Wang, Ziliang, et al.
Pubblicazione: (2025)
di: Wang, Ziliang, et al.
Pubblicazione: (2025)
Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs
di: Li, Wu, et al.
Pubblicazione: (2026)
di: Li, Wu, et al.
Pubblicazione: (2026)
Intelligent Model Update Strategy for Sequential Recommendation
di: Lv, Zheqi, et al.
Pubblicazione: (2023)
di: Lv, Zheqi, et al.
Pubblicazione: (2023)
Documenti analoghi
-
A Rolling Stone Gathers No Moss: Adaptive Policy Optimization for Stable Self-Evaluation in Large Multimodal Models
di: Wang, Wenkai, et al.
Pubblicazione: (2025) -
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
di: Lv, Zheqi, et al.
Pubblicazione: (2025) -
ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation
di: Tang, Zihao, et al.
Pubblicazione: (2024) -
Collaboration of Large Language Models and Small Recommendation Models for Device-Cloud Recommendation
di: Lv, Zheqi, et al.
Pubblicazione: (2025) -
MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference
di: Li, Kunxi, et al.
Pubblicazione: (2025)