Meta-Reflection: A Feedback-Free Reflection Learning Framework
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Yaoke, Zhu, Yun, Bao, Xintong, Zhang, Wenqiao, Dai, Suyang, Chen, Kehan, Li, Wenqiang, Huang, Gang, Tang, Siliang, Zhuang, Yueting |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bridging Local Details and Global Context in Text-Attributed Graphs
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
DuetRAG: Collaborative Retrieval-Augmented Generation
di: Jiao, Dian, et al.
Pubblicazione: (2024)
di: Jiao, Dian, et al.
Pubblicazione: (2024)
IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization
di: Cao, Jie, et al.
Pubblicazione: (2024)
di: Cao, Jie, et al.
Pubblicazione: (2024)
Efficient Tuning and Inference for Large Language Models on Textual Graphs
di: Zhu, Yun, et al.
Pubblicazione: (2024)
di: Zhu, Yun, et al.
Pubblicazione: (2024)
Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization
di: Zhang, Wenqi, et al.
Pubblicazione: (2024)
di: Zhang, Wenqi, et al.
Pubblicazione: (2024)
MoA: Heterogeneous Mixture of Adapters for Parameter-Efficient Fine-Tuning of Large Language Models
di: Cao, Jie, et al.
Pubblicazione: (2025)
di: Cao, Jie, et al.
Pubblicazione: (2025)
TeamLoRA: Boosting Low-Rank Adaptation with Expert Collaboration and Competition
di: Lin, Tianwei, et al.
Pubblicazione: (2024)
di: Lin, Tianwei, et al.
Pubblicazione: (2024)
Self-Contrast: Better Reflection Through Inconsistent Solving Perspectives
di: Zhang, Wenqi, et al.
Pubblicazione: (2024)
di: Zhang, Wenqi, et al.
Pubblicazione: (2024)
Graft: Integrating the Domain Knowledge via Efficient Parameter Synergy for MLLMs
di: Dai, Yang, et al.
Pubblicazione: (2025)
di: Dai, Yang, et al.
Pubblicazione: (2025)
MetaReflection: Learning Instructions for Language Agents using Past Reflections
di: Gupta, Priyanshu, et al.
Pubblicazione: (2024)
di: Gupta, Priyanshu, et al.
Pubblicazione: (2024)
Improving Large Models with Small models: Lower Costs and Better Performance
di: Chen, Dong, et al.
Pubblicazione: (2024)
di: Chen, Dong, et al.
Pubblicazione: (2024)
Chart-HQA: A Benchmark for Hypothetical Question Answering in Charts
di: Chen, Xiangnan, et al.
Pubblicazione: (2025)
di: Chen, Xiangnan, et al.
Pubblicazione: (2025)
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback
di: Yun, Taewon, et al.
Pubblicazione: (2025)
di: Yun, Taewon, et al.
Pubblicazione: (2025)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
di: Lee, Kyungjae, et al.
Pubblicazione: (2024)
di: Lee, Kyungjae, et al.
Pubblicazione: (2024)
Draft-Thinking: Learning Efficient Reasoning in Long Chain-of-Thought LLMs
di: Cao, Jie, et al.
Pubblicazione: (2026)
di: Cao, Jie, et al.
Pubblicazione: (2026)
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
di: Huang, Yue, et al.
Pubblicazione: (2025)
di: Huang, Yue, et al.
Pubblicazione: (2025)
CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark
di: Wang, Wei, et al.
Pubblicazione: (2026)
di: Wang, Wei, et al.
Pubblicazione: (2026)
LASER: Tuning-Free LLM-Driven Attention Control for Efficient Text-conditioned Image-to-Animation
di: Zheng, Haoyu, et al.
Pubblicazione: (2024)
di: Zheng, Haoyu, et al.
Pubblicazione: (2024)
DUAL-REFLECT: Enhancing Large Language Models for Reflective Translation through Dual Learning Feedback Mechanisms
di: Chen, Andong, et al.
Pubblicazione: (2024)
di: Chen, Andong, et al.
Pubblicazione: (2024)
Towards Meta-Cognitive Knowledge Editing for Multimodal LLMs
di: Fan, Zhaoyu, et al.
Pubblicazione: (2025)
di: Fan, Zhaoyu, et al.
Pubblicazione: (2025)
HyperLLaVA: Dynamic Visual and Language Expert Tuning for Multimodal Large Language Models
di: Zhang, Wenqiao, et al.
Pubblicazione: (2024)
di: Zhang, Wenqiao, et al.
Pubblicazione: (2024)
Debate, Reflect, and Distill: Multi-Agent Feedback with Tree-Structured Preference Optimization for Efficient Language Model Enhancement
di: Zhou, Xiaofeng, et al.
Pubblicazione: (2025)
di: Zhou, Xiaofeng, et al.
Pubblicazione: (2025)
KCM: KAN-Based Collaboration Models Enhance Pretrained Large Models
di: Dai, Guangyu, et al.
Pubblicazione: (2025)
di: Dai, Guangyu, et al.
Pubblicazione: (2025)
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
di: Ren, Houxing, et al.
Pubblicazione: (2024)
di: Ren, Houxing, et al.
Pubblicazione: (2024)
ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework
di: Qin, Kai, et al.
Pubblicazione: (2026)
di: Qin, Kai, et al.
Pubblicazione: (2026)
Fast Thinking for Large Language Models
di: Zheng, Haoyu, et al.
Pubblicazione: (2025)
di: Zheng, Haoyu, et al.
Pubblicazione: (2025)
ReflectSumm: A Benchmark for Course Reflection Summarization
di: Zhong, Yang, et al.
Pubblicazione: (2024)
di: Zhong, Yang, et al.
Pubblicazione: (2024)
Think Twice, Click Once: Enhancing GUI Grounding via Fast and Slow Systems
di: Tang, Fei, et al.
Pubblicazione: (2025)
di: Tang, Fei, et al.
Pubblicazione: (2025)
Stop Unnecessary Reflection: Training LRMs for Efficient Reasoning with Adaptive Reflection and Length Coordinated Penalty
di: Yu, Zewei, et al.
Pubblicazione: (2026)
di: Yu, Zewei, et al.
Pubblicazione: (2026)
MAIGO: Mitigating Lost-in-Conversation with History-Cleaned On-Policy Self-Distillation
di: Zheng, Haoyu, et al.
Pubblicazione: (2026)
di: Zheng, Haoyu, et al.
Pubblicazione: (2026)
Reflect then Learn: Active Prompting for Information Extraction Guided by Introspective Confusion
di: Zhao, Dong, et al.
Pubblicazione: (2025)
di: Zhao, Dong, et al.
Pubblicazione: (2025)
PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models
di: Zheng, Haoyu, et al.
Pubblicazione: (2026)
di: Zheng, Haoyu, et al.
Pubblicazione: (2026)
ReflectivePrompt: Reflective evolution in autoprompting algorithms
di: Zhuravlev, Viktor N., et al.
Pubblicazione: (2025)
di: Zhuravlev, Viktor N., et al.
Pubblicazione: (2025)
Dual Optimal: Make Your LLM Peer-like with Dignity
di: Wang, Xiangqi, et al.
Pubblicazione: (2026)
di: Wang, Xiangqi, et al.
Pubblicazione: (2026)
CAMEL: Confidence-Gated Reflection for Reward Modeling
di: Zhu, Zirui, et al.
Pubblicazione: (2026)
di: Zhu, Zirui, et al.
Pubblicazione: (2026)
Enhancing Financial Question Answering with a Multi-Agent Reflection Framework
di: Fatemi, Sorouralsadat, et al.
Pubblicazione: (2024)
di: Fatemi, Sorouralsadat, et al.
Pubblicazione: (2024)
Reflection-Window Decoding: Text Generation with Selective Refinement
di: Tang, Zeyu, et al.
Pubblicazione: (2025)
di: Tang, Zeyu, et al.
Pubblicazione: (2025)
SOYO: A Tuning-Free Approach for Video Style Morphing via Style-Adaptive Interpolation in Diffusion Models
di: Zheng, Haoyu, et al.
Pubblicazione: (2025)
di: Zheng, Haoyu, et al.
Pubblicazione: (2025)
Logic Distillation: Learning from Code Function by Function for Decision-making Tasks
di: Chen, Dong, et al.
Pubblicazione: (2024)
di: Chen, Dong, et al.
Pubblicazione: (2024)
Learning to Learn from Language Feedback with Social Meta-Learning
di: Cook, Jonathan, et al.
Pubblicazione: (2026)
di: Cook, Jonathan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Bridging Local Details and Global Context in Text-Attributed Graphs
di: Wang, Yaoke, et al.
Pubblicazione: (2024) -
DuetRAG: Collaborative Retrieval-Augmented Generation
di: Jiao, Dian, et al.
Pubblicazione: (2024) -
IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization
di: Cao, Jie, et al.
Pubblicazione: (2024) -
Efficient Tuning and Inference for Large Language Models on Textual Graphs
di: Zhu, Yun, et al.
Pubblicazione: (2024) -
Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization
di: Zhang, Wenqi, et al.
Pubblicazione: (2024)