D.Va: Validate Your Demonstration First Before You Use It
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Qi, Xiao, Zhiqing, Xiao, Ruixuan, Gao, Lirong, Zhao, Junbo |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models
par: Chen, Hao, et autres
Publié: (2025)
par: Chen, Hao, et autres
Publié: (2025)
DORY: Deliberative Prompt Recovery for LLM
par: Gao, Lirong, et autres
Publié: (2024)
par: Gao, Lirong, et autres
Publié: (2024)
On LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A Survey
par: Long, Lin, et autres
Publié: (2024)
par: Long, Lin, et autres
Publié: (2024)
FLoE: Fisher-Based Layer Selection for Efficient Sparse Adaptation of Low-Rank Experts
par: Wang, Xinyi, et autres
Publié: (2025)
par: Wang, Xinyi, et autres
Publié: (2025)
Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report
par: Cacioli, Jon-Paul
Publié: (2026)
par: Cacioli, Jon-Paul
Publié: (2026)
Look Before You Leap: Towards Decision-Aware and Generalizable Tool-Usage for Large Language Models
par: Gui, Anchun, et autres
Publié: (2024)
par: Gui, Anchun, et autres
Publié: (2024)
Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
par: Luo, Jiani, et autres
Publié: (2026)
par: Luo, Jiani, et autres
Publié: (2026)
FlowBench: Revisiting and Benchmarking Workflow-Guided Planning for LLM-based Agents
par: Xiao, Ruixuan, et autres
Publié: (2024)
par: Xiao, Ruixuan, et autres
Publié: (2024)
Think Thrice Before You Act: Progressive Thought Refinement in Large Language Models
par: Du, Chengyu, et autres
Publié: (2024)
par: Du, Chengyu, et autres
Publié: (2024)
Your Students Don't Use LLMs Like You Wish They Did
par: Kobler, Sebastian, et autres
Publié: (2026)
par: Kobler, Sebastian, et autres
Publié: (2026)
Look Before You Leap: Autonomous Exploration for LLM Agents
par: Ye, Ziang, et autres
Publié: (2026)
par: Ye, Ziang, et autres
Publié: (2026)
Screen Before You Interpret: A Portable Validity Protocol for Benchmark-Based LLM Confidence Signals
par: Cacioli, Jon-Paul
Publié: (2026)
par: Cacioli, Jon-Paul
Publié: (2026)
Can LLMs Act as Historians? Evaluating Historical Research Capabilities of LLMs via the Chinese Imperial Examination
par: Gao, Lirong, et autres
Publié: (2026)
par: Gao, Lirong, et autres
Publié: (2026)
Stop Unnecessary Reflection: Training LRMs for Efficient Reasoning with Adaptive Reflection and Length Coordinated Penalty
par: Yu, Zewei, et autres
Publié: (2026)
par: Yu, Zewei, et autres
Publié: (2026)
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM
par: Zhang, Shaoqing, et autres
Publié: (2024)
par: Zhang, Shaoqing, et autres
Publié: (2024)
Teaching Language Models to Self-Improve through Interactive Demonstrations
par: Yu, Xiao, et autres
Publié: (2023)
par: Yu, Xiao, et autres
Publié: (2023)
RECOST: External Knowledge Guided Data-efficient Instruction Tuning
par: Zhang, Qi, et autres
Publié: (2024)
par: Zhang, Qi, et autres
Publié: (2024)
LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization
par: Zhang, Qi, et autres
Publié: (2025)
par: Zhang, Qi, et autres
Publié: (2025)
Should You Use Your Large Language Model to Explore or Exploit?
par: Harris, Keegan, et autres
Publié: (2025)
par: Harris, Keegan, et autres
Publié: (2025)
From Parameters to Data: A Task-Parameter-Guided Fine-Tuning Pipeline for Efficient LLM Alignment
par: Chen, Hao, et autres
Publié: (2026)
par: Chen, Hao, et autres
Publié: (2026)
CrossICL: Cross-Task In-Context Learning via Unsupervised Demonstration Transfer
par: Gao, Jinglong, et autres
Publié: (2025)
par: Gao, Jinglong, et autres
Publié: (2025)
Think Before You Prune: Selective Self-Generated Calibration for Pruning Large Reasoning Models
par: Xiang, Yang, et autres
Publié: (2025)
par: Xiang, Yang, et autres
Publié: (2025)
Thinking Before You Speak: A Proactive Test-time Scaling Approach
par: Liu, Cong, et autres
Publié: (2025)
par: Liu, Cong, et autres
Publié: (2025)
Walk Before You Run! Concise LLM Reasoning via Reinforcement Learning
par: Song, Mingyang, et autres
Publié: (2025)
par: Song, Mingyang, et autres
Publié: (2025)
Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training
par: Nomura, Yoshinori
Publié: (2026)
par: Nomura, Yoshinori
Publié: (2026)
Self-Train Before You Transcribe
par: Flynn, Robert, et autres
Publié: (2024)
par: Flynn, Robert, et autres
Publié: (2024)
Talk Before You Retrieve: Agent-Led Discussions for Better RAG in Medical QA
par: Dong, Xuanzhao, et autres
Publié: (2025)
par: Dong, Xuanzhao, et autres
Publié: (2025)
Cite Before You Speak: Enhancing Context-Response Grounding in E-commerce Conversational LLM-Agents
par: Zeng, Jingying, et autres
Publié: (2025)
par: Zeng, Jingying, et autres
Publié: (2025)
Watch Before You Answer: Learning from Visually Grounded Post-Training
par: Zhang, Yuxuan, et autres
Publié: (2026)
par: Zhang, Yuxuan, et autres
Publié: (2026)
Shorten After You're Right: Lazy Length Penalties for Reasoning RL
par: Yuan, Danlong, et autres
Publié: (2025)
par: Yuan, Danlong, et autres
Publié: (2025)
SnapKV: LLM Knows What You are Looking for Before Generation
par: Li, Yuhong, et autres
Publié: (2024)
par: Li, Yuhong, et autres
Publié: (2024)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
par: Wang, Ziyan, et autres
Publié: (2025)
par: Wang, Ziyan, et autres
Publié: (2025)
Are Your LLMs Capable of Stable Reasoning?
par: Liu, Junnan, et autres
Publié: (2024)
par: Liu, Junnan, et autres
Publié: (2024)
LLMs Corrupt Your Documents When You Delegate
par: Laban, Philippe, et autres
Publié: (2026)
par: Laban, Philippe, et autres
Publié: (2026)
Think Twice Before You Judge: Mixture of Dual Reasoning Experts for Multimodal Sarcasm Detection
par: Jana, Soumyadeep, et autres
Publié: (2025)
par: Jana, Soumyadeep, et autres
Publié: (2025)
Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety
par: Bonagiri, Vamshi Krishna, et autres
Publié: (2025)
par: Bonagiri, Vamshi Krishna, et autres
Publié: (2025)
Divide-or-Conquer? Which Part Should You Distill Your LLM?
par: Wu, Zhuofeng, et autres
Publié: (2024)
par: Wu, Zhuofeng, et autres
Publié: (2024)
RPDR: A Round-trip Prediction-Based Data Augmentation Framework for Long-Tail Question Answering
par: Zhang, Yiming, et autres
Publié: (2026)
par: Zhang, Yiming, et autres
Publié: (2026)
Think Before You Act: Decision Transformers with Working Memory
par: Kang, Jikun, et autres
Publié: (2023)
par: Kang, Jikun, et autres
Publié: (2023)
Think Before You Lie: How Reasoning Leads to Honesty
par: Yuan, Ann, et autres
Publié: (2026)
par: Yuan, Ann, et autres
Publié: (2026)
Documents similaires
-
ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models
par: Chen, Hao, et autres
Publié: (2025) -
DORY: Deliberative Prompt Recovery for LLM
par: Gao, Lirong, et autres
Publié: (2024) -
On LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A Survey
par: Long, Lin, et autres
Publié: (2024) -
FLoE: Fisher-Based Layer Selection for Efficient Sparse Adaptation of Low-Rank Experts
par: Wang, Xinyi, et autres
Publié: (2025) -
Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report
par: Cacioli, Jon-Paul
Publié: (2026)