Utilize the Flow before Stepping into the Same River Twice: Certainty Represented Knowledge Flow for Refusal-Aware Instruction Tuning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhu, Runchuan, Ma, Zhipeng, Wu, Jiang, Gao, Junyuan, Wang, Jiaqi, Lin, Dahua, He, Conghui |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
par: Zhu, Runchuan, et autres
Publié: (2025)
par: Zhu, Runchuan, et autres
Publié: (2025)
Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error
par: Tang, Chenming, et autres
Publié: (2025)
par: Tang, Chenming, et autres
Publié: (2025)
Evaluating Large Language Model with Knowledge Oriented Language Specific Simple Question Answering
par: Jiang, Bowen, et autres
Publié: (2025)
par: Jiang, Bowen, et autres
Publié: (2025)
PM4Bench: Benchmarking Large Vision-Language Models with Parallel Multilingual Multi-Modal Multi-task Corpus
par: Gao, Junyuan, et autres
Publié: (2025)
par: Gao, Junyuan, et autres
Publié: (2025)
Parenting: Optimizing Knowledge Selection of Retrieval-Augmented Language Models with Parameter Decoupling and Tailored Tuning
par: Xu, Yongxin, et autres
Publié: (2024)
par: Xu, Yongxin, et autres
Publié: (2024)
AdaptFlow: Adaptive Workflow Optimization via Meta-Learning
par: Zhu, Runchuan, et autres
Publié: (2025)
par: Zhu, Runchuan, et autres
Publié: (2025)
BLINK-Twice: You see, but do you observe? A Reasoning Benchmark on Visual Perception
par: Ye, Junyan, et autres
Publié: (2025)
par: Ye, Junyan, et autres
Publié: (2025)
Look Twice before You Leap: A Rational Framework for Localized Adversarial Anonymization
par: Duan, Donghang, et autres
Publié: (2025)
par: Duan, Donghang, et autres
Publié: (2025)
Learning to Refuse: Refusal-Aware Reinforcement Fine-Tuning for Hard-Irrelevant Queries in Video Temporal Grounding
par: Lee, Jin-Seop, et autres
Publié: (2025)
par: Lee, Jin-Seop, et autres
Publié: (2025)
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks
par: Pan, Wenbo, et autres
Publié: (2025)
par: Pan, Wenbo, et autres
Publié: (2025)
Rethinking Traffic Flow Forecasting: From Transition to Generatation
par: Shijiao, Li, et autres
Publié: (2025)
par: Shijiao, Li, et autres
Publié: (2025)
Reflecting Twice before Speaking with Empathy: Self-Reflective Alternating Inference for Empathy-Aware End-to-End Spoken Dialogue
par: Jia, Yuhang, et autres
Publié: (2026)
par: Jia, Yuhang, et autres
Publié: (2026)
FoundaBench: Evaluating Chinese Fundamental Knowledge Capabilities of Large Language Models
par: Li, Wei, et autres
Publié: (2024)
par: Li, Wei, et autres
Publié: (2024)
Never the Same Storytime Twice: An Exploration of the Nature and Role of Reflection in Public Library Storytime Assessment
par: J. Elizabeth Mills
Publié: (2021)
par: J. Elizabeth Mills
Publié: (2021)
Trajectory Consistency for One-Step Generation on Euler Mean Flows
par: Li, Zhiqi, et autres
Publié: (2026)
par: Li, Zhiqi, et autres
Publié: (2026)
Process-Supervised LLM Recommenders via Flow-guided Tuning
par: Gao, Chongming, et autres
Publié: (2025)
par: Gao, Chongming, et autres
Publié: (2025)
SlimFlow: Training Smaller One-Step Diffusion Models with Rectified Flow
par: Zhu, Yuanzhi, et autres
Publié: (2024)
par: Zhu, Yuanzhi, et autres
Publié: (2024)
A First Step Toward Legality and Certainty / Alberto Anaya
par: Anaya, Alberto
par: Anaya, Alberto
HiFlow: Training-free High-Resolution Image Generation with Flow-Aligned Guidance
par: Bu, Jiazi, et autres
Publié: (2025)
par: Bu, Jiazi, et autres
Publié: (2025)
Flow-Modulated Scoring for Semantic-Aware Knowledge Graph Completion
par: Li, Siyuan, et autres
Publié: (2025)
par: Li, Siyuan, et autres
Publié: (2025)
MeanFlow-TSE: One-Step Generative Target Speaker Extraction with Mean Flow
par: Shimizu, Riki, et autres
Publié: (2025)
par: Shimizu, Riki, et autres
Publié: (2025)
HACK: Hallucinations Along Certainty and Knowledge Axes
par: Simhi, Adi, et autres
Publié: (2025)
par: Simhi, Adi, et autres
Publié: (2025)
T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
par: Chen, Zehui, et autres
Publié: (2023)
par: Chen, Zehui, et autres
Publié: (2023)
KnowPO: Knowledge-aware Preference Optimization for Controllable Knowledge Selection in Retrieval-Augmented Language Models
par: Zhang, Ruizhe, et autres
Publié: (2024)
par: Zhang, Ruizhe, et autres
Publié: (2024)
Certainty in Uncertainty: Reasoning over Uncertain Knowledge Graphs with Statistical Guarantees
par: Zhu, Yuqicheng, et autres
Publié: (2025)
par: Zhu, Yuqicheng, et autres
Publié: (2025)
One-Step Flow Policy Mirror Descent
par: Chen, Tianyi, et autres
Publié: (2025)
par: Chen, Tianyi, et autres
Publié: (2025)
DSFlow: Dual Supervision and Step-Aware Architecture for One-Step Flow Matching Speech Synthesis
par: Lin, Bin, et autres
Publié: (2026)
par: Lin, Bin, et autres
Publié: (2026)
Watching The River Flow
par: João Carlos Winck
Publié: (2011)
par: João Carlos Winck
Publié: (2011)
Importance-Aware Data Selection for Efficient LLM Instruction Tuning
par: Jiang, Tingyu, et autres
Publié: (2025)
par: Jiang, Tingyu, et autres
Publié: (2025)
AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation
par: Gu, Yuchao, et autres
Publié: (2026)
par: Gu, Yuchao, et autres
Publié: (2026)
Knowledge Graph Embedding by Normalizing Flows
par: Xiao, Changyi, et autres
Publié: (2024)
par: Xiao, Changyi, et autres
Publié: (2024)
Mean-Flow based One-Step Vision-Language-Action
par: Chen, Yang, et autres
Publié: (2026)
par: Chen, Yang, et autres
Publié: (2026)
MeanFlowSE: One-Step Generative Speech Enhancement via MeanFlow
par: Zhu, Yike, et autres
Publié: (2025)
par: Zhu, Yike, et autres
Publié: (2025)
Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
par: Zhao, Zhiyuan, et autres
Publié: (2023)
par: Zhao, Zhiyuan, et autres
Publié: (2023)
Think Twice: Branch-and-Rethink Reasoning Reward Model
par: Jiao, Yizhu, et autres
Publié: (2025)
par: Jiao, Yizhu, et autres
Publié: (2025)
Representing Flow Fields with Divergence-Free Kernels for Reconstruction
par: Ni, Xingyu, et autres
Publié: (2025)
par: Ni, Xingyu, et autres
Publié: (2025)
Refusal and Aporia: At the Limits of Anthropological Knowledge
par: Cory‐Alice André‐Johnson
Publié: (2026)
par: Cory‐Alice André‐Johnson
Publié: (2026)
FlowDrive: Energy Flow Field for End-to-End Autonomous Driving
par: Jiang, Hao, et autres
Publié: (2025)
par: Jiang, Hao, et autres
Publié: (2025)
Unearthing Large Scale Domain-Specific Knowledge from Public Corpora
par: Fei, Zhaoye, et autres
Publié: (2024)
par: Fei, Zhaoye, et autres
Publié: (2024)
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge
par: Fu, Jinlan, et autres
Publié: (2024)
par: Fu, Jinlan, et autres
Publié: (2024)
Documents similaires
-
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
par: Zhu, Runchuan, et autres
Publié: (2025) -
Do Not Step Into the Same River Twice: Learning to Reason from Trial and Error
par: Tang, Chenming, et autres
Publié: (2025) -
Evaluating Large Language Model with Knowledge Oriented Language Specific Simple Question Answering
par: Jiang, Bowen, et autres
Publié: (2025) -
PM4Bench: Benchmarking Large Vision-Language Models with Parallel Multilingual Multi-Modal Multi-task Corpus
par: Gao, Junyuan, et autres
Publié: (2025) -
Parenting: Optimizing Knowledge Selection of Retrieval-Augmented Language Models with Parameter Decoupling and Tailored Tuning
par: Xu, Yongxin, et autres
Publié: (2024)