Detecting Distillation Data from Reasoning Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Hengxiang, Choi, Hyeong Kyu, Li, Sharon, Wei, Hongxin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Fine-tuning can Help Detect Pretraining Data from Large Language Models
di: Zhang, Hengxiang, et al.
Pubblicazione: (2024)
di: Zhang, Hengxiang, et al.
Pubblicazione: (2024)
ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2026)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2026)
How Contaminated Is Your Benchmark? Quantifying Dataset Leakage in Large Language Models with Kernel Divergence
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
PICLe: Eliciting Diverse Behaviors from Large Language Models with Persona In-Context Learning
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
Safety-Aware Fine-Tuning of Large Language Models
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation
di: Park, Seongheon, et al.
Pubblicazione: (2026)
di: Park, Seongheon, et al.
Pubblicazione: (2026)
When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
Mixed Distillation Helps Smaller Language Model Better Reasoning
di: Li, Chenglin, et al.
Pubblicazione: (2023)
di: Li, Chenglin, et al.
Pubblicazione: (2023)
Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation
di: Dai, Chengwei, et al.
Pubblicazione: (2024)
di: Dai, Chengwei, et al.
Pubblicazione: (2024)
Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning
di: Hu, Qiang, et al.
Pubblicazione: (2024)
di: Hu, Qiang, et al.
Pubblicazione: (2024)
Distilling Mathematical Reasoning Capabilities into Small Language Models
di: Zhu, Xunyu, et al.
Pubblicazione: (2024)
di: Zhu, Xunyu, et al.
Pubblicazione: (2024)
Distilling the Essence: Efficient Reasoning Distillation via Sequence Truncation
di: Chen, Wei-Rui, et al.
Pubblicazione: (2025)
di: Chen, Wei-Rui, et al.
Pubblicazione: (2025)
Reliable Reasoning Path: Distilling Effective Guidance for LLM Reasoning with Knowledge Graphs
di: Xiao, Yilin, et al.
Pubblicazione: (2025)
di: Xiao, Yilin, et al.
Pubblicazione: (2025)
Debate or Vote: Which Yields Better Decisions in Multi-Agent Large Language Models?
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
Key-Point-Driven Mathematical Reasoning Distillation of Large Language Model
di: Zhu, Xunyu, et al.
Pubblicazione: (2024)
di: Zhu, Xunyu, et al.
Pubblicazione: (2024)
Improve Student's Reasoning Generalizability through Cascading Decomposed CoTs Distillation
di: Dai, Chengwei, et al.
Pubblicazione: (2024)
di: Dai, Chengwei, et al.
Pubblicazione: (2024)
Thinking with DistilQwen: A Tale of Four Distilled Reasoning and Reward Model Series
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
Detecting RLVR Training Data via Structural Convergence of Reasoning
di: Zhang, Hongbo, et al.
Pubblicazione: (2026)
di: Zhang, Hongbo, et al.
Pubblicazione: (2026)
Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation
di: Zhu, Xunyu, et al.
Pubblicazione: (2024)
di: Zhu, Xunyu, et al.
Pubblicazione: (2024)
LLM-MRD: LLM-Guided Multi-View Reasoning Distillation for Fake News Detection
di: Zhou, Weilin, et al.
Pubblicazione: (2026)
di: Zhou, Weilin, et al.
Pubblicazione: (2026)
ReasoningShield: Safety Detection over Reasoning Traces of Large Reasoning Models
di: Li, Changyi, et al.
Pubblicazione: (2025)
di: Li, Changyi, et al.
Pubblicazione: (2025)
Clean First, Align Later: Benchmarking Preference Data Cleaning for Reliable LLM Alignment
di: Yeh, Samuel, et al.
Pubblicazione: (2025)
di: Yeh, Samuel, et al.
Pubblicazione: (2025)
Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation
di: He, Bowei, et al.
Pubblicazione: (2026)
di: He, Bowei, et al.
Pubblicazione: (2026)
Learning from Committee: Reasoning Distillation from a Mixture of Teachers with Peer-Review
di: Li, Zhuochun, et al.
Pubblicazione: (2024)
di: Li, Zhuochun, et al.
Pubblicazione: (2024)
MiniLLM: On-Policy Distillation of Large Language Models
di: Gu, Yuxian, et al.
Pubblicazione: (2023)
di: Gu, Yuxian, et al.
Pubblicazione: (2023)
From Reasoning to Answer: Empirical, Attention-Based and Mechanistic Insights into Distilled DeepSeek R1 Models
di: Zhang, Jue, et al.
Pubblicazione: (2025)
di: Zhang, Jue, et al.
Pubblicazione: (2025)
SAGE-32B: Agentic Reasoning via Iterative Distillation
di: Jha, Basab, et al.
Pubblicazione: (2026)
di: Jha, Basab, et al.
Pubblicazione: (2026)
Document Intelligence in the Era of Large Language Models: A Survey
di: Wang, Weishi, et al.
Pubblicazione: (2025)
di: Wang, Weishi, et al.
Pubblicazione: (2025)
An Efficient and Precise Training Data Construction Framework for Process-supervised Reward Model in Mathematical Reasoning
di: Sun, Wei, et al.
Pubblicazione: (2025)
di: Sun, Wei, et al.
Pubblicazione: (2025)
Style over Substance: Distilled Language Models Reason Via Stylistic Replication
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners
di: Paliotta, Daniele, et al.
Pubblicazione: (2025)
di: Paliotta, Daniele, et al.
Pubblicazione: (2025)
Black-Box On-Policy Distillation of Large Language Models
di: Ye, Tianzhu, et al.
Pubblicazione: (2025)
di: Ye, Tianzhu, et al.
Pubblicazione: (2025)
Information-Theoretic Distillation for Reference-less Summarization
di: Jung, Jaehun, et al.
Pubblicazione: (2024)
di: Jung, Jaehun, et al.
Pubblicazione: (2024)
Chunk-Distilled Language Modeling
di: Li, Yanhong, et al.
Pubblicazione: (2024)
di: Li, Yanhong, et al.
Pubblicazione: (2024)
Self-Correction Distillation for Structured Data Question Answering
di: Zhu, Yushan, et al.
Pubblicazione: (2025)
di: Zhu, Yushan, et al.
Pubblicazione: (2025)
Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?
di: Zhang, Yudi, et al.
Pubblicazione: (2025)
di: Zhang, Yudi, et al.
Pubblicazione: (2025)
Structured Chemistry Reasoning with Large Language Models
di: Ouyang, Siru, et al.
Pubblicazione: (2023)
di: Ouyang, Siru, et al.
Pubblicazione: (2023)
Distilling an End-to-End Voice Assistant Without Instruction Training Data
di: Held, William, et al.
Pubblicazione: (2024)
di: Held, William, et al.
Pubblicazione: (2024)
Efficient Mathematical Reasoning Models via Dynamic Pruning and Knowledge Distillation
di: Yu, Fengming, et al.
Pubblicazione: (2025)
di: Yu, Fengming, et al.
Pubblicazione: (2025)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
di: Huang, Jiazhen, et al.
Pubblicazione: (2026)
di: Huang, Jiazhen, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Fine-tuning can Help Detect Pretraining Data from Large Language Models
di: Zhang, Hengxiang, et al.
Pubblicazione: (2024) -
ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2026) -
How Contaminated Is Your Benchmark? Quantifying Dataset Leakage in Large Language Models with Kernel Divergence
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025) -
PICLe: Eliciting Diverse Behaviors from Large Language Models with Persona In-Context Learning
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024) -
Safety-Aware Fine-Tuning of Large Language Models
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)