MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Jiarui, Zhang, Lechen, Yang, Yongjin, He, Yinghui, Wang, Yingheng, Xuan, Weihao, Jin, Zhijing, Diab, Mona |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Taming Object Hallucinations with Verified Atomic Confidence Estimation
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
di: Liu, Jiarui, et al.
Pubblicazione: (2024)
di: Liu, Jiarui, et al.
Pubblicazione: (2024)
Towards Global AI Inclusivity: A Large-Scale Multilingual Terminology Dataset (GIST)
di: Liu, Jiarui, et al.
Pubblicazione: (2024)
di: Liu, Jiarui, et al.
Pubblicazione: (2024)
CORE: Measuring Multi-Agent LLM Interaction Quality under Game-Theoretic Pressures
di: Pandey, Punya Syon, et al.
Pubblicazione: (2025)
di: Pandey, Punya Syon, et al.
Pubblicazione: (2025)
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
di: He, Yinghui, et al.
Pubblicazione: (2026)
di: He, Yinghui, et al.
Pubblicazione: (2026)
Analyzing the Role of Semantic Representations in the Era of Large Language Models
di: Jin, Zhijing, et al.
Pubblicazione: (2024)
di: Jin, Zhijing, et al.
Pubblicazione: (2024)
LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Utilization
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
di: Liu, Jiarui, et al.
Pubblicazione: (2025)
Can Large Language Models Infer Causation from Correlation?
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
Sentipolis: Emotion-Aware Agents for Social Simulations
di: Fu, Chiyuan, et al.
Pubblicazione: (2026)
di: Fu, Chiyuan, et al.
Pubblicazione: (2026)
Humanizing Machines: Rethinking LLM Anthropomorphism Through a Multi-Level Framework of Design
di: Xiao, Yunze, et al.
Pubblicazione: (2025)
di: Xiao, Yunze, et al.
Pubblicazione: (2025)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
di: Choi, Younwoo, et al.
Pubblicazione: (2025)
di: Choi, Younwoo, et al.
Pubblicazione: (2025)
Skill-Aware Data Selection and Fine-Tuning for Data-Efficient Reasoning Distillation
di: Zhang, Lechen, et al.
Pubblicazione: (2026)
di: Zhang, Lechen, et al.
Pubblicazione: (2026)
Efficient Knowledge Injection in LLMs via Self-Distillation
di: Kujanpää, Kalle, et al.
Pubblicazione: (2024)
di: Kujanpää, Kalle, et al.
Pubblicazione: (2024)
BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data
di: Li, Wenkai, et al.
Pubblicazione: (2024)
di: Li, Wenkai, et al.
Pubblicazione: (2024)
SD$^2$: Self-Distilled Sparse Drafters
di: Lasby, Mike, et al.
Pubblicazione: (2025)
di: Lasby, Mike, et al.
Pubblicazione: (2025)
A Note on Bias to Complete
di: Xu, Jia, et al.
Pubblicazione: (2024)
di: Xu, Jia, et al.
Pubblicazione: (2024)
UniSD: Towards a Unified Self-Distillation Framework for Large Language Models
di: Jin, Yiqiao, et al.
Pubblicazione: (2026)
di: Jin, Yiqiao, et al.
Pubblicazione: (2026)
Towards Valid Student Simulation with Large Language Models
di: Yuan, Zhihao, et al.
Pubblicazione: (2026)
di: Yuan, Zhihao, et al.
Pubblicazione: (2026)
Evaluating Large Language Model Biases in Persona-Steered Generation
di: Liu, Andy, et al.
Pubblicazione: (2024)
di: Liu, Andy, et al.
Pubblicazione: (2024)
SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning
di: Ma, Yufei, et al.
Pubblicazione: (2026)
di: Ma, Yufei, et al.
Pubblicazione: (2026)
Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
di: Wang, Hao, et al.
Pubblicazione: (2026)
di: Wang, Hao, et al.
Pubblicazione: (2026)
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
di: Wang, Yanting, et al.
Pubblicazione: (2025)
di: Wang, Yanting, et al.
Pubblicazione: (2025)
HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents
di: Yeo, Woongyeng, et al.
Pubblicazione: (2026)
di: Yeo, Woongyeng, et al.
Pubblicazione: (2026)
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
di: Salama, Rana, et al.
Pubblicazione: (2025)
di: Salama, Rana, et al.
Pubblicazione: (2025)
Enhancing Cross-Tokenizer Knowledge Distillation with Contextual Dynamical Mapping
di: Chen, Yijie, et al.
Pubblicazione: (2025)
di: Chen, Yijie, et al.
Pubblicazione: (2025)
Massive Values in Self-Attention Modules are the Key to Contextual Knowledge Understanding
di: Jin, Mingyu, et al.
Pubblicazione: (2025)
di: Jin, Mingyu, et al.
Pubblicazione: (2025)
Corrupted by Reasoning: Reasoning Language Models Become Free-Riders in Public Goods Games
di: Piedrahita, David Guzman, et al.
Pubblicazione: (2025)
di: Piedrahita, David Guzman, et al.
Pubblicazione: (2025)
SE-Bench: Benchmarking Self-Evolution with Knowledge Internalization
di: Yuan, Jiarui, et al.
Pubblicazione: (2026)
di: Yuan, Jiarui, et al.
Pubblicazione: (2026)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
di: Alqahtani, Amal, et al.
Pubblicazione: (2025)
di: Alqahtani, Amal, et al.
Pubblicazione: (2025)
DWTSumm: Discrete Wavelet Transform for Document Summarization
di: Salama, Rana, et al.
Pubblicazione: (2026)
di: Salama, Rana, et al.
Pubblicazione: (2026)
Semantic Compression for Word and Sentence Embeddings using Discrete Wavelet Transform
di: Salama, Rana Aref, et al.
Pubblicazione: (2025)
di: Salama, Rana Aref, et al.
Pubblicazione: (2025)
Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap
di: Qi, Xuan, et al.
Pubblicazione: (2025)
di: Qi, Xuan, et al.
Pubblicazione: (2025)
Mixed Distillation Helps Smaller Language Model Better Reasoning
di: Li, Chenglin, et al.
Pubblicazione: (2023)
di: Li, Chenglin, et al.
Pubblicazione: (2023)
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge
di: Li, Yangning, et al.
Pubblicazione: (2022)
di: Li, Yangning, et al.
Pubblicazione: (2022)
Beyond Mimicry to Contextual Guidance: Knowledge Distillation for Interactive AI
di: Wang, Tong, et al.
Pubblicazione: (2024)
di: Wang, Tong, et al.
Pubblicazione: (2024)
Revealing Hidden Mechanisms of Cross-Country Content Moderation with Natural Language Processing
di: Yadav, Neemesh, et al.
Pubblicazione: (2025)
di: Yadav, Neemesh, et al.
Pubblicazione: (2025)
Contextualization Distillation from Large Language Model for Knowledge Graph Completion
di: Li, Dawei, et al.
Pubblicazione: (2024)
di: Li, Dawei, et al.
Pubblicazione: (2024)
Decoding Dark Matter: Specialized Sparse Autoencoders for Interpreting Rare Concepts in Foundation Models
di: Muhamed, Aashiq, et al.
Pubblicazione: (2024)
di: Muhamed, Aashiq, et al.
Pubblicazione: (2024)
Emotion Classification in Low and Moderate Resource Languages
di: Tafreshi, Shabnam, et al.
Pubblicazione: (2024)
di: Tafreshi, Shabnam, et al.
Pubblicazione: (2024)
SimBA: Simplifying Benchmark Analysis Using Performance Matrices Alone
di: Subramani, Nishant, et al.
Pubblicazione: (2025)
di: Subramani, Nishant, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Taming Object Hallucinations with Verified Atomic Confidence Estimation
di: Liu, Jiarui, et al.
Pubblicazione: (2025) -
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
di: Liu, Jiarui, et al.
Pubblicazione: (2024) -
Towards Global AI Inclusivity: A Large-Scale Multilingual Terminology Dataset (GIST)
di: Liu, Jiarui, et al.
Pubblicazione: (2024) -
CORE: Measuring Multi-Agent LLM Interaction Quality under Game-Theoretic Pressures
di: Pandey, Punya Syon, et al.
Pubblicazione: (2025) -
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
di: He, Yinghui, et al.
Pubblicazione: (2026)