Circuit Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Wadhwa, Somin, Amir, Silvio, Wallace, Byron C. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Investigating Mysteries of CoT-Augmented Distillation
by: Wadhwa, Somin, et al.
Published: (2024)
by: Wadhwa, Somin, et al.
Published: (2024)
Who Taught You That? Tracing Teachers in Model Distillation
by: Wadhwa, Somin, et al.
Published: (2025)
by: Wadhwa, Somin, et al.
Published: (2025)
Revisiting Relation Extraction in the era of Large Language Models
by: Wadhwa, Somin, et al.
Published: (2023)
by: Wadhwa, Somin, et al.
Published: (2023)
Learning from Natural Language Explanations for Generalizable Entity Matching
by: Wadhwa, Somin, et al.
Published: (2024)
by: Wadhwa, Somin, et al.
Published: (2024)
Distilling Event Sequence Knowledge From Large Language Models
by: Wadhwa, Somin, et al.
Published: (2024)
by: Wadhwa, Somin, et al.
Published: (2024)
Elucidating Mechanisms of Demographic Bias in LLMs for Healthcare
by: Ahsan, Hiba, et al.
Published: (2025)
by: Ahsan, Hiba, et al.
Published: (2025)
On-the-fly Definition Augmentation of LLMs for Biomedical NER
by: Munnangi, Monica, et al.
Published: (2024)
by: Munnangi, Monica, et al.
Published: (2024)
Retrieving Evidence from EHRs with LLMs: Possibilities and Challenges
by: Ahsan, Hiba, et al.
Published: (2023)
by: Ahsan, Hiba, et al.
Published: (2023)
Can SAEs reveal and mitigate racial biases of LLMs in healthcare?
by: Ahsan, Hiba, et al.
Published: (2025)
by: Ahsan, Hiba, et al.
Published: (2025)
Do Automatic Factuality Metrics Measure Factuality? A Critical Evaluation
by: Ramprasad, Sanjana, et al.
Published: (2024)
by: Ramprasad, Sanjana, et al.
Published: (2024)
Open (Clinical) LLMs are Sensitive to Instruction Phrasings
by: Arroyo, Alberto Mario Ceballos, et al.
Published: (2024)
by: Arroyo, Alberto Mario Ceballos, et al.
Published: (2024)
Vector Arithmetic in Concept and Token Subspaces
by: Feucht, Sheridan, et al.
Published: (2025)
by: Feucht, Sheridan, et al.
Published: (2025)
Leveraging ChatGPT in Pharmacovigilance Event Extraction: An Empirical Study
by: Sun, Zhaoyue, et al.
Published: (2024)
by: Sun, Zhaoyue, et al.
Published: (2024)
Measuring AI "Slop" in Text
by: Shaib, Chantal, et al.
Published: (2025)
by: Shaib, Chantal, et al.
Published: (2025)
Detection and Measurement of Syntactic Templates in Generated Text
by: Shaib, Chantal, et al.
Published: (2024)
by: Shaib, Chantal, et al.
Published: (2024)
SkillFactory: Self-Distillation For Learning Cognitive Behaviors
by: Sprague, Zayne, et al.
Published: (2025)
by: Sprague, Zayne, et al.
Published: (2025)
Compared to What? Baselines and Metrics for Counterfactual Prompting
by: Yang, Zihao, et al.
Published: (2026)
by: Yang, Zihao, et al.
Published: (2026)
Don't Pay Attention, PLANT It: Pretraining Attention via Learning-to-Rank
by: Roy, Debjyoti Saha, et al.
Published: (2024)
by: Roy, Debjyoti Saha, et al.
Published: (2024)
Do Multi-Document Summarization Models Synthesize?
by: DeYoung, Jay, et al.
Published: (2023)
by: DeYoung, Jay, et al.
Published: (2023)
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
by: Shaib, Chantal, et al.
Published: (2025)
by: Shaib, Chantal, et al.
Published: (2025)
How Much Annotation is Needed to Compare Summarization Models?
by: Shaib, Chantal, et al.
Published: (2024)
by: Shaib, Chantal, et al.
Published: (2024)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
by: Ramprasad, Sanjana, et al.
Published: (2024)
by: Ramprasad, Sanjana, et al.
Published: (2024)
Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
by: Yun, Hye Sun, et al.
Published: (2024)
by: Yun, Hye Sun, et al.
Published: (2024)
Future Lens: Anticipating Subsequent Tokens from a Single Hidden State
by: Pal, Koyena, et al.
Published: (2023)
by: Pal, Koyena, et al.
Published: (2023)
Punctuation-aware treebank tree binarization
by: Klinger, Eitan, et al.
Published: (2025)
by: Klinger, Eitan, et al.
Published: (2025)
The Dual-Route Model of Induction
by: Feucht, Sheridan, et al.
Published: (2025)
by: Feucht, Sheridan, et al.
Published: (2025)
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
by: Feucht, Sheridan, et al.
Published: (2024)
by: Feucht, Sheridan, et al.
Published: (2024)
Do Activation Verbalization Methods Convey Privileged Information?
by: Li, Millicent, et al.
Published: (2025)
by: Li, Millicent, et al.
Published: (2025)
BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation
by: Pang, Bo, et al.
Published: (2025)
by: Pang, Bo, et al.
Published: (2025)
Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence
by: Mo, Kaijie, et al.
Published: (2026)
by: Mo, Kaijie, et al.
Published: (2026)
Standardizing the Measurement of Text Diversity: A Tool and a Comparative Analysis of Scores
by: Shaib, Chantal, et al.
Published: (2024)
by: Shaib, Chantal, et al.
Published: (2024)
GenAudit: Fixing Factual Errors in Language Model Outputs with Evidence
by: Krishna, Kundan, et al.
Published: (2024)
by: Krishna, Kundan, et al.
Published: (2024)
Domain-Adaptive Small Language Models for Structured Tax Code Prediction
by: Nath, Souvik, et al.
Published: (2025)
by: Nath, Souvik, et al.
Published: (2025)
InfoLossQA: Characterizing and Recovering Information Loss in Text Simplification
by: Trienes, Jan, et al.
Published: (2024)
by: Trienes, Jan, et al.
Published: (2024)
Using Natural Language Explanations to Rescale Human Judgments
by: Wadhwa, Manya, et al.
Published: (2023)
by: Wadhwa, Manya, et al.
Published: (2023)
Learning to Refine with Fine-Grained Natural Language Feedback
by: Wadhwa, Manya, et al.
Published: (2024)
by: Wadhwa, Manya, et al.
Published: (2024)
Function Vectors in Large Language Models
by: Todd, Eric, et al.
Published: (2023)
by: Todd, Eric, et al.
Published: (2023)
Better heads do not guarantee better binarized constituency parsing
by: Qi, Zeyao, et al.
Published: (2026)
by: Qi, Zeyao, et al.
Published: (2026)
FactPICO: Factuality Evaluation for Plain Language Summarization of Medical Evidence
by: Joseph, Sebastian Antony, et al.
Published: (2024)
by: Joseph, Sebastian Antony, et al.
Published: (2024)
Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?
by: Yun, Hye Sun, et al.
Published: (2025)
by: Yun, Hye Sun, et al.
Published: (2025)
Similar Items
-
Investigating Mysteries of CoT-Augmented Distillation
by: Wadhwa, Somin, et al.
Published: (2024) -
Who Taught You That? Tracing Teachers in Model Distillation
by: Wadhwa, Somin, et al.
Published: (2025) -
Revisiting Relation Extraction in the era of Large Language Models
by: Wadhwa, Somin, et al.
Published: (2023) -
Learning from Natural Language Explanations for Generalizable Entity Matching
by: Wadhwa, Somin, et al.
Published: (2024) -
Distilling Event Sequence Knowledge From Large Language Models
by: Wadhwa, Somin, et al.
Published: (2024)