TraceNAS: Zero-shot LLM Pruning via Gradient Trace Correlation
Fuente:
arXiv
Saved in:
| Main Authors: | Malettira, Prajna G., Nagaraj, Manish, Roy, Arjun, Negi, Shubham, Roy, Kaushik |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
by: Roy, Arjun, et al.
Published: (2026)
by: Roy, Arjun, et al.
Published: (2026)
TSkips: Efficiency Through Explicit Temporal Delay Connections in Spiking Neural Networks
by: Malettira, Prajna G., et al.
Published: (2024)
by: Malettira, Prajna G., et al.
Published: (2024)
TRIM: Token-wise Attention-Derived Saliency for Data-Efficient Instruction Tuning
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
Coresets from Trajectories: Selecting Data via Correlation of Loss Differences
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
TOFU: Towards Obfuscated Federated Updates by Encoding Weight Updates into Gradients from Proxy Data
by: Garg, Isha, et al.
Published: (2022)
by: Garg, Isha, et al.
Published: (2022)
Prompt-Based Bias Calibration for Better Zero/Few-Shot Learning of Language Models
by: He, Kang, et al.
Published: (2024)
by: He, Kang, et al.
Published: (2024)
SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution
by: He, Kang, et al.
Published: (2026)
by: He, Kang, et al.
Published: (2026)
TOFFE -- Temporally-binned Object Flow from Events for High-speed and Energy-Efficient Object Detection and Tracking
by: Kosta, Adarsh Kumar, et al.
Published: (2025)
by: Kosta, Adarsh Kumar, et al.
Published: (2025)
Where Did It Go Wrong? Attributing Undesirable LLM Behaviors via Representation Gradient Tracing
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
On Pruning State-Space LLMs
by: Ghattas, Tamer, et al.
Published: (2025)
by: Ghattas, Tamer, et al.
Published: (2025)
Best of Both Worlds: Hybrid SNN-ANN Architecture for Event-based Optical Flow Estimation
by: Negi, Shubham, et al.
Published: (2023)
by: Negi, Shubham, et al.
Published: (2023)
DCT-CryptoNets: Scaling Private Inference in the Frequency Domain
by: Roy, Arjun, et al.
Published: (2024)
by: Roy, Arjun, et al.
Published: (2024)
LogicTree: Structured Proof Exploration for Coherent and Rigorous Logical Reasoning with Large Language Models
by: He, Kang, et al.
Published: (2025)
by: He, Kang, et al.
Published: (2025)
DART-ing Through the Drift: Dynamic Tracing of Knowledge Neurons for Adaptive Inference-Time Pruning
by: Tyagi, Abhishek, et al.
Published: (2026)
by: Tyagi, Abhishek, et al.
Published: (2026)
Finding the Muses: Identifying Coresets through Loss Trajectories
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries
by: Hicke, Rebecca M. M., et al.
Published: (2026)
by: Hicke, Rebecca M. M., et al.
Published: (2026)
SpiDR: A Reconfigurable Digital Compute-in-Memory Spiking Neural Network Accelerator for Event-based Perception
by: Sharma, Deepika, et al.
Published: (2024)
by: Sharma, Deepika, et al.
Published: (2024)
ResQ: Mixed-Precision Quantization of Large Language Models with Low-Rank Residuals
by: Saxena, Utkarsh, et al.
Published: (2024)
by: Saxena, Utkarsh, et al.
Published: (2024)
FEDORA: Flying Event Dataset fOr Reactive behAvior
by: Joshi, Amogh, et al.
Published: (2023)
by: Joshi, Amogh, et al.
Published: (2023)
TraceDet: Hallucination Detection from the Decoding Trace of Diffusion Large Language Models
by: Chang, Shenxu, et al.
Published: (2025)
by: Chang, Shenxu, et al.
Published: (2025)
Value Drifts: Tracing Value Alignment During LLM Post-Training
by: Bhatia, Mehar, et al.
Published: (2025)
by: Bhatia, Mehar, et al.
Published: (2025)
Zero-shot and Few-shot Generation Strategies for Artificial Clinical Records
by: Frayling, Erlend, et al.
Published: (2024)
by: Frayling, Erlend, et al.
Published: (2024)
Variance Control via Weight Rescaling in LLM Pre-training
by: Owen, Louis, et al.
Published: (2025)
by: Owen, Louis, et al.
Published: (2025)
PORT: Preference Optimization on Reasoning Traces
by: Lahlou, Salem, et al.
Published: (2024)
by: Lahlou, Salem, et al.
Published: (2024)
Domain-Specific Pruning of Large Mixture-of-Experts Models with Few-shot Demonstrations
by: Dong, Zican, et al.
Published: (2025)
by: Dong, Zican, et al.
Published: (2025)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
by: Parekh, Tanmay, et al.
Published: (2025)
by: Parekh, Tanmay, et al.
Published: (2025)
PLAY2PROMPT: Zero-shot Tool Instruction Optimization for LLM Agents via Tool Play
by: Fang, Wei, et al.
Published: (2025)
by: Fang, Wei, et al.
Published: (2025)
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
by: Chen, Jianhui, et al.
Published: (2026)
by: Chen, Jianhui, et al.
Published: (2026)
Exploring Fusion Techniques in Multimodal AI-Based Recruitment: Insights from FairCVdb
by: Swati, Swati, et al.
Published: (2024)
by: Swati, Swati, et al.
Published: (2024)
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
by: Deng, Xinle, et al.
Published: (2026)
by: Deng, Xinle, et al.
Published: (2026)
Wanda++: Pruning Large Language Models via Regional Gradients
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Enhancing Authorship Attribution through Embedding Fusion: A Novel Approach with Masked and Encoder-Decoder Language Models
by: Kaushik, Arjun Ramesh, et al.
Published: (2024)
by: Kaushik, Arjun Ramesh, et al.
Published: (2024)
Zero-shot Factual Consistency Evaluation Across Domains
by: Agarwal, Raunak
Published: (2024)
by: Agarwal, Raunak
Published: (2024)
Eigen Attention: Attention in Low-Rank Space for KV Cache Compression
by: Saxena, Utkarsh, et al.
Published: (2024)
by: Saxena, Utkarsh, et al.
Published: (2024)
Beyond Behavioural Trade-Offs: Mechanistic Tracing of Pain-Pleasure Decisions in an LLM
by: Bianco, Francesca, et al.
Published: (2026)
by: Bianco, Francesca, et al.
Published: (2026)
Are You Being Tracked? Discover the Power of Zero-Shot Trajectory Tracing with LLMs!
by: Yang, Huanqi, et al.
Published: (2024)
by: Yang, Huanqi, et al.
Published: (2024)
Bypass Back-propagation: Optimization-based Structural Pruning for Large Language Models via Policy Gradient
by: Gao, Yuan, et al.
Published: (2024)
by: Gao, Yuan, et al.
Published: (2024)
Tracing Computation Density in LLMs
by: Kervadec, Corentin, et al.
Published: (2026)
by: Kervadec, Corentin, et al.
Published: (2026)
LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddings
by: Wang, Duo, et al.
Published: (2024)
by: Wang, Duo, et al.
Published: (2024)
Knowing When to Defer: Selective Prediction for Responsible Knowledge Tracing
by: Mitton, Joshua, et al.
Published: (2025)
by: Mitton, Joshua, et al.
Published: (2025)
Similar Items
-
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
by: Roy, Arjun, et al.
Published: (2026) -
TSkips: Efficiency Through Explicit Temporal Delay Connections in Spiking Neural Networks
by: Malettira, Prajna G., et al.
Published: (2024) -
TRIM: Token-wise Attention-Derived Saliency for Data-Efficient Instruction Tuning
by: Nagaraj, Manish, et al.
Published: (2025) -
Coresets from Trajectories: Selecting Data via Correlation of Loss Differences
by: Nagaraj, Manish, et al.
Published: (2025) -
TOFU: Towards Obfuscated Federated Updates by Encoding Weight Updates into Gradients from Proxy Data
by: Garg, Isha, et al.
Published: (2022)