Decomposing and Editing Predictions by Modeling Model Computation
Fuente:
arXiv
Saved in:
| Main Authors: | Shah, Harshay, Ilyas, Andrew, Madry, Aleksander |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ContextCite: Attributing Model Generation to Context
by: Cohen-Wang, Benjamin, et al.
Published: (2024)
by: Cohen-Wang, Benjamin, et al.
Published: (2024)
Optimizing ML Training with Metagradient Descent
by: Engstrom, Logan, et al.
Published: (2025)
by: Engstrom, Logan, et al.
Published: (2025)
Parameters vs FLOPs: Scaling Laws for Optimal Sparsity for Mixture-of-Experts Language Models
by: Abnar, Samira, et al.
Published: (2025)
by: Abnar, Samira, et al.
Published: (2025)
User Strategization and Trustworthy Algorithms
by: Cen, Sarah H., et al.
Published: (2023)
by: Cen, Sarah H., et al.
Published: (2023)
Why Inference in Large Models Becomes Decomposable After Training
by: Jin, Jidong
Published: (2026)
by: Jin, Jidong
Published: (2026)
A Decomposable Forward Process in Diffusion Models for Time-Series Forecasting
by: Caldas, Francisco, et al.
Published: (2026)
by: Caldas, Francisco, et al.
Published: (2026)
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
by: Oulkadda, Ilyas, et al.
Published: (2025)
by: Oulkadda, Ilyas, et al.
Published: (2025)
DsDm: Model-Aware Dataset Selection with Datamodels
by: Engstrom, Logan, et al.
Published: (2024)
by: Engstrom, Logan, et al.
Published: (2024)
Small-to-Large Generalization: Data Influences Models Consistently Across Scale
by: Khaddaj, Alaa, et al.
Published: (2025)
by: Khaddaj, Alaa, et al.
Published: (2025)
NeMo: A Neuron-Level Modularizing-While-Training Approach for Decomposing DNN Models
by: Bi, Xiaohan, et al.
Published: (2025)
by: Bi, Xiaohan, et al.
Published: (2025)
Multiplicative Orthogonal Sequential Editing for Language Models
by: Xu, Hao-Xiang, et al.
Published: (2026)
by: Xu, Hao-Xiang, et al.
Published: (2026)
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
by: Shah, Sanket, et al.
Published: (2023)
by: Shah, Sanket, et al.
Published: (2023)
Counterfactual Training: Teaching Models Plausible and Actionable Explanations
by: Altmeyer, Patrick, et al.
Published: (2026)
by: Altmeyer, Patrick, et al.
Published: (2026)
Efficiently Editing Mixture-of-Experts Models with Compressed Experts
by: He, Yifei, et al.
Published: (2025)
by: He, Yifei, et al.
Published: (2025)
Constraining Sequential Model Editing with Editing Anchor Compression
by: Xu, Hao-Xiang, et al.
Published: (2025)
by: Xu, Hao-Xiang, et al.
Published: (2025)
Predicting Compact Phrasal Rewrites with Large Language Models for ASR Post Editing
by: Zhang, Hao, et al.
Published: (2025)
by: Zhang, Hao, et al.
Published: (2025)
SDQ: Sparse Decomposed Quantization for LLM Inference
by: Jeong, Geonhwa, et al.
Published: (2024)
by: Jeong, Geonhwa, et al.
Published: (2024)
Decomposing Epistemic Uncertainty for Causal Decision Making
by: Rahman, Md Musfiqur, et al.
Published: (2026)
by: Rahman, Md Musfiqur, et al.
Published: (2026)
Data Debiasing with Datamodels (D3M): Improving Subgroup Robustness via Data Selection
by: Jain, Saachi, et al.
Published: (2024)
by: Jain, Saachi, et al.
Published: (2024)
Large-Scale, Longitudinal Study of Large Language Models During the 2024 US Election Season
by: Cen, Sarah H., et al.
Published: (2025)
by: Cen, Sarah H., et al.
Published: (2025)
K-Edit: Language Model Editing with Contextual Knowledge Awareness
by: Markowitz, Elan, et al.
Published: (2025)
by: Markowitz, Elan, et al.
Published: (2025)
Too Long, Didn't Model: Decomposing LLM Long-Context Understanding With Novels
by: Hamilton, Sil, et al.
Published: (2025)
by: Hamilton, Sil, et al.
Published: (2025)
Model Merging for Knowledge Editing
by: Fu, Zichuan, et al.
Published: (2025)
by: Fu, Zichuan, et al.
Published: (2025)
ConformaDecompose: Explaining Uncertainty via Calibration Localization
by: Yapicioglu, Fatima Rabia, et al.
Published: (2026)
by: Yapicioglu, Fatima Rabia, et al.
Published: (2026)
Research on Disease Prediction Model Construction Based on Computer AI deep Learning Technology
by: Lin, Yang, et al.
Published: (2024)
by: Lin, Yang, et al.
Published: (2024)
Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models
by: Samadi, Mehrzad, et al.
Published: (2025)
by: Samadi, Mehrzad, et al.
Published: (2025)
FuncGenFoil: Airfoil Generation and Editing Model in Function Space
by: Zhang, Jinouwen, et al.
Published: (2025)
by: Zhang, Jinouwen, et al.
Published: (2025)
DataMIL: Selecting Data for Robot Imitation Learning with Datamodels
by: Dass, Shivin, et al.
Published: (2025)
by: Dass, Shivin, et al.
Published: (2025)
How Robust is Model Editing after Fine-Tuning? An Empirical Study on Text-to-Image Diffusion Models
by: He, Feng, et al.
Published: (2025)
by: He, Feng, et al.
Published: (2025)
TimeMixer: Decomposable Multiscale Mixing for Time Series Forecasting
by: Wang, Shiyu, et al.
Published: (2024)
by: Wang, Shiyu, et al.
Published: (2024)
Purifying Shampoo: Investigating Shampoo's Heuristics by Decomposing its Preconditioner
by: Eschenhagen, Runa, et al.
Published: (2025)
by: Eschenhagen, Runa, et al.
Published: (2025)
A Unified Framework for Model Editing
by: Gupta, Akshat, et al.
Published: (2024)
by: Gupta, Akshat, et al.
Published: (2024)
Model Editing by Standard Fine-Tuning
by: Gangadhar, Govind, et al.
Published: (2024)
by: Gangadhar, Govind, et al.
Published: (2024)
Generalizable Multimodal Large Language Model Editing via Invariant Trajectory Learning
by: Su, Jiajie, et al.
Published: (2026)
by: Su, Jiajie, et al.
Published: (2026)
Latent Knowledge Scalpel: Precise and Massive Knowledge Editing for Large Language Models
by: Liu, Xin, et al.
Published: (2025)
by: Liu, Xin, et al.
Published: (2025)
Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport
by: Shah, Shaan, et al.
Published: (2025)
by: Shah, Shaan, et al.
Published: (2025)
Decomposing and Measuring Evaluation Awareness
by: Li, Changling, et al.
Published: (2026)
by: Li, Changling, et al.
Published: (2026)
LaMDA: Large Model Fine-Tuning via Spectrally Decomposed Low-Dimensional Adaptation
by: Azizi, Seyedarmin, et al.
Published: (2024)
by: Azizi, Seyedarmin, et al.
Published: (2024)
Decomposing Behavioral Phase Transitions in LLMs: Order Parameters for Emergent Misalignment
by: Arnold, Julian, et al.
Published: (2025)
by: Arnold, Julian, et al.
Published: (2025)
AI Supply Chains: An Emerging Ecosystem of AI Actors, Products, and Services
by: Hopkins, Aspen, et al.
Published: (2025)
by: Hopkins, Aspen, et al.
Published: (2025)
Similar Items
-
ContextCite: Attributing Model Generation to Context
by: Cohen-Wang, Benjamin, et al.
Published: (2024) -
Optimizing ML Training with Metagradient Descent
by: Engstrom, Logan, et al.
Published: (2025) -
Parameters vs FLOPs: Scaling Laws for Optimal Sparsity for Mixture-of-Experts Language Models
by: Abnar, Samira, et al.
Published: (2025) -
User Strategization and Trustworthy Algorithms
by: Cen, Sarah H., et al.
Published: (2023) -
Why Inference in Large Models Becomes Decomposable After Training
by: Jin, Jidong
Published: (2026)