Iterative Finetuning is Mostly Idempotent
Fuente:
arXiv
Saved in:
| Main Authors: | Roe, Zephaniah, Sanderson, Jack, Nguyen, Dang, Huang, Julian, Nief, Todd, Shrivastava, Aryan, Tan, Chenhao, Holtzman, Ari |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Linearly Decoding Refused Knowledge in Aligned Language Models
by: Shrivastava, Aryan, et al.
Published: (2025)
by: Shrivastava, Aryan, et al.
Published: (2025)
Know Thyself? On the Incapability and Implications of AI Self-Recognition
by: Bai, Xiaoyan, et al.
Published: (2025)
by: Bai, Xiaoyan, et al.
Published: (2025)
The Text Uncanny Valley: Non-Monotonic Performance Degradation in LLM Information Retrieval
by: Tong, Zekai, et al.
Published: (2026)
by: Tong, Zekai, et al.
Published: (2026)
Subliminal Learning is a LoRA Artifact
by: Nief, Todd, et al.
Published: (2026)
by: Nief, Todd, et al.
Published: (2026)
Dynamic Weight Grafting: Localizing Finetuned Factual Knowledge in Transformers
by: Nief, Todd, et al.
Published: (2025)
by: Nief, Todd, et al.
Published: (2025)
Moral Mazes in the Era of LLMs
by: Nguyen, Dang, et al.
Published: (2026)
by: Nguyen, Dang, et al.
Published: (2026)
The Story is Not the Science: Execution-Grounded Evaluation of Mechanistic Interpretability Research
by: Bai, Xiaoyan, et al.
Published: (2026)
by: Bai, Xiaoyan, et al.
Published: (2026)
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
by: Nguyen, Dang, et al.
Published: (2025)
by: Nguyen, Dang, et al.
Published: (2025)
Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing
by: Holtzman, Ari, et al.
Published: (2026)
by: Holtzman, Ari, et al.
Published: (2026)
AI as Entertainment
by: Kommers, Cody, et al.
Published: (2026)
by: Kommers, Cody, et al.
Published: (2026)
Prompting as Scientific Inquiry
by: Holtzman, Ari, et al.
Published: (2025)
by: Holtzman, Ari, et al.
Published: (2025)
RATE: Causal Explainability of Reward Models with Imperfect Counterfactuals
by: Reber, David, et al.
Published: (2024)
by: Reber, David, et al.
Published: (2024)
The Information Geometry of Softmax: Probing and Steering
by: Park, Kiho, et al.
Published: (2026)
by: Park, Kiho, et al.
Published: (2026)
AbsenceBench: Language Models Can't Tell What's Missing
by: Fu, Harvey Yiyun, et al.
Published: (2025)
by: Fu, Harvey Yiyun, et al.
Published: (2025)
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
by: Yang, Chenghao, et al.
Published: (2025)
by: Yang, Chenghao, et al.
Published: (2025)
Measuring Free-Form Decision-Making Inconsistency of Language Models in Military Crisis Simulations
by: Shrivastava, Aryan, et al.
Published: (2024)
by: Shrivastava, Aryan, et al.
Published: (2024)
Forking Paths in Neural Text Generation
by: Bigelow, Eric, et al.
Published: (2024)
by: Bigelow, Eric, et al.
Published: (2024)
GPT-4V Cannot Generate Radiology Reports Yet
by: Jiang, Yuyang, et al.
Published: (2024)
by: Jiang, Yuyang, et al.
Published: (2024)
Predicting vs. Acting: A Trade-off Between World Modeling & Agent Modeling
by: Li, Margaret, et al.
Published: (2024)
by: Li, Margaret, et al.
Published: (2024)
Score-based Idempotent Distillation of Diffusion Models
by: Zaman, Shehtab, et al.
Published: (2025)
by: Zaman, Shehtab, et al.
Published: (2025)
Mapping Overlaps in Benchmarks through Perplexity in the Wild
by: Wu, Siyang, et al.
Published: (2025)
by: Wu, Siyang, et al.
Published: (2025)
Conditional Idempotent Generative Networks
by: Ronchetti, Niccolò
Published: (2024)
by: Ronchetti, Niccolò
Published: (2024)
Actor-Critic based Online Data Mixing For Language Model Pre-Training
by: Ma, Jing, et al.
Published: (2025)
by: Ma, Jing, et al.
Published: (2025)
LL3M: Large Language 3D Modelers
by: Lu, Sining, et al.
Published: (2025)
by: Lu, Sining, et al.
Published: (2025)
Deep Model Merging: The Sister of Neural Network Interpretability -- A Survey
by: Khan, Arham, et al.
Published: (2024)
by: Khan, Arham, et al.
Published: (2024)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
by: Teo, Rachel S. Y., et al.
Published: (2025)
by: Teo, Rachel S. Y., et al.
Published: (2025)
CARE-RFT: Confidence-Anchored Reinforcement Finetuning for Reliable Reasoning in Large Language Models
by: Li, Shuozhe, et al.
Published: (2026)
by: Li, Shuozhe, et al.
Published: (2026)
Automatically Generating Hard Math Problems from Hypothesis-Driven Error Analysis
by: Fu, Jiayu, et al.
Published: (2026)
by: Fu, Jiayu, et al.
Published: (2026)
Idempotent Unsupervised Representation Learning for Skeleton-Based Action Recognition
by: Lin, Lilang, et al.
Published: (2024)
by: Lin, Lilang, et al.
Published: (2024)
On The Finetuning of MLIPs Through the Lens of Iterated Maps With BPTT
by: Dramko, Evan, et al.
Published: (2025)
by: Dramko, Evan, et al.
Published: (2025)
Generative AI and Agency in Education: A Critical Scoping Review and Thematic Analysis
by: Roe, Jasper, et al.
Published: (2024)
by: Roe, Jasper, et al.
Published: (2024)
Deepfakes and Higher Education: A Research Agenda and Scoping Review of Synthetic Media
by: Roe, Jasper, et al.
Published: (2024)
by: Roe, Jasper, et al.
Published: (2024)
TCIA: A Task-Centric Instruction Augmentation Method for Instruction Finetuning
by: Ma, Simin, et al.
Published: (2025)
by: Ma, Simin, et al.
Published: (2025)
Generative Kaleidoscopic Networks
by: Shrivastava, Harsh
Published: (2024)
by: Shrivastava, Harsh
Published: (2024)
Narrow Finetuning Leaves Clearly Readable Traces in Activation Differences
by: Minder, Julian, et al.
Published: (2025)
by: Minder, Julian, et al.
Published: (2025)
Finetuning LLMs for Automatic Form Interaction on Web-Browser in Selenium Testing Framework
by: Le, Nguyen-Khang, et al.
Published: (2025)
by: Le, Nguyen-Khang, et al.
Published: (2025)
Zero Data Retention in LLM-based Enterprise AI Assistants: A Comparative Study of Market Leading Agentic AI Products
by: Gupta, Komal, et al.
Published: (2025)
by: Gupta, Komal, et al.
Published: (2025)
One Leak Away: How Pretrained Model Exposure Amplifies Jailbreak Risks in Finetuned LLMs
by: Tan, Yixin, et al.
Published: (2025)
by: Tan, Yixin, et al.
Published: (2025)
Generative AI Tools in Academic Research: Applications and Implications for Qualitative and Quantitative Research Methodologies
by: Perkins, Mike, et al.
Published: (2024)
by: Perkins, Mike, et al.
Published: (2024)
Harnessing the Power of Beta Scoring in Deep Active Learning for Multi-Label Text Classification
by: Tan, Wei, et al.
Published: (2024)
by: Tan, Wei, et al.
Published: (2024)
Similar Items
-
Linearly Decoding Refused Knowledge in Aligned Language Models
by: Shrivastava, Aryan, et al.
Published: (2025) -
Know Thyself? On the Incapability and Implications of AI Self-Recognition
by: Bai, Xiaoyan, et al.
Published: (2025) -
The Text Uncanny Valley: Non-Monotonic Performance Degradation in LLM Information Retrieval
by: Tong, Zekai, et al.
Published: (2026) -
Subliminal Learning is a LoRA Artifact
by: Nief, Todd, et al.
Published: (2026) -
Dynamic Weight Grafting: Localizing Finetuned Factual Knowledge in Transformers
by: Nief, Todd, et al.
Published: (2025)