Saved in:
| Main Authors: | Fan, Ying, Lee, Kangwook |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2301.13362 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Memorization Capacity for Additive Fine-Tuning with Small ReLU Networks
by: Sohn, Jy-yong, et al.
Published: (2024)
by: Sohn, Jy-yong, et al.
Published: (2024)
Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models
by: Lee, Chungpa, et al.
Published: (2026)
by: Lee, Chungpa, et al.
Published: (2026)
Parameter-Efficient Fine-Tuning of State Space Models
by: Galim, Kevin, et al.
Published: (2024)
by: Galim, Kevin, et al.
Published: (2024)
Looped Transformers for Length Generalization
by: Fan, Ying, et al.
Published: (2024)
by: Fan, Ying, et al.
Published: (2024)
Dual Operating Modes of In-Context Learning
by: Lin, Ziqian, et al.
Published: (2024)
by: Lin, Ziqian, et al.
Published: (2024)
Tilt Matching for Scalable Sampling and Fine-Tuning
by: Potaptchik, Peter, et al.
Published: (2025)
by: Potaptchik, Peter, et al.
Published: (2025)
The Expressive Power of Low-Rank Adaptation
by: Zeng, Yuchen, et al.
Published: (2023)
by: Zeng, Yuchen, et al.
Published: (2023)
Shortcut Mitigation via Spurious-Positive Samples
by: Le, Phuong Quynh, et al.
Published: (2026)
by: Le, Phuong Quynh, et al.
Published: (2026)
In-Context Learning with Hypothesis-Class Guidance
by: Lin, Ziqian, et al.
Published: (2025)
by: Lin, Ziqian, et al.
Published: (2025)
TabDDPM: Modelling Tabular Data with Diffusion Models
by: Kotelnikov, Akim, et al.
Published: (2022)
by: Kotelnikov, Akim, et al.
Published: (2022)
Optimal Convergence Analysis of DDPM for General Distributions
by: Jiao, Yuchen, et al.
Published: (2025)
by: Jiao, Yuchen, et al.
Published: (2025)
DDPM Score Matching and Distribution Learning
by: Chewi, Sinho, et al.
Published: (2025)
by: Chewi, Sinho, et al.
Published: (2025)
Low-Rank Curvature for Zeroth-Order Optimization in LLM Fine-Tuning
by: Seung, Hyunseok, et al.
Published: (2025)
by: Seung, Hyunseok, et al.
Published: (2025)
Transformers Learn the Optimal DDPM Denoiser for Multi-Token GMMs
by: Li, Hongkang, et al.
Published: (2026)
by: Li, Hongkang, et al.
Published: (2026)
Zero-Order Optimization for LLM Fine-Tuning via Learnable Direction Sampling
by: Parfenov, Valery, et al.
Published: (2026)
by: Parfenov, Valery, et al.
Published: (2026)
Thompson Sampling via Fine-Tuning of LLMs
by: Menet, Nicolas, et al.
Published: (2025)
by: Menet, Nicolas, et al.
Published: (2025)
Variation Spaces for Multi-Output Neural Networks: Insights on Multi-Task Learning and Network Compression
by: Shenouda, Joseph, et al.
Published: (2023)
by: Shenouda, Joseph, et al.
Published: (2023)
An Edit Friendly DDPM Noise Space: Inversion and Manipulations
by: Huberman-Spiegelglas, Inbar, et al.
Published: (2023)
by: Huberman-Spiegelglas, Inbar, et al.
Published: (2023)
Models Know Their Shortcuts: Deployment-Time Shortcut Mitigation
by: Li, Jiayi, et al.
Published: (2026)
by: Li, Jiayi, et al.
Published: (2026)
Looped Transformers are Better at Learning Learning Algorithms
by: Yang, Liu, et al.
Published: (2023)
by: Yang, Liu, et al.
Published: (2023)
Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding
by: Yang, Seongjun, et al.
Published: (2023)
by: Yang, Seongjun, et al.
Published: (2023)
ShortcutProbe: Probing Prediction Shortcuts for Learning Robust Models
by: Zheng, Guangtao, et al.
Published: (2025)
by: Zheng, Guangtao, et al.
Published: (2025)
Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges
by: Lee, Nayoung, et al.
Published: (2025)
by: Lee, Nayoung, et al.
Published: (2025)
Task Vectors in In-Context Learning: Emergence, Formation, and Benefit
by: Yang, Liu, et al.
Published: (2025)
by: Yang, Liu, et al.
Published: (2025)
Upweighting Easy Samples in Fine-Tuning Mitigates Forgetting
by: Sanyal, Sunny, et al.
Published: (2025)
by: Sanyal, Sunny, et al.
Published: (2025)
An Optimization Framework for Differentially Private Sparse Fine-Tuning
by: Makni, Mehdi, et al.
Published: (2025)
by: Makni, Mehdi, et al.
Published: (2025)
Prior-Informed Zeroth-Order Optimization with Adaptive Direction Alignment for Memory-Efficient LLM Fine-Tuning
by: Jin, Feihu, et al.
Published: (2026)
by: Jin, Feihu, et al.
Published: (2026)
TransConv-DDPM: Enhanced Diffusion Model for Generating Time-Series Data in Healthcare
by: Kabir, Md Shahriar, et al.
Published: (2026)
by: Kabir, Md Shahriar, et al.
Published: (2026)
Infected Smallville: How Disease Threat Shapes Sociality in LLM Agents
by: Choi, Soyeon, et al.
Published: (2025)
by: Choi, Soyeon, et al.
Published: (2025)
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data
by: Xiong, Zheyang, et al.
Published: (2024)
by: Xiong, Zheyang, et al.
Published: (2024)
Sparse Logit Sampling: Accelerating Knowledge Distillation in LLMs
by: Anshumann, et al.
Published: (2025)
by: Anshumann, et al.
Published: (2025)
SEMRes-DDPM: Residual Network Based Diffusion Modelling Applied to Imbalanced Data
by: Zheng, Ming, et al.
Published: (2024)
by: Zheng, Ming, et al.
Published: (2024)
Learning a Zeroth-Order Optimizer for Fine-Tuning LLMs
by: Zhang, Kairun, et al.
Published: (2025)
by: Zhang, Kairun, et al.
Published: (2025)
Theoretical Insights into Fine-Tuning Attention Mechanism: Generalization and Optimization
by: Yao, Xinhao, et al.
Published: (2024)
by: Yao, Xinhao, et al.
Published: (2024)
Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling
by: Huang, Zeyu, et al.
Published: (2025)
by: Huang, Zeyu, et al.
Published: (2025)
Polyp-DDPM: Diffusion-Based Semantic Polyp Synthesis for Enhanced Segmentation
by: Dorjsembe, Zolnamar, et al.
Published: (2024)
by: Dorjsembe, Zolnamar, et al.
Published: (2024)
Short-circuiting Shortcuts: Mechanistic Investigation of Shortcuts in Text Classification
by: Eshuijs, Leon, et al.
Published: (2025)
by: Eshuijs, Leon, et al.
Published: (2025)
ENTP: Encoder-only Next Token Prediction
by: Ewer, Ethan, et al.
Published: (2024)
by: Ewer, Ethan, et al.
Published: (2024)
How to Correctly Report LLM-as-a-Judge Evaluations
by: Lee, Chungpa, et al.
Published: (2025)
by: Lee, Chungpa, et al.
Published: (2025)
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
by: Yi, Hongzhu, et al.
Published: (2026)
by: Yi, Hongzhu, et al.
Published: (2026)
Similar Items
-
Memorization Capacity for Additive Fine-Tuning with Small ReLU Networks
by: Sohn, Jy-yong, et al.
Published: (2024) -
Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models
by: Lee, Chungpa, et al.
Published: (2026) -
Parameter-Efficient Fine-Tuning of State Space Models
by: Galim, Kevin, et al.
Published: (2024) -
Looped Transformers for Length Generalization
by: Fan, Ying, et al.
Published: (2024) -
Dual Operating Modes of In-Context Learning
by: Lin, Ziqian, et al.
Published: (2024)