Half-order Fine-Tuning for Diffusion Model: A Recursive Likelihood Ratio Optimizer
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Tao, Zhang, Zishi, Jiang, Jingyang, Li, Zehao, Qin, Shentao, Zheng, Yi, Li, Guanghao, Sun, Qianyou, Li, Yan, Liang, Jiafeng, Li, Xinping, Peng, Yijie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal low-rank stochastic gradient estimation for LLM training
by: Li, Zehao, et al.
Published: (2026)
by: Li, Zehao, et al.
Published: (2026)
Nonparametric Bayesian Optimization for General Rewards
by: Zhang, Zishi, et al.
Published: (2026)
by: Zhang, Zishi, et al.
Published: (2026)
RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training
by: Ren, Tao, et al.
Published: (2025)
by: Ren, Tao, et al.
Published: (2025)
FLOPS: Forward Learning with OPtimal Sampling
by: Ren, Tao, et al.
Published: (2024)
by: Ren, Tao, et al.
Published: (2024)
Sample-Efficient "Clustering and Conquer" Procedures for Parallel Large-Scale Ranking and Selection
by: Zhang, Zishi, et al.
Published: (2024)
by: Zhang, Zishi, et al.
Published: (2024)
A New Stochastic Approximation Method for Gradient-based Simulated Parameter Estimation
by: Li, Zehao, et al.
Published: (2025)
by: Li, Zehao, et al.
Published: (2025)
LLM-Inspired Pretrain-Then-Finetune for Small-Data, Large-Scale Optimization
by: Zhang, Zishi, et al.
Published: (2026)
by: Zhang, Zishi, et al.
Published: (2026)
Dual-Agent Deep Reinforcement Learning for Dynamic Pricing and Replenishment
by: Zheng, Yi, et al.
Published: (2024)
by: Zheng, Yi, et al.
Published: (2024)
Fine-Tuning Flow Matching via Maximum Likelihood Estimation of Reconstructions
by: Li, Zhaoyi, et al.
Published: (2025)
by: Li, Zhaoyi, et al.
Published: (2025)
Beyond likelihood ratio bias: Nested multi-time-scale stochastic approximation for likelihood-free parameter estimation
by: Li, Zehao, et al.
Published: (2024)
by: Li, Zehao, et al.
Published: (2024)
Cold-Start Forecasting of New Product Life-Cycles via Conditional Diffusion Models
by: Zhou, Ruihan, et al.
Published: (2026)
by: Zhou, Ruihan, et al.
Published: (2026)
Task-Centric Policy Optimization from Misaligned Motion Priors
by: Zheng, Ziang, et al.
Published: (2026)
by: Zheng, Ziang, et al.
Published: (2026)
Approximated Likelihood Ratio: A Forward-Only and Parallel Framework for Boosting Neural Network Training
by: Zhang, Zeliang, et al.
Published: (2024)
by: Zhang, Zeliang, et al.
Published: (2024)
InterCorpRel-LLM: Enhancing Financial Relational Understanding with Graph-Language Models
by: Sun, Qianyou, et al.
Published: (2025)
by: Sun, Qianyou, et al.
Published: (2025)
RiskMiner: Discovering Formulaic Alphas via Risk Seeking Monte Carlo Tree Search
by: Ren, Tao, et al.
Published: (2024)
by: Ren, Tao, et al.
Published: (2024)
On Group Relative Policy Optimization Collapse in Agent Search: The Lazy Likelihood-Displacement
by: Deng, Wenlong, et al.
Published: (2025)
by: Deng, Wenlong, et al.
Published: (2025)
A Novel Adaptive Fine-Tuning Algorithm for Multimodal Models: Self-Optimizing Classification and Selection of High-Quality Datasets in Remote Sensing
by: Ren, Yi, et al.
Published: (2024)
by: Ren, Yi, et al.
Published: (2024)
PIRF: Physics-Informed Reward Fine-Tuning for Diffusion Models
by: Yuan, Mingze, et al.
Published: (2025)
by: Yuan, Mingze, et al.
Published: (2025)
Differentially Private Fine-Tuning of Diffusion Models
by: Tsai, Yu-Lin, et al.
Published: (2024)
by: Tsai, Yu-Lin, et al.
Published: (2024)
Improved Asymptotic Formulae for Statistical Interpretation Based on Likelihood Ratio Tests
by: Xia, Li-Gang, et al.
Published: (2021)
by: Xia, Li-Gang, et al.
Published: (2021)
DexEMG: Towards Dexterous Teleoperation System via EMG2Pose Generalization
by: Zhao, Qianyou, et al.
Published: (2026)
by: Zhao, Qianyou, et al.
Published: (2026)
Banach Hardy-Sobolev Spaces on the Upper Half-plane and Operator Theory
by: Liang, Haoxian, et al.
Published: (2026)
by: Liang, Haoxian, et al.
Published: (2026)
A Dense Reward View on Aligning Text-to-Image Diffusion with Preference
by: Yang, Shentao, et al.
Published: (2024)
by: Yang, Shentao, et al.
Published: (2024)
FinePOSE: Fine-Grained Prompt-Driven 3D Human Pose Estimation via Diffusion Models
by: Xu, Jinglin, et al.
Published: (2024)
by: Xu, Jinglin, et al.
Published: (2024)
Hypertensive Uveitis After Intravitreal Faricimab: Clinical and Methodological Considerations
by: Guanghao Qin
Published: (2026)
by: Guanghao Qin
Published: (2026)
HFT: Half Fine-Tuning for Large Language Models
by: Hui, Tingfeng, et al.
Published: (2024)
by: Hui, Tingfeng, et al.
Published: (2024)
Generalized Averaging Method for Power Electronics Modeling from DC to above Half the Switching Frequency
by: Li, Hongchang, et al.
Published: (2024)
by: Li, Hongchang, et al.
Published: (2024)
Gradual Learning: Optimizing Fine-Tuning with Partially Mastered Knowledge in Large Language Models
by: Li, Bozhou, et al.
Published: (2024)
by: Li, Bozhou, et al.
Published: (2024)
Lookahead Path Likelihood Optimization for Diffusion LLMs
by: Liu, Xuejie, et al.
Published: (2026)
by: Liu, Xuejie, et al.
Published: (2026)
Fine-Tuning Code Language Models to Detect Cross-Language Bugs
by: Li, Zengyang, et al.
Published: (2025)
by: Li, Zengyang, et al.
Published: (2025)
Fine-Tuning Diffusion-Based Recommender Systems via Reinforcement Learning with Reward Function Optimization
by: Hou, Yu, et al.
Published: (2025)
by: Hou, Yu, et al.
Published: (2025)
ACD: Direct Conditional Control for Video Diffusion Models via Attention Supervision
by: Li, Weiqi, et al.
Published: (2025)
by: Li, Weiqi, et al.
Published: (2025)
Distribution Analysis of Pre-Qin Copper Smelting Sites in Key Regions of China Using Maximum Entropy Modelling
by: Li, Zehao
Published: (2025)
by: Li, Zehao
Published: (2025)
Ratio-Variance Regularized Policy Optimization for Efficient LLM Fine-tuning
by: Luo, Yu, et al.
Published: (2026)
by: Luo, Yu, et al.
Published: (2026)
LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs
by: Wang, Xiaodong, et al.
Published: (2025)
by: Wang, Xiaodong, et al.
Published: (2025)
MUC-G4: Minimal Unsat Core-Guided Incremental Verification for Deep Neural Network Compression
by: Li, Jingyang, et al.
Published: (2025)
by: Li, Jingyang, et al.
Published: (2025)
ASFT: Aligned Supervised Fine-Tuning through Absolute Likelihood
by: Wang, Ruoyu, et al.
Published: (2024)
by: Wang, Ruoyu, et al.
Published: (2024)
Gradient Surgery for Safe LLM Fine-Tuning
by: Yi, Biao, et al.
Published: (2025)
by: Yi, Biao, et al.
Published: (2025)
Energy-Efficient Split Learning for Fine-Tuning Large Language Models in Edge Networks
by: Li, Zuguang, et al.
Published: (2024)
by: Li, Zuguang, et al.
Published: (2024)
Implicit Federated In-context Learning For Task-Specific LLM Fine-Tuning
by: Li, Dongcheng, et al.
Published: (2025)
by: Li, Dongcheng, et al.
Published: (2025)
Similar Items
-
Optimal low-rank stochastic gradient estimation for LLM training
by: Li, Zehao, et al.
Published: (2026) -
Nonparametric Bayesian Optimization for General Rewards
by: Zhang, Zishi, et al.
Published: (2026) -
RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training
by: Ren, Tao, et al.
Published: (2025) -
FLOPS: Forward Learning with OPtimal Sampling
by: Ren, Tao, et al.
Published: (2024) -
Sample-Efficient "Clustering and Conquer" Procedures for Parallel Large-Scale Ranking and Selection
by: Zhang, Zishi, et al.
Published: (2024)