Few-Shot Test-Time Optimization Without Retraining for Semiconductor Recipe Generation and Beyond
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Shangding, Ying, Donghao, Jin, Ming, Lu, Yu Joe, Wang, Jun, Lavaei, Javad, Spanos, Costas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Don't Trade Off Safety: Diffusion Regularization for Constrained Offline RL
by: Guo, Junyu, et al.
Published: (2025)
by: Guo, Junyu, et al.
Published: (2025)
StyleBench: Evaluating thinking styles in Large Language Models
by: Guo, Junyu, et al.
Published: (2025)
by: Guo, Junyu, et al.
Published: (2025)
LLMs Should Express Uncertainty Explicitly
by: Guo, Junyu, et al.
Published: (2026)
by: Guo, Junyu, et al.
Published: (2026)
AccidentBench: Benchmarking Multimodal Understanding and Reasoning in Vehicle Accidents and Beyond
by: Gu, Shangding, et al.
Published: (2025)
by: Gu, Shangding, et al.
Published: (2025)
Enhancing Efficiency of Safe Reinforcement Learning via Sample Manipulation
by: Gu, Shangding, et al.
Published: (2024)
by: Gu, Shangding, et al.
Published: (2024)
Robust Gymnasium: A Unified Modular Benchmark for Robust Reinforcement Learning
by: Gu, Shangding, et al.
Published: (2025)
by: Gu, Shangding, et al.
Published: (2025)
TeaMs-RL: Teaching LLMs to Generate Better Instruction Datasets via Reinforcement Learning
by: Gu, Shangding, et al.
Published: (2024)
by: Gu, Shangding, et al.
Published: (2024)
Data Uniformity Improves Training Efficiency and More, with a Convergence Framework Beyond the NTK Regime
by: Wang, Yuqing, et al.
Published: (2025)
by: Wang, Yuqing, et al.
Published: (2025)
Safe Continual Domain Adaptation after Sim2Real Transfer of Reinforcement Learning Policies in Robotics
by: Josifovski, Josip, et al.
Published: (2025)
by: Josifovski, Josip, et al.
Published: (2025)
Long Context, Less Focus: A Scaling Gap in LLMs Revealed through Privacy and Personalization
by: Gu, Shangding
Published: (2026)
by: Gu, Shangding
Published: (2026)
Mutual Enhancement of Large Language and Reinforcement Learning Models through Bi-Directional Feedback Mechanisms: A Planning Case Study
by: Gu, Shangding
Published: (2024)
by: Gu, Shangding
Published: (2024)
Safe Multi-Agent Reinforcement Learning with Bilevel Optimization in Autonomous Driving
by: Zheng, Zhi, et al.
Published: (2024)
by: Zheng, Zhi, et al.
Published: (2024)
Policy-based Primal-Dual Methods for Concave CMDP with Variance Reduction
by: Ying, Donghao, et al.
Published: (2022)
by: Ying, Donghao, et al.
Published: (2022)
Pausing Policy Learning in Non-stationary Reinforcement Learning
by: Lee, Hyunin, et al.
Published: (2024)
by: Lee, Hyunin, et al.
Published: (2024)
Huber-based Robust System Identification with Near-Optimal Guarantees Across Independent and Adversarial Regimes
by: Kim, Jihun, et al.
Published: (2026)
by: Kim, Jihun, et al.
Published: (2026)
System Identification from Partial Observations under Adversarial Attacks
by: Kim, Jihun, et al.
Published: (2025)
by: Kim, Jihun, et al.
Published: (2025)
Bridging Batch and Streaming Estimations to System Identification under Adversarial Attacks
by: Kim, Jihun, et al.
Published: (2025)
by: Kim, Jihun, et al.
Published: (2025)
On the Necessity of Two-Stage Estimation for Learning Dynamical Systems under Both Noise and Node-Wise Attacks
by: Kim, Jihun, et al.
Published: (2026)
by: Kim, Jihun, et al.
Published: (2026)
Prevailing against Adversarial Noncentral Disturbances: Exact Recovery of Linear Systems with the $l_1$-norm Estimator
by: Kim, Jihun, et al.
Published: (2024)
by: Kim, Jihun, et al.
Published: (2024)
Online Bandit Nonlinear Control with Dynamic Batch Length and Adaptive Learning Rate
by: Kim, Jihun, et al.
Published: (2024)
by: Kim, Jihun, et al.
Published: (2024)
Explainable Sentiment Analysis with DeepSeek-R1: Performance, Efficiency, and Few-Shot Learning
by: Huang, Donghao, et al.
Published: (2025)
by: Huang, Donghao, et al.
Published: (2025)
Balance Reward and Safety Optimization for Safe Reinforcement Learning: A Perspective of Gradient Manipulation
by: Gu, Shangding, et al.
Published: (2024)
by: Gu, Shangding, et al.
Published: (2024)
Beyond Exact Gradients: Convergence of Stochastic Soft-Max Policy Gradient Methods with Entropy Regularization
by: Ding, Yuhao, et al.
Published: (2021)
by: Ding, Yuhao, et al.
Published: (2021)
RLBenchNet: The Right Network for the Right Reinforcement Learning Task
by: Smirnov, Ivan, et al.
Published: (2025)
by: Smirnov, Ivan, et al.
Published: (2025)
A CMDP-within-online framework for Meta-Safe Reinforcement Learning
by: Khattar, Vanshaj, et al.
Published: (2024)
by: Khattar, Vanshaj, et al.
Published: (2024)
Adapting In-Domain Few-Shot Segmentation to New Domains without Source Domain Retraining
by: Fan, Qi, et al.
Published: (2025)
by: Fan, Qi, et al.
Published: (2025)
Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models
by: Parmar, Jupinder, et al.
Published: (2024)
by: Parmar, Jupinder, et al.
Published: (2024)
Structural Correspondence and Universal Approximation in Diagonal plus Low-Rank Neural Networks
by: Chen, Ying, et al.
Published: (2026)
by: Chen, Ying, et al.
Published: (2026)
Absence of spurious solutions far from ground truth: A low-rank analysis with high-order losses
by: Ma, Ziye, et al.
Published: (2024)
by: Ma, Ziye, et al.
Published: (2024)
The landscape of deterministic and stochastic optimal control problems: One-shot Optimization versus Dynamic Programming
by: Kim, Jihun, et al.
Published: (2024)
by: Kim, Jihun, et al.
Published: (2024)
Learning to Adapt Frozen CLIP for Few-Shot Test-Time Domain Adaptation
by: Chi, Zhixiang, et al.
Published: (2025)
by: Chi, Zhixiang, et al.
Published: (2025)
Sustainable Machine Learning Retraining: Optimizing Energy Efficiency Without Compromising Accuracy
by: Poenaru-Olaru, Lorena, et al.
Published: (2025)
by: Poenaru-Olaru, Lorena, et al.
Published: (2025)
Beyond Normal References: Discriminative Few-Shot Anomaly Detection
by: Wang, Huan, et al.
Published: (2026)
by: Wang, Huan, et al.
Published: (2026)
On the Sharp Input-Output Analysis of Nonlinear Systems under Adversarial Attacks
by: Kim, Jihun, et al.
Published: (2025)
by: Kim, Jihun, et al.
Published: (2025)
Why is Normalization Preferred? A Worst-Case Complexity Theory for Stochastically Preconditioned SGD under Heavy-Tailed Noise
by: Fang, Yuchen, et al.
Published: (2026)
by: Fang, Yuchen, et al.
Published: (2026)
High Probability Complexity Bounds of Trust-Region Stochastic Sequential Quadratic Programming with Heavy-Tailed Noise
by: Fang, Yuchen, et al.
Published: (2025)
by: Fang, Yuchen, et al.
Published: (2025)
Subgradient Method for System Identification with Non-Smooth Objectives
by: Yalcin, Baturalp, et al.
Published: (2025)
by: Yalcin, Baturalp, et al.
Published: (2025)
TRSVR: An Adaptive Stochastic Trust-Region Method with Variance Reduction
by: Fang, Yuchen, et al.
Published: (2026)
by: Fang, Yuchen, et al.
Published: (2026)
All-Pass Fractional OPF: A Solver-Friendly, Physics-Preserving Approximation of AC OPF
by: Hasanzadeh, Milad, et al.
Published: (2026)
by: Hasanzadeh, Milad, et al.
Published: (2026)
FSL-Rectifier: Rectify Outliers in Few-Shot Learning via Test-Time Augmentation
by: Bai, Yunwei, et al.
Published: (2024)
by: Bai, Yunwei, et al.
Published: (2024)
Similar Items
-
Don't Trade Off Safety: Diffusion Regularization for Constrained Offline RL
by: Guo, Junyu, et al.
Published: (2025) -
StyleBench: Evaluating thinking styles in Large Language Models
by: Guo, Junyu, et al.
Published: (2025) -
LLMs Should Express Uncertainty Explicitly
by: Guo, Junyu, et al.
Published: (2026) -
AccidentBench: Benchmarking Multimodal Understanding and Reasoning in Vehicle Accidents and Beyond
by: Gu, Shangding, et al.
Published: (2025) -
Enhancing Efficiency of Safe Reinforcement Learning via Sample Manipulation
by: Gu, Shangding, et al.
Published: (2024)