Path-Guided Particle-based Sampling
Fuente:
arXiv
Guardado en:
| Autores principales: | Fan, Mingzhou, Zhou, Ruida, Tian, Chao, Qian, Xiaoning |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Understanding Uncertainty-based Active Learning Under Model Mismatch
por: Rahmati, Amir Hossein, et al.
Publicado: (2024)
por: Rahmati, Amir Hossein, et al.
Publicado: (2024)
Towards Physics-Guided Foundation Models
por: Farhadloo, Majid, et al.
Publicado: (2025)
por: Farhadloo, Majid, et al.
Publicado: (2025)
The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling
por: Nguyen, Tu, et al.
Publicado: (2026)
por: Nguyen, Tu, et al.
Publicado: (2026)
On the Learn-to-Optimize Capabilities of Transformers in In-Context Sparse Recovery
por: Liu, Renpu, et al.
Publicado: (2024)
por: Liu, Renpu, et al.
Publicado: (2024)
Towards Fine-Tuning-Based Site Calibration for Knowledge-Guided Machine Learning: A Summary of Results
por: Zeng, Ruolei, et al.
Publicado: (2025)
por: Zeng, Ruolei, et al.
Publicado: (2025)
Beyond Binary Preferences: A Principled Framework for Reward Modeling with Ordinal Feedback
por: Afsharrad, Amirhossein, et al.
Publicado: (2026)
por: Afsharrad, Amirhossein, et al.
Publicado: (2026)
InvarGC: Invariant Granger Causality for Heterogeneous Interventional Time Series under Latent Confounding
por: Zhang, Ziyi, et al.
Publicado: (2025)
por: Zhang, Ziyi, et al.
Publicado: (2025)
Reframing Data Value for Large Language Models Through the Lens of Plausibility
por: Rammal, Mohamad Rida, et al.
Publicado: (2024)
por: Rammal, Mohamad Rida, et al.
Publicado: (2024)
Fairness Without Harm: An Influence-Guided Active Sampling Approach
por: Pang, Jinlong, et al.
Publicado: (2024)
por: Pang, Jinlong, et al.
Publicado: (2024)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
por: Wang, Xinchen, et al.
Publicado: (2025)
por: Wang, Xinchen, et al.
Publicado: (2025)
Path Planning for Masked Diffusion Model Sampling
por: Peng, Fred Zhangzhi, et al.
Publicado: (2025)
por: Peng, Fred Zhangzhi, et al.
Publicado: (2025)
Path-Guided Flow Matching for Dataset Distillation
por: Li, Xuhui, et al.
Publicado: (2026)
por: Li, Xuhui, et al.
Publicado: (2026)
GFlowNet Training by Policy Gradients
por: Niu, Puhua, et al.
Publicado: (2024)
por: Niu, Puhua, et al.
Publicado: (2024)
Sparse-VQ Transformer: An FFN-Free Framework with Vector Quantization for Enhanced Time Series Forecasting
por: Zhao, Yanjun, et al.
Publicado: (2024)
por: Zhao, Yanjun, et al.
Publicado: (2024)
Integer-only Quantized Transformers for Embedded FPGA-based Time-series Forecasting in AIoT
por: Ling, Tianheng, et al.
Publicado: (2024)
por: Ling, Tianheng, et al.
Publicado: (2024)
HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents
por: Peng, Jiangweizhi, et al.
Publicado: (2026)
por: Peng, Jiangweizhi, et al.
Publicado: (2026)
GAR: Generative Adversarial Reinforcement Learning for Formal Theorem Proving
por: Wang, Ruida, et al.
Publicado: (2025)
por: Wang, Ruida, et al.
Publicado: (2025)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
por: Gao, Fengyu, et al.
Publicado: (2024)
por: Gao, Fengyu, et al.
Publicado: (2024)
Geometry of Drifting MDPs with Path-Integral Stability Certificates
por: Zhang, Zuyuan, et al.
Publicado: (2026)
por: Zhang, Zuyuan, et al.
Publicado: (2026)
FASTER: Value-Guided Sampling for Fast RL
por: Dong, Perry, et al.
Publicado: (2026)
por: Dong, Perry, et al.
Publicado: (2026)
AIS: Adaptive Importance Sampling for Quantized RL
por: Zhou, Jiajun, et al.
Publicado: (2026)
por: Zhou, Jiajun, et al.
Publicado: (2026)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
por: Liu, Zongkai, et al.
Publicado: (2024)
por: Liu, Zongkai, et al.
Publicado: (2024)
Data-Augmented Few-Shot Neural Emulator for Computer-Model System Identification
por: Jantre, Sanket, et al.
Publicado: (2025)
por: Jantre, Sanket, et al.
Publicado: (2025)
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
por: Hu, Ruida, et al.
Publicado: (2025)
por: Hu, Ruida, et al.
Publicado: (2025)
NeuroPathNet: Dynamic Path Trajectory Learning for Brain Functional Connectivity Analysis
por: Guo, Tianqi, et al.
Publicado: (2025)
por: Guo, Tianqi, et al.
Publicado: (2025)
Contraction Actor-Critic: Contraction Metric-Guided Reinforcement Learning for Robust Path Tracking
por: Cho, Minjae, et al.
Publicado: (2025)
por: Cho, Minjae, et al.
Publicado: (2025)
Sharpening the Spear: Adaptive Expert-Guided Adversarial Attack Against DRL-based Autonomous Driving Policies
por: Fan, Junchao, et al.
Publicado: (2025)
por: Fan, Junchao, et al.
Publicado: (2025)
Preference as Reward, Maximum Preference Optimization with Importance Sampling
por: Jiang, Zaifan, et al.
Publicado: (2023)
por: Jiang, Zaifan, et al.
Publicado: (2023)
Monte Carlo Tree Search based Space Transfer for Black-box Optimization
por: Wang, Shukuan, et al.
Publicado: (2024)
por: Wang, Shukuan, et al.
Publicado: (2024)
Geometric Prior-Guided Federated Prompt Calibration
por: Luo, Fei, et al.
Publicado: (2025)
por: Luo, Fei, et al.
Publicado: (2025)
LEASE: Offline Preference-based Reinforcement Learning with High Sample Efficiency
por: Liu, Xiao-Yin, et al.
Publicado: (2024)
por: Liu, Xiao-Yin, et al.
Publicado: (2024)
OffLight: An Offline Multi-Agent Reinforcement Learning Framework for Traffic Signal Control
por: Bokade, Rohit, et al.
Publicado: (2024)
por: Bokade, Rohit, et al.
Publicado: (2024)
QUCE: The Minimisation and Quantification of Path-Based Uncertainty for Generative Counterfactual Explanations
por: Duell, Jamie, et al.
Publicado: (2024)
por: Duell, Jamie, et al.
Publicado: (2024)
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Guiding Diffusion Models with Reinforcement Learning for Stable Molecule Generation
por: Zhou, Zhijian, et al.
Publicado: (2025)
por: Zhou, Zhijian, et al.
Publicado: (2025)
TimeSAF: Towards LLM-Guided Semantic Asynchronous Fusion for Time Series Forecasting
por: Zhang, Fan, et al.
Publicado: (2026)
por: Zhang, Fan, et al.
Publicado: (2026)
Particle-Guided Diffusion for Gas-Phase Reaction Kinetics
por: Millard, Andrew, et al.
Publicado: (2026)
por: Millard, Andrew, et al.
Publicado: (2026)
Fast and Robust Likelihood-Guided Diffusion Posterior Sampling with Amortized Variational Inference
por: Zheng, Léon, et al.
Publicado: (2026)
por: Zheng, Léon, et al.
Publicado: (2026)
Clear Preferences Leave Traces: Reference Model-Guided Sampling for Preference Learning
por: Diwan, Nirav, et al.
Publicado: (2025)
por: Diwan, Nirav, et al.
Publicado: (2025)
GuidedSampling: Steering LLMs Towards Diverse Candidate Solutions at Inference-Time
por: Handa, Divij, et al.
Publicado: (2025)
por: Handa, Divij, et al.
Publicado: (2025)
Ejemplares similares
-
Understanding Uncertainty-based Active Learning Under Model Mismatch
por: Rahmati, Amir Hossein, et al.
Publicado: (2024) -
Towards Physics-Guided Foundation Models
por: Farhadloo, Majid, et al.
Publicado: (2025) -
The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling
por: Nguyen, Tu, et al.
Publicado: (2026) -
On the Learn-to-Optimize Capabilities of Transformers in In-Context Sparse Recovery
por: Liu, Renpu, et al.
Publicado: (2024) -
Towards Fine-Tuning-Based Site Calibration for Knowledge-Guided Machine Learning: A Summary of Results
por: Zeng, Ruolei, et al.
Publicado: (2025)