Online Posterior Sampling with a Diffusion Prior
Fuente:
arXiv
Guardado en:
| Autores principales: | Kveton, Branislav, Oreshkin, Boris, Park, Youngsuk, Deshmukh, Aniket, Song, Rui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LLM-as-Judge on a Budget
por: Saha, Aadirupa, et al.
Publicado: (2026)
por: Saha, Aadirupa, et al.
Publicado: (2026)
Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024)
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024)
Experimental Design for Active Transductive Inference in Large Language Models
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024)
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024)
Optimal Design for Human Preference Elicitation
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024)
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024)
Learning from a single labeled face and a stream of unlabeled data
por: Kveton, Branislav, et al.
Publicado: (2026)
por: Kveton, Branislav, et al.
Publicado: (2026)
Online semi-supervised perception: Real-time learning without explicit feedback
por: Kveton, Branislav, et al.
Publicado: (2026)
por: Kveton, Branislav, et al.
Publicado: (2026)
Cross-Validated Off-Policy Evaluation
por: Cief, Matej, et al.
Publicado: (2024)
por: Cief, Matej, et al.
Publicado: (2024)
Efficient and Interpretable Bandit Algorithms
por: Mukherjee, Subhojyoti, et al.
Publicado: (2023)
por: Mukherjee, Subhojyoti, et al.
Publicado: (2023)
Divide-and-Conquer Posterior Sampling for Denoising Diffusion Priors
por: Janati, Yazid, et al.
Publicado: (2024)
por: Janati, Yazid, et al.
Publicado: (2024)
MODL: Multilearner Online Deep Learning
por: Valkanas, Antonios, et al.
Publicado: (2024)
por: Valkanas, Antonios, et al.
Publicado: (2024)
Amortized Posterior Sampling with Diffusion Prior Distillation
por: Mammadov, Abbas, et al.
Publicado: (2024)
por: Mammadov, Abbas, et al.
Publicado: (2024)
Language-Model Prior Overcomes Cold-Start Items
por: Wang, Shiyu, et al.
Publicado: (2024)
por: Wang, Shiyu, et al.
Publicado: (2024)
Pessimistic Off-Policy Optimization for Learning to Rank
por: Cief, Matej, et al.
Publicado: (2022)
por: Cief, Matej, et al.
Publicado: (2022)
Semi-supervised learning with max-margin graph cuts
por: Kveton, Branislav, et al.
Publicado: (2026)
por: Kveton, Branislav, et al.
Publicado: (2026)
Spectral bandits for smooth graph functions
por: Valko, Michal, et al.
Publicado: (2026)
por: Valko, Michal, et al.
Publicado: (2026)
Off-Policy Evaluation from Logged Human Feedback
por: Bhargava, Aniruddha, et al.
Publicado: (2024)
por: Bhargava, Aniruddha, et al.
Publicado: (2024)
Spectral bandits for smooth graph functions with applications in recommender systems
por: Kocák, Tomáš, et al.
Publicado: (2026)
por: Kocák, Tomáš, et al.
Publicado: (2026)
Evidence-based anomaly detection in clinical domains
por: Hauskrecht, Milos, et al.
Publicado: (2026)
por: Hauskrecht, Milos, et al.
Publicado: (2026)
Finite-Time Logarithmic Bayes Regret Upper Bounds
por: Atsidakou, Alexia, et al.
Publicado: (2023)
por: Atsidakou, Alexia, et al.
Publicado: (2023)
Probabilistic Pretraining for Neural Regression
por: Oreshkin, Boris N., et al.
Publicado: (2025)
por: Oreshkin, Boris N., et al.
Publicado: (2025)
Partial Policy Gradients for RL in LLMs
por: Mathur, Puneet, et al.
Publicado: (2026)
por: Mathur, Puneet, et al.
Publicado: (2026)
Stochastic Rounding for LLM Training: Theory and Practice
por: Ozkara, Kaan, et al.
Publicado: (2025)
por: Ozkara, Kaan, et al.
Publicado: (2025)
Training LLMs with MXFP4
por: Tseng, Albert, et al.
Publicado: (2025)
por: Tseng, Albert, et al.
Publicado: (2025)
Flexible Bayesian Last Layer Models Using Implicit Priors and Diffusion Posterior Sampling
por: Xu, Jian, et al.
Publicado: (2024)
por: Xu, Jian, et al.
Publicado: (2024)
Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement Learning
por: Mutti, Mirco, et al.
Publicado: (2023)
por: Mutti, Mirco, et al.
Publicado: (2023)
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs
por: Bian, Song, et al.
Publicado: (2025)
por: Bian, Song, et al.
Publicado: (2025)
Conditional anomaly detection with soft harmonic functions
por: Valko, Michal, et al.
Publicado: (2026)
por: Valko, Michal, et al.
Publicado: (2026)
Conditional anomaly detection using soft harmonic functions: An application to clinical alerting
por: Valko, Michal, et al.
Publicado: (2026)
por: Valko, Michal, et al.
Publicado: (2026)
ML-Tool-Bench: Tool-Augmented Planning for ML Tasks
por: Chittepu, Yaswanth, et al.
Publicado: (2025)
por: Chittepu, Yaswanth, et al.
Publicado: (2025)
Spectral bandits
por: Kocák, Tomáš, et al.
Publicado: (2026)
por: Kocák, Tomáš, et al.
Publicado: (2026)
Split Gibbs Discrete Diffusion Posterior Sampling
por: Chu, Wenda, et al.
Publicado: (2025)
por: Chu, Wenda, et al.
Publicado: (2025)
Plug-and-Play Posterior Sampling under Mismatched Measurement and Prior Models
por: Renaud, Marien, et al.
Publicado: (2023)
por: Renaud, Marien, et al.
Publicado: (2023)
Variational Diffusion Posterior Sampling with Midpoint Guidance
por: Moufad, Badr, et al.
Publicado: (2024)
por: Moufad, Badr, et al.
Publicado: (2024)
Agentic Planning with Reasoning for Image Styling via Offline RL
por: Mukherjee, Subhojyoti, et al.
Publicado: (2026)
por: Mukherjee, Subhojyoti, et al.
Publicado: (2026)
An Efficient Plugin Method for Metric Optimization of Black-Box Models
por: Devic, Siddartha, et al.
Publicado: (2025)
por: Devic, Siddartha, et al.
Publicado: (2025)
RADAR: Reasoning-Ability and Difficulty-Aware Routing for Reasoning LLMs
por: Fernandez, Nigel, et al.
Publicado: (2025)
por: Fernandez, Nigel, et al.
Publicado: (2025)
Any-Quantile Probabilistic Forecasting of Short-Term Electricity Demand
por: Smyl, Slawek, et al.
Publicado: (2024)
por: Smyl, Slawek, et al.
Publicado: (2024)
Learning Posterior Predictive Distributions for Node Classification from Synthetic Graph Priors
por: Choi, Jeongwhan, et al.
Publicado: (2026)
por: Choi, Jeongwhan, et al.
Publicado: (2026)
Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey
por: Janati, Yazid, et al.
Publicado: (2025)
por: Janati, Yazid, et al.
Publicado: (2025)
Diffusion Posterior Sampling is Computationally Intractable
por: Gupta, Shivam, et al.
Publicado: (2024)
por: Gupta, Shivam, et al.
Publicado: (2024)
Ejemplares similares
-
LLM-as-Judge on a Budget
por: Saha, Aadirupa, et al.
Publicado: (2026) -
Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024) -
Experimental Design for Active Transductive Inference in Large Language Models
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024) -
Optimal Design for Human Preference Elicitation
por: Mukherjee, Subhojyoti, et al.
Publicado: (2024) -
Learning from a single labeled face and a stream of unlabeled data
por: Kveton, Branislav, et al.
Publicado: (2026)