Distributional Active Inference
Fuente:
arXiv
Guardado en:
| Autores principales: | Akgül, Abdullah, Baykal, Gulcin, Haußmann, Manuel, Çelikok, Mustafa Mert, Kandemir, Melih |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Overcoming Non-stationary Dynamics with Evidential Proximal Policy Optimization
por: Akgül, Abdullah, et al.
Publicado: (2025)
por: Akgül, Abdullah, et al.
Publicado: (2025)
A Measure-Theoretic Finite-Sample Theory for Adaptive-Data Fitted Q-Iteration
por: Haussmann, Manuel, et al.
Publicado: (2026)
por: Haussmann, Manuel, et al.
Publicado: (2026)
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies
por: Boock, Magnus Victor, et al.
Publicado: (2026)
por: Boock, Magnus Victor, et al.
Publicado: (2026)
Deterministic Uncertainty Propagation for Improved Model-Based Offline Reinforcement Learning
por: Akgül, Abdullah, et al.
Publicado: (2024)
por: Akgül, Abdullah, et al.
Publicado: (2024)
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
por: Baykal, Gulcin, et al.
Publicado: (2025)
por: Baykal, Gulcin, et al.
Publicado: (2025)
Disentanglement with Factor Quantized Variational Autoencoders
por: Baykal, Gulcin, et al.
Publicado: (2024)
por: Baykal, Gulcin, et al.
Publicado: (2024)
EdVAE: Mitigating Codebook Collapse with Evidential Discrete Variational Autoencoders
por: Baykal, Gulcin, et al.
Publicado: (2023)
por: Baykal, Gulcin, et al.
Publicado: (2023)
PAC-Bayesian Soft Actor-Critic Learning
por: Tasdighi, Bahareh, et al.
Publicado: (2023)
por: Tasdighi, Bahareh, et al.
Publicado: (2023)
Continual Learning of Multi-modal Dynamics with External Memory
por: Akgül, Abdullah, et al.
Publicado: (2022)
por: Akgül, Abdullah, et al.
Publicado: (2022)
Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits
por: Werge, Nicklas, et al.
Publicado: (2023)
por: Werge, Nicklas, et al.
Publicado: (2023)
Adaptive Ensemble Aggregation for Actor-Critics
por: Werge, Nicklas, et al.
Publicado: (2025)
por: Werge, Nicklas, et al.
Publicado: (2025)
Deep Exploration with PAC-Bayes
por: Tasdighi, Bahareh, et al.
Publicado: (2024)
por: Tasdighi, Bahareh, et al.
Publicado: (2024)
Calibrating Bayesian UNet++ for Sub-Seasonal Forecasting
por: Asan, Busra, et al.
Publicado: (2024)
por: Asan, Busra, et al.
Publicado: (2024)
Deep Actor-Critics with Tight Risk Certificates
por: Tasdighi, Bahareh, et al.
Publicado: (2025)
por: Tasdighi, Bahareh, et al.
Publicado: (2025)
Social Cooperation in Conversational AI Agents
por: Çelikok, Mustafa Mert, et al.
Publicado: (2025)
por: Çelikok, Mustafa Mert, et al.
Publicado: (2025)
On the Complexity of Learning to Cooperate with Populations of Socially Rational Agents
por: Loftin, Robert, et al.
Publicado: (2024)
por: Loftin, Robert, et al.
Publicado: (2024)
Distributed Influence-Augmented Local Simulators for Parallel MARL in Large Networked Systems
por: Suau, Miguel, et al.
Publicado: (2022)
por: Suau, Miguel, et al.
Publicado: (2022)
Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning
por: Vincent, Théo, et al.
Publicado: (2025)
por: Vincent, Théo, et al.
Publicado: (2025)
Policy-based Tuning of Autoregressive Image Models with Instance- and Distribution-Level Rewards
por: Baran, Orhun Buğra, et al.
Publicado: (2026)
por: Baran, Orhun Buğra, et al.
Publicado: (2026)
Improved Algorithms for Stochastic Linear Bandits Using Tail Bounds for Martingale Mixtures
por: Flynn, Hamish, et al.
Publicado: (2023)
por: Flynn, Hamish, et al.
Publicado: (2023)
Inverse Concave-Utility Reinforcement Learning is Inverse Game Theory
por: Çelikok, Mustafa Mert, et al.
Publicado: (2024)
por: Çelikok, Mustafa Mert, et al.
Publicado: (2024)
Uncoupled Learning of Differential Stackelberg Equilibria with Commitments
por: Loftin, Robert, et al.
Publicado: (2023)
por: Loftin, Robert, et al.
Publicado: (2023)
Improving Actor-Critic Training with Steerable Action-Value Approximation Errors
por: Tasdighi, Bahareh, et al.
Publicado: (2024)
por: Tasdighi, Bahareh, et al.
Publicado: (2024)
Value Improved Actor Critic Algorithms
por: Oren, Yaniv, et al.
Publicado: (2024)
por: Oren, Yaniv, et al.
Publicado: (2024)
GRU-D Characterizes Age-Specific Temporal Missingness in MIMIC-IV
por: Giesa, Niklas, et al.
Publicado: (2024)
por: Giesa, Niklas, et al.
Publicado: (2024)
Latent variable model for high-dimensional point process with structured missingness
por: Sinelnikov, Maksim, et al.
Publicado: (2024)
por: Sinelnikov, Maksim, et al.
Publicado: (2024)
The Challenge of Achieving Attributability in Multilingual Table-to-Text Generation with Question-Answer Blueprints
por: Haussmann, Aden
Publicado: (2025)
por: Haussmann, Aden
Publicado: (2025)
AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference
por: Feng, Yilin, et al.
Publicado: (2026)
por: Feng, Yilin, et al.
Publicado: (2026)
Latent mixed-effect models for high-dimensional longitudinal data
por: Ong, Priscilla, et al.
Publicado: (2024)
por: Ong, Priscilla, et al.
Publicado: (2024)
Enhancing Text2Cypher with Schema Filtering
por: Ozsoy, Makbule Gulcin
Publicado: (2025)
por: Ozsoy, Makbule Gulcin
Publicado: (2025)
Text2Cypher: Data Pruning using Hard Example Selection
por: Ozsoy, Makbule Gulcin
Publicado: (2025)
por: Ozsoy, Makbule Gulcin
Publicado: (2025)
Multi-Margin Cosine Loss: Proposal and Application in Recommender Systems
por: Ozsoy, Makbule Gulcin
Publicado: (2024)
por: Ozsoy, Makbule Gulcin
Publicado: (2024)
Adaptive Inference: Theoretical Limits and Unexplored Opportunities
por: Hor, Soheil, et al.
Publicado: (2024)
por: Hor, Soheil, et al.
Publicado: (2024)
Contrastive Active Inference
por: Mazzaglia, Pietro, et al.
Publicado: (2021)
por: Mazzaglia, Pietro, et al.
Publicado: (2021)
AdvSecureNet: A Python Toolkit for Adversarial Machine Learning
por: Catal, Melih, et al.
Publicado: (2024)
por: Catal, Melih, et al.
Publicado: (2024)
BaKlaVa -- Budgeted Allocation of KV cache for Long-context Inference
por: Gulhan, Ahmed Burak, et al.
Publicado: (2025)
por: Gulhan, Ahmed Burak, et al.
Publicado: (2025)
Probabilistic Shapley Value Modeling and Inference
por: Ketenci, Mert, et al.
Publicado: (2024)
por: Ketenci, Mert, et al.
Publicado: (2024)
Active Statistical Inference
por: Zrnic, Tijana, et al.
Publicado: (2024)
por: Zrnic, Tijana, et al.
Publicado: (2024)
Active Learning of Deep Neural Networks via Gradient-Free Cutting Planes
por: Zhang, Erica, et al.
Publicado: (2024)
por: Zhang, Erica, et al.
Publicado: (2024)
Accurate and Scalable Stochastic Gaussian Process Regression via Learnable Coreset-based Variational Inference
por: Ketenci, Mert, et al.
Publicado: (2023)
por: Ketenci, Mert, et al.
Publicado: (2023)
Ejemplares similares
-
Overcoming Non-stationary Dynamics with Evidential Proximal Policy Optimization
por: Akgül, Abdullah, et al.
Publicado: (2025) -
A Measure-Theoretic Finite-Sample Theory for Adaptive-Data Fitted Q-Iteration
por: Haussmann, Manuel, et al.
Publicado: (2026) -
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies
por: Boock, Magnus Victor, et al.
Publicado: (2026) -
Deterministic Uncertainty Propagation for Improved Model-Based Offline Reinforcement Learning
por: Akgül, Abdullah, et al.
Publicado: (2024) -
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
por: Baykal, Gulcin, et al.
Publicado: (2025)