How Hard is it to Confuse a World Model?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Radji, Waris, Maillard, Odalric-Ambrym |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Confusing Instance Principle for Online Linear Quadratic Control
von: Radji, Waris, et al.
Veröffentlicht: (2025)
von: Radji, Waris, et al.
Veröffentlicht: (2025)
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025)
Leveraging priors on distribution functions for multi-arm bandits
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025)
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025)
The regret lower bound for communicating Markov Decision Processes
von: Boone, Victor, et al.
Veröffentlicht: (2025)
von: Boone, Victor, et al.
Veröffentlicht: (2025)
How to Shrink Confidence Sets for Many Equivalent Discrete Distributions?
von: Maillard, Odalric-Ambrym, et al.
Veröffentlicht: (2024)
von: Maillard, Odalric-Ambrym, et al.
Veröffentlicht: (2024)
A Continual Offline Reinforcement Learning Benchmark for Navigation Tasks
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
Pliable rejection sampling
von: Erraqabi, Akram, et al.
Veröffentlicht: (2026)
von: Erraqabi, Akram, et al.
Veröffentlicht: (2026)
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
von: Kobanda, Anthony, et al.
Veröffentlicht: (2024)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2024)
Intrinsic-Energy Joint Embedding Predictive Architectures Induce Quasimetric Spaces
von: Kobanda, Anthony, et al.
Veröffentlicht: (2026)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2026)
Provably Efficient Exploration in Reward Machines with Low Regret
von: Bourel, Hippolyte, et al.
Veröffentlicht: (2024)
von: Bourel, Hippolyte, et al.
Veröffentlicht: (2024)
Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX
von: Radji, Waris, et al.
Veröffentlicht: (2025)
von: Radji, Waris, et al.
Veröffentlicht: (2025)
Interpolation pour l'augmentation de donnees : Application à la gestion des adventices de la canne a sucre a la Reunion
von: Ferber, Frederick Fabre, et al.
Veröffentlicht: (2025)
von: Ferber, Frederick Fabre, et al.
Veröffentlicht: (2025)
AdaStop: adaptive statistical testing for sound comparisons of Deep RL agents
von: Mathieu, Timothée, et al.
Veröffentlicht: (2023)
von: Mathieu, Timothée, et al.
Veröffentlicht: (2023)
Power Mean Estimation in Stochastic Monte-Carlo Tree_Search
von: Dam, Tuan, et al.
Veröffentlicht: (2024)
von: Dam, Tuan, et al.
Veröffentlicht: (2024)
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
von: Kohler, Hector, et al.
Veröffentlicht: (2025)
von: Kohler, Hector, et al.
Veröffentlicht: (2025)
Kriging and Gaussian Process Interpolation for Georeferenced Data Augmentation
von: Ferber, Frédérick Fabre, et al.
Veröffentlicht: (2025)
von: Ferber, Frédérick Fabre, et al.
Veröffentlicht: (2025)
Stacked Confusion Reject Plots (SCORE)
von: Hasler, Stephan, et al.
Veröffentlicht: (2024)
von: Hasler, Stephan, et al.
Veröffentlicht: (2024)
Reasoning in Diffusion Large Language Models is Concentrated in Dynamic Confusion Zones
von: Chen, Ranfei, et al.
Veröffentlicht: (2025)
von: Chen, Ranfei, et al.
Veröffentlicht: (2025)
How Hard Can It Be? Hardness-Aware Multi-Objective Unlearning
von: Chen, Jiangwei, et al.
Veröffentlicht: (2026)
von: Chen, Jiangwei, et al.
Veröffentlicht: (2026)
The Confusion is Real: GRAPHIC -- A Network Science Approach to Confusion Matrices in Deep Learning
von: Fröhlich, Johanna S., et al.
Veröffentlicht: (2026)
von: Fröhlich, Johanna S., et al.
Veröffentlicht: (2026)
Improving Deep Ensembles by Estimating Confusion Matrices
von: Kuzin, Danil, et al.
Veröffentlicht: (2025)
von: Kuzin, Danil, et al.
Veröffentlicht: (2025)
On the Normalization of Confusion Matrices: Methods and Geometric Interpretations
von: Erbani, Johan, et al.
Veröffentlicht: (2025)
von: Erbani, Johan, et al.
Veröffentlicht: (2025)
DarkStream: real-time speech anonymization with low latency
von: Quamer, Waris, et al.
Veröffentlicht: (2025)
von: Quamer, Waris, et al.
Veröffentlicht: (2025)
End-to-end streaming model for low-latency speech anonymization
von: Quamer, Waris, et al.
Veröffentlicht: (2024)
von: Quamer, Waris, et al.
Veröffentlicht: (2024)
Fusion or Confusion? Multimodal Complexity Is Not All You Need
von: Rheude, Tillmann, et al.
Veröffentlicht: (2025)
von: Rheude, Tillmann, et al.
Veröffentlicht: (2025)
A Noise Sensitivity Exponent Controls Large Statistical-to-Computational Gaps in Single- and Multi-Index Models
von: Defilippis, Leonardo, et al.
Veröffentlicht: (2026)
von: Defilippis, Leonardo, et al.
Veröffentlicht: (2026)
Toward Adaptive Grid Resilience: A Gradient-Free Meta-RL Framework for Critical Load Restoration
von: Abdeen, Zain ul, et al.
Veröffentlicht: (2026)
von: Abdeen, Zain ul, et al.
Veröffentlicht: (2026)
Reducing Class-wise Confusion for Incremental Learning with Disentangled Manifolds
von: Chen, Huitong, et al.
Veröffentlicht: (2025)
von: Chen, Huitong, et al.
Veröffentlicht: (2025)
ProToken: Token-Level Attribution for Federated Large Language Models
von: Gill, Waris, et al.
Veröffentlicht: (2026)
von: Gill, Waris, et al.
Veröffentlicht: (2026)
Information-theoretic Distinctions Between Deception and Confusion
von: Young, Robin
Veröffentlicht: (2025)
von: Young, Robin
Veröffentlicht: (2025)
How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness
von: Gordienko, Polina, et al.
Veröffentlicht: (2026)
von: Gordienko, Polina, et al.
Veröffentlicht: (2026)
Task Confusion and Catastrophic Forgetting in Class-Incremental Learning: A Mathematical Framework for Discriminative and Generative Modelings
von: Nori, Milad Khademi, et al.
Veröffentlicht: (2024)
von: Nori, Milad Khademi, et al.
Veröffentlicht: (2024)
Exploring and Addressing Reward Confusion in Offline Preference Learning
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Clarify Confused Nodes via Separated Learning
von: Zhou, Jiajun, et al.
Veröffentlicht: (2023)
von: Zhou, Jiajun, et al.
Veröffentlicht: (2023)
BinaryShield: Cross-Service Threat Intelligence in LLM Services using Privacy-Preserving Fingerprints
von: Gill, Waris, et al.
Veröffentlicht: (2025)
von: Gill, Waris, et al.
Veröffentlicht: (2025)
Optimal scaling laws in learning hierarchical multi-index models
von: Defilippis, Leonardo, et al.
Veröffentlicht: (2026)
von: Defilippis, Leonardo, et al.
Veröffentlicht: (2026)
How to Train Your Latent Control Barrier Function: Smooth Safety Filtering Under Hard-to-Model Constraints
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
From Confusion to Clarity: ProtoScore -- A Framework for Evaluating Prototype-Based XAI
von: Monke, Helena, et al.
Veröffentlicht: (2025)
von: Monke, Helena, et al.
Veröffentlicht: (2025)
Improving Generalization Ability of Robotic Imitation Learning by Resolving Causal Confusion in Observations
von: Chen, Yifei, et al.
Veröffentlicht: (2025)
von: Chen, Yifei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Confusing Instance Principle for Online Linear Quadratic Control
von: Radji, Waris, et al.
Veröffentlicht: (2025) -
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025) -
Leveraging priors on distribution functions for multi-arm bandits
von: Vashishtha, Sumit, et al.
Veröffentlicht: (2025) -
The regret lower bound for communicating Markov Decision Processes
von: Boone, Victor, et al.
Veröffentlicht: (2025) -
How to Shrink Confidence Sets for Many Equivalent Discrete Distributions?
von: Maillard, Odalric-Ambrym, et al.
Veröffentlicht: (2024)