Intrinsic Rewards for Exploration without Harm from Observational Noise: A Simulation Study Based on the Free Energy Principle
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tinker, Theodore Jerome, Doya, Kenji, Tani, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Curiosity-Driven Development of Action and Language in Robots Through Self-Exploration
von: Tinker, Theodore Jerome, et al.
Veröffentlicht: (2025)
von: Tinker, Theodore Jerome, et al.
Veröffentlicht: (2025)
The Selective G-Bispectrum and its Inversion: Applications to G-Invariant Networks
von: Mataigne, Simon, et al.
Veröffentlicht: (2024)
von: Mataigne, Simon, et al.
Veröffentlicht: (2024)
Strategic inputs: feature selection from game-theoretic perspective
von: Zhao, Chi, et al.
Veröffentlicht: (2025)
von: Zhao, Chi, et al.
Veröffentlicht: (2025)
Machine Learning for Physical Simulation Challenge Results and Retrospective Analysis: Power Grid Use Case
von: Leyli-Abadi, Milad, et al.
Veröffentlicht: (2025)
von: Leyli-Abadi, Milad, et al.
Veröffentlicht: (2025)
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
von: Xu, Zhi-Qin John, et al.
Veröffentlicht: (2019)
von: Xu, Zhi-Qin John, et al.
Veröffentlicht: (2019)
AI-based modular warning machine for risk identification in proximity healthcare
von: Razzetta, Chiara, et al.
Veröffentlicht: (2025)
von: Razzetta, Chiara, et al.
Veröffentlicht: (2025)
Activity Detection for Massive Connectivity in Cell-free Networks with Unknown Large-scale Fading, Channel Statistics, Noise Variance, and Activity Probability: A Bayesian Approach
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
Evaluating AI Grading on Real-World Handwritten College Mathematics: A Large-Scale Study Toward a Benchmark
von: Yu, Zhiqi, et al.
Veröffentlicht: (2026)
von: Yu, Zhiqi, et al.
Veröffentlicht: (2026)
Differential Parity: Relative Fairness Between Two Sets of Decisions
von: Yu, Zhe, et al.
Veröffentlicht: (2021)
von: Yu, Zhe, et al.
Veröffentlicht: (2021)
Freeze, Diffuse, Decode: Geometry-Aware Adaptation of Pretrained Transformer Embeddings for Antimicrobial Peptide Design
von: Gawade, Pankhil, et al.
Veröffentlicht: (2025)
von: Gawade, Pankhil, et al.
Veröffentlicht: (2025)
ReMIA: a Powerful and Efficient Alternative to Membership Inference Attacks against Synthetic Data Generators
von: Scassola, Davide, et al.
Veröffentlicht: (2026)
von: Scassola, Davide, et al.
Veröffentlicht: (2026)
Physics-Informed Neural Networks and Neural Operators for Parametric PDEs
von: Zhang, Zhuo, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuo, et al.
Veröffentlicht: (2025)
On the Optimal Memorization Capacity of Transformers
von: Kajitsuka, Tokio, et al.
Veröffentlicht: (2024)
von: Kajitsuka, Tokio, et al.
Veröffentlicht: (2024)
Representation Integrity in Temporal Graph Learning Methods
von: Kooshafar, Elahe
Veröffentlicht: (2025)
von: Kooshafar, Elahe
Veröffentlicht: (2025)
A Data-Driven Measure of Relative Uncertainty for Misclassification Detection
von: Dadalto, Eduardo, et al.
Veröffentlicht: (2023)
von: Dadalto, Eduardo, et al.
Veröffentlicht: (2023)
Modeling Membrane Degradation in PEM Electrolyzers with Physics-Informed Neural Networks
von: Polo-Molina, Alejandro, et al.
Veröffentlicht: (2025)
von: Polo-Molina, Alejandro, et al.
Veröffentlicht: (2025)
Dynamic Observation Policies in Observation Cost-Sensitive Reinforcement Learning
von: Bellinger, Colin, et al.
Veröffentlicht: (2023)
von: Bellinger, Colin, et al.
Veröffentlicht: (2023)
A theoretical basis for model collapse in recursive training
von: Borkar, Vivek Shripad
Veröffentlicht: (2025)
von: Borkar, Vivek Shripad
Veröffentlicht: (2025)
Optimal Bound for PCA with Outliers using Higher-Degree Voronoi Diagrams
von: Hashemian, Sajjad, et al.
Veröffentlicht: (2024)
von: Hashemian, Sajjad, et al.
Veröffentlicht: (2024)
Learning from Complementary Features
von: Sugiyama, Kosuke, et al.
Veröffentlicht: (2024)
von: Sugiyama, Kosuke, et al.
Veröffentlicht: (2024)
Learning to Control the Smoothness of Graph Convolutional Network Features
von: Wang, Shih-Hsin, et al.
Veröffentlicht: (2024)
von: Wang, Shih-Hsin, et al.
Veröffentlicht: (2024)
Learning from Similarity-Confidence and Confidence-Difference
von: Tate, Tomoya, et al.
Veröffentlicht: (2025)
von: Tate, Tomoya, et al.
Veröffentlicht: (2025)
Reproducibility and Geometric Intrinsic Dimensionality: An Investigation on Graph Neural Network Research
von: Hille, Tobias, et al.
Veröffentlicht: (2024)
von: Hille, Tobias, et al.
Veröffentlicht: (2024)
Learning from Hard Labels with Additional Supervision on Non-Hard-Labeled Classes
von: Sugiyama, Kosuke, et al.
Veröffentlicht: (2025)
von: Sugiyama, Kosuke, et al.
Veröffentlicht: (2025)
Learning Through Noise: Why Subliminal Learning Works and When It Fails
von: Brockers, Vincent C., et al.
Veröffentlicht: (2026)
von: Brockers, Vincent C., et al.
Veröffentlicht: (2026)
Balanced Group Convolution: An Improved Group Convolution Based on Approximability Estimates
von: Lee, Youngkyu, et al.
Veröffentlicht: (2023)
von: Lee, Youngkyu, et al.
Veröffentlicht: (2023)
Tape: A Cellular Automata Benchmark for Evaluating Rule-Shift Generalization in Reinforcement Learning
von: Pan, Enze
Veröffentlicht: (2026)
von: Pan, Enze
Veröffentlicht: (2026)
seqme: a Python library for evaluating biological sequence design
von: Møller-Larsen, Rasmus, et al.
Veröffentlicht: (2025)
von: Møller-Larsen, Rasmus, et al.
Veröffentlicht: (2025)
The Llama 4 Herd: Architecture, Training, Evaluation, and Deployment Notes
von: arXiv, Redacted by
Veröffentlicht: (2026)
von: arXiv, Redacted by
Veröffentlicht: (2026)
RASP-Tuner: Retrieval-Augmented Soft Prompts for Context-Aware Black-Box Optimization in Non-Stationary Environments
von: Pan, Enze
Veröffentlicht: (2026)
von: Pan, Enze
Veröffentlicht: (2026)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
Stabilizing Machine Learning for Reproducible and Explainable Results: A Novel Validation Approach to Subject-Specific Insights
von: Vos, Gideon, et al.
Veröffentlicht: (2024)
von: Vos, Gideon, et al.
Veröffentlicht: (2024)
Differentially Private Prototypes for Imbalanced Transfer Learning
von: Wahdany, Dariush, et al.
Veröffentlicht: (2024)
von: Wahdany, Dariush, et al.
Veröffentlicht: (2024)
Predictors of disease outbreaks at continentalscale in the African region: Insights and predictions with geospatial artificial intelligence using earth observations and routine disease surveillance data
von: Pezanowski, Scott, et al.
Veröffentlicht: (2024)
von: Pezanowski, Scott, et al.
Veröffentlicht: (2024)
A Human-In-The-Loop Approach for Improving Fairness in Predictive Business Process Monitoring
von: Käppel, Martin, et al.
Veröffentlicht: (2025)
von: Käppel, Martin, et al.
Veröffentlicht: (2025)
An Efficient Transport-Based Dissimilarity Measure for Time Series Classification under Warping Distortions
von: Aldroubi, Akram, et al.
Veröffentlicht: (2025)
von: Aldroubi, Akram, et al.
Veröffentlicht: (2025)
Fast, close, non-singular and property-preserving approximations of entropic measures
von: Horenko, Illia, et al.
Veröffentlicht: (2025)
von: Horenko, Illia, et al.
Veröffentlicht: (2025)
Sequence Transferability and Task Order Selection in Continual Learning
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
von: Aksoy, Sinan G., et al.
Veröffentlicht: (2026)
von: Aksoy, Sinan G., et al.
Veröffentlicht: (2026)
ECG-FM: An Open Electrocardiogram Foundation Model
von: McKeen, Kaden, et al.
Veröffentlicht: (2024)
von: McKeen, Kaden, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Curiosity-Driven Development of Action and Language in Robots Through Self-Exploration
von: Tinker, Theodore Jerome, et al.
Veröffentlicht: (2025) -
The Selective G-Bispectrum and its Inversion: Applications to G-Invariant Networks
von: Mataigne, Simon, et al.
Veröffentlicht: (2024) -
Strategic inputs: feature selection from game-theoretic perspective
von: Zhao, Chi, et al.
Veröffentlicht: (2025) -
Machine Learning for Physical Simulation Challenge Results and Retrospective Analysis: Power Grid Use Case
von: Leyli-Abadi, Milad, et al.
Veröffentlicht: (2025) -
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
von: Xu, Zhi-Qin John, et al.
Veröffentlicht: (2019)