MAME: Multidimensional Adaptive Metamer Exploration with Human Perceptual Feedback
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kamao, Mina, Ono, Hayato, Yamashita, Ayumu, Amano, Kaoru, Sawayama, Masataka |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BrainCodec: Neural fMRI codec for the decoding of cognitive brain states
von: Nishimura, Yuto, et al.
Veröffentlicht: (2024)
von: Nishimura, Yuto, et al.
Veröffentlicht: (2024)
Model Metamers Reveal Invariances in Graph Neural Networks
von: Xu, Wei, et al.
Veröffentlicht: (2025)
von: Xu, Wei, et al.
Veröffentlicht: (2025)
Prediction of Alpha Power Using Multiple Subjective Measures and Autonomic Responses
von: Yuting Xu, et al.
Veröffentlicht: (2025)
von: Yuting Xu, et al.
Veröffentlicht: (2025)
Data-dependent Exploration for Online Reinforcement Learning from Human Feedback
von: Zhang, Zhen-Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Zhen-Yu, et al.
Veröffentlicht: (2026)
Sampling Method for Generalized Graph Signals with Pre-selected Vertices via DC Optimization
von: Yamashita, Keitaro, et al.
Veröffentlicht: (2025)
von: Yamashita, Keitaro, et al.
Veröffentlicht: (2025)
Stick to your Role! Stability of Personal Values Expressed in Large Language Models
von: Kovač, Grgur, et al.
Veröffentlicht: (2024)
von: Kovač, Grgur, et al.
Veröffentlicht: (2024)
Pure Exploration with Feedback Graphs
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
Aligning Generative Speech Enhancement with Perceptual Feedback
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
Pure Exploration under Mediators' Feedback
von: Poiani, Riccardo, et al.
Veröffentlicht: (2023)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2023)
Adaptive Preference Scaling for Reinforcement Learning with Human Feedback
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
Understanding Cross-Model Perceptual Invariances Through Ensemble Metamers
von: Boehm, Lukas, et al.
Veröffentlicht: (2025)
von: Boehm, Lukas, et al.
Veröffentlicht: (2025)
Adaptive Querying for Reward Learning from Human Feedback
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
Scalable predictive processing framework for multitask caregiving robots
von: Idei, Hayato, et al.
Veröffentlicht: (2025)
von: Idei, Hayato, et al.
Veröffentlicht: (2025)
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
von: Yamada, Daisuke, et al.
Veröffentlicht: (2025)
von: Yamada, Daisuke, et al.
Veröffentlicht: (2025)
Pure Exploration Beyond Reward Feedback: The Role of Post-Action Context
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback
von: Li, Zitian, et al.
Veröffentlicht: (2026)
von: Li, Zitian, et al.
Veröffentlicht: (2026)
Human-Flow Digital Twin for Predicting the Effects of Mobility Introduction on Visitor Circulation
von: Shima, Chiharu, et al.
Veröffentlicht: (2026)
von: Shima, Chiharu, et al.
Veröffentlicht: (2026)
Reinforcement Learning from Human Feedback
von: Lambert, Nathan
Veröffentlicht: (2025)
von: Lambert, Nathan
Veröffentlicht: (2025)
Out-of-Distribution Learning with Human Feedback
von: Bai, Haoyue, et al.
Veröffentlicht: (2024)
von: Bai, Haoyue, et al.
Veröffentlicht: (2024)
VarteX: Enhancing Weather Forecast through Distributed Variable Representation
von: Ueyama, Ayumu, et al.
Veröffentlicht: (2024)
von: Ueyama, Ayumu, et al.
Veröffentlicht: (2024)
Guided Diffusion Sampling for Precipitation Forecast Interventions
von: Ueyama, Ayumu, et al.
Veröffentlicht: (2026)
von: Ueyama, Ayumu, et al.
Veröffentlicht: (2026)
Complex System Exploration with Interactive Human Guidance
von: Morel, Bastien, et al.
Veröffentlicht: (2025)
von: Morel, Bastien, et al.
Veröffentlicht: (2025)
Strategyproof Reinforcement Learning from Human Feedback
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2025)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2025)
Capture-Calibrate-Coach: A Graph-Based Framework for Knowledge Monitoring Estimation and Adaptive Feedback
von: Li, Gen, et al.
Veröffentlicht: (2026)
von: Li, Gen, et al.
Veröffentlicht: (2026)
Boundary Exploration for Bayesian Optimization With Unknown Physical Constraints
von: Tian, Yunsheng, et al.
Veröffentlicht: (2024)
von: Tian, Yunsheng, et al.
Veröffentlicht: (2024)
Adaptive Bounded Exploration and Intermediate Actions for Data Debiasing
von: Yang, Yifan, et al.
Veröffentlicht: (2025)
von: Yang, Yifan, et al.
Veröffentlicht: (2025)
Human and AI Perceptual Differences in Image Classification Errors
von: Liu, Minghao, et al.
Veröffentlicht: (2023)
von: Liu, Minghao, et al.
Veröffentlicht: (2023)
Prompt Optimization with Human Feedback
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2024)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2024)
Proximal Policy Optimization with Adaptive Exploration
von: Lixandru, Andrei
Veröffentlicht: (2024)
von: Lixandru, Andrei
Veröffentlicht: (2024)
Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback
von: Yang, Yongjin, et al.
Veröffentlicht: (2025)
von: Yang, Yongjin, et al.
Veröffentlicht: (2025)
Proximal Point Nash Learning from Human Feedback
von: Tiapkin, Daniil, et al.
Veröffentlicht: (2025)
von: Tiapkin, Daniil, et al.
Veröffentlicht: (2025)
Off-Policy Evaluation from Logged Human Feedback
von: Bhargava, Aniruddha, et al.
Veröffentlicht: (2024)
von: Bhargava, Aniruddha, et al.
Veröffentlicht: (2024)
Robust Reinforcement Learning from Corrupted Human Feedback
von: Bukharin, Alexander, et al.
Veröffentlicht: (2024)
von: Bukharin, Alexander, et al.
Veröffentlicht: (2024)
DiBA: Diagonal and Binary Matrix Approximation for Neural Network Weight Compression
von: Ono, Nobutaka
Veröffentlicht: (2026)
von: Ono, Nobutaka
Veröffentlicht: (2026)
ARISE: Adaptive Reinforcement Integrated with Swarm Exploration
von: M, Rajiv Chaitanya, et al.
Veröffentlicht: (2026)
von: M, Rajiv Chaitanya, et al.
Veröffentlicht: (2026)
Learning Representations for CSI Adaptive Quantization and Feedback
von: Rizzello, Valentina, et al.
Veröffentlicht: (2022)
von: Rizzello, Valentina, et al.
Veröffentlicht: (2022)
Reinforcement Learning from Multi-level and Episodic Human Feedback
von: Elahi, Muhammad Qasim, et al.
Veröffentlicht: (2025)
von: Elahi, Muhammad Qasim, et al.
Veröffentlicht: (2025)
Reinforcement Learning from Human Feedback: A Statistical Perspective
von: Liu, Pangpang, et al.
Veröffentlicht: (2026)
von: Liu, Pangpang, et al.
Veröffentlicht: (2026)
Dual Active Learning for Reinforcement Learning from Human Feedback
von: Liu, Pangpang, et al.
Veröffentlicht: (2024)
von: Liu, Pangpang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
BrainCodec: Neural fMRI codec for the decoding of cognitive brain states
von: Nishimura, Yuto, et al.
Veröffentlicht: (2024) -
Model Metamers Reveal Invariances in Graph Neural Networks
von: Xu, Wei, et al.
Veröffentlicht: (2025) -
Prediction of Alpha Power Using Multiple Subjective Measures and Autonomic Responses
von: Yuting Xu, et al.
Veröffentlicht: (2025) -
Data-dependent Exploration for Online Reinforcement Learning from Human Feedback
von: Zhang, Zhen-Yu, et al.
Veröffentlicht: (2026) -
Sampling Method for Generalized Graph Signals with Pre-selected Vertices via DC Optimization
von: Yamashita, Keitaro, et al.
Veröffentlicht: (2025)