KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cherepanov, Egor, Zelezetsky, Daniil, Kovalev, Alexey K., Panov, Aleksandr I. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Re:Frame -- Retrieving Experience From Associative Memory
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
Symbolic Disentangled Representations for Images
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)
Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding
von: Patratskiy, Maxim A., et al.
Veröffentlicht: (2025)
von: Patratskiy, Maxim A., et al.
Veröffentlicht: (2025)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025)
Recurrent Action Transformer with Memory
von: Cherepanov, Egor, et al.
Veröffentlicht: (2023)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2023)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)
von: Cherepanov, Egor, et al.
Veröffentlicht: (2024)
Accelerating Transformers in Online RL
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025)
MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning
von: Nesterova, Maria, et al.
Veröffentlicht: (2026)
von: Nesterova, Maria, et al.
Veröffentlicht: (2026)
Visual Grounding for Object-Level Generalization in Reinforcement Learning
von: Jiang, Haobin, et al.
Veröffentlicht: (2024)
von: Jiang, Haobin, et al.
Veröffentlicht: (2024)
Exploring adversarial robustness of JPEG AI: methodology, comparison and new methods
von: Kovalev, Egor, et al.
Veröffentlicht: (2024)
von: Kovalev, Egor, et al.
Veröffentlicht: (2024)
Enhancing Radiology Report Generation and Visual Grounding using Reinforcement Learning
von: Gundersen, Benjamin, et al.
Veröffentlicht: (2025)
von: Gundersen, Benjamin, et al.
Veröffentlicht: (2025)
Redefining Generalization in Visual Domains: A Two-Axis Framework for Fake Image Detection with FusionDetect
von: Amanzadi, Amirtaha, et al.
Veröffentlicht: (2025)
von: Amanzadi, Amirtaha, et al.
Veröffentlicht: (2025)
VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots
von: Grigorev, Danil S., et al.
Veröffentlicht: (2025)
von: Grigorev, Danil S., et al.
Veröffentlicht: (2025)
ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners?
von: Han, Haonan, et al.
Veröffentlicht: (2026)
von: Han, Haonan, et al.
Veröffentlicht: (2026)
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
von: Jia, Zi-Yi, et al.
Veröffentlicht: (2026)
von: Jia, Zi-Yi, et al.
Veröffentlicht: (2026)
MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence
von: Chen, Yifan, et al.
Veröffentlicht: (2026)
von: Chen, Yifan, et al.
Veröffentlicht: (2026)
Dream to Generalize: Zero-Shot Model-Based Reinforcement Learning for Unseen Visual Distractions
von: Ha, Jeongsoo, et al.
Veröffentlicht: (2025)
von: Ha, Jeongsoo, et al.
Veröffentlicht: (2025)
VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents
von: Yi, Hongzhu, et al.
Veröffentlicht: (2026)
von: Yi, Hongzhu, et al.
Veröffentlicht: (2026)
Leveraging Vision-Language Models for Visual Grounding and Analysis of Automotive UI
von: Ernhofer, Benjamin Raphael, et al.
Veröffentlicht: (2025)
von: Ernhofer, Benjamin Raphael, et al.
Veröffentlicht: (2025)
Relational Retrieval: Leveraging Known-Novel Interactions for Generalized Category Discovery
von: Xu, Yulin, et al.
Veröffentlicht: (2026)
von: Xu, Yulin, et al.
Veröffentlicht: (2026)
Poivre: Self-Refining Visual Pointing with Reinforcement Learning
von: Yang, Wenjie, et al.
Veröffentlicht: (2025)
von: Yang, Wenjie, et al.
Veröffentlicht: (2025)
RELO: Reinforcement Learning to Localize for Visual Object Tracking
von: Chen, Xin, et al.
Veröffentlicht: (2026)
von: Chen, Xin, et al.
Veröffentlicht: (2026)
Object-level Visual Prompts for Compositional Image Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
Intelligent Known and Novel Aircraft Recognition -- A Shift from Classification to Similarity Learning for Combat Identification
von: Saeed, Ahmad, et al.
Veröffentlicht: (2024)
von: Saeed, Ahmad, et al.
Veröffentlicht: (2024)
TurtleBench: A Visual Programming Benchmark in Turtle Geometry
von: Rismanchian, Sina, et al.
Veröffentlicht: (2024)
von: Rismanchian, Sina, et al.
Veröffentlicht: (2024)
A Comprehensive Survey of Data Augmentation in Visual Reinforcement Learning
von: Ma, Guozheng, et al.
Veröffentlicht: (2022)
von: Ma, Guozheng, et al.
Veröffentlicht: (2022)
Saliency-Guided Representation with Consistency Policy Learning for Visual Unsupervised Reinforcement Learning
von: Sun, Jingbo, et al.
Veröffentlicht: (2026)
von: Sun, Jingbo, et al.
Veröffentlicht: (2026)
AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sanjoy, et al.
Veröffentlicht: (2025)
VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation
von: Jiang, Longteng, et al.
Veröffentlicht: (2026)
von: Jiang, Longteng, et al.
Veröffentlicht: (2026)
Neural network task specialization via domain constraining
von: Malashin, Roman, et al.
Veröffentlicht: (2025)
von: Malashin, Roman, et al.
Veröffentlicht: (2025)
DiffuSyn Bench: Evaluating Vision-Language Models on Real-World Complexities with Diffusion-Generated Synthetic Benchmarks
von: Zhou, Haokun, et al.
Veröffentlicht: (2024)
von: Zhou, Haokun, et al.
Veröffentlicht: (2024)
CompareBench: A Benchmark for Visual Comparison Reasoning in Vision-Language Models
von: Cai, Jie, et al.
Veröffentlicht: (2025)
von: Cai, Jie, et al.
Veröffentlicht: (2025)
VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)
A-Bench: Are LMMs Masters at Evaluating AI-generated Images?
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
LocateBench: Evaluating the Locating Ability of Vision Language Models
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024)
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Re:Frame -- Retrieving Experience From Associative Memory
von: Zelezetsky, Daniil, et al.
Veröffentlicht: (2025) -
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
von: Kachaev, Nikita, et al.
Veröffentlicht: (2025) -
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023) -
Symbolic Disentangled Representations for Images
von: Korchemnyi, Alexandr, et al.
Veröffentlicht: (2024) -
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
von: Shchendrigin, Oleg, et al.
Veröffentlicht: (2026)