SegDAC: Visual Generalization in Reinforcement Learning via Dynamic Object Tokens
Fuente:
arXiv
Saved in:
| Main Authors: | Brown, Alexandre, Berseth, Glen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SegXAL: Explainable Active Learning for Semantic Segmentation in Driving Scene Scenarios
by: Mandalika, Sriram, et al.
Published: (2024)
by: Mandalika, Sriram, et al.
Published: (2024)
When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?
by: Mu, Tongzhou, et al.
Published: (2024)
by: Mu, Tongzhou, et al.
Published: (2024)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025)
by: Liu, Shaowei, et al.
Published: (2025)
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
by: Bendikas, Rokas, et al.
Published: (2025)
by: Bendikas, Rokas, et al.
Published: (2025)
Learning Equivariant Neural-Augmented Object Dynamics From Few Interactions
by: Orozco, Sergio, et al.
Published: (2026)
by: Orozco, Sergio, et al.
Published: (2026)
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning
by: Huang, Chi-Pin, et al.
Published: (2025)
by: Huang, Chi-Pin, et al.
Published: (2025)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
Adaptive Length Image Tokenization via Recurrent Allocation
by: Duggal, Shivam, et al.
Published: (2024)
by: Duggal, Shivam, et al.
Published: (2024)
Solving Physics Olympiad via Reinforcement Learning on Physics Simulators
by: Prabhudesai, Mihir, et al.
Published: (2026)
by: Prabhudesai, Mihir, et al.
Published: (2026)
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
by: Lee, Seokmin, et al.
Published: (2026)
by: Lee, Seokmin, et al.
Published: (2026)
ObjectReact: Learning Object-Relative Control for Visual Navigation
by: Garg, Sourav, et al.
Published: (2025)
by: Garg, Sourav, et al.
Published: (2025)
Learning Visual Feature-Based World Models via Residual Latent Action
by: Zhang, Xinyu, et al.
Published: (2026)
by: Zhang, Xinyu, et al.
Published: (2026)
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
by: Fekri, Pedram, et al.
Published: (2025)
by: Fekri, Pedram, et al.
Published: (2025)
MAPS: Preserving Vision-Language Representations via Module-Wise Proximity Scheduling for Better Vision-Language-Action Generalization
by: Huang, Chengyue, et al.
Published: (2025)
by: Huang, Chengyue, et al.
Published: (2025)
Learning Human-Humanoid Coordination for Collaborative Object Carrying
by: Du, Yushi, et al.
Published: (2025)
by: Du, Yushi, et al.
Published: (2025)
Temporally Consistent Object-Centric Learning by Contrasting Slots
by: Manasyan, Anna, et al.
Published: (2024)
by: Manasyan, Anna, et al.
Published: (2024)
Learning to Visually Connect Actions and their Effects
by: Parmar, Paritosh, et al.
Published: (2024)
by: Parmar, Paritosh, et al.
Published: (2024)
TTF-VLA: Temporal Token Fusion via Pixel-Attention Integration for Vision-Language-Action Models
by: Liu, Chenghao, et al.
Published: (2025)
by: Liu, Chenghao, et al.
Published: (2025)
A Survey of Embodied Learning for Object-Centric Robotic Manipulation
by: Zheng, Ying, et al.
Published: (2024)
by: Zheng, Ying, et al.
Published: (2024)
Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation
by: Xing, Eliot, et al.
Published: (2024)
by: Xing, Eliot, et al.
Published: (2024)
GraphSeg: Segmented 3D Representations via Graph Edge Addition and Contraction
by: Tang, Haozhan, et al.
Published: (2025)
by: Tang, Haozhan, et al.
Published: (2025)
PH-Dreamer: A Physics-Driven World Model via Port-Hamiltonian Generative Dynamics
by: Luan, Xueyu, et al.
Published: (2026)
by: Luan, Xueyu, et al.
Published: (2026)
Visual Representation Learning with Stochastic Frame Prediction
by: Jang, Huiwon, et al.
Published: (2024)
by: Jang, Huiwon, et al.
Published: (2024)
unMORE: Unsupervised Multi-Object Segmentation via Center-Boundary Reasoning
by: Yang, Yafei, et al.
Published: (2025)
by: Yang, Yafei, et al.
Published: (2025)
GrabS: Generative Embodied Agent for 3D Object Segmentation without Scene Supervision
by: Zhang, Zihui, et al.
Published: (2025)
by: Zhang, Zihui, et al.
Published: (2025)
Reconstruction by Generation: 3D Multi-Object Scene Reconstruction from Sparse Observations
by: Zadaianchuk, Andrii, et al.
Published: (2026)
by: Zadaianchuk, Andrii, et al.
Published: (2026)
HGSFusion: Radar-Camera Fusion with Hybrid Generation and Synchronization for 3D Object Detection
by: Gu, Zijian, et al.
Published: (2024)
by: Gu, Zijian, et al.
Published: (2024)
Learned Visual Navigation for Under-Canopy Agricultural Robots
by: Sivakumar, Arun Narenthiran, et al.
Published: (2021)
by: Sivakumar, Arun Narenthiran, et al.
Published: (2021)
Focus On What Matters: Separated Models For Visual-Based RL Generalization
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
FALCON: Future-Aware Learning with Contextual Object-Centric Pretraining for UAV Action Recognition
by: Xian, Ruiqi, et al.
Published: (2024)
by: Xian, Ruiqi, et al.
Published: (2024)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Subtask-Aware Visual Reward Learning from Segmented Demonstrations
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
PhysInOne: Visual Physics Learning and Reasoning in One Suite
by: Zhou, Siyuan, et al.
Published: (2026)
by: Zhou, Siyuan, et al.
Published: (2026)
Synchronous vs Asynchronous Reinforcement Learning in a Real World Robot
by: Parsaee, Ali, et al.
Published: (2025)
by: Parsaee, Ali, et al.
Published: (2025)
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation
by: Zhou, Zihan, et al.
Published: (2024)
by: Zhou, Zihan, et al.
Published: (2024)
LiDAR-EDIT: LiDAR Data Generation by Editing the Object Layouts in Real-World Scenes
by: Ho, Shing-Hei, et al.
Published: (2024)
by: Ho, Shing-Hei, et al.
Published: (2024)
DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning
by: Jiang, Zhenyu, et al.
Published: (2024)
by: Jiang, Zhenyu, et al.
Published: (2024)
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation
by: Wang, Yufei, et al.
Published: (2023)
by: Wang, Yufei, et al.
Published: (2023)
DeFIX: Detecting and Fixing Failure Scenarios with Reinforcement Learning in Imitation Learning Based Autonomous Driving
by: Dagdanov, Resul, et al.
Published: (2022)
by: Dagdanov, Resul, et al.
Published: (2022)
EC-IoU: Orienting Safety for Object Detectors via Ego-Centric Intersection-over-Union
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2024)
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2024)
Similar Items
-
SegXAL: Explainable Active Learning for Semantic Segmentation in Driving Scene Scenarios
by: Mandalika, Sriram, et al.
Published: (2024) -
When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?
by: Mu, Tongzhou, et al.
Published: (2024) -
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025) -
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
by: Bendikas, Rokas, et al.
Published: (2025) -
Learning Equivariant Neural-Augmented Object Dynamics From Few Interactions
by: Orozco, Sergio, et al.
Published: (2026)