IGen: Scalable Data Generation for Robot Learning from Open-World Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gu, Chenghao, Kang, Haolan, Lin, Junchao, Wang, Jinghe, Wu, Duo, Xie, Shuzhao, Huang, Fanding, Ge, Junchen, Gong, Ziyang, Li, Letian, Zheng, Hongying, Lv, Changwei, Wang, Zhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tuning-Free Visual Customization via View Iterative Self-Attention Control
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
VGGT-DP: Generalizable Robot Control via Vision Foundation Models
von: Ge, Shijia, et al.
Veröffentlicht: (2025)
von: Ge, Shijia, et al.
Veröffentlicht: (2025)
IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
von: Tang, Yinghao, et al.
Veröffentlicht: (2026)
von: Tang, Yinghao, et al.
Veröffentlicht: (2026)
Lungmix: A Mixup-Based Strategy for Generalization in Respiratory Sound Classification
von: Ge, Shijia, et al.
Veröffentlicht: (2024)
von: Ge, Shijia, et al.
Veröffentlicht: (2024)
Robotic Visual Instruction
von: Li, Yanbang, et al.
Veröffentlicht: (2025)
von: Li, Yanbang, et al.
Veröffentlicht: (2025)
Enhancing Implicit Neural Representations via Symmetric Power Transformation
von: Zhang, Weixiang, et al.
Veröffentlicht: (2024)
von: Zhang, Weixiang, et al.
Veröffentlicht: (2024)
CATP-LLM: Empowering Large Language Models for Cost-Aware Tool Planning
von: Wu, Duo, et al.
Veröffentlicht: (2024)
von: Wu, Duo, et al.
Veröffentlicht: (2024)
Learning with Challenges: Adaptive Difficulty-Aware Data Generation for Mobile GUI Agent Training
von: Kang, Linjia, et al.
Veröffentlicht: (2026)
von: Kang, Linjia, et al.
Veröffentlicht: (2026)
Expansive Supervision for Neural Radiance Field
von: Zhang, Weixiang, et al.
Veröffentlicht: (2024)
von: Zhang, Weixiang, et al.
Veröffentlicht: (2024)
SD-GS: Structured Deformable 3D Gaussians for Efficient Dynamic Scene Reconstruction
von: Yao, Wei, et al.
Veröffentlicht: (2025)
von: Yao, Wei, et al.
Veröffentlicht: (2025)
OpenNav: Open-World Navigation with Multimodal Large Language Models
von: Yuan, Mingfeng, et al.
Veröffentlicht: (2025)
von: Yuan, Mingfeng, et al.
Veröffentlicht: (2025)
Collaborative Belief Reasoning with LLMs for Efficient Multi-Agent Collaboration
von: Wang, Zhimin, et al.
Veröffentlicht: (2025)
von: Wang, Zhimin, et al.
Veröffentlicht: (2025)
EVOS: Efficient Implicit Neural Training via EVOlutionary Selector
von: Zhang, Weixiang, et al.
Veröffentlicht: (2024)
von: Zhang, Weixiang, et al.
Veröffentlicht: (2024)
MesonGS: Post-training Compression of 3D Gaussians via Efficient Attribute Transformation
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
Multimodal Adversarial Quality Policy for Safe Grasping
von: Xie, Kunlin, et al.
Veröffentlicht: (2026)
von: Xie, Kunlin, et al.
Veröffentlicht: (2026)
MesonGS++: Post-training Compression of 3D Gaussian Splatting with Hyperparameter Searching
von: Xie, Shuzhao, et al.
Veröffentlicht: (2026)
von: Xie, Shuzhao, et al.
Veröffentlicht: (2026)
LEGO‐like Origami Robots Standardize Structure Design of Soft Robots
von: Zheng Wang, et al.
Veröffentlicht: (2025)
von: Zheng Wang, et al.
Veröffentlicht: (2025)
English Medium Instruction as a Local Practice
von: Han, Jinghe
Veröffentlicht: (2022)
von: Han, Jinghe
Veröffentlicht: (2022)
Knowledge Distillation for Underwater Feature Extraction and Matching via GAN-synthesized Images
von: Yang, Jinghe, et al.
Veröffentlicht: (2025)
von: Yang, Jinghe, et al.
Veröffentlicht: (2025)
A Joint Approach to Local Updating and Gradient Compression for Efficient Asynchronous Federated Learning
von: Song, Jiajun, et al.
Veröffentlicht: (2024)
von: Song, Jiajun, et al.
Veröffentlicht: (2024)
NetLLM: Adapting Large Language Models for Networking
von: Wu, Duo, et al.
Veröffentlicht: (2024)
von: Wu, Duo, et al.
Veröffentlicht: (2024)
WATER-GS: Toward Copyright Protection for 3D Gaussian Splatting via Universal Watermarking
von: Tan, Yuqi, et al.
Veröffentlicht: (2024)
von: Tan, Yuqi, et al.
Veröffentlicht: (2024)
KAN We Flow? Advancing Robotic Manipulation with 3D Flow Matching via KAN & RWKV
von: Chen, Zhihao, et al.
Veröffentlicht: (2026)
von: Chen, Zhihao, et al.
Veröffentlicht: (2026)
Grounding Large Language Models In Embodied Environment With Imperfect World Models
von: Liu, Haolan, et al.
Veröffentlicht: (2024)
von: Liu, Haolan, et al.
Veröffentlicht: (2024)
SR-SLAM: Scene-reliability Based RGB-D SLAM in Diverse Environments
von: Zhang, Haolan, et al.
Veröffentlicht: (2025)
von: Zhang, Haolan, et al.
Veröffentlicht: (2025)
SmoothTurn: Learning to Turn Smoothly for Agile Navigation with Quadrupedal Robots
von: You, Zunzhi, et al.
Veröffentlicht: (2026)
von: You, Zunzhi, et al.
Veröffentlicht: (2026)
OpenViewer: Openness-Aware Multi-View Learning
von: Du, Shide, et al.
Veröffentlicht: (2024)
von: Du, Shide, et al.
Veröffentlicht: (2024)
Adaptive Prior Scene-Object SLAM for Dynamic Environments
von: Zhang, Haolan, et al.
Veröffentlicht: (2025)
von: Zhang, Haolan, et al.
Veröffentlicht: (2025)
The Discursive Construction of Literature Review: An Examination of Chinese PhD Students' Information Behavior
von: Liao, Jiadong, et al.
Veröffentlicht: (2012)
von: Liao, Jiadong, et al.
Veröffentlicht: (2012)
SizeGS: Size-aware Compression of 3D Gaussian Splatting via Mixed Integer Programming
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
WeatherFormer: Empowering Global Numerical Weather Forecasting with Space-Time Transformer
von: Gong, Junchao, et al.
Veröffentlicht: (2024)
von: Gong, Junchao, et al.
Veröffentlicht: (2024)
RoboMatrix: A Skill-centric Hierarchical Framework for Scalable Robot Task Planning and Execution in Open-World
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
Learning Primitive Embodied World Models: Towards Scalable Robotic Learning
von: Sun, Qiao, et al.
Veröffentlicht: (2025)
von: Sun, Qiao, et al.
Veröffentlicht: (2025)
DUViN: Diffusion-Based Underwater Visual Navigation via Knowledge-Transferred Depth Features
von: Yang, Jinghe, et al.
Veröffentlicht: (2025)
von: Yang, Jinghe, et al.
Veröffentlicht: (2025)
RoboStream: Weaving Spatio-Temporal Reasoning with Memory in Vision-Language Models for Robotics
von: Huang, Yuzhi, et al.
Veröffentlicht: (2026)
von: Huang, Yuzhi, et al.
Veröffentlicht: (2026)
COSMIC: Clique-Oriented Semantic Multi-space Integration for Robust CLIP Test-Time Adaptation
von: Huang, Fanding, et al.
Veröffentlicht: (2025)
von: Huang, Fanding, et al.
Veröffentlicht: (2025)
Music-Aligned Holistic 3D Dance Generation via Hierarchical Motion Modeling
von: Li, Xiaojie, et al.
Veröffentlicht: (2025)
von: Li, Xiaojie, et al.
Veröffentlicht: (2025)
Retraining-free Model Quantization via One-Shot Weight-Coupling Learning
von: Tang, Chen, et al.
Veröffentlicht: (2024)
von: Tang, Chen, et al.
Veröffentlicht: (2024)
DTBS: Dual-Teacher Bi-directional Self-training for Domain Adaptation in Nighttime Semantic Segmentation
von: Huang, Fanding, et al.
Veröffentlicht: (2024)
von: Huang, Fanding, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Tuning-Free Visual Customization via View Iterative Self-Attention Control
von: Li, Xiaojie, et al.
Veröffentlicht: (2024) -
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
von: Gu, Chenghao, et al.
Veröffentlicht: (2024) -
VGGT-DP: Generalizable Robot Control via Vision Foundation Models
von: Ge, Shijia, et al.
Veröffentlicht: (2025) -
IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
von: Tang, Yinghao, et al.
Veröffentlicht: (2026) -
Lungmix: A Mixup-Based Strategy for Generalization in Respiratory Sound Classification
von: Ge, Shijia, et al.
Veröffentlicht: (2024)