From Scene to Object: Text-Guided Dual-Gaze Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ke, Zehong, Jiang, Yanbo, Li, Jinhao, Liu, Zhiyuan, Tu, Yiqian, Meng, Qingwen, Huang, Heye, Wang, Jianqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
D2E-An Autonomous Decision-making Dataset involving Driver States and Human Evaluation
von: Ke, Zehong, et al.
Veröffentlicht: (2024)
von: Ke, Zehong, et al.
Veröffentlicht: (2024)
Guided Diffusion-based Generation of Adversarial Objects for Real-World Monocular Depth Estimation Attacks
von: Chen, Yongtao, et al.
Veröffentlicht: (2025)
von: Chen, Yongtao, et al.
Veröffentlicht: (2025)
DeFlow: Decoder of Scene Flow Network in Autonomous Driving
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026)
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026)
TeFlow: Enabling Multi-frame Supervision for Self-Supervised Feed-forward Scene Flow Estimation
von: Zhang, Qingwen, et al.
Veröffentlicht: (2026)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2026)
SeFlow: A Self-Supervised Scene Flow Method in Autonomous Driving
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
GSGTrack: Gaussian Splatting-Guided Object Pose Tracking from RGB Videos
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2024)
Scene-Agnostic Traversability Labeling and Estimation via a Multimodal Self-supervised Framework
von: Fang, Zipeng, et al.
Veröffentlicht: (2025)
von: Fang, Zipeng, et al.
Veröffentlicht: (2025)
DriveCode: Domain Specific Numerical Encoding for LLM-Based Autonomous Driving
von: Wang, Zhiye, et al.
Veröffentlicht: (2026)
von: Wang, Zhiye, et al.
Veröffentlicht: (2026)
DeltaFlow: An Efficient Multi-frame Scene Flow Estimation Method
von: Zhang, Qingwen, et al.
Veröffentlicht: (2025)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2025)
Learning to Tune Like an Expert: Interpretable and Scene-Aware Navigation via MLLM Reasoning and CVAE-Based Adaptation
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
InVDriver: Intra-Instance Aware Vectorized Query-Based Autonomous Driving Transformer
von: Zhang, Bo, et al.
Veröffentlicht: (2025)
von: Zhang, Bo, et al.
Veröffentlicht: (2025)
MGTR: Multi-Granular Transformer for Motion Prediction with LiDAR
von: Gan, Yiqian, et al.
Veröffentlicht: (2023)
von: Gan, Yiqian, et al.
Veröffentlicht: (2023)
TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification
von: Tao, Huaqi, et al.
Veröffentlicht: (2025)
von: Tao, Huaqi, et al.
Veröffentlicht: (2025)
Belief Scene Graphs: Expanding Partial Scenes with Objects through Computation of Expectation
von: Saucedo, Mario A. V., et al.
Veröffentlicht: (2024)
von: Saucedo, Mario A. V., et al.
Veröffentlicht: (2024)
DeformGS: Scene Flow in Highly Deformable Scenes for Deformable Object Manipulation
von: Duisterhof, Bardienus P., et al.
Veröffentlicht: (2023)
von: Duisterhof, Bardienus P., et al.
Veröffentlicht: (2023)
Gaze-Guided 3D Hand Motion Prediction for Detecting Intent in Egocentric Grasping Tasks
von: He, Yufei, et al.
Veröffentlicht: (2025)
von: He, Yufei, et al.
Veröffentlicht: (2025)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
HiMo: High-Speed Objects Motion Compensation in Point Clouds
von: Zhang, Qingwen, et al.
Veröffentlicht: (2025)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2025)
From Local Matches to Global Masks: Template-Guided Instance Detection and Segmentation in Open-World Scenes
von: Zhang, Qifan, et al.
Veröffentlicht: (2026)
von: Zhang, Qifan, et al.
Veröffentlicht: (2026)
SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs
von: Zhai, Guangyao, et al.
Veröffentlicht: (2023)
von: Zhai, Guangyao, et al.
Veröffentlicht: (2023)
Zero-shot Reconstruction of In-Scene Object Manipulation from Video
von: Lin, Dixuan, et al.
Veröffentlicht: (2025)
von: Lin, Dixuan, et al.
Veröffentlicht: (2025)
SceneMotion: From Agent-Centric Embeddings to Scene-Wide Forecasts
von: Wagner, Royden, et al.
Veröffentlicht: (2024)
von: Wagner, Royden, et al.
Veröffentlicht: (2024)
APEX: A Decoupled Memory-based Explorer for Asynchronous Aerial Object Goal Navigation
von: Zhang, Daoxuan, et al.
Veröffentlicht: (2026)
von: Zhang, Daoxuan, et al.
Veröffentlicht: (2026)
PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes
von: Abdelreheem, Ahmed, et al.
Veröffentlicht: (2025)
von: Abdelreheem, Ahmed, et al.
Veröffentlicht: (2025)
3D Object Visibility Prediction in Autonomous Driving
von: Luo, Chuanyu, et al.
Veröffentlicht: (2024)
von: Luo, Chuanyu, et al.
Veröffentlicht: (2024)
GEMINUS: Dual-aware Global and Scene-Adaptive Mixture-of-Experts for End-to-End Autonomous Driving
von: Wan, Chi, et al.
Veröffentlicht: (2025)
von: Wan, Chi, et al.
Veröffentlicht: (2025)
SALT: A Flexible Semi-Automatic Labeling Tool for General LiDAR Point Clouds with Cross-Scene Adaptability and 4D Consistency
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents
von: Xiao, Jianqiang, et al.
Veröffentlicht: (2025)
von: Xiao, Jianqiang, et al.
Veröffentlicht: (2025)
Online 3D Scene Reconstruction Using Neural Object Priors
von: Chabal, Thomas, et al.
Veröffentlicht: (2025)
von: Chabal, Thomas, et al.
Veröffentlicht: (2025)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
von: Lee, Andrew, et al.
Veröffentlicht: (2025)
von: Lee, Andrew, et al.
Veröffentlicht: (2025)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
von: Jeong, Seunghoon, et al.
Veröffentlicht: (2026)
von: Jeong, Seunghoon, et al.
Veröffentlicht: (2026)
Memorize What Matters: Emergent Scene Decomposition from Multitraverse
von: Li, Yiming, et al.
Veröffentlicht: (2024)
von: Li, Yiming, et al.
Veröffentlicht: (2024)
LLM-enhanced Scene Graph Learning for Household Rearrangement
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
von: Rim, Patrick, et al.
Veröffentlicht: (2026)
von: Rim, Patrick, et al.
Veröffentlicht: (2026)
MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting
von: Xing, Yining, et al.
Veröffentlicht: (2026)
von: Xing, Yining, et al.
Veröffentlicht: (2026)
MASSTAR: A Multi-Modal and Large-Scale Scene Dataset with a Versatile Toolchain for Surface Prediction and Completion
von: Zheng, Guiyong, et al.
Veröffentlicht: (2024)
von: Zheng, Guiyong, et al.
Veröffentlicht: (2024)
SymDrive: Realistic and Controllable Driving Simulator via Symmetric Auto-regressive Online Restoration
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
D2E-An Autonomous Decision-making Dataset involving Driver States and Human Evaluation
von: Ke, Zehong, et al.
Veröffentlicht: (2024) -
Guided Diffusion-based Generation of Adversarial Objects for Real-World Monocular Depth Estimation Attacks
von: Chen, Yongtao, et al.
Veröffentlicht: (2025) -
DeFlow: Decoder of Scene Flow Network in Autonomous Driving
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024) -
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026) -
TeFlow: Enabling Multi-frame Supervision for Self-Supervised Feed-forward Scene Flow Estimation
von: Zhang, Qingwen, et al.
Veröffentlicht: (2026)