OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Lingdong, Liu, Youquan, Ng, Lai Xing, Cottereau, Benoit R., Ooi, Wei Tsang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EventFly: Event Camera Perception from Ground to the Sky
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
Talk2Event: Grounded Understanding of Dynamic Scenes from Event Cameras
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
Visual Grounding from Event Cameras
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
Perspective-Invariant 3D Object Detection
by: Liang, Ao, et al.
Published: (2025)
by: Liang, Ao, et al.
Published: (2025)
FlexEvent: Towards Flexible Event-Frame Object Detection at Varying Operational Frequencies
by: Lu, Dongyue, et al.
Published: (2024)
by: Lu, Dongyue, et al.
Published: (2024)
Learning to Generate 4D LiDAR Sequences
by: Liang, Ao, et al.
Published: (2025)
by: Liang, Ao, et al.
Published: (2025)
LiDARCrafter: Dynamic 4D World Modeling from LiDAR Sequences
by: Liang, Ao, et al.
Published: (2025)
by: Liang, Ao, et al.
Published: (2025)
Learning to Remove Lens Flare in Event Camera
by: Han, Haiqian, et al.
Published: (2025)
by: Han, Haiqian, et al.
Published: (2025)
Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving
by: Kong, Lingdong, et al.
Published: (2024)
by: Kong, Lingdong, et al.
Published: (2024)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024)
by: Li, Rong, et al.
Published: (2024)
Is Your Driving World Model an All-Around Player?
by: Kong, Lingdong, et al.
Published: (2026)
by: Kong, Lingdong, et al.
Published: (2026)
Open-Vocabulary Online Semantic Mapping for SLAM
by: Martins, Tomas Berriel, et al.
Published: (2024)
by: Martins, Tomas Berriel, et al.
Published: (2024)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
by: Jiang, Haochen, et al.
Published: (2024)
by: Jiang, Haochen, et al.
Published: (2024)
E-TIDE: Fast, Structure-Preserving Motion Forecasting from Event Sequences
by: Sen, Biswadeep, et al.
Published: (2026)
by: Sen, Biswadeep, et al.
Published: (2026)
Monocular Semantic Scene Completion via Masked Recurrent Networks
by: Wang, Xuzhi, et al.
Published: (2025)
by: Wang, Xuzhi, et al.
Published: (2025)
Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis
by: Xu, Xiang, et al.
Published: (2026)
by: Xu, Xiang, et al.
Published: (2026)
Is Your LiDAR Placement Optimized for 3D Scene Understanding?
by: Li, Ye, et al.
Published: (2024)
by: Li, Ye, et al.
Published: (2024)
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding
by: Wu, Yanmin, et al.
Published: (2024)
by: Wu, Yanmin, et al.
Published: (2024)
SPIRAL: Semantic-Aware Progressive LiDAR Scene Generation and Understanding
by: Zhu, Dekai, et al.
Published: (2025)
by: Zhu, Dekai, et al.
Published: (2025)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
by: Jiang, Jiajun, et al.
Published: (2025)
by: Jiang, Jiajun, et al.
Published: (2025)
The Bare Necessities: Designing Simple, Effective Open-Vocabulary Scene Graphs
by: Kassab, Christina, et al.
Published: (2024)
by: Kassab, Christina, et al.
Published: (2024)
OpenLex3D: A Tiered Evaluation Benchmark for Open-Vocabulary 3D Scene Representations
by: Kassab, Christina, et al.
Published: (2025)
by: Kassab, Christina, et al.
Published: (2025)
Open Vocabulary Semantic Scene Sketch Understanding
by: Bourouis, Ahmed, et al.
Published: (2023)
by: Bourouis, Ahmed, et al.
Published: (2023)
U4D: Uncertainty-Aware 4D World Modeling from LiDAR Sequences
by: Xu, Xiang, et al.
Published: (2025)
by: Xu, Xiang, et al.
Published: (2025)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026)
by: Hu, Xinggang, et al.
Published: (2026)
La La LiDAR: Large-Scale Layout Generation from LiDAR Data
by: Liu, Youquan, et al.
Published: (2025)
by: Liu, Youquan, et al.
Published: (2025)
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
by: Qiu, Dicong, et al.
Published: (2025)
by: Qiu, Dicong, et al.
Published: (2025)
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
by: Deng, Yinan, et al.
Published: (2024)
by: Deng, Yinan, et al.
Published: (2024)
SpikeCLR: Contrastive Self-Supervised Learning for Few-Shot Event-Based Vision using Spiking Neural Networks
by: Vaillant, Maxime, et al.
Published: (2026)
by: Vaillant, Maxime, et al.
Published: (2026)
LOVON: Legged Open-Vocabulary Object Navigator
by: Peng, Daojie, et al.
Published: (2025)
by: Peng, Daojie, et al.
Published: (2025)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
by: Wang, Zhaowei, et al.
Published: (2024)
by: Wang, Zhaowei, et al.
Published: (2024)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
by: Ishaq, Ayesha, et al.
Published: (2024)
by: Ishaq, Ayesha, et al.
Published: (2024)
Calib3D: Calibrating Model Preferences for Reliable 3D Scene Understanding
by: Kong, Lingdong, et al.
Published: (2024)
by: Kong, Lingdong, et al.
Published: (2024)
DynamicCity: Large-Scale 4D Occupancy Generation from Dynamic Scenes
by: Bian, Hengwei, et al.
Published: (2024)
by: Bian, Hengwei, et al.
Published: (2024)
GROVE: A Generalized Reward for Learning Open-Vocabulary Physical Skill
by: Cui, Jieming, et al.
Published: (2025)
by: Cui, Jieming, et al.
Published: (2025)
WildOS: Open-Vocabulary Object Search in the Wild
by: Shah, Hardik, et al.
Published: (2026)
by: Shah, Hardik, et al.
Published: (2026)
SNOW: Spatio-Temporal Scene Understanding with World Knowledge for Open-World Embodied Reasoning
by: Sohn, Tin Stribor, et al.
Published: (2025)
by: Sohn, Tin Stribor, et al.
Published: (2025)
OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies
by: Chen, Runnan, et al.
Published: (2024)
by: Chen, Runnan, et al.
Published: (2024)
OpenSGA: Efficient 3D Scene Graph Alignment in the Open World
by: Chen, Gang, et al.
Published: (2026)
by: Chen, Gang, et al.
Published: (2026)
3EED: Ground Everything Everywhere in 3D
by: Li, Rong, et al.
Published: (2025)
by: Li, Rong, et al.
Published: (2025)
Similar Items
-
EventFly: Event Camera Perception from Ground to the Sky
by: Kong, Lingdong, et al.
Published: (2025) -
Talk2Event: Grounded Understanding of Dynamic Scenes from Event Cameras
by: Kong, Lingdong, et al.
Published: (2025) -
Visual Grounding from Event Cameras
by: Kong, Lingdong, et al.
Published: (2025) -
Perspective-Invariant 3D Object Detection
by: Liang, Ao, et al.
Published: (2025) -
FlexEvent: Towards Flexible Event-Frame Object Detection at Varying Operational Frequencies
by: Lu, Dongyue, et al.
Published: (2024)