ESAM++: Efficient Online 3D Perception on the Edge
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Qin, Aggarwal, Lavisha, Bandyopadhyay, Saptarashmi, Bahirwani, Vikas, Niethammer, Marc, Adeli, Ehsan, Colaco, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Videos to Conversations: Egocentric Instructions for Task Assistance
by: Aggarwal, Lavisha, et al.
Published: (2026)
by: Aggarwal, Lavisha, et al.
Published: (2026)
Generating Dialogues from Egocentric Instructional Videos for Task Assistance: Dataset, Method and Benchmark
by: Aggarwal, Lavisha, et al.
Published: (2025)
by: Aggarwal, Lavisha, et al.
Published: (2025)
YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks
by: Bandyopadhyay, Saptarashmi, et al.
Published: (2025)
by: Bandyopadhyay, Saptarashmi, et al.
Published: (2025)
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
by: Tian, Junjiao, et al.
Published: (2023)
by: Tian, Junjiao, et al.
Published: (2023)
Zero-shot Domain Generalization of Foundational Models for 3D Medical Image Segmentation: An Experimental Study
by: Chattopadhyay, Soumitri, et al.
Published: (2025)
by: Chattopadhyay, Soumitri, et al.
Published: (2025)
Towards Robust 3D Pose Transfer with Adversarial Learning
by: Chen, Haoyu, et al.
Published: (2024)
by: Chen, Haoyu, et al.
Published: (2024)
Rethinking Interactive Image Segmentation with Low Latency, High Quality, and Diverse Prompts
by: Liu, Qin, et al.
Published: (2024)
by: Liu, Qin, et al.
Published: (2024)
Repurposing 2D Diffusion Models for 3D Shape Completion
by: He, Yao, et al.
Published: (2025)
by: He, Yao, et al.
Published: (2025)
EgoTrigger: Toward Audio-Driven Image Capture for Human Memory Enhancement in All-Day Energy-Efficient Smart Glasses
by: Paruchuri, Akshay, et al.
Published: (2025)
by: Paruchuri, Akshay, et al.
Published: (2025)
SOE: SO(3)-Equivariant 3D MRI Encoding
by: He, Shizhe, et al.
Published: (2024)
by: He, Shizhe, et al.
Published: (2024)
On The Robustness of Foundational 3D Medical Image Segmentation Models Against Imprecise Visual Prompts
by: Chattopadhyay, Soumitri, et al.
Published: (2026)
by: Chattopadhyay, Soumitri, et al.
Published: (2026)
NFL-BA: Near-Field Light Bundle Adjustment for SLAM in Dynamic Lighting
by: Beltran, Andrea Dunn, et al.
Published: (2024)
by: Beltran, Andrea Dunn, et al.
Published: (2024)
Artist-Created Mesh Generation from Raw Observation
by: He, Yao, et al.
Published: (2025)
by: He, Yao, et al.
Published: (2025)
Efficient Universal Perception Encoder
by: Zhu, Chenchen, et al.
Published: (2026)
by: Zhu, Chenchen, et al.
Published: (2026)
Repurposing 2D Diffusion Models with Gaussian Atlas for 3D Generation
by: Xiang, Tiange, et al.
Published: (2025)
by: Xiang, Tiange, et al.
Published: (2025)
Investigating Demographic Bias in Brain MRI Segmentation: A Comparative Study of Deep-Learning and Non-Deep-Learning Methods
by: Danaee, Ghazal, et al.
Published: (2025)
by: Danaee, Ghazal, et al.
Published: (2025)
Guiding Registration with Emergent Similarity from Pre-Trained Diffusion Models
by: Tursynbek, Nurislam, et al.
Published: (2025)
by: Tursynbek, Nurislam, et al.
Published: (2025)
Exploring Cycle Consistency Learning in Interactive Volume Segmentation
by: Liu, Qin, et al.
Published: (2023)
by: Liu, Qin, et al.
Published: (2023)
LiVOS: Light Video Object Segmentation with Gated Linear Matching
by: Liu, Qin, et al.
Published: (2024)
by: Liu, Qin, et al.
Published: (2024)
A Unified Model for Longitudinal Multi-Modal Multi-View Prediction with Missingness
by: Chen, Boqi, et al.
Published: (2024)
by: Chen, Boqi, et al.
Published: (2024)
PPS-Ctrl: Controllable Sim-to-Real Translation for Colonoscopy Depth Estimation
by: Xiong, Xinqi, et al.
Published: (2025)
by: Xiong, Xinqi, et al.
Published: (2025)
AdaVid: Adaptive Video-Language Pretraining
by: Patel, Chaitanya, et al.
Published: (2025)
by: Patel, Chaitanya, et al.
Published: (2025)
3D Neural Edge Reconstruction
by: Li, Lei, et al.
Published: (2024)
by: Li, Lei, et al.
Published: (2024)
Anatomically Guided Latent Diffusion for Brain MRI Progression Modeling
by: Wan, Cheng, et al.
Published: (2026)
by: Wan, Cheng, et al.
Published: (2026)
AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
by: Li, Hongjie, et al.
Published: (2026)
by: Li, Hongjie, et al.
Published: (2026)
ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body
by: Zhang, Juze, et al.
Published: (2025)
by: Zhang, Juze, et al.
Published: (2025)
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
by: Chen, Changan, et al.
Published: (2024)
by: Chen, Changan, et al.
Published: (2024)
Memory-based Adapters for Online 3D Scene Perception
by: Xu, Xiuwei, et al.
Published: (2024)
by: Xu, Xiuwei, et al.
Published: (2024)
TeDiO: Temporal Diagonal Optimization for Training-Free Coherent Video Diffusion
by: Tursynbek, Nurislam, et al.
Published: (2026)
by: Tursynbek, Nurislam, et al.
Published: (2026)
Integrating Anatomical Priors into a Causal Diffusion Model
by: Li, Binxu, et al.
Published: (2025)
by: Li, Binxu, et al.
Published: (2025)
OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
by: Sun, Adam, et al.
Published: (2024)
by: Sun, Adam, et al.
Published: (2024)
WASABI: A Metric for Evaluating Morphometric Plausibility of Synthetic Brain MRIs
by: Jafrasteh, Bahram, et al.
Published: (2025)
by: Jafrasteh, Bahram, et al.
Published: (2025)
Enforcing Conditional Independence for Fair Representation Learning and Causal Image Generation
by: Hwa, Jensen, et al.
Published: (2024)
by: Hwa, Jensen, et al.
Published: (2024)
GenFusion: Feed-forward Human Performance Capture via Progressive Canonical Space Updates
by: Kwon, Youngjoong, et al.
Published: (2026)
by: Kwon, Youngjoong, et al.
Published: (2026)
PE3R: Perception-Efficient 3D Reconstruction
by: Hu, Jie, et al.
Published: (2025)
by: Hu, Jie, et al.
Published: (2025)
Understanding Model Behavior in Monocular Polyp Sizing
by: Xiong, Xinqi, et al.
Published: (2026)
by: Xiong, Xinqi, et al.
Published: (2026)
NAISR: A 3D Neural Additive Model for Interpretable Shape Representation
by: Jiao, Yining, et al.
Published: (2023)
by: Jiao, Yining, et al.
Published: (2023)
A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding
by: Liu, Christina, et al.
Published: (2025)
by: Liu, Christina, et al.
Published: (2025)
$\texttt{NePhi}$: Neural Deformation Fields for Approximately Diffeomorphic Medical Image Registration
by: Tian, Lin, et al.
Published: (2023)
by: Tian, Lin, et al.
Published: (2023)
Wild2Avatar: Rendering Humans Behind Occlusions
by: Xiang, Tiange, et al.
Published: (2023)
by: Xiang, Tiange, et al.
Published: (2023)
Similar Items
-
From Videos to Conversations: Egocentric Instructions for Task Assistance
by: Aggarwal, Lavisha, et al.
Published: (2026) -
Generating Dialogues from Egocentric Instructional Videos for Task Assistance: Dataset, Method and Benchmark
by: Aggarwal, Lavisha, et al.
Published: (2025) -
YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks
by: Bandyopadhyay, Saptarashmi, et al.
Published: (2025) -
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
by: Tian, Junjiao, et al.
Published: (2023) -
Zero-shot Domain Generalization of Foundational Models for 3D Medical Image Segmentation: An Experimental Study
by: Chattopadhyay, Soumitri, et al.
Published: (2025)