SemLT3D: Semantic-Guided Expert Distillation for Camera-only Long-Tailed 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Vo, Hao, Vo, Khoa, Phan, Thinh, Cuong, Ngo Xuan, Doretto, Gianfranco, Nguyen, Hien, Nguyen, Anh, Le, Ngan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Z-GMOT: Zero-shot Generic Multiple Object Tracking
by: Tran, Kim Hoang, et al.
Published: (2023)
by: Tran, Kim Hoang, et al.
Published: (2023)
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model
by: Vo, Khoa, et al.
Published: (2024)
by: Vo, Khoa, et al.
Published: (2024)
ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
Amodal Instance Segmentation with Diffusion Shape Prior Estimation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling
by: Pham, Trong-Thang, et al.
Published: (2025)
by: Pham, Trong-Thang, et al.
Published: (2025)
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation
by: Pham, Trong Thang, et al.
Published: (2024)
by: Pham, Trong Thang, et al.
Published: (2024)
Synthetic Swarm Mosquito Dataset for Acoustic Classification: A Proof of Concept
by: Dinh, Thai-Duy, et al.
Published: (2025)
by: Dinh, Thai-Duy, et al.
Published: (2025)
VlogQA: Task, Dataset, and Baseline Models for Vietnamese Spoken-Based Machine Reading Comprehension
by: Ngo, Thinh Phuoc, et al.
Published: (2024)
by: Ngo, Thinh Phuoc, et al.
Published: (2024)
DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving
by: Vo, Hao, et al.
Published: (2026)
by: Vo, Hao, et al.
Published: (2026)
Twin electroweak bubble nucleation and gravitational wave under the $S_3$ symmetry of two-Higgs-doublet model
by: Phong, Vo Quoc, et al.
Published: (2024)
by: Phong, Vo Quoc, et al.
Published: (2024)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
by: Nguyen, Phuc, et al.
Published: (2024)
by: Nguyen, Phuc, et al.
Published: (2024)
Early Outcomes of Totally Endoscopic Repair of Partial Atrioventricular Septal Defect in Adults Using A 3D Visualization System
by: Nguyen Sinh Hien, et al.
Published: (2026)
by: Nguyen Sinh Hien, et al.
Published: (2026)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
by: Nguyen, Phuc D. A., et al.
Published: (2023)
by: Nguyen, Phuc D. A., et al.
Published: (2023)
Enhanced Multimodal Video Retrieval System: Integrating Query Expansion and Cross-modal Temporal Event Retrieval
by: Vo, Van-Thinh, et al.
Published: (2025)
by: Vo, Van-Thinh, et al.
Published: (2025)
Language-driven Grasp Detection with Mask-guided Attention
by: Van Vo, Tuan, et al.
Published: (2024)
by: Van Vo, Tuan, et al.
Published: (2024)
Rethinking Progression of Memory State in Robotic Manipulation: An Object-Centric Perspective
by: Chung, Nhat, et al.
Published: (2025)
by: Chung, Nhat, et al.
Published: (2025)
SAMURAI: Shape-Aware Multimodal Retrieval for 3D Object Identification
by: Vo, Dinh-Khoi, et al.
Published: (2025)
by: Vo, Dinh-Khoi, et al.
Published: (2025)
Balancing Knowledge Distillation for Imbalance Learning with Bilevel Optimization
by: Nguyen, Anh B. H., et al.
Published: (2026)
by: Nguyen, Anh B. H., et al.
Published: (2026)
MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
by: Pham, Trong-Thang, et al.
Published: (2026)
by: Pham, Trong-Thang, et al.
Published: (2026)
Geospatial Data Clustering in Network Space: A Survey
by: Loan T. T. Nguyen, et al.
Published: (2025)
by: Loan T. T. Nguyen, et al.
Published: (2025)
Blurry-Consistency Segmentation Framework with Selective Stacking on Differential Interference Contrast 3D Breast Cancer Spheroid
by: Nguyen, Thanh-Huy, et al.
Published: (2024)
by: Nguyen, Thanh-Huy, et al.
Published: (2024)
Global Englishes in English Language Teaching: Evaluating Coursebooks Used in an Undergraduate Program in Vietnam
by: Nu Anh Vo, et al.
Published: (2025)
by: Nu Anh Vo, et al.
Published: (2025)
Neural Network‐Driven Adaptive Control for Swing‐Free Trajectory Tracking in Double‐Link Overhead Cranes With Uncertain Dynamics
by: Manh Cuong Nguyen, et al.
Published: (2026)
by: Manh Cuong Nguyen, et al.
Published: (2026)
Lightweight Language-driven Grasp Detection using Conditional Consistency Model
by: Nguyen, Nghia, et al.
Published: (2024)
by: Nguyen, Nghia, et al.
Published: (2024)
CodeGraphVLP: Code-as-Planner Meets Semantic-Graph State for Non-Markovian Vision-Language-Action Models
by: Vo, Khoa, et al.
Published: (2026)
by: Vo, Khoa, et al.
Published: (2026)
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation
by: Hanyu, Taisei, et al.
Published: (2025)
by: Hanyu, Taisei, et al.
Published: (2025)
Tractable Approximation of Labeled Multi-Object Posterior Densities
by: Nguyen, Thi Hong Thai, et al.
Published: (2025)
by: Nguyen, Thi Hong Thai, et al.
Published: (2025)
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
by: Nguyen, Phuc D. A., et al.
Published: (2024)
by: Nguyen, Phuc D. A., et al.
Published: (2024)
Sustainability reporting quality and firm value in ASEAN+3: A series moderation model
by: Hien Vo Van, et al.
Published: (2024)
by: Hien Vo Van, et al.
Published: (2024)
Evaluating the accuracy of DFT functionals applying for determining geometries of coumarin derivatives using the averaged Cartesian atomic coordinates
by: Nguyen Khoa Hien, et al.
Published: (2025)
by: Nguyen Khoa Hien, et al.
Published: (2025)
The Mean of Multi-Object Trajectories
by: Nguyen, Tran Thien Dat, et al.
Published: (2025)
by: Nguyen, Tran Thien Dat, et al.
Published: (2025)
Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding
by: Vo, Khoa, et al.
Published: (2025)
by: Vo, Khoa, et al.
Published: (2025)
Learning Human Motion with Temporally Conditional Mamba
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
by: Nguyen, Thinh, et al.
Published: (2025)
by: Nguyen, Thinh, et al.
Published: (2025)
PoC-Adapt: Semantic-Aware Automated Vulnerability Reproduction with LLM Multi-Agents and Reinforcement Learning-Driven Adaptive Policy
by: Duy, Phan The, et al.
Published: (2026)
by: Duy, Phan The, et al.
Published: (2026)
Vietnam's Resilient Diplomacy: Navigating Global Shifts in the Post‐ COVID Era
by: Nguyễn Anh Cường
Published: (2026)
by: Nguyễn Anh Cường
Published: (2026)
Track Initialization and Re-Identification for~3D Multi-View Multi-Object Tracking
by: Van Ma, Linh, et al.
Published: (2024)
by: Van Ma, Linh, et al.
Published: (2024)
MMA: A Momentum Mamba Architecture for Human Activity Recognition with Inertial Sensors
by: Nguyen, Thai-Khanh, et al.
Published: (2025)
by: Nguyen, Thai-Khanh, et al.
Published: (2025)
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing
by: Nguyen, Trong-Tung, et al.
Published: (2024)
by: Nguyen, Trong-Tung, et al.
Published: (2024)
SAM3D: Segment Anything Model in Volumetric Medical Images
by: Bui, Nhat-Tan, et al.
Published: (2023)
by: Bui, Nhat-Tan, et al.
Published: (2023)
Similar Items
-
Z-GMOT: Zero-shot Generic Multiple Object Tracking
by: Tran, Kim Hoang, et al.
Published: (2023) -
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model
by: Vo, Khoa, et al.
Published: (2024) -
ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation
by: Tran, Minh, et al.
Published: (2024) -
Amodal Instance Segmentation with Diffusion Shape Prior Estimation
by: Tran, Minh, et al.
Published: (2024) -
CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling
by: Pham, Trong-Thang, et al.
Published: (2025)