NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Shiyu, Shan, Lianlei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LESV: Language Embedded Sparse Voxel Fusion for Open-Vocabulary 3D Scene Understanding
by: Wang, Fusang, et al.
Published: (2026)
by: Wang, Fusang, et al.
Published: (2026)
VIAFormer: Voxel-Image Alignment Transformer for High-Fidelity Voxel Refinement
by: Fang, Tiancheng, et al.
Published: (2026)
by: Fang, Tiancheng, et al.
Published: (2026)
DSA-SRGS: Super-Resolution Gaussian Splatting for Dynamic Sparse-View DSA Reconstruction
by: Zhang, Shiyu, et al.
Published: (2026)
by: Zhang, Shiyu, et al.
Published: (2026)
SoccerNet-v3D: Leveraging Sports Broadcast Replays for 3D Scene Understanding
by: Gutiérrez-Pérez, Marc, et al.
Published: (2025)
by: Gutiérrez-Pérez, Marc, et al.
Published: (2025)
PnLCalib: Sports Field Registration via Points and Lines Optimization
by: Gutiérrez-Pérez, Marc, et al.
Published: (2024)
by: Gutiérrez-Pérez, Marc, et al.
Published: (2024)
M3LEO: A Multi-Modal, Multi-Label Earth Observation Dataset Integrating Interferometric SAR and Multispectral Data
by: Allen, Matthew J, et al.
Published: (2024)
by: Allen, Matthew J, et al.
Published: (2024)
LATTE: Latent Trajectory Embedding for Diffusion-Generated Image Detection
by: Vasilcoiu, Ana, et al.
Published: (2025)
by: Vasilcoiu, Ana, et al.
Published: (2025)
Just a Few Glances: Open-Set Visual Perception with Image Prompt Paradigm
by: Zhang, Jinrong, et al.
Published: (2024)
by: Zhang, Jinrong, et al.
Published: (2024)
Intracoronary Optical Coherence Tomography Image Processing and Vessel Classification Using Machine Learning
by: Lahchim, Amal, et al.
Published: (2026)
by: Lahchim, Amal, et al.
Published: (2026)
IDDR-NGP: Incorporating Detectors for Distractor Removal with Instant Neural Radiance Field
by: Huang, Xianliang, et al.
Published: (2026)
by: Huang, Xianliang, et al.
Published: (2026)
LiqD: A Dynamic Liquid Level Detection Model under Tricky Small Containers
by: Ma, Yukun, et al.
Published: (2024)
by: Ma, Yukun, et al.
Published: (2024)
Manual Labelling Artificially Inflates Deep Learning-Based Segmentation Performance on RGB Images of Closed Canopy: Validation Using TLS
by: Allen, Matthew J., et al.
Published: (2025)
by: Allen, Matthew J., et al.
Published: (2025)
From Latent to Engine Manifolds: Analyzing ImageBind's Multimodal Embedding Space
by: Hamara, Andrew, et al.
Published: (2024)
by: Hamara, Andrew, et al.
Published: (2024)
Boosting Few-Shot Learning with Disentangled Self-Supervised Learning and Meta-Learning for Medical Image Classification
by: Pachetti, Eva, et al.
Published: (2024)
by: Pachetti, Eva, et al.
Published: (2024)
NOCTIS: Novel Object Cyclic Threshold based Instance Segmentation
by: Gandyra, Max, et al.
Published: (2025)
by: Gandyra, Max, et al.
Published: (2025)
JotlasNet: Joint Tensor Low-Rank and Attention-based Sparse Unrolling Network for Accelerating Dynamic MRI
by: Zhang, Yinghao, et al.
Published: (2025)
by: Zhang, Yinghao, et al.
Published: (2025)
Grounding Synthetic Data Generation With Vision and Language Models
by: Çağlar, Ümit Mert, et al.
Published: (2026)
by: Çağlar, Ümit Mert, et al.
Published: (2026)
OnlineSplatter: Pose-Free Online 3D Reconstruction for Free-Moving Objects
by: Huang, Mark He, et al.
Published: (2025)
by: Huang, Mark He, et al.
Published: (2025)
LiftAvatar: Kinematic-Space Completion for Expression-Controlled 3D Gaussian Avatar Animation
by: Wei, Hualiang, et al.
Published: (2026)
by: Wei, Hualiang, et al.
Published: (2026)
VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models
by: Bastien, JF, et al.
Published: (2026)
by: Bastien, JF, et al.
Published: (2026)
A Computer Vision Pipeline for Iterative Bullet Hole Tracking in Rifle Zeroing
by: Belcher, Robert M., et al.
Published: (2026)
by: Belcher, Robert M., et al.
Published: (2026)
RealHD: A High-Quality Dataset for Robust Detection of State-of-the-Art AI-Generated Images
by: Yu, Hanzhe, et al.
Published: (2026)
by: Yu, Hanzhe, et al.
Published: (2026)
I Can't Believe TTA Is Not Better: When Test-Time Augmentation Hurts Medical Image Classification
by: Medeiros, Daniel Nobrega
Published: (2026)
by: Medeiros, Daniel Nobrega
Published: (2026)
A Novel Global Context-aware Deep Neural Network for Enhanced Brain Tumor Segmentation using Magnetic Resonance Images
by: Mukherjee, Sourjya, et al.
Published: (2026)
by: Mukherjee, Sourjya, et al.
Published: (2026)
RPCASSM: Robust PCA State Space Model For Infrared Small Target Detection
by: Liu, Pingping, et al.
Published: (2026)
by: Liu, Pingping, et al.
Published: (2026)
CANSURF: An ASV-View Can Dataset and Benchmark for Detection and Tracking of Surface-Level Debris
by: Aljundi, Zaid, et al.
Published: (2026)
by: Aljundi, Zaid, et al.
Published: (2026)
Detecting AI-Generated Videos with Spiking Neural Networks
by: Jang, Minsuk, et al.
Published: (2026)
by: Jang, Minsuk, et al.
Published: (2026)
Few and Fewer: Learning Better from Few Examples Using Fewer Base Classes
by: Lafargue, Raphael, et al.
Published: (2024)
by: Lafargue, Raphael, et al.
Published: (2024)
Conjuring Positive Pairs for Efficient Unification of Representation Learning and Image Synthesis
by: Estepa, Imanol G., et al.
Published: (2025)
by: Estepa, Imanol G., et al.
Published: (2025)
NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection
by: Chaowakarn, Krittin, et al.
Published: (2025)
by: Chaowakarn, Krittin, et al.
Published: (2025)
Unified Local and Global Attention Interaction Modeling for Vision Transformers
by: Nguyen, Tan, et al.
Published: (2024)
by: Nguyen, Tan, et al.
Published: (2024)
Perceptual Influence: Improving the Perceptual Loss Design for Low-Dose CT Enhancement
by: Viana, Gabriel A., et al.
Published: (2025)
by: Viana, Gabriel A., et al.
Published: (2025)
HY-Himmel Technical Report: Hierarchical Interleaved Multi-stream Motion Encoding for Long Video Understanding
by: Jin, Haopeng, et al.
Published: (2026)
by: Jin, Haopeng, et al.
Published: (2026)
Learnings from Scaling Visual Tokenizers for Reconstruction and Generation
by: Hansen-Estruch, Philippe, et al.
Published: (2025)
by: Hansen-Estruch, Philippe, et al.
Published: (2025)
Mask-Conditioned Voxel Diffusion for Joint Geometry and Color Inpainting
by: Sumuk, Aarya
Published: (2026)
by: Sumuk, Aarya
Published: (2026)
Efficient Temporally-Aware DeepFake Detection using H.264 Motion Vectors
by: Grönquist, Peter, et al.
Published: (2023)
by: Grönquist, Peter, et al.
Published: (2023)
MetaErr: Towards Predicting Error Patterns in Deep Neural Networks
by: Totakura, Varun, et al.
Published: (2026)
by: Totakura, Varun, et al.
Published: (2026)
Semi-Supervised Segmentation via Embedding Matching
by: Xie, Weiyi, et al.
Published: (2024)
by: Xie, Weiyi, et al.
Published: (2024)
CUBIC: Concept Embeddings for Unsupervised Bias Identification using VLMs
by: Méndez, David, et al.
Published: (2025)
by: Méndez, David, et al.
Published: (2025)
Two-step Authentication: Multi-biometric System Using Voice and Facial Recognition
by: Chen, Kuan Wei, et al.
Published: (2026)
by: Chen, Kuan Wei, et al.
Published: (2026)
Similar Items
-
LESV: Language Embedded Sparse Voxel Fusion for Open-Vocabulary 3D Scene Understanding
by: Wang, Fusang, et al.
Published: (2026) -
VIAFormer: Voxel-Image Alignment Transformer for High-Fidelity Voxel Refinement
by: Fang, Tiancheng, et al.
Published: (2026) -
DSA-SRGS: Super-Resolution Gaussian Splatting for Dynamic Sparse-View DSA Reconstruction
by: Zhang, Shiyu, et al.
Published: (2026) -
SoccerNet-v3D: Leveraging Sports Broadcast Replays for 3D Scene Understanding
by: Gutiérrez-Pérez, Marc, et al.
Published: (2025) -
PnLCalib: Sports Field Registration via Points and Lines Optimization
by: Gutiérrez-Pérez, Marc, et al.
Published: (2024)