RepVF: A Unified Vector Fields Representation for Multi-task 3D Perception
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Chunliang, Han, Wencheng, Yin, Junbo, Zhao, Sanyuan, Shen, Jianbing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Depth Adaptive Efficient Visual Autoregressive Modeling
di: Li, Chunliang, et al.
Pubblicazione: (2026)
di: Li, Chunliang, et al.
Pubblicazione: (2026)
High-Precision Self-Supervised Monocular Depth Estimation with Rich-Resource Prior
di: Han, Wencheng, et al.
Pubblicazione: (2024)
di: Han, Wencheng, et al.
Pubblicazione: (2024)
DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving
di: Han, Wencheng, et al.
Pubblicazione: (2024)
di: Han, Wencheng, et al.
Pubblicazione: (2024)
RAWMamba: Unified sRGB-to-RAW De-rendering With State Space Model
di: Chen, Hongjun, et al.
Pubblicazione: (2024)
di: Chen, Hongjun, et al.
Pubblicazione: (2024)
Rethinking the Encoding and Annotating of 3D Bounding Box: Corner-Aware 3D Object Detection from Point Clouds
di: Meng, Qinghao, et al.
Pubblicazione: (2025)
di: Meng, Qinghao, et al.
Pubblicazione: (2025)
ALOcc: Adaptive Lifting-Based 3D Semantic Occupancy and Cost Volume-Based Flow Predictions
di: Chen, Dubing, et al.
Pubblicazione: (2024)
di: Chen, Dubing, et al.
Pubblicazione: (2024)
Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution
di: Zheng, Huan, et al.
Pubblicazione: (2024)
di: Zheng, Huan, et al.
Pubblicazione: (2024)
AdaOcc: Adaptive Forward View Transformation and Flow Modeling for 3D Occupancy and Flow Prediction
di: Chen, Dubing, et al.
Pubblicazione: (2024)
di: Chen, Dubing, et al.
Pubblicazione: (2024)
Towards High-Fidelity 3D Portrait Generation with Rich Details by Cross-View Prior-Aware Diffusion
di: Wei, Haoran, et al.
Pubblicazione: (2024)
di: Wei, Haoran, et al.
Pubblicazione: (2024)
Breaking Down Monocular Ambiguity: Exploiting Temporal Evolution for 3D Lane Detection
di: Zheng, Huan, et al.
Pubblicazione: (2025)
di: Zheng, Huan, et al.
Pubblicazione: (2025)
VF-NeRF: Learning Neural Vector Fields for Indoor Scene Reconstruction
di: Puigjaner, Albert Gassol, et al.
Pubblicazione: (2024)
di: Puigjaner, Albert Gassol, et al.
Pubblicazione: (2024)
Bridging Scene Generation and Planning: Driving with World Model via Unifying Vision and Motion Representation
di: Gui, Xingtai, et al.
Pubblicazione: (2026)
di: Gui, Xingtai, et al.
Pubblicazione: (2026)
Self-Rewarding Large Vision-Language Models for Optimizing Prompts in Text-to-Image Generation
di: Yang, Hongji, et al.
Pubblicazione: (2025)
di: Yang, Hongji, et al.
Pubblicazione: (2025)
DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion Models
di: Yang, Hongji, et al.
Pubblicazione: (2025)
di: Yang, Hongji, et al.
Pubblicazione: (2025)
Towards Geometry-Aware and Motion-Guided Video Human Mesh Recovery
di: Chen, Hongjun, et al.
Pubblicazione: (2026)
di: Chen, Hongjun, et al.
Pubblicazione: (2026)
HanMoVLM: Large Vision-Language Models for Professional Artistic Painting Evaluation
di: Yang, Hongji, et al.
Pubblicazione: (2026)
di: Yang, Hongji, et al.
Pubblicazione: (2026)
IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object Detection
di: Yin, Junbo, et al.
Pubblicazione: (2024)
di: Yin, Junbo, et al.
Pubblicazione: (2024)
Towards Better Cephalometric Landmark Detection with Diffusion Data Generation
di: Guo, Dongqian, et al.
Pubblicazione: (2025)
di: Guo, Dongqian, et al.
Pubblicazione: (2025)
TrajDiff: End-to-end Autonomous Driving without Perception Annotation
di: Gui, Xingtai, et al.
Pubblicazione: (2025)
di: Gui, Xingtai, et al.
Pubblicazione: (2025)
TextFormer: A Query-based End-to-End Text Spotter with Mixed Supervision
di: Zhai, Yukun, et al.
Pubblicazione: (2023)
di: Zhai, Yukun, et al.
Pubblicazione: (2023)
Reducing CT Metal Artifacts by Learning Latent Space Alignment with Gemstone Spectral Imaging Data
di: Han, Wencheng, et al.
Pubblicazione: (2025)
di: Han, Wencheng, et al.
Pubblicazione: (2025)
VF-NeRF: Viewshed Fields for Rigid NeRF Registration
di: Segre, Leo, et al.
Pubblicazione: (2024)
di: Segre, Leo, et al.
Pubblicazione: (2024)
HiCoGen: Hierarchical Compositional Text-to-Image Generation in Diffusion Models via Reinforcement Learning
di: Yang, Hongji, et al.
Pubblicazione: (2025)
di: Yang, Hongji, et al.
Pubblicazione: (2025)
MambaVF: State Space Model for Efficient Video Fusion
di: Zhao, Zixiang, et al.
Pubblicazione: (2026)
di: Zhao, Zixiang, et al.
Pubblicazione: (2026)
DenoiseRep: Denoising Model for Representation Learning
di: Xu, Zhengrui, et al.
Pubblicazione: (2024)
di: Xu, Zhengrui, et al.
Pubblicazione: (2024)
Language Prompt for Autonomous Driving
di: Wu, Dongming, et al.
Pubblicazione: (2023)
di: Wu, Dongming, et al.
Pubblicazione: (2023)
RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation
di: Yan, Tianyi, et al.
Pubblicazione: (2025)
di: Yan, Tianyi, et al.
Pubblicazione: (2025)
DrivingSphere: Building a High-fidelity 4D World for Closed-loop Simulation
di: Yan, Tianyi, et al.
Pubblicazione: (2024)
di: Yan, Tianyi, et al.
Pubblicazione: (2024)
USDRL: Unified Skeleton-Based Dense Representation Learning with Multi-Grained Feature Decorrelation
di: Weng, Wanjiang, et al.
Pubblicazione: (2024)
di: Weng, Wanjiang, et al.
Pubblicazione: (2024)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
di: Zou, Jian, et al.
Pubblicazione: (2023)
di: Zou, Jian, et al.
Pubblicazione: (2023)
LiDARFormer: A Unified Transformer-based Multi-task Network for LiDAR Perception
di: Zhou, Zixiang, et al.
Pubblicazione: (2023)
di: Zhou, Zixiang, et al.
Pubblicazione: (2023)
OLiDM: Object-aware LiDAR Diffusion Models for Autonomous Driving
di: Yan, Tianyi, et al.
Pubblicazione: (2024)
di: Yan, Tianyi, et al.
Pubblicazione: (2024)
SignRep: Enhancing Self-Supervised Sign Representations
di: Wong, Ryan, et al.
Pubblicazione: (2025)
di: Wong, Ryan, et al.
Pubblicazione: (2025)
HENet: Hybrid Encoding for End-to-end Multi-task 3D Perception from Multi-view Cameras
di: Xia, Zhongyu, et al.
Pubblicazione: (2024)
di: Xia, Zhongyu, et al.
Pubblicazione: (2024)
Rep-MTL: Unleashing the Power of Representation-level Task Saliency for Multi-Task Learning
di: Wang, Zedong, et al.
Pubblicazione: (2025)
di: Wang, Zedong, et al.
Pubblicazione: (2025)
Rethinking Temporal Fusion with a Unified Gradient Descent View for 3D Semantic Occupancy Prediction
di: Chen, Dubing, et al.
Pubblicazione: (2025)
di: Chen, Dubing, et al.
Pubblicazione: (2025)
CLIP-GS: Unifying Vision-Language Representation with 3D Gaussian Splatting
di: Jiao, Siyu, et al.
Pubblicazione: (2024)
di: Jiao, Siyu, et al.
Pubblicazione: (2024)
HENet++: Hybrid Encoding and Multi-task Learning for 3D Perception and End-to-end Autonomous Driving
di: Xia, Zhongyu, et al.
Pubblicazione: (2025)
di: Xia, Zhongyu, et al.
Pubblicazione: (2025)
Vector Quantized Feature Fields for Fast 3D Semantic Lifting
di: Tang, George, et al.
Pubblicazione: (2025)
di: Tang, George, et al.
Pubblicazione: (2025)
3DeepRep: 3D Deep Low-rank Tensor Representation for Hyperspectral Image Inpainting
di: Li, Yunshan, et al.
Pubblicazione: (2025)
di: Li, Yunshan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Depth Adaptive Efficient Visual Autoregressive Modeling
di: Li, Chunliang, et al.
Pubblicazione: (2026) -
High-Precision Self-Supervised Monocular Depth Estimation with Rich-Resource Prior
di: Han, Wencheng, et al.
Pubblicazione: (2024) -
DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving
di: Han, Wencheng, et al.
Pubblicazione: (2024) -
RAWMamba: Unified sRGB-to-RAW De-rendering With State Space Model
di: Chen, Hongjun, et al.
Pubblicazione: (2024) -
Rethinking the Encoding and Annotating of 3D Bounding Box: Corner-Aware 3D Object Detection from Point Clouds
di: Meng, Qinghao, et al.
Pubblicazione: (2025)