Saved in:
| Main Authors: | Xu, Hongbin, Huang, Junduan, Ma, Yuer, Li, Zifeng, Kang, Wenxiong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.09582 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ControLRM: Fast and Controllable 3D Generation via Large Reconstruction Model
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
SeCG: Semantic-Enhanced 3D Visual Grounding via Cross-modal Graph Attention
by: Xiao, Feng, et al.
Published: (2024)
by: Xiao, Feng, et al.
Published: (2024)
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
by: Guo, Mengqi, et al.
Published: (2025)
by: Guo, Mengqi, et al.
Published: (2025)
B2N3D: Progressive Learning from Binary to N-ary Relationships for 3D Object Grounding
by: Xiao, Feng, et al.
Published: (2025)
by: Xiao, Feng, et al.
Published: (2025)
Generalizable Sensor-Based Activity Recognition via Categorical Concept Invariant Learning
by: Xiong, Di, et al.
Published: (2024)
by: Xiong, Di, et al.
Published: (2024)
StructGS: Adaptive Spherical Harmonics and Rendering Enhancements for Superior 3D Gaussian Splatting
by: Huang, Zexu, et al.
Published: (2025)
by: Huang, Zexu, et al.
Published: (2025)
LP-LLM: End-to-End Real-World Degraded License Plate Text Recognition via Large Multimodal Models
by: Gong, Haoyan, et al.
Published: (2026)
by: Gong, Haoyan, et al.
Published: (2026)
TextSplat: Text-Guided Semantic Fusion for Generalizable Gaussian Splatting
by: Wu, Zhicong, et al.
Published: (2025)
by: Wu, Zhicong, et al.
Published: (2025)
MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
by: Gao, Ruiyuan, et al.
Published: (2024)
by: Gao, Ruiyuan, et al.
Published: (2024)
LSVG: Language-Guided Scene Graphs with 2D-Assisted Multi-Modal Encoding for 3D Visual Grounding
by: Xiao, Feng, et al.
Published: (2025)
by: Xiao, Feng, et al.
Published: (2025)
Generalizable Facial Expression Recognition
by: Zhang, Yuhang, et al.
Published: (2024)
by: Zhang, Yuhang, et al.
Published: (2024)
CoDe-NeRF: Neural Rendering via Dynamic Coefficient Decomposition
by: Xing, Wenpeng, et al.
Published: (2025)
by: Xing, Wenpeng, et al.
Published: (2025)
RenderWorld: World Model with Self-Supervised 3D Label
by: Yan, Ziyang, et al.
Published: (2024)
by: Yan, Ziyang, et al.
Published: (2024)
LinkTo-Anime: A 2D Animation Optical Flow Dataset from 3D Model Rendering
by: Feng, Xiaoyi, et al.
Published: (2025)
by: Feng, Xiaoyi, et al.
Published: (2025)
StyleDyRF: Zero-shot 4D Style Transfer for Dynamic Neural Radiance Fields
by: Xu, Hongbin, et al.
Published: (2024)
by: Xu, Hongbin, et al.
Published: (2024)
Yo'City: Personalized and Boundless 3D Realistic City Scene Generation via Self-Critic Expansion
by: Lu, Keyang, et al.
Published: (2025)
by: Lu, Keyang, et al.
Published: (2025)
Robin3D: Improving 3D Large Language Model via Robust Instruction Tuning
by: Kang, Weitai, et al.
Published: (2024)
by: Kang, Weitai, et al.
Published: (2024)
GARF: Learning Generalizable 3D Reassembly for Real-World Fractures
by: Li, Sihang, et al.
Published: (2025)
by: Li, Sihang, et al.
Published: (2025)
DynaRend: Learning 3D Dynamics via Masked Future Rendering for Robotic Manipulation
by: Tian, Jingyi, et al.
Published: (2025)
by: Tian, Jingyi, et al.
Published: (2025)
GSNeRF: Generalizable Semantic Neural Radiance Fields with Enhanced 3D Scene Understanding
by: Chou, Zi-Ting, et al.
Published: (2024)
by: Chou, Zi-Ting, et al.
Published: (2024)
ProtoGS: Efficient and High-Quality Rendering with 3D Gaussian Prototypes
by: Gao, Zhengqing, et al.
Published: (2025)
by: Gao, Zhengqing, et al.
Published: (2025)
Generalizable Object Re-Identification via Visual In-Context Prompting
by: Huang, Zhizhong, et al.
Published: (2025)
by: Huang, Zhizhong, et al.
Published: (2025)
3D Scene Rendering with Multimodal Gaussian Splatting
by: Gau, Chi-Shiang, et al.
Published: (2026)
by: Gau, Chi-Shiang, et al.
Published: (2026)
Explainable Face Recognition via Improved Localization
by: Shadman, Rashik, et al.
Published: (2025)
by: Shadman, Rashik, et al.
Published: (2025)
3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography
by: Zhu, Weicheng, et al.
Published: (2025)
by: Zhu, Weicheng, et al.
Published: (2025)
PA-FAS: Towards Interpretable and Generalizable Multimodal Face Anti-Spoofing via Path-Augmented Reinforcement Learning
by: Ma, Yingjie, et al.
Published: (2025)
by: Ma, Yingjie, et al.
Published: (2025)
GLEAM: Learning Generalizable Exploration Policy for Active Mapping in Complex 3D Indoor Scenes
by: Chen, Xiao, et al.
Published: (2025)
by: Chen, Xiao, et al.
Published: (2025)
Online Hand Gesture Recognition Using 3D Convolutional Neural Networks
by: Qin, Yinghao, et al.
Published: (2026)
by: Qin, Yinghao, et al.
Published: (2026)
GAME: Learning Multimodal Interactions via Graph Structures for Personality Trait Estimation
by: Wang, Kangsheng, et al.
Published: (2025)
by: Wang, Kangsheng, et al.
Published: (2025)
LVD-GS: Gaussian Splatting SLAM for Dynamic Scenes via Hierarchical Explicit-Implicit Representation Collaboration Rendering
by: Zhu, Wenkai, et al.
Published: (2025)
by: Zhu, Wenkai, et al.
Published: (2025)
SparseOIT: Improving Order-Independent Transparency 3DGS via Active Set Method
by: Yang, Wentao, et al.
Published: (2026)
by: Yang, Wentao, et al.
Published: (2026)
Improving Generalization of Deep Learning for Brain Metastases Segmentation Across Institutions
by: Yang, Yuchen, et al.
Published: (2026)
by: Yang, Yuchen, et al.
Published: (2026)
Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D Data
by: Ma, Zhiyuan, et al.
Published: (2025)
by: Ma, Zhiyuan, et al.
Published: (2025)
Computation-Efficient and Recognition-Friendly 3D Point Cloud Privacy Protection
by: Ma, Haotian, et al.
Published: (2025)
by: Ma, Haotian, et al.
Published: (2025)
IllumiNeRF: 3D Relighting Without Inverse Rendering
by: Zhao, Xiaoming, et al.
Published: (2024)
by: Zhao, Xiaoming, et al.
Published: (2024)
Align before Adapt: Leveraging Entity-to-Region Alignments for Generalizable Video Action Recognition
by: Chen, Yifei, et al.
Published: (2023)
by: Chen, Yifei, et al.
Published: (2023)
Obtaining Optimal Spiking Neural Network in Sequence Learning via CRNN-SNN Conversion
by: Su, Jiahao, et al.
Published: (2024)
by: Su, Jiahao, et al.
Published: (2024)
Depth Map Denoising Network and Lightweight Fusion Network for Enhanced 3D Face Recognition
by: Xu, Ruizhuo, et al.
Published: (2024)
by: Xu, Ruizhuo, et al.
Published: (2024)
Learning from Loss Landscape: Generalizable Mixed-Precision Quantization via Adaptive Sharpness-Aware Gradient Aligning
by: Ma, Lianbo, et al.
Published: (2025)
by: Ma, Lianbo, et al.
Published: (2025)
Similar Items
-
ControLRM: Fast and Controllable 3D Generation via Large Reconstruction Model
by: Xu, Hongbin, et al.
Published: (2024) -
SeCG: Semantic-Enhanced 3D Visual Grounding via Cross-modal Graph Attention
by: Xiao, Feng, et al.
Published: (2024) -
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
by: Guo, Mengqi, et al.
Published: (2025) -
B2N3D: Progressive Learning from Binary to N-ary Relationships for 3D Object Grounding
by: Xiao, Feng, et al.
Published: (2025) -
Generalizable Sensor-Based Activity Recognition via Categorical Concept Invariant Learning
by: Xiong, Di, et al.
Published: (2024)