Multi-View Large Reconstruction Model via Geometry-Aware Positional Encoding and Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Mengfei, Long, Xiaoxiao, Liang, Yixun, Li, Weiyu, Liu, Yuan, Li, Peng, Luo, Wenhan, Wang, Wenping, Guo, Yike |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Era3D: High-Resolution Multiview Diffusion using Efficient Row-wise Attention
von: Li, Peng, et al.
Veröffentlicht: (2024)
von: Li, Peng, et al.
Veröffentlicht: (2024)
CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner
von: Li, Weiyu, et al.
Veröffentlicht: (2024)
von: Li, Weiyu, et al.
Veröffentlicht: (2024)
UniSH: Unifying Scene and Human Reconstruction in a Feed-Forward Pass
von: Li, Mengfei, et al.
Veröffentlicht: (2026)
von: Li, Mengfei, et al.
Veröffentlicht: (2026)
Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders
von: Chen, Rui, et al.
Veröffentlicht: (2024)
von: Chen, Rui, et al.
Veröffentlicht: (2024)
UniRecGen: Unifying Multi-View 3D Reconstruction and Generation
von: Huang, Zhisheng, et al.
Veröffentlicht: (2026)
von: Huang, Zhisheng, et al.
Veröffentlicht: (2026)
Weakly-Supervised Emotion Transition Learning for Diverse 3D Co-speech Gesture Generation
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
Geometry-Aware Attenuation Learning for Sparse-View CBCT Reconstruction
von: Liu, Zhentao, et al.
Veröffentlicht: (2023)
von: Liu, Zhentao, et al.
Veröffentlicht: (2023)
CMD: Controllable Multiview Diffusion for 3D Editing and Progressive Generation
von: Li, Peng, et al.
Veröffentlicht: (2025)
von: Li, Peng, et al.
Veröffentlicht: (2025)
GVGS: Gaussian Visibility-Aware Multi-View Geometry for Accurate Surface Reconstruction
von: Su, Mai, et al.
Veröffentlicht: (2026)
von: Su, Mai, et al.
Veröffentlicht: (2026)
X-LRM: X-ray Large Reconstruction Model for Extremely Sparse-View Computed Tomography Recovery in One Second
von: Zhang, Guofeng, et al.
Veröffentlicht: (2025)
von: Zhang, Guofeng, et al.
Veröffentlicht: (2025)
Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models
von: Wang, Zhiqiang, et al.
Veröffentlicht: (2026)
von: Wang, Zhiqiang, et al.
Veröffentlicht: (2026)
PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation
von: Yan, Hongyu, et al.
Veröffentlicht: (2025)
von: Yan, Hongyu, et al.
Veröffentlicht: (2025)
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
von: Li, Yan, et al.
Veröffentlicht: (2024)
von: Li, Yan, et al.
Veröffentlicht: (2024)
DecoRec: Decomposed 3D Scene Reconstruction from Single-View Images via Object-Level Diffusion
von: Ping, Yuhan, et al.
Veröffentlicht: (2026)
von: Ping, Yuhan, et al.
Veröffentlicht: (2026)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
Part123: Part-aware 3D Reconstruction from a Single-view Image
von: Liu, Anran, et al.
Veröffentlicht: (2024)
von: Liu, Anran, et al.
Veröffentlicht: (2024)
GausSurf: Geometry-Guided 3D Gaussian Splatting for Surface Reconstruction
von: Wang, Jiepeng, et al.
Veröffentlicht: (2024)
von: Wang, Jiepeng, et al.
Veröffentlicht: (2024)
UniTEX: Universal High Fidelity Generative Texturing for 3D Shapes
von: Liang, Yixun, et al.
Veröffentlicht: (2025)
von: Liang, Yixun, et al.
Veröffentlicht: (2025)
PSHuman: Photorealistic Single-image 3D Human Reconstruction using Cross-Scale Multiview Diffusion and Explicit Remeshing
von: Li, Peng, et al.
Veröffentlicht: (2024)
von: Li, Peng, et al.
Veröffentlicht: (2024)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
von: Miyato, Takeru, et al.
Veröffentlicht: (2023)
von: Miyato, Takeru, et al.
Veröffentlicht: (2023)
RaDe-GS: Rasterizing Depth in Gaussian Splatting
von: Zhang, Baowen, et al.
Veröffentlicht: (2024)
von: Zhang, Baowen, et al.
Veröffentlicht: (2024)
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
von: Qi, Xingqun, et al.
Veröffentlicht: (2024)
von: Qi, Xingqun, et al.
Veröffentlicht: (2024)
CADDreamer: CAD Object Generation from Single-view Images
von: Li, Yuan, et al.
Veröffentlicht: (2025)
von: Li, Yuan, et al.
Veröffentlicht: (2025)
GeoQuery: Geometry-Query Diffusion for Sparse-View Reconstruction
von: Cao, Xiao, et al.
Veröffentlicht: (2026)
von: Cao, Xiao, et al.
Veröffentlicht: (2026)
SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization
von: Li, Deming, et al.
Veröffentlicht: (2026)
von: Li, Deming, et al.
Veröffentlicht: (2026)
SolidGS: Consolidating Gaussian Surfel Splatting for Sparse-View Surface Reconstruction
von: Shen, Zhuowen, et al.
Veröffentlicht: (2024)
von: Shen, Zhuowen, et al.
Veröffentlicht: (2024)
Dynamic Gaussian Scene Reconstruction from Unsynchronized Videos
von: Xu, Zhixin, et al.
Veröffentlicht: (2025)
von: Xu, Zhixin, et al.
Veröffentlicht: (2025)
CogniEdit: Dense Gradient Flow Optimization for Fine-Grained Image Editing
von: Li, Yan, et al.
Veröffentlicht: (2025)
von: Li, Yan, et al.
Veröffentlicht: (2025)
Positional Encoding Field
von: Bai, Yunpeng, et al.
Veröffentlicht: (2025)
von: Bai, Yunpeng, et al.
Veröffentlicht: (2025)
Attention-Aware Multi-View Pedestrian Tracking
von: Alturki, Reef, et al.
Veröffentlicht: (2025)
von: Alturki, Reef, et al.
Veröffentlicht: (2025)
Seeing the Unseen: Mask-Driven Positional Encoding and Strip-Convolution Context Modeling for Cross-View Object Geo-Localization
von: Hu, Shuhan, et al.
Veröffentlicht: (2025)
von: Hu, Shuhan, et al.
Veröffentlicht: (2025)
Sparse2DGS: Geometry-Prioritized Gaussian Splatting for Surface Reconstruction from Sparse Views
von: Wu, Jiang, et al.
Veröffentlicht: (2025)
von: Wu, Jiang, et al.
Veröffentlicht: (2025)
MCGA: Mixture of Codebooks Hyperspectral Reconstruction via Grayscale-Aware Attention
von: Yang, Zhanjiang, et al.
Veröffentlicht: (2025)
von: Yang, Zhanjiang, et al.
Veröffentlicht: (2025)
Foundation Cures Personalization: Improving Personalized Models' Prompt Consistency via Hidden Foundation Knowledge
von: Cai, Yiyang, et al.
Veröffentlicht: (2024)
von: Cai, Yiyang, et al.
Veröffentlicht: (2024)
AFUNet: Cross-Iterative Alignment-Fusion Synergy for HDR Reconstruction via Deep Unfolding Paradigm
von: Li, Xinyue, et al.
Veröffentlicht: (2025)
von: Li, Xinyue, et al.
Veröffentlicht: (2025)
SA-GS: Semantic-Aware Gaussian Splatting for Large Scene Reconstruction with Geometry Constrain
von: Xiong, Butian, et al.
Veröffentlicht: (2024)
von: Xiong, Butian, et al.
Veröffentlicht: (2024)
PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing
von: Xiong, Peilin, et al.
Veröffentlicht: (2025)
von: Xiong, Peilin, et al.
Veröffentlicht: (2025)
A 2D Semantic-Aware Position Encoding for Vision Transformers
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
Semantic Human Mesh Reconstruction with Textures
von: Zhan, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhan, Xiaoyu, et al.
Veröffentlicht: (2024)
EfficientMonoHair: Fast Strand-Level Reconstruction from Monocular Video via Multi-View Direction Fusion
von: Li, Da, et al.
Veröffentlicht: (2026)
von: Li, Da, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Era3D: High-Resolution Multiview Diffusion using Efficient Row-wise Attention
von: Li, Peng, et al.
Veröffentlicht: (2024) -
CraftsMan3D: High-fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner
von: Li, Weiyu, et al.
Veröffentlicht: (2024) -
UniSH: Unifying Scene and Human Reconstruction in a Feed-Forward Pass
von: Li, Mengfei, et al.
Veröffentlicht: (2026) -
Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders
von: Chen, Rui, et al.
Veröffentlicht: (2024) -
UniRecGen: Unifying Multi-View 3D Reconstruction and Generation
von: Huang, Zhisheng, et al.
Veröffentlicht: (2026)