Saved in:
| Main Authors: | Yang, Mengyu, Yang, Yanming, Xu, Chenyi, Song, Chenxi, Zuo, Yufan, Zhao, Tong, Li, Ruibo, Zhang, Chi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.22533 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
by: Song, Chenxi, et al.
Published: (2025)
by: Song, Chenxi, et al.
Published: (2025)
FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing
by: Li, Guangzhao, et al.
Published: (2025)
by: Li, Guangzhao, et al.
Published: (2025)
Hash3D: Training-free Acceleration for 3D Generation
by: Yang, Xingyi, et al.
Published: (2024)
by: Yang, Xingyi, et al.
Published: (2024)
DyWeight: Dynamic Gradient Weighting for Few-Step Diffusion Sampling
by: Zhao, Tong, et al.
Published: (2026)
by: Zhao, Tong, et al.
Published: (2026)
Bimanual Grasp Synthesis for Dexterous Robot Hands
by: Shao, Yanming, et al.
Published: (2024)
by: Shao, Yanming, et al.
Published: (2024)
MAISI-v2: Accelerated 3D High-Resolution Medical Image Synthesis with Rectified Flow and Region-specific Contrastive Loss
by: Zhao, Can, et al.
Published: (2025)
by: Zhao, Can, et al.
Published: (2025)
SwitchCraft: Training-Free Multi-Event Video Generation with Attention Controls
by: Xu, Qianxun, et al.
Published: (2026)
by: Xu, Qianxun, et al.
Published: (2026)
Training-free Diffusion Acceleration with Bottleneck Sampling
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
Diffusion Models are Geometry Critics: Single Image 3D Editing Using Pre-Trained Diffusion Priors
by: Wang, Ruicheng, et al.
Published: (2024)
by: Wang, Ruicheng, et al.
Published: (2024)
FastVGGT: Training-Free Acceleration of Visual Geometry Transformer
by: Shen, You, et al.
Published: (2025)
by: Shen, You, et al.
Published: (2025)
Auto3DSeg for Brain Tumor Segmentation from 3D MRI in BraTS 2023 Challenge
by: Myronenko, Andriy, et al.
Published: (2025)
by: Myronenko, Andriy, et al.
Published: (2025)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
by: Yang, Feng, et al.
Published: (2025)
by: Yang, Feng, et al.
Published: (2025)
A Prediction-as-Perception Framework for 3D Object Detection
by: Zhang, Song, et al.
Published: (2026)
by: Zhang, Song, et al.
Published: (2026)
Single-Slice-to-3D Reconstruction in Medical Imaging and Natural Objects: A Comparative Benchmark with SAM 3D
by: Luo, Yan, et al.
Published: (2026)
by: Luo, Yan, et al.
Published: (2026)
FastAvatar: Towards Unified and Fast 3D Avatar Reconstruction with Large Gaussian Reconstruction Transformers
by: Wu, Yue, et al.
Published: (2025)
by: Wu, Yue, et al.
Published: (2025)
Graph-based 3D Human Pose Estimation using WiFi Signals
by: Chen, Jichao, et al.
Published: (2025)
by: Chen, Jichao, et al.
Published: (2025)
Fast SAM 3D Body: Accelerating SAM 3D Body for Real-Time Full-Body Human Mesh Recovery
by: Yang, Timing, et al.
Published: (2026)
by: Yang, Timing, et al.
Published: (2026)
Tail-Aware Post-Training Quantization for 3D Geometry Models
by: Pan, Sicheng, et al.
Published: (2026)
by: Pan, Sicheng, et al.
Published: (2026)
QuadBox: Accelerating 3D Gaussian Splatting with Geometry-Aware Boxes
by: Li, Xinze, et al.
Published: (2026)
by: Li, Xinze, et al.
Published: (2026)
3D Geometry-aware Deformable Gaussian Splatting for Dynamic View Synthesis
by: Lu, Zhicheng, et al.
Published: (2024)
by: Lu, Zhicheng, et al.
Published: (2024)
TrajVG: 3D Trajectory-Coupled Visual Geometry Learning
by: Miao, Xingyu, et al.
Published: (2026)
by: Miao, Xingyu, et al.
Published: (2026)
Geometry-as-context: Modulating Explicit 3D in Scene-consistent Video Generation to Geometry Context
by: Hu, JiaKui, et al.
Published: (2026)
by: Hu, JiaKui, et al.
Published: (2026)
Pathformer3D: A 3D Scanpath Transformer for 360° Images
by: Quan, Rong, et al.
Published: (2024)
by: Quan, Rong, et al.
Published: (2024)
Leveraging Previous Steps: A Training-free Fast Solver for Flow Diffusion
by: Song, Kaiyu, et al.
Published: (2024)
by: Song, Kaiyu, et al.
Published: (2024)
A Simple Baseline for Spoken Language to Sign Language Translation with 3D Avatars
by: Zuo, Ronglai, et al.
Published: (2024)
by: Zuo, Ronglai, et al.
Published: (2024)
Fast Person Detection Using YOLOX With AI Accelerator For Train Station Safety
by: Achmadiah, Mas Nurul, et al.
Published: (2026)
by: Achmadiah, Mas Nurul, et al.
Published: (2026)
BrightDreamer: Generic 3D Gaussian Generative Framework for Fast Text-to-3D Synthesis
by: Jiang, Lutao, et al.
Published: (2024)
by: Jiang, Lutao, et al.
Published: (2024)
ARM3D: Attention-based relation module for indoor 3D object detection
by: Lan, Yuqing, et al.
Published: (2022)
by: Lan, Yuqing, et al.
Published: (2022)
Lighting Every Darkness with 3DGS: Fast Training and Real-Time Rendering for HDR View Synthesis
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
by: Xu, Shaoqing, et al.
Published: (2024)
by: Xu, Shaoqing, et al.
Published: (2024)
VISTA3D: A Unified Segmentation Foundation Model For 3D Medical Imaging
by: He, Yufan, et al.
Published: (2024)
by: He, Yufan, et al.
Published: (2024)
A Short Review and Evaluation of SAM2's Performance in 3D CT Image Segmentation
by: He, Yufan, et al.
Published: (2024)
by: He, Yufan, et al.
Published: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
by: Xu, Wan, et al.
Published: (2023)
by: Xu, Wan, et al.
Published: (2023)
Reflections Unlock: Geometry-Aware Reflection Disentanglement in 3D Gaussian Splatting for Photorealistic Scenes Rendering
by: Song, Jiayi, et al.
Published: (2025)
by: Song, Jiayi, et al.
Published: (2025)
Reg3D: Reconstructive Geometry Instruction Tuning for 3D Scene Understanding
by: Zheng, Hongpei, et al.
Published: (2025)
by: Zheng, Hongpei, et al.
Published: (2025)
Fed3D: Federated 3D Object Detection
by: Dai, Suyan, et al.
Published: (2026)
by: Dai, Suyan, et al.
Published: (2026)
ActiveNeRF: Learning Accurate 3D Geometry by Active Pattern Projection
by: Tao, Jianyu, et al.
Published: (2024)
by: Tao, Jianyu, et al.
Published: (2024)
Beyond Geometry: Artistic Disparity Synthesis for Immersive 2D-to-3D
by: Chen, Ping, et al.
Published: (2026)
by: Chen, Ping, et al.
Published: (2026)
Training-free Token Reduction for Vision Mamba
by: Ma, Qiankun, et al.
Published: (2025)
by: Ma, Qiankun, et al.
Published: (2025)
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding
by: Huang, Wencan, et al.
Published: (2025)
by: Huang, Wencan, et al.
Published: (2025)
Similar Items
-
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
by: Song, Chenxi, et al.
Published: (2025) -
FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing
by: Li, Guangzhao, et al.
Published: (2025) -
Hash3D: Training-free Acceleration for 3D Generation
by: Yang, Xingyi, et al.
Published: (2024) -
DyWeight: Dynamic Gradient Weighting for Few-Step Diffusion Sampling
by: Zhao, Tong, et al.
Published: (2026) -
Bimanual Grasp Synthesis for Dexterous Robot Hands
by: Shao, Yanming, et al.
Published: (2024)