FB-4D: Spatial-Temporal Coherent Dynamic 3D Content Generation with Feature Banks
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Jinwei, Gao, Huan-ang, Li, Wenyi, Chi, Haohan, Liu, Chenyu, Du, Chenxi, Liu, Yiqian, Gao, Mingju, Zhang, Guiyu, Zhang, Zongzheng, Yi, Li, Yao, Yao, Zhao, Jingwei, Li, Hongyang, Wang, Yikai, Zhao, Hao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FairDiff: Fair Segmentation with Point-Image Diffusion
di: Li, Wenyi, et al.
Pubblicazione: (2024)
di: Li, Wenyi, et al.
Pubblicazione: (2024)
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
di: Gao, Mingju, et al.
Pubblicazione: (2025)
di: Gao, Mingju, et al.
Pubblicazione: (2025)
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
di: Li, Wenyi, et al.
Pubblicazione: (2026)
di: Li, Wenyi, et al.
Pubblicazione: (2026)
Training-Free Model Merging for Multi-target Domain Adaptation
di: Li, Wenyi, et al.
Pubblicazione: (2024)
di: Li, Wenyi, et al.
Pubblicazione: (2024)
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
di: Gao, Huan-ang, et al.
Pubblicazione: (2024)
di: Gao, Huan-ang, et al.
Pubblicazione: (2024)
Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward Modeling
di: Zhang, Guiyu, et al.
Pubblicazione: (2024)
di: Zhang, Guiyu, et al.
Pubblicazione: (2024)
Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models
di: Chi, Haohan, et al.
Pubblicazione: (2025)
di: Chi, Haohan, et al.
Pubblicazione: (2025)
Diffusion-based Visual Anagram as Multi-task Learning
di: Xu, Zhiyuan, et al.
Pubblicazione: (2024)
di: Xu, Zhiyuan, et al.
Pubblicazione: (2024)
Alias-free 4D Gaussian Splatting
di: Chen, Zilong, et al.
Pubblicazione: (2025)
di: Chen, Zilong, et al.
Pubblicazione: (2025)
PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
di: Gao, Mingju, et al.
Pubblicazione: (2026)
di: Gao, Mingju, et al.
Pubblicazione: (2026)
Challenger: Affordable Adversarial Driving Video Generation
di: Xu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Xu, Zhiyuan, et al.
Pubblicazione: (2025)
Delving into Mapping Uncertainty for Mapless Trajectory Prediction
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
Reusing Attention for One-stage Lane Topology Understanding
di: Li, Yang, et al.
Pubblicazione: (2025)
di: Li, Yang, et al.
Pubblicazione: (2025)
3D StreetUnveiler with Semantic-aware 2DGS -- a simple baseline
di: Xu, Jingwei, et al.
Pubblicazione: (2024)
di: Xu, Jingwei, et al.
Pubblicazione: (2024)
TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
PreAfford: Universal Affordance-Based Pre-Grasping for Diverse Objects and Environments
di: Ding, Kairui, et al.
Pubblicazione: (2024)
di: Ding, Kairui, et al.
Pubblicazione: (2024)
DiffVL: Diffusion-Based Visual Localization on 2D Maps via BEV-Conditioned GPS Denoising
di: Gao, Li, et al.
Pubblicazione: (2025)
di: Gao, Li, et al.
Pubblicazione: (2025)
Dual-frame Fluid Motion Estimation with Test-time Optimization and Zero-divergence Loss
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
Coca-Splat: Collaborative Optimization for Camera Parameters and 3D Gaussians
di: Wu, Jiamin, et al.
Pubblicazione: (2025)
di: Wu, Jiamin, et al.
Pubblicazione: (2025)
Straight-Line Diffusion Model for Efficient 3D Molecular Generation
di: Ni, Yuyan, et al.
Pubblicazione: (2025)
di: Ni, Yuyan, et al.
Pubblicazione: (2025)
Dexora: Open-source VLA for High-DoF Bimanual Dexterity
di: Zhang, Zongzheng, et al.
Pubblicazione: (2026)
di: Zhang, Zongzheng, et al.
Pubblicazione: (2026)
Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving
di: Ding, Kairui, et al.
Pubblicazione: (2024)
di: Ding, Kairui, et al.
Pubblicazione: (2024)
Synthesis and biological evaluation of mirror isomers of β-(1 → 3)-glucans as immune modulators.
di: Sun, Ao, et al.
Pubblicazione: (2025)
di: Sun, Ao, et al.
Pubblicazione: (2025)
Detect Anything 3D in the Wild
di: Zhang, Hanxue, et al.
Pubblicazione: (2025)
di: Zhang, Hanxue, et al.
Pubblicazione: (2025)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
di: Zhu, Xiaomeng, et al.
Pubblicazione: (2025)
di: Zhu, Xiaomeng, et al.
Pubblicazione: (2025)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
di: Song, Chenxi, et al.
Pubblicazione: (2025)
di: Song, Chenxi, et al.
Pubblicazione: (2025)
SA-GS: Scale-Adaptive Gaussian Splatting for Training-Free Anti-Aliasing
di: Song, Xiaowei, et al.
Pubblicazione: (2024)
di: Song, Xiaowei, et al.
Pubblicazione: (2024)
Progressive Correspondence Regenerator for Robust 3D Registration
di: Zhao, Guiyu, et al.
Pubblicazione: (2025)
di: Zhao, Guiyu, et al.
Pubblicazione: (2025)
Ultraman: Single Image 3D Human Reconstruction with Ultra Speed and Detail
di: Chen, Mingjin, et al.
Pubblicazione: (2024)
di: Chen, Mingjin, et al.
Pubblicazione: (2024)
OVSeg3R: Learn Open-vocabulary Instance Segmentation from 2D via 3D Reconstruction
di: Li, Hongyang, et al.
Pubblicazione: (2025)
di: Li, Hongyang, et al.
Pubblicazione: (2025)
Sustainable Porous Carbon from Waste Fungal Substrates of Gastrodia Elata : A Dual‐Function Material for High‐Performance Supercapacitors and Lithium Batteries
di: Changle Li, et al.
Pubblicazione: (2025)
di: Changle Li, et al.
Pubblicazione: (2025)
Principles of designing electrocatalysts to boost C–N coupling reactions for urea synthesis
di: Jingwei Li, et al.
Pubblicazione: (2024)
di: Jingwei Li, et al.
Pubblicazione: (2024)
Combinatorial Optimization with Automated Graph Neural Networks
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Incidence and risk factors of de novo hepatitis E virus infection after receiving liver transplantation
di: Ping He, et al.
Pubblicazione: (2024)
di: Ping He, et al.
Pubblicazione: (2024)
PIONM: A Generalized Approach to Solving Density-Constrained Mean-Field Games Equilibrium under Modified Boundary Conditions
di: Liu, Jinwei, et al.
Pubblicazione: (2025)
di: Liu, Jinwei, et al.
Pubblicazione: (2025)
SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation
di: Zhang, Guiyu, et al.
Pubblicazione: (2026)
di: Zhang, Guiyu, et al.
Pubblicazione: (2026)
Towards Understanding Adversarial Transferability in Federated Learning
di: Li, Yijiang, et al.
Pubblicazione: (2023)
di: Li, Yijiang, et al.
Pubblicazione: (2023)
4D SlingBAG: spatial-temporal coupled Gaussian ball for large-scale dynamic 3D photoacoustic iterative reconstruction
di: Li, Shuang, et al.
Pubblicazione: (2024)
di: Li, Shuang, et al.
Pubblicazione: (2024)
Relightable 3D Gaussians: Realistic Point Cloud Relighting with BRDF Decomposition and Ray Tracing
di: Gao, Jian, et al.
Pubblicazione: (2023)
di: Gao, Jian, et al.
Pubblicazione: (2023)
Documenti analoghi
-
FairDiff: Fair Segmentation with Point-Image Diffusion
di: Li, Wenyi, et al.
Pubblicazione: (2024) -
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
di: Gao, Mingju, et al.
Pubblicazione: (2025) -
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
di: Li, Wenyi, et al.
Pubblicazione: (2026) -
Training-Free Model Merging for Multi-target Domain Adaptation
di: Li, Wenyi, et al.
Pubblicazione: (2024) -
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
di: Gao, Huan-ang, et al.
Pubblicazione: (2024)