OrthoPhys: Physically Plausible Video Generation with Orthogonal-View Geometry Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Cong, Zhu, Hanxin, Tang, Xiao, Luo, Jiayi, Jin, Xin, Chen, Long, Chen, Zhibo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026)
CMC: Few-shot Novel View Synthesis via Cross-view Multiplane Consistency
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
Is Vanilla MLP in Neural Radiance Field Enough for Few-shot View Synthesis?
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
Compositional 3D-aware Video Generation with LLM Director
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
Future Forcing: Future-aware Training-free KV Cache Policy for Autoregressive Video Generation
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
TIV-Diffusion: Towards Object-Centric Movement for Text-driven Image to Video Generation
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
Embody4D: A Generalist 4D World Model for Embodied AI
von: Tu, Peiyan, et al.
Veröffentlicht: (2026)
von: Tu, Peiyan, et al.
Veröffentlicht: (2026)
AR4D: Autoregressive 4D Generation from Monocular Videos
von: Zhu, Hanxin, et al.
Veröffentlicht: (2025)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2025)
Light Field Compression Based on Implicit Neural Representation
von: Wang, Henan, et al.
Veröffentlicht: (2024)
von: Wang, Henan, et al.
Veröffentlicht: (2024)
PhysPart: Physically Plausible Part Completion for Interactable Objects
von: Luo, Rundong, et al.
Veröffentlicht: (2024)
von: Luo, Rundong, et al.
Veröffentlicht: (2024)
GaussianSR: 3D Gaussian Super-Resolution with 2D Diffusion Priors
von: Yu, Xiqian, et al.
Veröffentlicht: (2024)
von: Yu, Xiqian, et al.
Veröffentlicht: (2024)
PhysVideoGenerator: Towards Physically Aware Video Generation via Latent Physics Guidance
von: Satish, Siddarth Nilol Kundur, et al.
Veröffentlicht: (2026)
von: Satish, Siddarth Nilol Kundur, et al.
Veröffentlicht: (2026)
Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility
von: Hao, Yutong, et al.
Veröffentlicht: (2025)
von: Hao, Yutong, et al.
Veröffentlicht: (2025)
Physically Plausible Human-Object Rendering from Sparse Views via 3D Gaussian Splatting
von: Wang, Weiquan, et al.
Veröffentlicht: (2025)
von: Wang, Weiquan, et al.
Veröffentlicht: (2025)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
UCIP: A Universal Framework for Compressed Image Super-Resolution using Dynamic Prompt
von: Li, Xin, et al.
Veröffentlicht: (2024)
von: Li, Xin, et al.
Veröffentlicht: (2024)
Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
von: Ji, Sihui, et al.
Veröffentlicht: (2025)
P-4DGS: Predictive 4D Gaussian Splatting with 90$\times$ Compression
von: Wang, Henan, et al.
Veröffentlicht: (2025)
von: Wang, Henan, et al.
Veröffentlicht: (2025)
GSemSplat: Generalizable Semantic 3D Gaussian Splatting from Uncalibrated Image Pairs
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
ACWM-Phys: Investigating Generalized Physical Interaction in Action-Conditioned Video World Models
von: Xue, Haotian, et al.
Veröffentlicht: (2026)
von: Xue, Haotian, et al.
Veröffentlicht: (2026)
Chain of Event-Centric Causal Thought for Physically Plausible Video Generation
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
MMPhysVideo: Scaling Physical Plausibility in Video Generation via Joint Multimodal Modeling
von: Lin, Shubo, et al.
Veröffentlicht: (2026)
von: Lin, Shubo, et al.
Veröffentlicht: (2026)
Training-free Camera Control for Video Generation
von: Hou, Chen, et al.
Veröffentlicht: (2024)
von: Hou, Chen, et al.
Veröffentlicht: (2024)
SeD: Semantic-Aware Discriminator for Image Super-Resolution
von: Li, Bingchen, et al.
Veröffentlicht: (2024)
von: Li, Bingchen, et al.
Veröffentlicht: (2024)
Cross-View Meets Diffusion: Aerial Image Synthesis with Geometry and Text Guidance
von: Arrabi, Ahmad, et al.
Veröffentlicht: (2024)
von: Arrabi, Ahmad, et al.
Veröffentlicht: (2024)
4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation Models
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
PhysReaction: Physically Plausible Real-Time Humanoid Reaction Synthesis via Forward Dynamics Guided 4D Imitation
von: Liu, Yunze, et al.
Veröffentlicht: (2024)
von: Liu, Yunze, et al.
Veröffentlicht: (2024)
OrthoEraser: Coupled-Neuron Orthogonal Projection for Concept Erasure
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2026)
CaliTex: Geometry-Calibrated Attention for View-Coherent 3D Texture Generation
von: Liu, Chenyu, et al.
Veröffentlicht: (2025)
von: Liu, Chenyu, et al.
Veröffentlicht: (2025)
PhysMoDPO: Physically-Plausible Humanoid Motion with Preference Optimization
von: Zhang, Yangsong, et al.
Veröffentlicht: (2026)
von: Zhang, Yangsong, et al.
Veröffentlicht: (2026)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
von: Zhang, Haoze, et al.
Veröffentlicht: (2025)
von: Zhang, Haoze, et al.
Veröffentlicht: (2025)
From Generated Human Videos to Physically Plausible Robot Trajectories
von: Ni, James, et al.
Veröffentlicht: (2025)
von: Ni, James, et al.
Veröffentlicht: (2025)
PhysHMR: Learning Humanoid Control Policies from Vision for Physically Plausible Human Motion Reconstruction
von: Feng, Qiao, et al.
Veröffentlicht: (2025)
von: Feng, Qiao, et al.
Veröffentlicht: (2025)
FlowMotion: Training-Free Flow Guidance for Video Motion Transfer
von: Wang, Zhen, et al.
Veröffentlicht: (2026)
von: Wang, Zhen, et al.
Veröffentlicht: (2026)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
PhysRVG: Physics-Aware Unified Reinforcement Learning for Video Generative Models
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2026)
Ortho-Hydra: Orthogonalized Experts for DiT LoRA
von: Ji, Seunghyun
Veröffentlicht: (2026)
von: Ji, Seunghyun
Veröffentlicht: (2026)
Ähnliche Einträge
-
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026) -
CMC: Few-shot Novel View Synthesis via Cross-view Multiplane Consistency
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024) -
Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering
von: Luo, Jiayi, et al.
Veröffentlicht: (2026) -
Is Vanilla MLP in Neural Radiance Field Enough for Few-shot View Synthesis?
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024) -
Compositional 3D-aware Video Generation with LLM Director
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)