FOLIAGE: Towards Physical Intelligence World Models Via Unbounded Surface Evolution
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Xiaoyi, Tang, Hao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Enhanced Image Generation Via Multi-modal Chain of Thought in Unified Generative Models
di: Wang, Yi, et al.
Pubblicazione: (2025)
di: Wang, Yi, et al.
Pubblicazione: (2025)
WorldFlow3D: Flowing Through 3D Distributions for Unbounded World Generation
di: Joshi, Amogh, et al.
Pubblicazione: (2026)
di: Joshi, Amogh, et al.
Pubblicazione: (2026)
InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
di: Lu, Yifan, et al.
Pubblicazione: (2024)
di: Lu, Yifan, et al.
Pubblicazione: (2024)
Thinking Ahead: Foresight Intelligence in MLLMs and World Models
di: Gong, Zhantao, et al.
Pubblicazione: (2025)
di: Gong, Zhantao, et al.
Pubblicazione: (2025)
XMorph: Explainable Brain Tumor Analysis Via LLM-Assisted Hybrid Deep Intelligence
di: Ghahfarokhi, Sepehr Salem, et al.
Pubblicazione: (2026)
di: Ghahfarokhi, Sepehr Salem, et al.
Pubblicazione: (2026)
PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World
di: Wang, Changpeng, et al.
Pubblicazione: (2026)
di: Wang, Changpeng, et al.
Pubblicazione: (2026)
Simulating the Visual World with Artificial Intelligence: A Roadmap
di: Yue, Jingtong, et al.
Pubblicazione: (2025)
di: Yue, Jingtong, et al.
Pubblicazione: (2025)
Interpreting Physics in Video World Models
di: Joseph, Sonia, et al.
Pubblicazione: (2026)
di: Joseph, Sonia, et al.
Pubblicazione: (2026)
Towards Visual Discrimination and Reasoning of Real-World Physical Dynamics: Physics-Grounded Anomaly Detection
di: Li, Wenqiao, et al.
Pubblicazione: (2025)
di: Li, Wenqiao, et al.
Pubblicazione: (2025)
Pandora: Towards General World Model with Natural Language Actions and Video States
di: Xiang, Jiannan, et al.
Pubblicazione: (2024)
di: Xiang, Jiannan, et al.
Pubblicazione: (2024)
Chain of World: World Model Thinking in Latent Motion
di: Yang, Fuxiang, et al.
Pubblicazione: (2026)
di: Yang, Fuxiang, et al.
Pubblicazione: (2026)
3DGS-Enhancer: Enhancing Unbounded 3D Gaussian Splatting with View-consistent 2D Diffusion Priors
di: Liu, Xi, et al.
Pubblicazione: (2024)
di: Liu, Xi, et al.
Pubblicazione: (2024)
Deepfake Detection Via Facial Feature Extraction and Modeling
di: Carter, Benjamin, et al.
Pubblicazione: (2025)
di: Carter, Benjamin, et al.
Pubblicazione: (2025)
Cabbage: A Differential Growth Framework for Open Surfaces
di: Liu, Xiaoyi, et al.
Pubblicazione: (2025)
di: Liu, Xiaoyi, et al.
Pubblicazione: (2025)
LatXGen: Towards Radiation-Free and Accurate Quantitative Analysis of Sagittal Spinal Alignment Via Cross-Modal Radiographic View Synthesis
di: Zhao, Moxin, et al.
Pubblicazione: (2025)
di: Zhao, Moxin, et al.
Pubblicazione: (2025)
CityX: Controllable Procedural Content Generation for Unbounded 3D Cities
di: Zhang, Shougao, et al.
Pubblicazione: (2024)
di: Zhang, Shougao, et al.
Pubblicazione: (2024)
Do Vision-Language Models Have Internal World Models? Towards an Atomic Evaluation
di: Gao, Qiyue, et al.
Pubblicazione: (2025)
di: Gao, Qiyue, et al.
Pubblicazione: (2025)
ReconViaGen: Towards Accurate Multi-view 3D Object Reconstruction via Generation
di: Chang, Jiahao, et al.
Pubblicazione: (2025)
di: Chang, Jiahao, et al.
Pubblicazione: (2025)
SparseWorld: A Flexible, Adaptive, and Efficient 4D Occupancy World Model Powered by Sparse and Dynamic Queries
di: Dang, Chenxu, et al.
Pubblicazione: (2025)
di: Dang, Chenxu, et al.
Pubblicazione: (2025)
Robot Learning from a Physical World Model
di: Mao, Jiageng, et al.
Pubblicazione: (2025)
di: Mao, Jiageng, et al.
Pubblicazione: (2025)
"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models
di: Gu, Jing, et al.
Pubblicazione: (2025)
di: Gu, Jing, et al.
Pubblicazione: (2025)
World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks
di: Lin, Zuyao, et al.
Pubblicazione: (2026)
di: Lin, Zuyao, et al.
Pubblicazione: (2026)
Towards Ancient Plant Seed Classification: A Benchmark Dataset and Baseline Model
di: Xing, Rui, et al.
Pubblicazione: (2025)
di: Xing, Rui, et al.
Pubblicazione: (2025)
Application of Multimodal Fusion Deep Learning Model in Disease Recognition
di: Liu, Xiaoyi, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyi, et al.
Pubblicazione: (2024)
DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks
di: Deng, Yueci, et al.
Pubblicazione: (2026)
di: Deng, Yueci, et al.
Pubblicazione: (2026)
One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy
di: Tang, Zuojin, et al.
Pubblicazione: (2026)
di: Tang, Zuojin, et al.
Pubblicazione: (2026)
Mirage2Matter: A Physically Grounded Gaussian World Model from Video
di: Gao, Zhengqing, et al.
Pubblicazione: (2026)
di: Gao, Zhengqing, et al.
Pubblicazione: (2026)
Towards Efficient and Intelligent Laser Weeding: Method and Dataset for Weed Stem Detection
di: Liu, Dingning, et al.
Pubblicazione: (2025)
di: Liu, Dingning, et al.
Pubblicazione: (2025)
RenderWorld: World Model with Self-Supervised 3D Label
di: Yan, Ziyang, et al.
Pubblicazione: (2024)
di: Yan, Ziyang, et al.
Pubblicazione: (2024)
SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence
di: Wu, Haoning, et al.
Pubblicazione: (2025)
di: Wu, Haoning, et al.
Pubblicazione: (2025)
Bridging the Gap: Toward Cognitive Autonomy in Artificial Intelligence
di: Golilarz, Noorbakhsh Amiri, et al.
Pubblicazione: (2025)
di: Golilarz, Noorbakhsh Amiri, et al.
Pubblicazione: (2025)
Learning Vision-Language-Action World Models for Autonomous Driving
di: Wang, Guoqing, et al.
Pubblicazione: (2026)
di: Wang, Guoqing, et al.
Pubblicazione: (2026)
WorldEdit: Towards Open-World Image Editing with a Knowledge-Informed Benchmark
di: Lin, Wang, et al.
Pubblicazione: (2026)
di: Lin, Wang, et al.
Pubblicazione: (2026)
DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding
di: Feng, Hao, et al.
Pubblicazione: (2023)
di: Feng, Hao, et al.
Pubblicazione: (2023)
MimeQA: Towards Socially-Intelligent Nonverbal Foundation Models
di: Li, Hengzhi, et al.
Pubblicazione: (2025)
di: Li, Hengzhi, et al.
Pubblicazione: (2025)
Foundation Models -- A Panacea for Artificial Intelligence in Pathology?
di: Mulliqi, Nita, et al.
Pubblicazione: (2025)
di: Mulliqi, Nita, et al.
Pubblicazione: (2025)
PhysicsMind: Sim and Real Mechanics Benchmarking for Physical Reasoning and Prediction in Foundational VLMs and World Models
di: Mak, Chak-Wing, et al.
Pubblicazione: (2026)
di: Mak, Chak-Wing, et al.
Pubblicazione: (2026)
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment
di: Tang, Jinzhou, et al.
Pubblicazione: (2025)
di: Tang, Jinzhou, et al.
Pubblicazione: (2025)
EMRA-proxy: Enhancing Multi-Class Region Semantic Segmentation in Remote Sensing Images with Attention Proxy
di: Yu, Yichun, et al.
Pubblicazione: (2025)
di: Yu, Yichun, et al.
Pubblicazione: (2025)
Towards Robust Physical-world Backdoor Attacks on Lane Detection
di: Zhang, Xinwei, et al.
Pubblicazione: (2024)
di: Zhang, Xinwei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Enhanced Image Generation Via Multi-modal Chain of Thought in Unified Generative Models
di: Wang, Yi, et al.
Pubblicazione: (2025) -
WorldFlow3D: Flowing Through 3D Distributions for Unbounded World Generation
di: Joshi, Amogh, et al.
Pubblicazione: (2026) -
InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
di: Lu, Yifan, et al.
Pubblicazione: (2024) -
Thinking Ahead: Foresight Intelligence in MLLMs and World Models
di: Gong, Zhantao, et al.
Pubblicazione: (2025) -
XMorph: Explainable Brain Tumor Analysis Via LLM-Assisted Hybrid Deep Intelligence
di: Ghahfarokhi, Sepehr Salem, et al.
Pubblicazione: (2026)