IRIS: A Real-World Benchmark for Inverse Recovery and Identification of Physical Dynamic Systems from Monocular Video
Fuente:
arXiv
Salvato in:
| Autori principali: | Khanbayov, Rasul, Barhdadi, Mohamed Rayan, Serpedin, Erchin, Kurban, Hasan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
4D Synchronized Fields: Motion-Language Gaussian Splatting for Temporal Scene Understanding
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2026)
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2026)
Stress-Testing Multimodal Foundation Models for Crystallographic Reasoning
di: Polat, Can, et al.
Pubblicazione: (2025)
di: Polat, Can, et al.
Pubblicazione: (2025)
QuantumCanvas: A Multimodal Benchmark for Visual Learning of Atomic Interactions
di: Polat, Can, et al.
Pubblicazione: (2025)
di: Polat, Can, et al.
Pubblicazione: (2025)
PhysicsNeRF: Physics-Guided 3D Reconstruction from Sparse Views
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)
EMPATHIA: Multi-Faceted Human-AI Collaboration for Refugee Integration
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)
C2NP: A Benchmark for Learning Scale-Dependent Geometric Invariances in 3D Materials Generation
di: Polat, Can, et al.
Pubblicazione: (2026)
di: Polat, Can, et al.
Pubblicazione: (2026)
Understanding the Capabilities of Molecular Graph Neural Networks in Materials Science Through Multimodal Learning and Physical Context Encoding
di: Polat, Can, et al.
Pubblicazione: (2025)
di: Polat, Can, et al.
Pubblicazione: (2025)
Beyond Atomic Geometry Representations in Materials Science: A Human-in-the-Loop Multimodal Framework
di: Polat, Can, et al.
Pubblicazione: (2025)
di: Polat, Can, et al.
Pubblicazione: (2025)
SCALAR: Quantifying Structural Hallucination, Consistency, and Reasoning Gaps in Materials Foundation Models
di: Polat, Can, et al.
Pubblicazione: (2026)
di: Polat, Can, et al.
Pubblicazione: (2026)
How Far Can You Grow? Characterizing the Extrapolation Frontier of Graph Generative Models for Materials Science
di: Polat, Can, et al.
Pubblicazione: (2026)
di: Polat, Can, et al.
Pubblicazione: (2026)
Automated Knot Detection and Pairing for Wood Analysis in the Timber Industry
di: Lin, Guohao, et al.
Pubblicazione: (2025)
di: Lin, Guohao, et al.
Pubblicazione: (2025)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
Knowing When Not to Answer: Abstention-Aware Scientific Reasoning
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
Multimodal AI for Body Fat Estimation: Computer Vision and Anthropometry with DEXA Benchmarks
di: Aldajani, Rayan
Pubblicazione: (2025)
di: Aldajani, Rayan
Pubblicazione: (2025)
Advanced Deep Learning and Large Language Models: Comprehensive Insights for Cancer Detection
di: Habchi, Yassine, et al.
Pubblicazione: (2025)
di: Habchi, Yassine, et al.
Pubblicazione: (2025)
Beyond Closed-Pool Video Retrieval: A Benchmark and Agent Framework for Real-World Video Search and Moment Localization
di: Yu, Tao, et al.
Pubblicazione: (2026)
di: Yu, Tao, et al.
Pubblicazione: (2026)
LGQ: Learning Discretization Geometry for Scalable and Stable Image Tokenization
di: Altun, Idil Bilge, et al.
Pubblicazione: (2026)
di: Altun, Idil Bilge, et al.
Pubblicazione: (2026)
Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression
di: Wang, Lirui, et al.
Pubblicazione: (2025)
di: Wang, Lirui, et al.
Pubblicazione: (2025)
IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray Tracing
di: Wang, Shaofei, et al.
Pubblicazione: (2023)
di: Wang, Shaofei, et al.
Pubblicazione: (2023)
CorVS: Person Identification via Video Trajectory-Sensor Correspondence in a Real-World Warehouse
di: Kano, Kazuma, et al.
Pubblicazione: (2025)
di: Kano, Kazuma, et al.
Pubblicazione: (2025)
Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos
di: Wei, Dingkun, et al.
Pubblicazione: (2026)
di: Wei, Dingkun, et al.
Pubblicazione: (2026)
Physics-guided Shape-from-Template: Monocular Video Perception through Neural Surrogate Models
di: Stotko, David, et al.
Pubblicazione: (2023)
di: Stotko, David, et al.
Pubblicazione: (2023)
Alice Benchmarks: Connecting Real World Re-Identification with the Synthetic
di: Sun, Xiaoxiao, et al.
Pubblicazione: (2023)
di: Sun, Xiaoxiao, et al.
Pubblicazione: (2023)
WorldTree: Towards 4D Dynamic Worlds from Monocular Video using Tree-Chains
di: Wang, Qisen, et al.
Pubblicazione: (2026)
di: Wang, Qisen, et al.
Pubblicazione: (2026)
Temporal Realism Evaluation of Generated Videos Using Compressed-Domain Motion Vectors
di: Cakiroglu, Mert Onur, et al.
Pubblicazione: (2025)
di: Cakiroglu, Mert Onur, et al.
Pubblicazione: (2025)
OSN: Infinite Representations of Dynamic 3D Scenes from Monocular Videos
di: Song, Ziyang, et al.
Pubblicazione: (2024)
di: Song, Ziyang, et al.
Pubblicazione: (2024)
xChemAgents: Agentic AI for Explainable Quantum Chemistry
di: Polat, Can, et al.
Pubblicazione: (2025)
di: Polat, Can, et al.
Pubblicazione: (2025)
SPARK: Self-supervised Personalized Real-time Monocular Face Capture
di: Baert, Kelian, et al.
Pubblicazione: (2024)
di: Baert, Kelian, et al.
Pubblicazione: (2024)
4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos
di: Xu, Zhen, et al.
Pubblicazione: (2025)
di: Xu, Zhen, et al.
Pubblicazione: (2025)
Theorem-of-Thought: A Multi-Agent Framework for Abductive, Deductive, and Inductive Reasoning in Language Models
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
Audit-of-Understanding: Posterior-Constrained Inference for Mathematical Reasoning in Language Models
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
Evaluating Multilingual and Code-Switched Alignment in LLMs via Synthetic Natural Language Inference
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2025)
Halluverse-M^3: A multitask multilingual benchmark for hallucination in LLMs
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
D-NPC: Dynamic Neural Point Clouds for Non-Rigid View Synthesis from Monocular Video
di: Kappel, Moritz, et al.
Pubblicazione: (2024)
di: Kappel, Moritz, et al.
Pubblicazione: (2024)
Decoupling Dynamic Monocular Videos for Dynamic View Synthesis
di: You, Meng, et al.
Pubblicazione: (2023)
di: You, Meng, et al.
Pubblicazione: (2023)
Object-Centric Learning for Real-World Videos by Predicting Temporal Feature Similarities
di: Zadaianchuk, Andrii, et al.
Pubblicazione: (2023)
di: Zadaianchuk, Andrii, et al.
Pubblicazione: (2023)
Surfel-based Gaussian Inverse Rendering for Fast and Relightable Dynamic Human Reconstruction from Monocular Video
di: Zhao, Yiqun, et al.
Pubblicazione: (2024)
di: Zhao, Yiqun, et al.
Pubblicazione: (2024)
Underwater Monocular Metric Depth Estimation: Real-World Benchmarks and Synthetic Fine-Tuning with Vision Foundation Models
di: Cai, Zijie, et al.
Pubblicazione: (2025)
di: Cai, Zijie, et al.
Pubblicazione: (2025)
IRIS: Inverse Rendering of Indoor Scenes from Low Dynamic Range Images
di: Lin, Chih-Hao, et al.
Pubblicazione: (2024)
di: Lin, Chih-Hao, et al.
Pubblicazione: (2024)
W-HMR: Monocular Human Mesh Recovery in World Space with Weak-Supervised Calibration
di: Yao, Wei, et al.
Pubblicazione: (2023)
di: Yao, Wei, et al.
Pubblicazione: (2023)
Documenti analoghi
-
4D Synchronized Fields: Motion-Language Gaussian Splatting for Temporal Scene Understanding
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2026) -
Stress-Testing Multimodal Foundation Models for Crystallographic Reasoning
di: Polat, Can, et al.
Pubblicazione: (2025) -
QuantumCanvas: A Multimodal Benchmark for Visual Learning of Atomic Interactions
di: Polat, Can, et al.
Pubblicazione: (2025) -
PhysicsNeRF: Physics-Guided 3D Reconstruction from Sparse Views
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025) -
EMPATHIA: Multi-Faceted Human-AI Collaboration for Refugee Integration
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)