Physics-Grounded Motion Forecasting via Equation Discovery for Trajectory-Guided Image-to-Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Feng, Tao, Zhao, Xianbing, Chen, Zhenhua, Wong, Tien Tsin, Rezatofighi, Hamid, Haffari, Gholamreza, Qu, Lizhen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
por: Feng, Tao, et al.
Publicado: (2025)
por: Feng, Tao, et al.
Publicado: (2025)
Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents
por: Huang, Sukai, et al.
Publicado: (2026)
por: Huang, Sukai, et al.
Publicado: (2026)
CausalScore: An Automatic Reference-Free Metric for Assessing Response Relevance in Open-Domain Dialogue Systems
por: Feng, Tao, et al.
Publicado: (2024)
por: Feng, Tao, et al.
Publicado: (2024)
On the Reliability of Large Language Models for Causal Discovery
por: Feng, Tao, et al.
Publicado: (2024)
por: Feng, Tao, et al.
Publicado: (2024)
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
por: Hua, Yuncheng, et al.
Publicado: (2024)
por: Hua, Yuncheng, et al.
Publicado: (2024)
Physics-based Scene Layout Generation from Human Motion
por: Li, Jianan, et al.
Publicado: (2024)
por: Li, Jianan, et al.
Publicado: (2024)
Learning to Control Physically-simulated 3D Characters via Generating and Mimicking 2D Motions
por: Li, Jianan, et al.
Publicado: (2025)
por: Li, Jianan, et al.
Publicado: (2025)
Towards Probing Speech-Specific Risks in Large Multimodal Models: A Taxonomy, Benchmark, and Insights
por: Yang, Hao, et al.
Publicado: (2024)
por: Yang, Hao, et al.
Publicado: (2024)
Zero-Shot Privacy-Aware Text Rewriting via Iterative Tree Search
por: Huang, Shuo, et al.
Publicado: (2025)
por: Huang, Shuo, et al.
Publicado: (2025)
Jigsaw Puzzles: Splitting Harmful Questions to Jailbreak Large Language Models
por: Yang, Hao, et al.
Publicado: (2024)
por: Yang, Hao, et al.
Publicado: (2024)
Audio Is the Achilles' Heel: Red Teaming Audio Large Multimodal Models
por: Yang, Hao, et al.
Publicado: (2024)
por: Yang, Hao, et al.
Publicado: (2024)
Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models
por: Yang, Hao, et al.
Publicado: (2025)
por: Yang, Hao, et al.
Publicado: (2025)
Causal Discovery Inspired Unsupervised Domain Adaptation for Emotion-Cause Pair Extraction
por: Hua, Yuncheng, et al.
Publicado: (2024)
por: Hua, Yuncheng, et al.
Publicado: (2024)
IMO: Greedy Layer-Wise Sparse Representation Learning for Out-of-Distribution Text Classification with Pre-trained Models
por: Feng, Tao, et al.
Publicado: (2024)
por: Feng, Tao, et al.
Publicado: (2024)
Evidence-based Distributional Alignment for Large Language Models
por: Pham, Viet-Thanh, et al.
Publicado: (2026)
por: Pham, Viet-Thanh, et al.
Publicado: (2026)
The Best of Both Worlds: Bridging Quality and Diversity in Data Selection with Bipartite Graph
por: Wu, Minghao, et al.
Publicado: (2024)
por: Wu, Minghao, et al.
Publicado: (2024)
Mixture-of-Skills: Learning to Optimize Data Usage for Fine-Tuning Large Language Models
por: Wu, Minghao, et al.
Publicado: (2024)
por: Wu, Minghao, et al.
Publicado: (2024)
MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation
por: Xing, Jinbo, et al.
Publicado: (2025)
por: Xing, Jinbo, et al.
Publicado: (2025)
Learning in Order! A Sequential Strategy to Learn Invariant Features for Multimodal Sentiment Analysis
por: Zhao, Xianbing, et al.
Publicado: (2024)
por: Zhao, Xianbing, et al.
Publicado: (2024)
Importance-Aware Data Augmentation for Document-Level Neural Machine Translation
por: Wu, Minghao, et al.
Publicado: (2024)
por: Wu, Minghao, et al.
Publicado: (2024)
Unbiased Sliced Wasserstein Kernels for High-Quality Audio Captioning
por: Luong, Manh, et al.
Publicado: (2025)
por: Luong, Manh, et al.
Publicado: (2025)
ACCESS : A Benchmark for Abstract Causal Event Discovery and Reasoning
por: Vo, Vy, et al.
Publicado: (2025)
por: Vo, Vy, et al.
Publicado: (2025)
Adapting Large Language Models for Document-Level Machine Translation
por: Wu, Minghao, et al.
Publicado: (2024)
por: Wu, Minghao, et al.
Publicado: (2024)
Resurfacing Paralinguistic Awareness in Large Audio Language Models
por: Yang, Hao, et al.
Publicado: (2026)
por: Yang, Hao, et al.
Publicado: (2026)
LiveCultureBench: a Multi-Agent, Multi-Cultural Benchmark for Large Language Models in Dynamic Social Simulations
por: Pham, Viet-Thanh, et al.
Publicado: (2026)
por: Pham, Viet-Thanh, et al.
Publicado: (2026)
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
por: Ma, Xin, et al.
Publicado: (2025)
por: Ma, Xin, et al.
Publicado: (2025)
Text-Guided Texturing by Synchronized Multi-View Diffusion
por: Liu, Yuxin, et al.
Publicado: (2023)
por: Liu, Yuxin, et al.
Publicado: (2023)
Synchronized Multi‐Frame Diffusion for Temporally Consistent Video Stylization
por: Minshan Xie, et al.
Publicado: (2025)
por: Minshan Xie, et al.
Publicado: (2025)
Training-free Stylized Text-to-Image Generation with Fast Inference
por: Ma, Xin, et al.
Publicado: (2025)
por: Ma, Xin, et al.
Publicado: (2025)
SCAR: Data Selection via Style Consistency-Aware Response Ranking for Efficient Instruction-Tuning of Large Language Models
por: Li, Zhuang, et al.
Publicado: (2024)
por: Li, Zhuang, et al.
Publicado: (2024)
RIDE: Enhancing Large Language Model Alignment through Restyled In-Context Learning Demonstration Exemplars
por: Hua, Yuncheng, et al.
Publicado: (2025)
por: Hua, Yuncheng, et al.
Publicado: (2025)
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
por: Ehsanpour, Mahsa, et al.
Publicado: (2024)
por: Ehsanpour, Mahsa, et al.
Publicado: (2024)
Taming Reversible Halftoning via Predictive Luminance
por: Lau, Cheuk-Kit, et al.
Publicado: (2023)
por: Lau, Cheuk-Kit, et al.
Publicado: (2023)
SituatedThinker: Grounding LLM Reasoning with Real-World through Situated Thinking
por: Liu, Junnan, et al.
Publicado: (2025)
por: Liu, Junnan, et al.
Publicado: (2025)
VIEW2SPACE: Studying Multi-View Visual Reasoning from Sparse Observations
por: Ke, Fucai, et al.
Publicado: (2026)
por: Ke, Fucai, et al.
Publicado: (2026)
GyroCopter: Differential Bearing Measuring Trajectory Planner for Tracking and Localizing Radio Frequency Sources
por: Chen, Fei, et al.
Publicado: (2024)
por: Chen, Fei, et al.
Publicado: (2024)
NAVER: A Neuro-Symbolic Compositional Automaton for Visual Grounding with Explicit Logic Reasoning
por: Cai, Zhixi, et al.
Publicado: (2025)
por: Cai, Zhixi, et al.
Publicado: (2025)
Let's Negotiate! A Survey of Negotiation Dialogue Systems
por: Zhan, Haolan, et al.
Publicado: (2024)
por: Zhan, Haolan, et al.
Publicado: (2024)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
por: Jahangard, Simindokht, et al.
Publicado: (2025)
por: Jahangard, Simindokht, et al.
Publicado: (2025)
VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior
por: Yang, Xindi, et al.
Publicado: (2025)
por: Yang, Xindi, et al.
Publicado: (2025)
Ejemplares similares
-
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
por: Feng, Tao, et al.
Publicado: (2025) -
Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents
por: Huang, Sukai, et al.
Publicado: (2026) -
CausalScore: An Automatic Reference-Free Metric for Assessing Response Relevance in Open-Domain Dialogue Systems
por: Feng, Tao, et al.
Publicado: (2024) -
On the Reliability of Large Language Models for Causal Discovery
por: Feng, Tao, et al.
Publicado: (2024) -
Assistive Large Language Model Agents for Socially-Aware Negotiation Dialogues
por: Hua, Yuncheng, et al.
Publicado: (2024)