PhysVid: Physics Aware Local Conditioning for Generative Video Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Pathak, Saurabh, Arani, Elahe, Pechenizkiy, Mykola, Zonooz, Bahram |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
por: Semenov, Andrei, et al.
Publicado: (2024)
por: Semenov, Andrei, et al.
Publicado: (2024)
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
por: Lapin, Mykyta, et al.
Publicado: (2025)
por: Lapin, Mykyta, et al.
Publicado: (2025)
FedWCM: Unleashing the Potential of Momentum-based Federated Learning in Long-Tailed Scenarios
por: Li, Tianle, et al.
Publicado: (2025)
por: Li, Tianle, et al.
Publicado: (2025)
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
por: Wu, Jason, et al.
Publicado: (2026)
por: Wu, Jason, et al.
Publicado: (2026)
CASE: Contrastive Activation for Saliency Estimation
por: Williamson, Dane, et al.
Publicado: (2025)
por: Williamson, Dane, et al.
Publicado: (2025)
Tricks and Plug-ins for Gradient Boosting in Image Classification
por: Fang, Biyi, et al.
Publicado: (2025)
por: Fang, Biyi, et al.
Publicado: (2025)
Multimodal Ensemble with Conditional Feature Fusion for Dysgraphia Diagnosis in Children from Handwriting Samples
por: Kunhoth, Jayakanth, et al.
Publicado: (2024)
por: Kunhoth, Jayakanth, et al.
Publicado: (2024)
Localizing Adversarial Attacks To Produces More Imperceptible Noise
por: Reddy, Pavan, et al.
Publicado: (2025)
por: Reddy, Pavan, et al.
Publicado: (2025)
Salient Concept-Aware Generative Data Augmentation
por: Zhao, Tianchen, et al.
Publicado: (2025)
por: Zhao, Tianchen, et al.
Publicado: (2025)
SpectralCA: Bi-Directional Cross-Attention for Next-Generation UAV Hyperspectral Vision
por: Brovko, D. V.
Publicado: (2025)
por: Brovko, D. V.
Publicado: (2025)
Butter: Frequency Consistency and Hierarchical Fusion for Autonomous Driving Object Detection
por: Lin, Xiaojian, et al.
Publicado: (2025)
por: Lin, Xiaojian, et al.
Publicado: (2025)
Object detection in adverse weather conditions for autonomous vehicles using Instruct Pix2Pix
por: Gurbindo, Unai, et al.
Publicado: (2025)
por: Gurbindo, Unai, et al.
Publicado: (2025)
Implementing Adaptations for Vision AutoRegressive Model
por: Shaikh, Kaif, et al.
Publicado: (2025)
por: Shaikh, Kaif, et al.
Publicado: (2025)
Classification of Cattle Behavior and Detection of Heat (Estrus) using Sensor Data
por: Dhakshinamoorthy, Druva, et al.
Publicado: (2025)
por: Dhakshinamoorthy, Druva, et al.
Publicado: (2025)
CytoNet: A Foundation Model for the Human Cerebral Cortex at Cellular Resolution
por: Schiffer, Christian, et al.
Publicado: (2025)
por: Schiffer, Christian, et al.
Publicado: (2025)
ZO-DARTS++: An Efficient and Size-Variable Zeroth-Order Neural Architecture Search Algorithm
por: Xie, Lunchen, et al.
Publicado: (2025)
por: Xie, Lunchen, et al.
Publicado: (2025)
TRIGS: Trojan Identification from Gradient-based Signatures
por: Hussein, Mohamed E., et al.
Publicado: (2023)
por: Hussein, Mohamed E., et al.
Publicado: (2023)
DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
por: Ho, Darryl, et al.
Publicado: (2025)
por: Ho, Darryl, et al.
Publicado: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
por: Gasparino, Mateus Valverde, et al.
Publicado: (2024)
por: Gasparino, Mateus Valverde, et al.
Publicado: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
por: Gupta, Sunny, et al.
Publicado: (2024)
por: Gupta, Sunny, et al.
Publicado: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
por: Kashyap, Pankhi, et al.
Publicado: (2024)
por: Kashyap, Pankhi, et al.
Publicado: (2024)
Lost in Latent Space: Disentangled Models and the Challenge of Combinatorial Generalisation
por: Montero, Milton L., et al.
Publicado: (2022)
por: Montero, Milton L., et al.
Publicado: (2022)
QoSGMAA: A Robust Multi-Order Graph Attention and Adversarial Framework for Sparse QoS Prediction
por: Du, Guanchen, et al.
Publicado: (2025)
por: Du, Guanchen, et al.
Publicado: (2025)
Robust Taxi Fare Prediction Under Noisy Conditions: A Comparative Study of GAT, TimesNet, and XGBoost
por: Moorthy, Padmavathi
Publicado: (2025)
por: Moorthy, Padmavathi
Publicado: (2025)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
por: Qesaraku, Bjorna, et al.
Publicado: (2025)
por: Qesaraku, Bjorna, et al.
Publicado: (2025)
RDPO: Real Data Preference Optimization for Physics Consistency Video Generation
por: Qian, Wenxu, et al.
Publicado: (2025)
por: Qian, Wenxu, et al.
Publicado: (2025)
Motion Attribution for Video Generation
por: Wu, Xindi, et al.
Publicado: (2026)
por: Wu, Xindi, et al.
Publicado: (2026)
NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection
por: Chaowakarn, Krittin, et al.
Publicado: (2025)
por: Chaowakarn, Krittin, et al.
Publicado: (2025)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
por: Yang, Shan
Publicado: (2026)
por: Yang, Shan
Publicado: (2026)
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
por: Ji, Binbin, et al.
Publicado: (2025)
por: Ji, Binbin, et al.
Publicado: (2025)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
por: Romero, Angel, et al.
Publicado: (2025)
por: Romero, Angel, et al.
Publicado: (2025)
FEDTAIL: Federated Long-Tailed Domain Generalization with Sharpness-Guided Gradient Matching
por: Gupta, Sunny, et al.
Publicado: (2025)
por: Gupta, Sunny, et al.
Publicado: (2025)
eStonefish-Scenes: A Sim-to-Real Validated and Robot-Centric Event-based Optical Flow Dataset for Underwater Vehicles
por: Mansour, Jad, et al.
Publicado: (2025)
por: Mansour, Jad, et al.
Publicado: (2025)
eCARLA-scenes: A synthetically generated dataset for event-based optical flow prediction
por: Mansour, Jad, et al.
Publicado: (2024)
por: Mansour, Jad, et al.
Publicado: (2024)
Learning Association via Track-Detection Matching for Multi-Object Tracking
por: Adžemović, Momir
Publicado: (2025)
por: Adžemović, Momir
Publicado: (2025)
AGOP as Explanation: From Feature Learning to Per-Sample Attribution in Image Classifiers
por: Katakam, Raj Kiran Gupta
Publicado: (2026)
por: Katakam, Raj Kiran Gupta
Publicado: (2026)
EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique
por: Zhu, Chenglin, et al.
Publicado: (2025)
por: Zhu, Chenglin, et al.
Publicado: (2025)
Canonical Space Representation for 4D Panoptic Segmentation of Articulated Objects
por: Gomes, Manuel, et al.
Publicado: (2025)
por: Gomes, Manuel, et al.
Publicado: (2025)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
por: Asanuma, Haruka, et al.
Publicado: (2025)
por: Asanuma, Haruka, et al.
Publicado: (2025)
Task Singular Vectors: Reducing Task Interference in Model Merging
por: Gargiulo, Antonio Andrea, et al.
Publicado: (2024)
por: Gargiulo, Antonio Andrea, et al.
Publicado: (2024)
Ejemplares similares
-
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
por: Semenov, Andrei, et al.
Publicado: (2024) -
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
por: Lapin, Mykyta, et al.
Publicado: (2025) -
FedWCM: Unleashing the Potential of Momentum-based Federated Learning in Long-Tailed Scenarios
por: Li, Tianle, et al.
Publicado: (2025) -
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
por: Wu, Jason, et al.
Publicado: (2026) -
CASE: Contrastive Activation for Saliency Estimation
por: Williamson, Dane, et al.
Publicado: (2025)