Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification
Fuente:
arXiv
Saved in:
| Main Authors: | Sharan, S P, Choi, Minkyu, Shah, Sahil, Goel, Harsh, Omama, Mohammad, Chinchali, Sandeep |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Neuro-Symbolic Video Understanding
by: Choi, Minkyu, et al.
Published: (2024)
by: Choi, Minkyu, et al.
Published: (2024)
We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback
by: Choi, Minkyu, et al.
Published: (2025)
by: Choi, Minkyu, et al.
Published: (2025)
NeuS-QA: Grounding Long-Form Video Understanding in Temporal Logic and Neuro-Symbolic Reasoning
by: Shah, Sahil, et al.
Published: (2025)
by: Shah, Sahil, et al.
Published: (2025)
ObjectAlign: Neuro-Symbolic Object Consistency Verification and Correction
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
LE-NeuS: Latency-Efficient Neuro-Symbolic Video Understanding via Adaptive Temporal Verification
by: Liang, Shawn, et al.
Published: (2026)
by: Liang, Shawn, et al.
Published: (2026)
A Challenge to Build Neuro-Symbolic Video Agents
by: Shah, Sahil, et al.
Published: (2025)
by: Shah, Sahil, et al.
Published: (2025)
SSR: A Generic Framework for Text-Aided Map Compression for Localization
by: Omama, Mohammad, et al.
Published: (2026)
by: Omama, Mohammad, et al.
Published: (2026)
SynDiff-AD: Improving Semantic Segmentation and End-to-End Autonomous Driving with Synthetic Data from Latent Diffusion Models
by: Goel, Harsh, et al.
Published: (2024)
by: Goel, Harsh, et al.
Published: (2024)
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
Real-Time Privacy Preservation for Robot Visual Perception
by: Choi, Minkyu, et al.
Published: (2025)
by: Choi, Minkyu, et al.
Published: (2025)
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
by: Li, Po-han, et al.
Published: (2026)
by: Li, Po-han, et al.
Published: (2026)
NEURO-GUARD: Neuro-Symbolic Generalization and Unbiased Adaptive Routing for Diagnostics -- Explainable Medical AI
by: Urooj, Midhat, et al.
Published: (2025)
by: Urooj, Midhat, et al.
Published: (2025)
VIRO: Robust and Efficient Neuro-Symbolic Reasoning with Verification for Referring Expression Comprehension
by: Park, Hyejin, et al.
Published: (2026)
by: Park, Hyejin, et al.
Published: (2026)
Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach
by: Urooj, Midhat, et al.
Published: (2025)
by: Urooj, Midhat, et al.
Published: (2025)
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
XAI-MeD: Explainable Knowledge Guided Neuro-Symbolic Framework for Domain Generalization and Rare Class Detection in Medical Imaging
by: Urooj, Midhat, et al.
Published: (2026)
by: Urooj, Midhat, et al.
Published: (2026)
CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
PhonemeFake: Redefining Deepfake Realism with Language-Driven Segmental Manipulation and Adaptive Bilevel Detection
by: Baser, Oguzhan, et al.
Published: (2025)
by: Baser, Oguzhan, et al.
Published: (2025)
LazyVLM: Neuro-Symbolic Approach to Video Analytics
by: Jian, Xiangru, et al.
Published: (2025)
by: Jian, Xiangru, et al.
Published: (2025)
ViLLa: A Neuro-Symbolic approach for Animal Monitoring
by: Koduri, Harsha
Published: (2025)
by: Koduri, Harsha
Published: (2025)
When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations
by: Lalai, Harsh Nishant, et al.
Published: (2026)
by: Lalai, Harsh Nishant, et al.
Published: (2026)
Neuro-Symbolic Concepts
by: Mao, Jiayuan, et al.
Published: (2025)
by: Mao, Jiayuan, et al.
Published: (2025)
Formal Reasoning About Confidence and Automated Verification of Neural Networks
by: Afzal, Mohammad, et al.
Published: (2025)
by: Afzal, Mohammad, et al.
Published: (2025)
Image Clustering Conditioned on Text Criteria
by: Kwon, Sehyun, et al.
Published: (2023)
by: Kwon, Sehyun, et al.
Published: (2023)
ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models
by: Rawte, Vipula, et al.
Published: (2024)
by: Rawte, Vipula, et al.
Published: (2024)
Neuro Symbolic Knowledge Reasoning for Procedural Video Question Answering
by: Fernando, Basura, et al.
Published: (2025)
by: Fernando, Basura, et al.
Published: (2025)
QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing
by: Mittal, Garvit Kumar, et al.
Published: (2026)
by: Mittal, Garvit Kumar, et al.
Published: (2026)
Improving Interpretability and Accuracy in Neuro-Symbolic Rule Extraction Using Class-Specific Sparse Filters
by: Padalkar, Parth, et al.
Published: (2025)
by: Padalkar, Parth, et al.
Published: (2025)
Can VLMs Reason Robustly? A Neuro-Symbolic Investigation
by: Chen, Weixin, et al.
Published: (2026)
by: Chen, Weixin, et al.
Published: (2026)
Robot-Enabled Machine Learning-Based Diagnosis of Gastric Cancer Polyps Using Partial Surface Tactile Imaging
by: Kapuria, Siddhartha, et al.
Published: (2024)
by: Kapuria, Siddhartha, et al.
Published: (2024)
Elucidating Optimal Reward-Diversity Tradeoffs in Text-to-Image Diffusion Models
by: Jena, Rohit, et al.
Published: (2024)
by: Jena, Rohit, et al.
Published: (2024)
AsymLoc: Towards Asymmetric Feature Matching for Efficient Visual Localization
by: Omama, Mohammad, et al.
Published: (2026)
by: Omama, Mohammad, et al.
Published: (2026)
Neuro-Symbolic Scene Graph Conditioning for Synthetic Image Dataset Generation
by: Savazzi, Giacomo, et al.
Published: (2025)
by: Savazzi, Giacomo, et al.
Published: (2025)
DynamicEval: Rethinking Evaluation for Dynamic Text-to-Video Synthesis
by: Babu, Nithin C., et al.
Published: (2025)
by: Babu, Nithin C., et al.
Published: (2025)
FIFO-Diffusion: Generating Infinite Videos from Text without Training
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models
by: Gu, Jing, et al.
Published: (2025)
by: Gu, Jing, et al.
Published: (2025)
VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance
by: Taesiri, Mohammad Reza, et al.
Published: (2025)
by: Taesiri, Mohammad Reza, et al.
Published: (2025)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
by: Lim, Youngsun, et al.
Published: (2024)
by: Lim, Youngsun, et al.
Published: (2024)
VisualPredicator: Learning Abstract World Models with Neuro-Symbolic Predicates for Robot Planning
by: Liang, Yichao, et al.
Published: (2024)
by: Liang, Yichao, et al.
Published: (2024)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
by: Jahangard, Simindokht, et al.
Published: (2025)
by: Jahangard, Simindokht, et al.
Published: (2025)
Similar Items
-
Towards Neuro-Symbolic Video Understanding
by: Choi, Minkyu, et al.
Published: (2024) -
We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback
by: Choi, Minkyu, et al.
Published: (2025) -
NeuS-QA: Grounding Long-Form Video Understanding in Temporal Logic and Neuro-Symbolic Reasoning
by: Shah, Sahil, et al.
Published: (2025) -
ObjectAlign: Neuro-Symbolic Object Consistency Verification and Correction
by: Munir, Mustafa, et al.
Published: (2025) -
LE-NeuS: Latency-Efficient Neuro-Symbolic Video Understanding via Adaptive Temporal Verification
by: Liang, Shawn, et al.
Published: (2026)