PropTest: Automatic Property Testing for Improved Visual Programming
Fuente:
arXiv
Saved in:
| Main Authors: | Koo, Jaywon, Yang, Ziyan, Cascante-Bonilla, Paola, Ray, Baishakhi, Ordonez, Vicente |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Text-to-Image and Text-to-Video Synthesis with a Conditional Fréchet Distance
by: Koo, Jaywon, et al.
Published: (2025)
by: Koo, Jaywon, et al.
Published: (2025)
ProxyThinker: Test-Time Guidance through Small Visual Reasoners
by: Xiao, Zilin, et al.
Published: (2025)
by: Xiao, Zilin, et al.
Published: (2025)
Learning from Synthetic Data for Visual Grounding
by: He, Ruozhen, et al.
Published: (2024)
by: He, Ruozhen, et al.
Published: (2024)
Agentic Discovery with Active Hypothesis Exploration for Visual Recognition
by: Koo, Jaywon, et al.
Published: (2026)
by: Koo, Jaywon, et al.
Published: (2026)
Beyond Referring Expressions: Scenario Comprehension Visual Grounding
by: He, Ruozhen, et al.
Published: (2026)
by: He, Ruozhen, et al.
Published: (2026)
Grounding Language Models for Visual Entity Recognition
by: Xiao, Zilin, et al.
Published: (2024)
by: Xiao, Zilin, et al.
Published: (2024)
Improving Visual Grounding by Encouraging Consistent Gradient-based Explanations
by: Yang, Ziyan, et al.
Published: (2022)
by: Yang, Ziyan, et al.
Published: (2022)
Do VLMs Need Vision Transformers? Evaluating State Space Models as Vision Encoders
by: Kuo, Shang-Jui Ray, et al.
Published: (2026)
by: Kuo, Shang-Jui Ray, et al.
Published: (2026)
SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis
by: Sengupta, Kathakoli, et al.
Published: (2026)
by: Sengupta, Kathakoli, et al.
Published: (2026)
Natural Language Inference Improves Compositionality in Vision-Language Models
by: Cascante-Bonilla, Paola, et al.
Published: (2024)
by: Cascante-Bonilla, Paola, et al.
Published: (2024)
Can Hallucination Correction Improve Video-Language Alignment?
by: Zhao, Lingjun, et al.
Published: (2025)
by: Zhao, Lingjun, et al.
Published: (2025)
EntityBench: Towards Entity-Consistent Long-Range Multi-Shot Video Generation
by: He, Ruozhen, et al.
Published: (2026)
by: He, Ruozhen, et al.
Published: (2026)
ViUniT: Visual Unit Tests for More Robust Visual Programming
by: Panagopoulou, Artemis, et al.
Published: (2024)
by: Panagopoulou, Artemis, et al.
Published: (2024)
NoiseShift: Resolution-Aware Noise Recalibration for Better Low-Resolution Image Generation
by: He, Ruozhen, et al.
Published: (2025)
by: He, Ruozhen, et al.
Published: (2025)
Generative Visual Instruction Tuning
by: Hernandez, Jefferson, et al.
Published: (2024)
by: Hernandez, Jefferson, et al.
Published: (2024)
SCoRD: Subject-Conditional Relation Detection with Text-Augmented Data
by: Yang, Ziyan, et al.
Published: (2023)
by: Yang, Ziyan, et al.
Published: (2023)
Visual Personalization Turing Test
by: Abdal, Rameen, et al.
Published: (2026)
by: Abdal, Rameen, et al.
Published: (2026)
TPC: Test-time Procrustes Calibration for Diffusion-based Human Image Animation
by: Yoon, Sunjae, et al.
Published: (2024)
by: Yoon, Sunjae, et al.
Published: (2024)
Beyond Blanket Masking: Examining Granularity for Privacy Protection in Images Captured by Blind and Low Vision Users
by: Murrugarra-LLerena, Jeffri, et al.
Published: (2025)
by: Murrugarra-LLerena, Jeffri, et al.
Published: (2025)
EgoGroups: A Benchmark For Detecting Social Groups of People in the Wild
by: Murrugarra-Llerena, Jeffri, et al.
Published: (2026)
by: Murrugarra-Llerena, Jeffri, et al.
Published: (2026)
Adversarial Testing for Visual Grounding via Image-Aware Property Reduction
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination
by: Dai, Ming, et al.
Published: (2025)
by: Dai, Ming, et al.
Published: (2025)
Benchmark Designers Should "Train on the Test Set" to Expose Exploitable Non-Visual Shortcuts
by: Brown, Ellis, et al.
Published: (2025)
by: Brown, Ellis, et al.
Published: (2025)
Improving Large Vision and Language Models by Learning from a Panel of Peers
by: Hernandez, Jefferson, et al.
Published: (2025)
by: Hernandez, Jefferson, et al.
Published: (2025)
MinBackProp -- Backpropagating through Minimal Solvers
by: Sungatullina, Diana, et al.
Published: (2024)
by: Sungatullina, Diana, et al.
Published: (2024)
Test-Time Conditioning with Representation-Aligned Visual Features
by: Sereyjol-Garros, Nicolas, et al.
Published: (2026)
by: Sereyjol-Garros, Nicolas, et al.
Published: (2026)
Automatically Generating Visual Hallucination Test Cases for Multimodal Large Language Models
by: Liu, Zhongye, et al.
Published: (2024)
by: Liu, Zhongye, et al.
Published: (2024)
ViDA: Homeostatic Visual Domain Adapter for Continual Test Time Adaptation
by: Liu, Jiaming, et al.
Published: (2023)
by: Liu, Jiaming, et al.
Published: (2023)
S$^3$-TTA: Scale-Style Selection for Test-Time Augmentation in Biomedical Image Segmentation
by: Xie, Kangxian, et al.
Published: (2023)
by: Xie, Kangxian, et al.
Published: (2023)
Towards Visually Explaining Statistical Tests with Applications in Biomedical Imaging
by: Javanbakhat, Masoumeh, et al.
Published: (2026)
by: Javanbakhat, Masoumeh, et al.
Published: (2026)
VACT: A Video Automatic Causal Testing System and a Benchmark
by: Yang, Haotong, et al.
Published: (2025)
by: Yang, Haotong, et al.
Published: (2025)
Can Test-Time Scaling Improve World Foundation Model?
by: Cong, Wenyan, et al.
Published: (2025)
by: Cong, Wenyan, et al.
Published: (2025)
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction
by: Xiao, Zilin, et al.
Published: (2025)
by: Xiao, Zilin, et al.
Published: (2025)
Test3R: Learning to Reconstruct 3D at Test Time
by: Yuan, Yuheng, et al.
Published: (2025)
by: Yuan, Yuheng, et al.
Published: (2025)
Test-time Correction: An Online 3D Detection System via Visual Prompting
by: Zhang, Hanxue, et al.
Published: (2024)
by: Zhang, Hanxue, et al.
Published: (2024)
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement
by: Khan, Zaid, et al.
Published: (2024)
by: Khan, Zaid, et al.
Published: (2024)
EntProp: High Entropy Propagation for Improving Accuracy and Robustness
by: Enomoto, Shohei
Published: (2024)
by: Enomoto, Shohei
Published: (2024)
Test-time Distribution Learning Adapter for Cross-modal Visual Reasoning
by: Zhang, Yi, et al.
Published: (2024)
by: Zhang, Yi, et al.
Published: (2024)
LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs
by: Lou, Haoran, et al.
Published: (2025)
by: Lou, Haoran, et al.
Published: (2025)
The Overlooked Value of Test-time Reference Sets in Visual Place Recognition
by: Zaffar, Mubariz, et al.
Published: (2025)
by: Zaffar, Mubariz, et al.
Published: (2025)
Similar Items
-
Evaluating Text-to-Image and Text-to-Video Synthesis with a Conditional Fréchet Distance
by: Koo, Jaywon, et al.
Published: (2025) -
ProxyThinker: Test-Time Guidance through Small Visual Reasoners
by: Xiao, Zilin, et al.
Published: (2025) -
Learning from Synthetic Data for Visual Grounding
by: He, Ruozhen, et al.
Published: (2024) -
Agentic Discovery with Active Hypothesis Exploration for Visual Recognition
by: Koo, Jaywon, et al.
Published: (2026) -
Beyond Referring Expressions: Scenario Comprehension Visual Grounding
by: He, Ruozhen, et al.
Published: (2026)