Saved in:
| Main Author: | Sarkar, Krisanu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.00832 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RLS3: RL-Based Synthetic Sample Selection to Enhance Spatial Reasoning in Vision-Language Models for Indoor Autonomous Perception
by: Waite, Joshua R., et al.
Published: (2025)
by: Waite, Joshua R., et al.
Published: (2025)
Synthetic Data is an Elegant GIFT for Continual Vision-Language Models
by: Wu, Bin, et al.
Published: (2025)
by: Wu, Bin, et al.
Published: (2025)
Training a Computer Vision Model for Commercial Bakeries with Primarily Synthetic Images
by: Schmitt, Thomas H., et al.
Published: (2024)
by: Schmitt, Thomas H., et al.
Published: (2024)
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
by: Sarkar, Sreetama, et al.
Published: (2025)
by: Sarkar, Sreetama, et al.
Published: (2025)
Block Selective Reprogramming for On-device Training of Vision Transformers
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
Leveraging Vision Language Models for Specialized Agricultural Tasks
by: Arshad, Muhammad Arbab, et al.
Published: (2024)
by: Arshad, Muhammad Arbab, et al.
Published: (2024)
Prompting Medical Large Vision-Language Models to Diagnose Pathologies by Visual Question Answering
by: Guo, Danfeng, et al.
Published: (2024)
by: Guo, Danfeng, et al.
Published: (2024)
SEVD: Synthetic Event-based Vision Dataset for Ego and Fixed Traffic Perception
by: Aliminati, Manideep Reddy, et al.
Published: (2024)
by: Aliminati, Manideep Reddy, et al.
Published: (2024)
The Perceptual Bandwidth Bottleneck in Vision-Language Models: Active Visual Reasoning via Sequential Experimental Design
by: Liu, Anjie, et al.
Published: (2026)
by: Liu, Anjie, et al.
Published: (2026)
ImageNet-Think-250K: A Large-Scale Synthetic Dataset for Multimodal Reasoning for Vision Language Models
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
StableSemantics: A Synthetic Language-Vision Dataset of Semantic Representations in Naturalistic Images
by: Zawar, Rushikesh, et al.
Published: (2024)
by: Zawar, Rushikesh, et al.
Published: (2024)
LangGap: Diagnosing and Closing the Language Gap in Vision-Language-Action Models
by: Hou, Yuchen, et al.
Published: (2026)
by: Hou, Yuchen, et al.
Published: (2026)
Utilizing Synthetic Data for Medical Vision-Language Pre-training: Bypassing the Need for Real Images
by: Liu, Che, et al.
Published: (2023)
by: Liu, Che, et al.
Published: (2023)
SynthVision -- Harnessing Minimal Input for Maximal Output in Computer Vision Models using Synthetic Image data
by: Kularathne, Yudara, et al.
Published: (2024)
by: Kularathne, Yudara, et al.
Published: (2024)
Powerful Design of Small Vision Transformer on CIFAR10
by: Wu, Gent
Published: (2025)
by: Wu, Gent
Published: (2025)
Provably Improving Generalization of Few-Shot Models with Synthetic Data
by: Nguyen, Lan-Cuong, et al.
Published: (2025)
by: Nguyen, Lan-Cuong, et al.
Published: (2025)
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
by: Cai, Rui, et al.
Published: (2025)
by: Cai, Rui, et al.
Published: (2025)
Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value
by: Xu, Yixian, et al.
Published: (2025)
by: Xu, Yixian, et al.
Published: (2025)
Diagnosing Shortcut-Induced Rigidity in Continual Learning: The Einstellung Rigidity Index (ERI)
by: Gu, Kai, et al.
Published: (2025)
by: Gu, Kai, et al.
Published: (2025)
Synthetically Enhanced: Unveiling Synthetic Data's Potential in Medical Imaging Research
by: Khosravi, Bardia, et al.
Published: (2023)
by: Khosravi, Bardia, et al.
Published: (2023)
Fusing Foveal Fixations Using Linear Retinal Transformations and Bayesian Experimental Design
by: Williams, Christopher K. I.
Published: (2025)
by: Williams, Christopher K. I.
Published: (2025)
Hybrid Synthetic Data Generation with Domain Randomization Enables Zero-Shot Vision-Based Part Inspection Under Extreme Class Imbalance
by: Mei, Ruo-Syuan, et al.
Published: (2025)
by: Mei, Ruo-Syuan, et al.
Published: (2025)
Bi-LORA: A Vision-Language Approach for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL
by: Feng, Yichen, et al.
Published: (2025)
by: Feng, Yichen, et al.
Published: (2025)
Synthetic Cardiac MRI Image Generation using Deep Generative Models
by: Kumarasinghe, Ishan, et al.
Published: (2026)
by: Kumarasinghe, Ishan, et al.
Published: (2026)
Why Prototypes Collapse: Diagnosing and Preventing Partial Collapse in Prototypical Self-Supervised Learning
by: Arteaga, Gabriel Y., et al.
Published: (2025)
by: Arteaga, Gabriel Y., et al.
Published: (2025)
Vision Language Models are Biased
by: Vo, An, et al.
Published: (2025)
by: Vo, An, et al.
Published: (2025)
Understanding the Failure Modes of Out-of-Distribution Generalization
by: Nagarajan, Vaishnavh, et al.
Published: (2020)
by: Nagarajan, Vaishnavh, et al.
Published: (2020)
How to Train your Text-to-Image Model: Evaluating Design Choices for Synthetic Training Captions
by: Brack, Manuel, et al.
Published: (2025)
by: Brack, Manuel, et al.
Published: (2025)
Left-Right Symmetry Breaking in CLIP-style Vision-Language Models Trained on Synthetic Spatial-Relation Data
by: Yamamoto, Takaki, et al.
Published: (2026)
by: Yamamoto, Takaki, et al.
Published: (2026)
Detect Fake with Fake: Leveraging Synthetic Data-driven Representation for Synthetic Image Detection
by: Otake, Hina, et al.
Published: (2024)
by: Otake, Hina, et al.
Published: (2024)
Building a General SimCLR Self-Supervised Foundation Model Across Neurological Diseases to Advance 3D Brain MRI Diagnoses
by: Kaczmarek, Emily, et al.
Published: (2025)
by: Kaczmarek, Emily, et al.
Published: (2025)
Coordinated Robustness Evaluation Framework for Vision-Language Models
by: Babu, Ashwin Ramesh, et al.
Published: (2025)
by: Babu, Ashwin Ramesh, et al.
Published: (2025)
SynGen-Vision: Synthetic Data Generation for training industrial vision models
by: Dubey, Alpana, et al.
Published: (2025)
by: Dubey, Alpana, et al.
Published: (2025)
Training Feature Attribution for Vision Models
by: Bacha, Aziz, et al.
Published: (2025)
by: Bacha, Aziz, et al.
Published: (2025)
Tutorial on Diffusion Models for Imaging and Vision
by: Chan, Stanley H.
Published: (2024)
by: Chan, Stanley H.
Published: (2024)
Benchmarking the Attribution Quality of Vision Models
by: Hesse, Robin, et al.
Published: (2024)
by: Hesse, Robin, et al.
Published: (2024)
On the Use of Anchoring for Training Vision Models
by: Narayanaswamy, Vivek, et al.
Published: (2024)
by: Narayanaswamy, Vivek, et al.
Published: (2024)
Failures to Find Transferable Image Jailbreaks Between Vision-Language Models
by: Schaeffer, Rylan, et al.
Published: (2024)
by: Schaeffer, Rylan, et al.
Published: (2024)
Similar Items
-
RLS3: RL-Based Synthetic Sample Selection to Enhance Spatial Reasoning in Vision-Language Models for Indoor Autonomous Perception
by: Waite, Joshua R., et al.
Published: (2025) -
Synthetic Data is an Elegant GIFT for Continual Vision-Language Models
by: Wu, Bin, et al.
Published: (2025) -
Training a Computer Vision Model for Commercial Bakeries with Primarily Synthetic Images
by: Schmitt, Thomas H., et al.
Published: (2024) -
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
by: Sarkar, Sreetama, et al.
Published: (2025) -
Block Selective Reprogramming for On-device Training of Vision Transformers
by: Sarkar, Sreetama, et al.
Published: (2024)