Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, William, Wu, Xindi, Deng, Zhiwei, Tureci, Esin, Russakovsky, Olga |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Compositional Tuning
by: Wu, Xindi, et al.
Published: (2025)
by: Wu, Xindi, et al.
Published: (2025)
Personalized Generative Models for Contextual Debiasing
by: Liang, Xinran, et al.
Published: (2026)
by: Liang, Xinran, et al.
Published: (2026)
Vision-Language Dataset Distillation
by: Wu, Xindi, et al.
Published: (2023)
by: Wu, Xindi, et al.
Published: (2023)
ICONS: Influence Consensus for Vision-Language Data Selection
by: Wu, Xindi, et al.
Published: (2024)
by: Wu, Xindi, et al.
Published: (2024)
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
by: Wu, Xindi, et al.
Published: (2024)
by: Wu, Xindi, et al.
Published: (2024)
Bias at the End of the Score
by: Magid, Salma Abdel, et al.
Published: (2026)
by: Magid, Salma Abdel, et al.
Published: (2026)
ImageNet-OOD: Deciphering Modern Out-of-Distribution Detection Algorithms
by: Yang, William, et al.
Published: (2023)
by: Yang, William, et al.
Published: (2023)
SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation
by: Pei, Yuhan, et al.
Published: (2024)
by: Pei, Yuhan, et al.
Published: (2024)
A Sampling-Based Domain Generalization Study with Diffusion Generative Models
by: Zhu, Ye, et al.
Published: (2023)
by: Zhu, Ye, et al.
Published: (2023)
Seeing Beyond the Scene: Analyzing and Mitigating Background Bias in Action Recognition
by: Zhou, Ellie, et al.
Published: (2025)
by: Zhou, Ellie, et al.
Published: (2025)
D2D: Detector-to-Differentiable Critic for Improved Numeracy in Text-to-Image Generation
by: Yoo, Nobline, et al.
Published: (2025)
by: Yoo, Nobline, et al.
Published: (2025)
The Silent Assistant: NoiseQuery as Implicit Guidance for Goal-Driven Image Generation
by: Wang, Ruoyu, et al.
Published: (2024)
by: Wang, Ruoyu, et al.
Published: (2024)
Video Models Reason Early: Exploiting Plan Commitment for Maze Solving
by: Newman, Kaleb, et al.
Published: (2026)
by: Newman, Kaleb, et al.
Published: (2026)
D$^3$: Scaling Up Deepfake Detection by Learning from Discrepancy
by: Yang, Yongqi, et al.
Published: (2024)
by: Yang, Yongqi, et al.
Published: (2024)
The Impact of Coreset Selection on Spurious Correlations and Group Robustness
by: Dharmasiri, Amaya, et al.
Published: (2025)
by: Dharmasiri, Amaya, et al.
Published: (2025)
Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents
by: Liu, Dayong, et al.
Published: (2025)
by: Liu, Dayong, et al.
Published: (2025)
Fine-Grained Prototypes Distillation for Few-Shot Object Detection
by: Wang, Zichen, et al.
Published: (2024)
by: Wang, Zichen, et al.
Published: (2024)
Beyond Global Alignment: Fine-Grained Motion-Language Retrieval via Pyramidal Shapley-Taylor Learning
by: Chen, Hanmo, et al.
Published: (2026)
by: Chen, Hanmo, et al.
Published: (2026)
Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation
by: Ng, Kam Woh, et al.
Published: (2025)
by: Ng, Kam Woh, et al.
Published: (2025)
Consistency Beyond Contrast: Enhancing Open-Vocabulary Object Detection Robustness via Contextual Consistency Learning
by: Li, Bozhao, et al.
Published: (2026)
by: Li, Bozhao, et al.
Published: (2026)
Fine-Grained Zero-Shot Object Detection
by: Ma, Hongxu, et al.
Published: (2025)
by: Ma, Hongxu, et al.
Published: (2025)
Fine-Grained Scene Image Classification with Modality-Agnostic Adapter
by: Wang, Yiqun, et al.
Published: (2024)
by: Wang, Yiqun, et al.
Published: (2024)
Attributed Synthetic Data Generation for Zero-shot Domain-specific Image Classification
by: Wang, Shijian, et al.
Published: (2025)
by: Wang, Shijian, et al.
Published: (2025)
Progressively Exploring and Exploiting Inference Data to Break Fine-Grained Classification Barrier
by: Zhao, Li-Jun, et al.
Published: (2024)
by: Zhao, Li-Jun, et al.
Published: (2024)
Modeling Fine-Grained Hand-Object Dynamics for Egocentric Video Representation Learning
by: Pei, Baoqi, et al.
Published: (2025)
by: Pei, Baoqi, et al.
Published: (2025)
SGIA: Enhancing Fine-Grained Visual Classification with Sequence Generative Image Augmentation
by: Liao, Qiyu, et al.
Published: (2024)
by: Liao, Qiyu, et al.
Published: (2024)
GeoDE: a Geographically Diverse Evaluation Dataset for Object Recognition
by: Ramaswamy, Vikram V., et al.
Published: (2023)
by: Ramaswamy, Vikram V., et al.
Published: (2023)
Glo-VLMs: Leveraging Vision-Language Models for Fine-Grained Diseased Glomerulus Classification
by: Guo, Zhenhao, et al.
Published: (2025)
by: Guo, Zhenhao, et al.
Published: (2025)
Cross-Hierarchical Bidirectional Consistency Learning for Fine-Grained Visual Classification
by: Gao, Pengxiang, et al.
Published: (2025)
by: Gao, Pengxiang, et al.
Published: (2025)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
by: Geigle, Gregor, et al.
Published: (2024)
by: Geigle, Gregor, et al.
Published: (2024)
FineXtrol: Controllable Motion Generation via Fine-Grained Text
by: Shen, Keming, et al.
Published: (2025)
by: Shen, Keming, et al.
Published: (2025)
Bidirectional Logits Tree: Pursuing Granularity Reconcilement in Fine-Grained Classification
by: Lu, Zhiguang, et al.
Published: (2024)
by: Lu, Zhiguang, et al.
Published: (2024)
Hybrid Feature Collaborative Reconstruction Network for Few-Shot Fine-Grained Image Classification
by: Qiu, Shulei, et al.
Published: (2024)
by: Qiu, Shulei, et al.
Published: (2024)
Boundless: Generating Photorealistic Synthetic Data for Object Detection in Urban Streetscapes
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)
by: Turkcan, Mehmet Kerem, et al.
Published: (2024)
Exploration of Class Center for Fine-Grained Visual Classification
by: Yao, Hang, et al.
Published: (2024)
by: Yao, Hang, et al.
Published: (2024)
FACTS: Fine-Grained Action Classification for Tactical Sports
by: Lai, Christopher, et al.
Published: (2024)
by: Lai, Christopher, et al.
Published: (2024)
Analyzing the Roles of Language and Vision in Learning from Limited Data
by: Chen, Allison, et al.
Published: (2024)
by: Chen, Allison, et al.
Published: (2024)
MCFNet: A Multimodal Collaborative Fusion Network for Fine-Grained Semantic Classification
by: Qiao, Yang, et al.
Published: (2025)
by: Qiao, Yang, et al.
Published: (2025)
Combining Discrepancy-Confusion Uncertainty and Calibration Diversity for Active Fine-Grained Image Classification
by: Jin, Yinghao, et al.
Published: (2025)
by: Jin, Yinghao, et al.
Published: (2025)
FiGO: Fine-Grained Object Counting without Annotations
by: D'Alessandro, Adriano, et al.
Published: (2025)
by: D'Alessandro, Adriano, et al.
Published: (2025)
Similar Items
-
Visual Compositional Tuning
by: Wu, Xindi, et al.
Published: (2025) -
Personalized Generative Models for Contextual Debiasing
by: Liang, Xinran, et al.
Published: (2026) -
Vision-Language Dataset Distillation
by: Wu, Xindi, et al.
Published: (2023) -
ICONS: Influence Consensus for Vision-Language Data Selection
by: Wu, Xindi, et al.
Published: (2024) -
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
by: Wu, Xindi, et al.
Published: (2024)