Gespeichert in:
| Hauptverfasser: | Wu, Xindi, Hwang, Hee Seung, Kirichenko, Polina, Tureci, Esin, Russakovsky, Olga |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2504.21850 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification
von: Yang, William, et al.
Veröffentlicht: (2025)
von: Yang, William, et al.
Veröffentlicht: (2025)
The Impact of Coreset Selection on Spurious Correlations and Group Robustness
von: Dharmasiri, Amaya, et al.
Veröffentlicht: (2025)
von: Dharmasiri, Amaya, et al.
Veröffentlicht: (2025)
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
von: Wu, Xindi, et al.
Veröffentlicht: (2024)
von: Wu, Xindi, et al.
Veröffentlicht: (2024)
Personalized Generative Models for Contextual Debiasing
von: Liang, Xinran, et al.
Veröffentlicht: (2026)
von: Liang, Xinran, et al.
Veröffentlicht: (2026)
Vision-Language Dataset Distillation
von: Wu, Xindi, et al.
Veröffentlicht: (2023)
von: Wu, Xindi, et al.
Veröffentlicht: (2023)
Bias at the End of the Score
von: Magid, Salma Abdel, et al.
Veröffentlicht: (2026)
von: Magid, Salma Abdel, et al.
Veröffentlicht: (2026)
ICONS: Influence Consensus for Vision-Language Data Selection
von: Wu, Xindi, et al.
Veröffentlicht: (2024)
von: Wu, Xindi, et al.
Veröffentlicht: (2024)
D2D: Detector-to-Differentiable Critic for Improved Numeracy in Text-to-Image Generation
von: Yoo, Nobline, et al.
Veröffentlicht: (2025)
von: Yoo, Nobline, et al.
Veröffentlicht: (2025)
Video Models Reason Early: Exploiting Plan Commitment for Maze Solving
von: Newman, Kaleb, et al.
Veröffentlicht: (2026)
von: Newman, Kaleb, et al.
Veröffentlicht: (2026)
ImageNet-OOD: Deciphering Modern Out-of-Distribution Detection Algorithms
von: Yang, William, et al.
Veröffentlicht: (2023)
von: Yang, William, et al.
Veröffentlicht: (2023)
Reinforced Fast Weights with Next-Sequence Prediction
von: Hwang, Hee Seung, et al.
Veröffentlicht: (2026)
von: Hwang, Hee Seung, et al.
Veröffentlicht: (2026)
Seeing Beyond the Scene: Analyzing and Mitigating Background Bias in Action Recognition
von: Zhou, Ellie, et al.
Veröffentlicht: (2025)
von: Zhou, Ellie, et al.
Veröffentlicht: (2025)
D$^3$: Scaling Up Deepfake Detection by Learning from Discrepancy
von: Yang, Yongqi, et al.
Veröffentlicht: (2024)
von: Yang, Yongqi, et al.
Veröffentlicht: (2024)
The Silent Assistant: NoiseQuery as Implicit Guidance for Goal-Driven Image Generation
von: Wang, Ruoyu, et al.
Veröffentlicht: (2024)
von: Wang, Ruoyu, et al.
Veröffentlicht: (2024)
What's in Common? Multimodal Models Hallucinate When Reasoning Across Scenes
von: Ross, Candace, et al.
Veröffentlicht: (2025)
von: Ross, Candace, et al.
Veröffentlicht: (2025)
Attention IoU: Examining Biases in CelebA using Attention Maps
von: Serianni, Aaron, et al.
Veröffentlicht: (2025)
von: Serianni, Aaron, et al.
Veröffentlicht: (2025)
Unifying Specialized Visual Encoders for Video Language Models
von: Chung, Jihoon, et al.
Veröffentlicht: (2025)
von: Chung, Jihoon, et al.
Veröffentlicht: (2025)
A Sampling-Based Domain Generalization Study with Diffusion Generative Models
von: Zhu, Ye, et al.
Veröffentlicht: (2023)
von: Zhu, Ye, et al.
Veröffentlicht: (2023)
SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation
von: Pei, Yuhan, et al.
Veröffentlicht: (2024)
von: Pei, Yuhan, et al.
Veröffentlicht: (2024)
Analyzing the Roles of Language and Vision in Learning from Limited Data
von: Chen, Allison, et al.
Veröffentlicht: (2024)
von: Chen, Allison, et al.
Veröffentlicht: (2024)
ECLIPSE: Efficient Continual Learning in Panoptic Segmentation with Visual Prompt Tuning
von: Kim, Beomyoung, et al.
Veröffentlicht: (2024)
von: Kim, Beomyoung, et al.
Veröffentlicht: (2024)
Understanding the Detrimental Class-level Effects of Data Augmentation
von: Kirichenko, Polina, et al.
Veröffentlicht: (2023)
von: Kirichenko, Polina, et al.
Veröffentlicht: (2023)
Explain Before You Answer: A Survey on Compositional Visual Reasoning
von: Ke, Fucai, et al.
Veröffentlicht: (2025)
von: Ke, Fucai, et al.
Veröffentlicht: (2025)
Interactivity x Explainability: Toward Understanding How Interactivity Can Improve Computer Vision Explanations
von: Panigrahi, Indu, et al.
Veröffentlicht: (2025)
von: Panigrahi, Indu, et al.
Veröffentlicht: (2025)
Visual Generation Tuning
von: Guo, Jiahao, et al.
Veröffentlicht: (2025)
von: Guo, Jiahao, et al.
Veröffentlicht: (2025)
Bayesian Approximation-Based Trajectory Prediction and Tracking with 4D Radar
von: Kim, Dong-In, et al.
Veröffentlicht: (2025)
von: Kim, Dong-In, et al.
Veröffentlicht: (2025)
GeoDE: a Geographically Diverse Evaluation Dataset for Object Recognition
von: Ramaswamy, Vikram V., et al.
Veröffentlicht: (2023)
von: Ramaswamy, Vikram V., et al.
Veröffentlicht: (2023)
Modeling Caption Diversity in Contrastive Vision-Language Pretraining
von: Lavoie, Samuel, et al.
Veröffentlicht: (2024)
von: Lavoie, Samuel, et al.
Veröffentlicht: (2024)
Availability-aware Sensor Fusion via Unified Canonical Space
von: Paek, Dong-Hee, et al.
Veröffentlicht: (2025)
von: Paek, Dong-Hee, et al.
Veröffentlicht: (2025)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
von: Zhang, Xiaoman, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaoman, et al.
Veröffentlicht: (2023)
4DR P2T: 4D Radar Tensor Synthesis with Point Clouds
von: Jung, Woo-Jin, et al.
Veröffentlicht: (2025)
von: Jung, Woo-Jin, et al.
Veröffentlicht: (2025)
Visual Spatial Tuning
von: Yang, Rui, et al.
Veröffentlicht: (2025)
von: Yang, Rui, et al.
Veröffentlicht: (2025)
Enhanced 3D Object Detection via Diverse Feature Representations of 4D Radar Tensor
von: Song, Seung-Hyun, et al.
Veröffentlicht: (2025)
von: Song, Seung-Hyun, et al.
Veröffentlicht: (2025)
Visual Fourier Prompt Tuning
von: Zeng, Runjia, et al.
Veröffentlicht: (2024)
von: Zeng, Runjia, et al.
Veröffentlicht: (2024)
Visual-RFT: Visual Reinforcement Fine-Tuning
von: Liu, Ziyu, et al.
Veröffentlicht: (2025)
von: Liu, Ziyu, et al.
Veröffentlicht: (2025)
Minimal Interaction Separated Tuning: A New Paradigm for Visual Adaptation
von: Tang, Ningyuan, et al.
Veröffentlicht: (2024)
von: Tang, Ningyuan, et al.
Veröffentlicht: (2024)
Personalized Visual Instruction Tuning
von: Pi, Renjie, et al.
Veröffentlicht: (2024)
von: Pi, Renjie, et al.
Veröffentlicht: (2024)
Generative Visual Instruction Tuning
von: Hernandez, Jefferson, et al.
Veröffentlicht: (2024)
von: Hernandez, Jefferson, et al.
Veröffentlicht: (2024)
Comparison Visual Instruction Tuning
von: Lin, Wei, et al.
Veröffentlicht: (2024)
von: Lin, Wei, et al.
Veröffentlicht: (2024)
FastTracker: Real-Time and Accurate Visual Tracking
von: Hashempoor, Hamidreza, et al.
Veröffentlicht: (2025)
von: Hashempoor, Hamidreza, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification
von: Yang, William, et al.
Veröffentlicht: (2025) -
The Impact of Coreset Selection on Spurious Correlations and Group Robustness
von: Dharmasiri, Amaya, et al.
Veröffentlicht: (2025) -
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
von: Wu, Xindi, et al.
Veröffentlicht: (2024) -
Personalized Generative Models for Contextual Debiasing
von: Liang, Xinran, et al.
Veröffentlicht: (2026) -
Vision-Language Dataset Distillation
von: Wu, Xindi, et al.
Veröffentlicht: (2023)