SLANT: Spurious Logo ANalysis Toolkit
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qraitem, Maan, Teterwak, Piotr, Saenko, Kate, Plummer, Bryan A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Web Artifact Attacks Disrupt Vision Language Models
von: Qraitem, Maan, et al.
Veröffentlicht: (2025)
von: Qraitem, Maan, et al.
Veröffentlicht: (2025)
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
von: Qraitem, Maan, et al.
Veröffentlicht: (2023)
von: Qraitem, Maan, et al.
Veröffentlicht: (2023)
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
von: Qraitem, Maan, et al.
Veröffentlicht: (2024)
von: Qraitem, Maan, et al.
Veröffentlicht: (2024)
OP-LoRA: The Blessing of Dimensionality
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
CLAMP: Contrastive LAnguage Model Prompt-tuning
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
ERM++: An Improved Baseline for Domain Generalization
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)
Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation
von: Qraitem, Maan, et al.
Veröffentlicht: (2026)
von: Qraitem, Maan, et al.
Veröffentlicht: (2026)
Tell Me What's Next: Textual Foresight for Generic UI Representations
von: Burns, Andrea, et al.
Veröffentlicht: (2024)
von: Burns, Andrea, et al.
Veröffentlicht: (2024)
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
von: Yiu, Eunice, et al.
Veröffentlicht: (2024)
von: Yiu, Eunice, et al.
Veröffentlicht: (2024)
Koala: Key frame-conditioned long video-LLM
von: Tan, Reuben, et al.
Veröffentlicht: (2024)
von: Tan, Reuben, et al.
Veröffentlicht: (2024)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
von: Liu, Aoming, et al.
Veröffentlicht: (2025)
von: Liu, Aoming, et al.
Veröffentlicht: (2025)
Mull-Tokens: Modality-Agnostic Latent Thinking
von: Ray, Arijit, et al.
Veröffentlicht: (2025)
von: Ray, Arijit, et al.
Veröffentlicht: (2025)
LNL+K: Enhancing Learning with Noisy Labels Through Noise Source Knowledge Integration
von: Wang, Siqi, et al.
Veröffentlicht: (2023)
von: Wang, Siqi, et al.
Veröffentlicht: (2023)
Concept Arithmetics for Circumventing Concept Inhibition in Diffusion Models
von: Petsiuk, Vitali, et al.
Veröffentlicht: (2024)
von: Petsiuk, Vitali, et al.
Veröffentlicht: (2024)
CodeSCAN: ScreenCast ANalysis for Video Programming Tutorials
von: Naumann, Alexander, et al.
Veröffentlicht: (2024)
von: Naumann, Alexander, et al.
Veröffentlicht: (2024)
SCRAMBLe : Enhancing Multimodal LLM Compositionality with Synthetic Preference Data
von: Mishra, Samarth, et al.
Veröffentlicht: (2025)
von: Mishra, Samarth, et al.
Veröffentlicht: (2025)
Decompose, Mix, Adapt: A Unified Framework for Parameter-Efficient Neural Network Recombination and Compression
von: Tasnim, Nazia, et al.
Veröffentlicht: (2026)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2026)
RECAST: Reparameterized, Compact weight Adaptation for Sequential Tasks
von: Tasnim, Nazia, et al.
Veröffentlicht: (2024)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2024)
Enhancing Feature Diversity Boosts Channel-Adaptive Vision Transformers
von: Pham, Chau, et al.
Veröffentlicht: (2024)
von: Pham, Chau, et al.
Veröffentlicht: (2024)
SPARC: Score Prompting and Adaptive Fusion for Zero-Shot Multi-Label Recognition in Vision-Language Models
von: Miller, Kevin, et al.
Veröffentlicht: (2025)
von: Miller, Kevin, et al.
Veröffentlicht: (2025)
Enhancing Virtual Try-On with Synthetic Pairs and Error-Aware Noise Scheduling
von: Li, Nannan, et al.
Veröffentlicht: (2025)
von: Li, Nannan, et al.
Veröffentlicht: (2025)
Federated Adversarial Domain Adaptation
von: Peng, Xingchao, et al.
Veröffentlicht: (2019)
von: Peng, Xingchao, et al.
Veröffentlicht: (2019)
LogoSticker: Inserting Logos into Diffusion Models for Customized Generation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
FuTCR: Future-Targeted Contrast and Repulsion for Continual Panoptic Segmentation
von: Ikechukwu, Nicholas, et al.
Veröffentlicht: (2026)
von: Ikechukwu, Nicholas, et al.
Veröffentlicht: (2026)
Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain Generalization
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
von: Pham, Chau, et al.
Veröffentlicht: (2025)
von: Pham, Chau, et al.
Veröffentlicht: (2025)
LogoDiffuser: Training-Free Multilingual Logo Generation and Stylization via Letter-Aware Attention Control
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
von: Kang, Mingyu, et al.
Veröffentlicht: (2026)
SynCDR : Training Cross Domain Retrieval Models with Synthetic Data
von: Mishra, Samarth, et al.
Veröffentlicht: (2023)
von: Mishra, Samarth, et al.
Veröffentlicht: (2023)
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
von: Ray, Arijit, et al.
Veröffentlicht: (2024)
von: Ray, Arijit, et al.
Veröffentlicht: (2024)
Walk and Read Less: Improving the Efficiency of Vision-and-Language Navigation via Tuning-Free Multimodal Token Pruning
von: Qin, Wenda, et al.
Veröffentlicht: (2025)
von: Qin, Wenda, et al.
Veröffentlicht: (2025)
Logo-VGR: Visual Grounded Reasoning for Open-world Logo Recognition
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
PanoFree: Tuning-Free Holistic Multi-view Image Generation with Cross-view Self-Guidance
von: Liu, Aoming, et al.
Veröffentlicht: (2024)
von: Liu, Aoming, et al.
Veröffentlicht: (2024)
Reducing Spurious Correlation for Federated Domain Generalization
von: Ma, Shuran, et al.
Veröffentlicht: (2024)
von: Ma, Shuran, et al.
Veröffentlicht: (2024)
SLIM: Spuriousness Mitigation with Minimal Human Annotations
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024)
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024)
Focusing Image Generation to Mitigate Spurious Correlations
von: Li, Xuewei, et al.
Veröffentlicht: (2024)
von: Li, Xuewei, et al.
Veröffentlicht: (2024)
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
von: Cao, Yuxin, et al.
Veröffentlicht: (2023)
Seeing Isn't Orienting: A Cognitively Grounded Benchmark Reveals Systematic Orientation Failures in MLLMs Supplementary
von: Tasnim, Nazia, et al.
Veröffentlicht: (2026)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2026)
Seeing Isn't Orienting: A Cognitively Grounded Benchmark Reveals Systematic Orientation Failures in MLLMs
von: Tasnim, Nazia, et al.
Veröffentlicht: (2025)
von: Tasnim, Nazia, et al.
Veröffentlicht: (2025)
Spuriousness-Aware Meta-Learning for Learning Robust Classifiers
von: Zheng, Guangtao, et al.
Veröffentlicht: (2024)
von: Zheng, Guangtao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Web Artifact Attacks Disrupt Vision Language Models
von: Qraitem, Maan, et al.
Veröffentlicht: (2025) -
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
von: Qraitem, Maan, et al.
Veröffentlicht: (2023) -
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
von: Qraitem, Maan, et al.
Veröffentlicht: (2024) -
OP-LoRA: The Blessing of Dimensionality
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024) -
CLAMP: Contrastive LAnguage Model Prompt-tuning
von: Teterwak, Piotr, et al.
Veröffentlicht: (2023)