SegDebias: Test-Time Bias Mitigation for ViT-Based CLIP via Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Fangyu, Cai, Yujun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ViT$^3$: Unlocking Test-Time Training in Vision
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
von: Puy, Gilles, et al.
Veröffentlicht: (2026)
von: Puy, Gilles, et al.
Veröffentlicht: (2026)
Applying ViT in Generalized Few-shot Semantic Segmentation
von: Geng, Liyuan, et al.
Veröffentlicht: (2024)
von: Geng, Liyuan, et al.
Veröffentlicht: (2024)
Test-Time Adaptation for Height Completion via Self-Supervised ViT Features and Monocular Foundation Models
von: Rafaeli, Osher, et al.
Veröffentlicht: (2026)
von: Rafaeli, Osher, et al.
Veröffentlicht: (2026)
Deeper Inside Deep ViT
von: Hong, Sungrae
Veröffentlicht: (2025)
von: Hong, Sungrae
Veröffentlicht: (2025)
SPAR: Single-Pass Any-Resolution ViT for Open-vocabulary Segmentation
von: Kombol, Naomi, et al.
Veröffentlicht: (2026)
von: Kombol, Naomi, et al.
Veröffentlicht: (2026)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
Few-Shot Class-Incremental Model Attribution Using Learnable Representation From CLIP-ViT Features
von: Lee, Hanbyul, et al.
Veröffentlicht: (2025)
von: Lee, Hanbyul, et al.
Veröffentlicht: (2025)
DeNAS-ViT: Data Efficient NAS-Optimized Vision Transformer for Ultrasound Image Segmentation
von: Chen, Renqi, et al.
Veröffentlicht: (2024)
von: Chen, Renqi, et al.
Veröffentlicht: (2024)
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024)
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024)
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
ViG-Bias: Visually Grounded Bias Discovery and Mitigation
von: Marani, Badr-Eddine, et al.
Veröffentlicht: (2024)
von: Marani, Badr-Eddine, et al.
Veröffentlicht: (2024)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
Partial CLIP is Enough: Chimera-Seg for Zero-shot Semantic Segmentation
von: Chen, Jialei, et al.
Veröffentlicht: (2025)
von: Chen, Jialei, et al.
Veröffentlicht: (2025)
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
von: Caramia, Donato, et al.
Veröffentlicht: (2025)
von: Caramia, Donato, et al.
Veröffentlicht: (2025)
Do Students Debias Like Teachers? On the Distillability of Bias Mitigation Methods
von: Cheng, Jiali, et al.
Veröffentlicht: (2025)
von: Cheng, Jiali, et al.
Veröffentlicht: (2025)
VidEoMT: Your ViT is Secretly Also a Video Segmentation Model
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)
Sub-token ViT Embedding via Stochastic Resonance Transformers
von: Lao, Dong, et al.
Veröffentlicht: (2023)
von: Lao, Dong, et al.
Veröffentlicht: (2023)
Rethinking Random Masking in Self-Distillation on ViT
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
ViT-5: Vision Transformers for The Mid-2020s
von: Wang, Feng, et al.
Veröffentlicht: (2026)
von: Wang, Feng, et al.
Veröffentlicht: (2026)
ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation
von: Wang, Jingyun, et al.
Veröffentlicht: (2024)
von: Wang, Jingyun, et al.
Veröffentlicht: (2024)
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer
von: Tai, Yu-Shan, et al.
Veröffentlicht: (2024)
von: Tai, Yu-Shan, et al.
Veröffentlicht: (2024)
ERVD: An Efficient and Robust ViT-Based Distillation Framework for Remote Sensing Image Retrieval
von: Dong, Le, et al.
Veröffentlicht: (2024)
von: Dong, Le, et al.
Veröffentlicht: (2024)
Alias-Free ViT: Fractional Shift Invariance via Linear Attention
von: Michaeli, Hagay, et al.
Veröffentlicht: (2025)
von: Michaeli, Hagay, et al.
Veröffentlicht: (2025)
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
von: Chew, Oscar, et al.
Veröffentlicht: (2026)
von: Chew, Oscar, et al.
Veröffentlicht: (2026)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
GradiSeg: Gradient-Guided Gaussian Segmentation with Enhanced 3D Boundary Precision
von: Li, Zehao, et al.
Veröffentlicht: (2024)
von: Li, Zehao, et al.
Veröffentlicht: (2024)
Colinearity Decay: Training Quantization-Friendly ViTs with Outlier Decay
von: Tong, Jin, et al.
Veröffentlicht: (2026)
von: Tong, Jin, et al.
Veröffentlicht: (2026)
ViTCAE: ViT-based Class-conditioned Autoencoder
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
APHQ-ViT: Post-Training Quantization with Average Perturbation Hessian Based Reconstruction for Vision Transformers
von: Wu, Zhuguanyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhuguanyu, et al.
Veröffentlicht: (2025)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
One-Shot Multilingual Font Generation Via ViT
von: Wang, Zhiheng, et al.
Veröffentlicht: (2024)
von: Wang, Zhiheng, et al.
Veröffentlicht: (2024)
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
FastPose-ViT: A Vision Transformer for Real-Time Spacecraft Pose Estimation
von: Ancey, Pierre, et al.
Veröffentlicht: (2025)
von: Ancey, Pierre, et al.
Veröffentlicht: (2025)
A Dual-Mode ViT-Conditioned Diffusion Framework with an Adaptive Conditioning Bridge for Breast Cancer Segmentation
von: Singh, Prateek, et al.
Veröffentlicht: (2025)
von: Singh, Prateek, et al.
Veröffentlicht: (2025)
UniRefiner: Teaching Pre-trained ViTs to Self-Dispose Dross via Contrastive Register
von: Qiu, Congpei, et al.
Veröffentlicht: (2026)
von: Qiu, Congpei, et al.
Veröffentlicht: (2026)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ViT$^3$: Unlocking Test-Time Training in Vision
von: Han, Dongchen, et al.
Veröffentlicht: (2025) -
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025) -
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
von: Puy, Gilles, et al.
Veröffentlicht: (2026) -
Applying ViT in Generalized Few-shot Semantic Segmentation
von: Geng, Liyuan, et al.
Veröffentlicht: (2024) -
Test-Time Adaptation for Height Completion via Self-Supervised ViT Features and Monocular Foundation Models
von: Rafaeli, Osher, et al.
Veröffentlicht: (2026)