A CLIP-Powered Framework for Robust and Generalizable Data Selection
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Suorong, Ye, Peng, Ouyang, Wanli, Zhou, Dongzhan, Shen, Furao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Dynamic Data Selection Meets Data Augmentation
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
Multimodal-Guided Dynamic Dataset Pruning for Robust and Efficient Data-Centric Learning
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
EntAugment: Entropy-Driven Adaptive Data Augmentation Framework for Image Classification
by: Yang, Suorong, et al.
Published: (2024)
by: Yang, Suorong, et al.
Published: (2024)
IPF-RDA: An Information-Preserving Framework for Robust Data Augmentation
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
RL-Selector: Reinforcement Learning-Guided Data Selection via Redundancy Assessment
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
$Δ$-AttnMask: Attention-Guided Masked Hidden States for Efficient Data Selection and Augmentation
by: Hu, Jucheng, et al.
Published: (2025)
by: Hu, Jucheng, et al.
Published: (2025)
AdaAugment: A Tuning-Free and Adaptive Approach to Enhance Data Augmentation
by: Yang, Suorong, et al.
Published: (2024)
by: Yang, Suorong, et al.
Published: (2024)
Data Agent: Learning to Select Data via End-to-End Dynamic Optimization
by: Yang, Suorong, et al.
Published: (2026)
by: Yang, Suorong, et al.
Published: (2026)
Structure-Level Disentangled Diffusion for Few-Shot Chinese Font Generation
by: Li, Jie, et al.
Published: (2026)
by: Li, Jie, et al.
Published: (2026)
HiSplat: Hierarchical 3D Gaussian Splatting for Generalizable Sparse-View Reconstruction
by: Tang, Shengji, et al.
Published: (2024)
by: Tang, Shengji, et al.
Published: (2024)
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs
by: Tu, Chongjun, et al.
Published: (2025)
by: Tu, Chongjun, et al.
Published: (2025)
Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
On-the-Fly Data Augmentation via Gradient-Guided and Sample-Aware Influence Estimation
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
CMT: A Cascade MAR with Topology Predictor for Multimodal Conditional CAD Generation
by: Wu, Jianyu, et al.
Published: (2025)
by: Wu, Jianyu, et al.
Published: (2025)
LOCR: Location-Guided Transformer for Optical Character Recognition
by: Sun, Yu, et al.
Published: (2024)
by: Sun, Yu, et al.
Published: (2024)
DeeCLIP: A Robust and Generalizable Transformer-Based Framework for Detecting AI-Generated Images
by: Keita, Mamadou, et al.
Published: (2025)
by: Keita, Mamadou, et al.
Published: (2025)
Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge
by: Chen, Ruiming, et al.
Published: (2025)
by: Chen, Ruiming, et al.
Published: (2025)
One-shot Optimized Steering Vector for Hallucination Mitigation for VLMs
by: Shi, Youxu, et al.
Published: (2026)
by: Shi, Youxu, et al.
Published: (2026)
Exposing Hallucinations To Suppress Them: VLMs Representation Editing With Generative Anchors
by: Shi, Youxu, et al.
Published: (2025)
by: Shi, Youxu, et al.
Published: (2025)
Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision
by: Li, Minglei, et al.
Published: (2024)
by: Li, Minglei, et al.
Published: (2024)
CLIP-DFGS: A Hard Sample Mining Method for CLIP in Generalizable Person Re-Identification
by: Zhao, Huazhong, et al.
Published: (2024)
by: Zhao, Huazhong, et al.
Published: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
by: Zhang, Yaning, et al.
Published: (2024)
by: Zhang, Yaning, et al.
Published: (2024)
Unlocking the Hidden Potential of CLIP in Generalizable Deepfake Detection
by: Yermakov, Andrii, et al.
Published: (2025)
by: Yermakov, Andrii, et al.
Published: (2025)
Generalizable Prompt Learning of CLIP: A Brief Overview
by: Cui, Fangming, et al.
Published: (2025)
by: Cui, Fangming, et al.
Published: (2025)
RGM: A Robust Generalizable Matching Model
by: Zhang, Songyan, et al.
Published: (2023)
by: Zhang, Songyan, et al.
Published: (2023)
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
by: Mei, Guofeng, et al.
Published: (2025)
by: Mei, Guofeng, et al.
Published: (2025)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
by: Zhang, Zilun, et al.
Published: (2022)
by: Zhang, Zilun, et al.
Published: (2022)
CLIP-Mamba: CLIP Pretrained Mamba Models with OOD and Hessian Evaluation
by: Huang, Weiquan, et al.
Published: (2024)
by: Huang, Weiquan, et al.
Published: (2024)
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
by: Xie, Jingyou, et al.
Published: (2024)
by: Xie, Jingyou, et al.
Published: (2024)
Improving Weakly Supervised Temporal Action Localization by Exploiting Multi-resolution Information in Temporal Domain
by: Su, Rui, et al.
Published: (2025)
by: Su, Rui, et al.
Published: (2025)
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
by: Su, Rui, et al.
Published: (2019)
by: Su, Rui, et al.
Published: (2019)
SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal Models
by: Zhang, Ruiyang, et al.
Published: (2026)
by: Zhang, Ruiyang, et al.
Published: (2026)
Seeing Before Reasoning: A Unified Framework for Generalizable and Explainable Fake Image Detection
by: Lin, Kaiqing, et al.
Published: (2025)
by: Lin, Kaiqing, et al.
Published: (2025)
PostCast: Generalizable Postprocessing for Precipitation Nowcasting via Unsupervised Blurriness Modeling
by: Gong, Junchao, et al.
Published: (2024)
by: Gong, Junchao, et al.
Published: (2024)
NeuRodin: A Two-stage Framework for High-Fidelity Neural Surface Reconstruction
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
Where Am I and What Will I See: An Auto-Regressive Model for Spatial Localization and View Prediction
by: Chen, Junyi, et al.
Published: (2024)
by: Chen, Junyi, et al.
Published: (2024)
TokenCarve: Information-Preserving Visual Token Compression in Multimodal Large Language Models
by: Tan, Xudong, et al.
Published: (2025)
by: Tan, Xudong, et al.
Published: (2025)
Transfer CLIP for Generalizable Image Denoising
by: Cheng, Jun, et al.
Published: (2024)
by: Cheng, Jun, et al.
Published: (2024)
Unlearning the Noisy Correspondence Makes CLIP More Robust
by: Han, Haochen, et al.
Published: (2025)
by: Han, Haochen, et al.
Published: (2025)
PolyReal: A Benchmark for Real-World Polymer Science Workflows
by: Liu, Wanhao, et al.
Published: (2026)
by: Liu, Wanhao, et al.
Published: (2026)
Similar Items
-
When Dynamic Data Selection Meets Data Augmentation
by: Yang, Suorong, et al.
Published: (2025) -
Multimodal-Guided Dynamic Dataset Pruning for Robust and Efficient Data-Centric Learning
by: Yang, Suorong, et al.
Published: (2025) -
EntAugment: Entropy-Driven Adaptive Data Augmentation Framework for Image Classification
by: Yang, Suorong, et al.
Published: (2024) -
IPF-RDA: An Information-Preserving Framework for Robust Data Augmentation
by: Yang, Suorong, et al.
Published: (2025) -
RL-Selector: Reinforcement Learning-Guided Data Selection via Redundancy Assessment
by: Yang, Suorong, et al.
Published: (2025)