Role-SynthCLIP: A Role Play Driven Diverse Synthetic Data Approach
Fuente:
arXiv
Salvato in:
| Autori principali: | Huangfu, Yuanxiang, Wang, Chaochao, Wang, Weilei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
SynthVerse: A Large-Scale Diverse Synthetic Dataset for Point Tracking
di: Zhao, Weiguang, et al.
Pubblicazione: (2026)
di: Zhao, Weiguang, et al.
Pubblicazione: (2026)
Language Plays a Pivotal Role in the Object-Attribute Compositional Generalization of CLIP
di: Abbasi, Reza, et al.
Pubblicazione: (2024)
di: Abbasi, Reza, et al.
Pubblicazione: (2024)
Synth-Align: Improving Trustworthiness in Vision-Language Model with Synthetic Preference Data Alignment
di: Wijaya, Robert, et al.
Pubblicazione: (2024)
di: Wijaya, Robert, et al.
Pubblicazione: (2024)
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels
di: Singh, Darshan, et al.
Pubblicazione: (2024)
di: Singh, Darshan, et al.
Pubblicazione: (2024)
SynPlay: Large-Scale Synthetic Human Data with Real-World Diversity for Aerial-View Perception
di: Yim, Jinsub, et al.
Pubblicazione: (2024)
di: Yim, Jinsub, et al.
Pubblicazione: (2024)
Synth It Like KITTI: Synthetic Data Generation for Object Detection in Driving Scenarios
di: Marcus, Richard, et al.
Pubblicazione: (2025)
di: Marcus, Richard, et al.
Pubblicazione: (2025)
SoccerSynth-Detection: A Synthetic Dataset for Soccer Player Detection
di: Qin, Haobin, et al.
Pubblicazione: (2025)
di: Qin, Haobin, et al.
Pubblicazione: (2025)
AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks
di: Li, You, et al.
Pubblicazione: (2024)
di: Li, You, et al.
Pubblicazione: (2024)
Know "No" Better: A Data-Driven Approach for Enhancing Negation Awareness in CLIP
di: Park, Junsung, et al.
Pubblicazione: (2025)
di: Park, Junsung, et al.
Pubblicazione: (2025)
Scene Graph Generation with Role-Playing Large Language Models
di: Chen, Guikun, et al.
Pubblicazione: (2024)
di: Chen, Guikun, et al.
Pubblicazione: (2024)
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions
di: Liu, Yanqing, et al.
Pubblicazione: (2024)
di: Liu, Yanqing, et al.
Pubblicazione: (2024)
AmodalSynthDrive: A Synthetic Amodal Perception Dataset for Autonomous Driving
di: Sekkat, Ahmed Rida, et al.
Pubblicazione: (2023)
di: Sekkat, Ahmed Rida, et al.
Pubblicazione: (2023)
SynthForensics: Benchmarking and Evaluating People-Centric Synthetic Video Deepfakes
di: Leotta, Roberto, et al.
Pubblicazione: (2026)
di: Leotta, Roberto, et al.
Pubblicazione: (2026)
SynthPID: P&ID digitization from Topology-Preserving Synthetic Data
di: Prasad, Suraj, et al.
Pubblicazione: (2026)
di: Prasad, Suraj, et al.
Pubblicazione: (2026)
SynthSeg-Agents: Multi-Agent Synthetic Data Generation for Zero-Shot Weakly Supervised Semantic Segmentation
di: Wu, Wangyu, et al.
Pubblicazione: (2025)
di: Wu, Wangyu, et al.
Pubblicazione: (2025)
Deciphering the Role of Representation Disentanglement: Investigating Compositional Generalization in CLIP Models
di: Abbasi, Reza, et al.
Pubblicazione: (2024)
di: Abbasi, Reza, et al.
Pubblicazione: (2024)
SynthRAR: Ring Artifacts Reduction in CT with Unrolled Network and Synthetic Data Training
di: Yang, Hongxu, et al.
Pubblicazione: (2026)
di: Yang, Hongxu, et al.
Pubblicazione: (2026)
Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent
di: Tang, Yihong, et al.
Pubblicazione: (2026)
di: Tang, Yihong, et al.
Pubblicazione: (2026)
CDUL: CLIP-Driven Unsupervised Learning for Multi-Label Image Classification
di: Abdelfattah, Rabab, et al.
Pubblicazione: (2023)
di: Abdelfattah, Rabab, et al.
Pubblicazione: (2023)
psPRF:Pansharpening Planar Neural Radiance Field for Generalized 3D Reconstruction Satellite Imagery
di: Zhang, Tongtong, et al.
Pubblicazione: (2024)
di: Zhang, Tongtong, et al.
Pubblicazione: (2024)
Cross-Lingual SynthDocs: A Large-Scale Synthetic Corpus for Any to Arabic OCR and Document Understanding
di: Al-Homoud, Haneen, et al.
Pubblicazione: (2025)
di: Al-Homoud, Haneen, et al.
Pubblicazione: (2025)
Zero-Shot Class Unlearning in CLIP with Synthetic Samples
di: Kravets, A., et al.
Pubblicazione: (2024)
di: Kravets, A., et al.
Pubblicazione: (2024)
CultureCLIP: Empowering CLIP with Cultural Awareness through Synthetic Images and Contextualized Captions
di: Huang, Yuchen, et al.
Pubblicazione: (2025)
di: Huang, Yuchen, et al.
Pubblicazione: (2025)
TripletCLIP: Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives
di: Patel, Maitreya, et al.
Pubblicazione: (2024)
di: Patel, Maitreya, et al.
Pubblicazione: (2024)
Data Distribution Distilled Generative Model for Generalized Zero-Shot Recognition
di: Wang, Yijie, et al.
Pubblicazione: (2024)
di: Wang, Yijie, et al.
Pubblicazione: (2024)
NeuroCLIP: Neuromorphic Data Understanding by CLIP and SNN
di: Guo, Yufei, et al.
Pubblicazione: (2023)
di: Guo, Yufei, et al.
Pubblicazione: (2023)
Harnessing Textual Semantic Priors for Knowledge Transfer and Refinement in CLIP-Driven Continual Learning
di: He, Lingfeng, et al.
Pubblicazione: (2025)
di: He, Lingfeng, et al.
Pubblicazione: (2025)
SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic Faces
di: Chaturvedi, Sumit, et al.
Pubblicazione: (2025)
di: Chaturvedi, Sumit, et al.
Pubblicazione: (2025)
ARM: A Learnable, Plug-and-Play Module for CLIP-based Open-vocabulary Semantic Segmentation
di: Liu, Ziquan, et al.
Pubblicazione: (2025)
di: Liu, Ziquan, et al.
Pubblicazione: (2025)
SuperCLIP: CLIP with Simple Classification Supervision
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts
di: Wang, Juan, et al.
Pubblicazione: (2026)
di: Wang, Juan, et al.
Pubblicazione: (2026)
GUESS:GradUally Enriching SyntheSis for Text-Driven Human Motion Generation
di: Gao, Xuehao, et al.
Pubblicazione: (2024)
di: Gao, Xuehao, et al.
Pubblicazione: (2024)
RoleMotion: A Large-Scale Dataset towards Robust Scene-Specific Role-Playing Motion Synthesis with Fine-grained Descriptions
di: Peng, Junran, et al.
Pubblicazione: (2025)
di: Peng, Junran, et al.
Pubblicazione: (2025)
ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation
di: Wang, Jingyun, et al.
Pubblicazione: (2024)
di: Wang, Jingyun, et al.
Pubblicazione: (2024)
Infrared and Visible Image Fusion with Language-Driven Loss in CLIP Embedding Space
di: Wang, Yuhao, et al.
Pubblicazione: (2024)
di: Wang, Yuhao, et al.
Pubblicazione: (2024)
SAM-MIL: A Spatial Contextual Aware Multiple Instance Learning Approach for Whole Slide Image Classification
di: Fang, Heng, et al.
Pubblicazione: (2024)
di: Fang, Heng, et al.
Pubblicazione: (2024)
Synth$^2$: Boosting Visual-Language Models with Synthetic Captions and Image Embeddings
di: Sharifzadeh, Sahand, et al.
Pubblicazione: (2024)
di: Sharifzadeh, Sahand, et al.
Pubblicazione: (2024)
CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
di: Kang, Bin, et al.
Pubblicazione: (2025)
di: Kang, Bin, et al.
Pubblicazione: (2025)
Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation
di: Fei, Yuanchen, et al.
Pubblicazione: (2026)
di: Fei, Yuanchen, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024) -
SynthVerse: A Large-Scale Diverse Synthetic Dataset for Point Tracking
di: Zhao, Weiguang, et al.
Pubblicazione: (2026) -
Language Plays a Pivotal Role in the Object-Attribute Compositional Generalization of CLIP
di: Abbasi, Reza, et al.
Pubblicazione: (2024) -
Synth-Align: Improving Trustworthiness in Vision-Language Model with Synthetic Preference Data Alignment
di: Wijaya, Robert, et al.
Pubblicazione: (2024) -
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels
di: Singh, Darshan, et al.
Pubblicazione: (2024)