Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Minho, Park, Sunghyun, Yun, Jooyeol, Choo, Jaegul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
von: Park, Minho, et al.
Veröffentlicht: (2025)
von: Park, Minho, et al.
Veröffentlicht: (2025)
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
von: Jo, Kyungmin, et al.
Veröffentlicht: (2025)
von: Jo, Kyungmin, et al.
Veröffentlicht: (2025)
CA-LoRA: Concept-Aware LoRA for Domain-Aligned Segmentation Dataset Generation
von: Park, Minho, et al.
Veröffentlicht: (2025)
von: Park, Minho, et al.
Veröffentlicht: (2025)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
von: Chung, Chaeyeon, et al.
Veröffentlicht: (2024)
von: Chung, Chaeyeon, et al.
Veröffentlicht: (2024)
Enabling Region-Specific Control via Lassos in Point-Based Colorization
von: Lee, Sanghyeon, et al.
Veröffentlicht: (2024)
von: Lee, Sanghyeon, et al.
Veröffentlicht: (2024)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
von: Kim, Jeongho, et al.
Veröffentlicht: (2025)
von: Kim, Jeongho, et al.
Veröffentlicht: (2025)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
Imagining the Unseen: Generative Location Modeling for Object Placement
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024)
Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
von: Hwang, Sungwon, et al.
Veröffentlicht: (2025)
von: Hwang, Sungwon, et al.
Veröffentlicht: (2025)
DesignLab: Designing Slides Through Iterative Detection and Correction
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
EgoX: Egocentric Video Generation from a Single Exocentric Video
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
von: Cho, Wonwoo, et al.
Veröffentlicht: (2024)
von: Cho, Wonwoo, et al.
Veröffentlicht: (2024)
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
von: Jo, Kyungmin, et al.
Veröffentlicht: (2024)
von: Jo, Kyungmin, et al.
Veröffentlicht: (2024)
Memory-Efficient Fine-Tuning Diffusion Transformers via Dynamic Patch Sampling and Block Skipping
von: Park, Sunghyun, et al.
Veröffentlicht: (2026)
von: Park, Sunghyun, et al.
Veröffentlicht: (2026)
Enhancing Intrinsic Features for Debiasing via Investigating Class-Discerning Common Attributes in Bias-Contrastive Pair
von: Park, Jeonghoon, et al.
Veröffentlicht: (2024)
von: Park, Jeonghoon, et al.
Veröffentlicht: (2024)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
von: Kim, Min-Jung, et al.
Veröffentlicht: (2025)
von: Kim, Min-Jung, et al.
Veröffentlicht: (2025)
Steering Guidance for Personalized Text-to-Image Diffusion Models
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
von: Kim, Donghu, et al.
Veröffentlicht: (2024)
von: Kim, Donghu, et al.
Veröffentlicht: (2024)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
von: Kim, Kinam, et al.
Veröffentlicht: (2025)
von: Kim, Kinam, et al.
Veröffentlicht: (2025)
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models
von: Song, Junha, et al.
Veröffentlicht: (2026)
von: Song, Junha, et al.
Veröffentlicht: (2026)
Fair Generation without Unfair Distortions: Debiasing Text-to-Image Generation with Entanglement-Free Attention
von: Park, Jeonghoon, et al.
Veröffentlicht: (2025)
von: Park, Jeonghoon, et al.
Veröffentlicht: (2025)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
von: Shim, Gyumin, et al.
Veröffentlicht: (2025)
von: Shim, Gyumin, et al.
Veröffentlicht: (2025)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
von: Choi, Jinho, et al.
Veröffentlicht: (2025)
von: Choi, Jinho, et al.
Veröffentlicht: (2025)
Memory-Efficient Personalization of Text-to-Image Diffusion Models via Selective Optimization Strategies
von: Choi, Seokeon, et al.
Veröffentlicht: (2025)
von: Choi, Seokeon, et al.
Veröffentlicht: (2025)
Towards Calibrated Robust Fine-Tuning of Vision-Language Models
von: Oh, Changdae, et al.
Veröffentlicht: (2023)
von: Oh, Changdae, et al.
Veröffentlicht: (2023)
Effective Rank Analysis and Regularization for Enhanced 3D Gaussian Splatting
von: Hyung, Junha, et al.
Veröffentlicht: (2024)
von: Hyung, Junha, et al.
Veröffentlicht: (2024)
Good Noise Makes Good Edits: A Training-Free Diffusion-Based Video Editing with Image and Text Prompts
von: Choi, Saemee, et al.
Veröffentlicht: (2025)
von: Choi, Saemee, et al.
Veröffentlicht: (2025)
Leveraging Learned Image Prior for 3D Gaussian Compression
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
Locality-aware Gaussian Compression for Fast and High-quality Rendering
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
von: Shin, Seungjoo, et al.
Veröffentlicht: (2025)
Diffusion Model Compression for Image-to-Image Translation
von: Kim, Geonung, et al.
Veröffentlicht: (2024)
von: Kim, Geonung, et al.
Veröffentlicht: (2024)
Color Names in Vision-Language Models
von: Gomez-Villa, Alexandra, et al.
Veröffentlicht: (2025)
von: Gomez-Villa, Alexandra, et al.
Veröffentlicht: (2025)
Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
von: Kim, Jinhee, et al.
Veröffentlicht: (2024)
von: Kim, Jinhee, et al.
Veröffentlicht: (2024)
Emergence of Text Readability in Vision Language Models
von: Park, Jaeyoo, et al.
Veröffentlicht: (2025)
von: Park, Jaeyoo, et al.
Veröffentlicht: (2025)
Zero-Shot Head Swapping in Real-World Scenarios
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
von: Kim, Min-Jung, et al.
Veröffentlicht: (2025)
von: Kim, Min-Jung, et al.
Veröffentlicht: (2025)
RL makes MLLMs see better than SFT
von: Song, Junha, et al.
Veröffentlicht: (2025)
von: Song, Junha, et al.
Veröffentlicht: (2025)
Do Vision-Language Models Understand Visual Persuasiveness?
von: Park, Gyuwon
Veröffentlicht: (2025)
von: Park, Gyuwon
Veröffentlicht: (2025)
MagiCapture: High-Resolution Multi-Concept Portrait Customization
von: Hyung, Junha, et al.
Veröffentlicht: (2023)
von: Hyung, Junha, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
von: Park, Minho, et al.
Veröffentlicht: (2025) -
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
von: Yun, Jooyeol, et al.
Veröffentlicht: (2024) -
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025) -
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
von: Jo, Kyungmin, et al.
Veröffentlicht: (2025) -
CA-LoRA: Concept-Aware LoRA for Domain-Aligned Segmentation Dataset Generation
von: Park, Minho, et al.
Veröffentlicht: (2025)