RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Oh, Yeongtak, Chung, Dohyun, Shin, Juhyeon, Park, Sangha, Barthelemy, Johan, Mok, Jisoo, Yoon, Sungroh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contextualized Visual Personalization in Vision-Language Models
by: Oh, Yeongtak, et al.
Published: (2026)
by: Oh, Yeongtak, et al.
Published: (2026)
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
by: Park, Sangha, et al.
Published: (2025)
by: Park, Sangha, et al.
Published: (2025)
Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection
by: Lee, Saehyung, et al.
Published: (2024)
by: Lee, Saehyung, et al.
Published: (2024)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
by: Park, Sangha, et al.
Published: (2025)
by: Park, Sangha, et al.
Published: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
by: Oh, Yeongtak, et al.
Published: (2026)
by: Oh, Yeongtak, et al.
Published: (2026)
On mitigating stability-plasticity dilemma in CLIP-guided image morphing via geodesic distillation loss
by: Oh, Yeongtak, et al.
Published: (2024)
by: Oh, Yeongtak, et al.
Published: (2024)
ControlDreamer: Blending Geometry and Style in Text-to-3D
by: Oh, Yeongtak, et al.
Published: (2023)
by: Oh, Yeongtak, et al.
Published: (2023)
CANDI: Curated Test-Time Adaptation for Multivariate Time-Series Anomaly Detection Under Distribution Shift
by: Kim, HyunGi, et al.
Published: (2026)
by: Kim, HyunGi, et al.
Published: (2026)
Style-Friendly SNR Sampler for Style-Driven Generation
by: Choi, Jooyoung, et al.
Published: (2024)
by: Choi, Jooyoung, et al.
Published: (2024)
SF(DA)$^2$: Source-free Domain Adaptation Through the Lens of Data Augmentation
by: Hwang, Uiwon, et al.
Published: (2024)
by: Hwang, Uiwon, et al.
Published: (2024)
Exploring the Potential of LLMs as Personalized Assistants: Dataset, Evaluation, and Analysis
by: Mok, Jisoo, et al.
Published: (2025)
by: Mok, Jisoo, et al.
Published: (2025)
Negative-Guided Subject Fidelity Optimization for Zero-Shot Subject-Driven Generation
by: Shin, Chaehun, et al.
Published: (2025)
by: Shin, Chaehun, et al.
Published: (2025)
Efficient Diffusion-Driven Corruption Editor for Test-Time Adaptation
by: Oh, Yeongtak, et al.
Published: (2024)
by: Oh, Yeongtak, et al.
Published: (2024)
Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning
by: Park, Sangkwon, et al.
Published: (2026)
by: Park, Sangkwon, et al.
Published: (2026)
STAG: Structural Test-time Alignment of Gradients for Online Adaptation
by: Shin, Juhyeon, et al.
Published: (2024)
by: Shin, Juhyeon, et al.
Published: (2024)
Entropy is not Enough for Test-Time Adaptation: From the Perspective of Disentangled Factors
by: Lee, Jonghyun, et al.
Published: (2024)
by: Lee, Jonghyun, et al.
Published: (2024)
Self-Supervised Time-Series Anomaly Detection Using Learnable Data Augmentation
by: Choi, Kukjin, et al.
Published: (2024)
by: Choi, Kukjin, et al.
Published: (2024)
Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation
by: Kim, HyunGi, et al.
Published: (2025)
by: Kim, HyunGi, et al.
Published: (2025)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
by: Lew, Jaihyun, et al.
Published: (2024)
by: Lew, Jaihyun, et al.
Published: (2024)
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
by: Jung, Mingi, et al.
Published: (2025)
by: Jung, Mingi, et al.
Published: (2025)
AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
by: Sung-Bin, Kim, et al.
Published: (2024)
by: Sung-Bin, Kim, et al.
Published: (2024)
CKNN: Cleansed k-Nearest Neighbor for Unsupervised Video Anomaly Detection
by: Yi, Jihun, et al.
Published: (2024)
by: Yi, Jihun, et al.
Published: (2024)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
by: Shin, Chaehun, et al.
Published: (2024)
by: Shin, Chaehun, et al.
Published: (2024)
LINGO-Space: Language-Conditioned Incremental Grounding for Space
by: Kim, Dohyun, et al.
Published: (2024)
by: Kim, Dohyun, et al.
Published: (2024)
Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
by: Lee, Saehyung, et al.
Published: (2024)
by: Lee, Saehyung, et al.
Published: (2024)
Improving Geometry in Sparse-View 3DGS via Reprojection-based DoF Separation
by: Kim, Yongsung, et al.
Published: (2024)
by: Kim, Yongsung, et al.
Published: (2024)
LLM-based Frameworks for API Argument Filling in Task-Oriented Conversational Systems
by: Mok, Jisoo, et al.
Published: (2024)
by: Mok, Jisoo, et al.
Published: (2024)
Domain Generalization for Person Re-identification: A Survey Towards Domain-Agnostic Person Matching
by: Lee, Hyeonseo, et al.
Published: (2025)
by: Lee, Hyeonseo, et al.
Published: (2025)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
by: Baek, Kanghyun, et al.
Published: (2025)
by: Baek, Kanghyun, et al.
Published: (2025)
Diagnosing and Correcting Concept Omission in Multimodal Diffusion Transformers
by: Baek, Kanghyun, et al.
Published: (2026)
by: Baek, Kanghyun, et al.
Published: (2026)
Disentangled Motion Modeling for Video Frame Interpolation
by: Lew, Jaihyun, et al.
Published: (2024)
by: Lew, Jaihyun, et al.
Published: (2024)
Merge-Friendly Post-Training Quantization for Multi-Target Domain Adaptation
by: Shin, Juncheol, et al.
Published: (2025)
by: Shin, Juncheol, et al.
Published: (2025)
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
by: Sun, Zhen, et al.
Published: (2025)
by: Sun, Zhen, et al.
Published: (2025)
Causality-Aware Contrastive Learning for Robust Multivariate Time-Series Anomaly Detection
by: Kim, HyunGi, et al.
Published: (2025)
by: Kim, HyunGi, et al.
Published: (2025)
Normality Addition via Normality Detection in Industrial Image Anomaly Detection Models
by: Yi, Jihun, et al.
Published: (2024)
by: Yi, Jihun, et al.
Published: (2024)
TLDR: Text Based Last-layer Retraining for Debiasing Image Classifiers
by: Park, Juhyeon, et al.
Published: (2023)
by: Park, Juhyeon, et al.
Published: (2023)
DEAL: Decoupled Classifier with Adaptive Linear Modulation for Group Robust Early Diagnosis of MCI to AD Conversion
by: Lee, Donggyu, et al.
Published: (2024)
by: Lee, Donggyu, et al.
Published: (2024)
Improving Diffusion-Based Generative Models via Approximated Optimal Transport
by: Kim, Daegyu, et al.
Published: (2024)
by: Kim, Daegyu, et al.
Published: (2024)
MMPB: It's Time for Multi-Modal Personalization
by: Kim, Jaeik, et al.
Published: (2025)
by: Kim, Jaeik, et al.
Published: (2025)
PaMM: Pose-aware Multi-shot Matching for Improving Person Re-identification
by: Cho, Yeong-Jun, et al.
Published: (2017)
by: Cho, Yeong-Jun, et al.
Published: (2017)
Similar Items
-
Contextualized Visual Personalization in Vision-Language Models
by: Oh, Yeongtak, et al.
Published: (2026) -
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
by: Park, Sangha, et al.
Published: (2025) -
Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection
by: Lee, Saehyung, et al.
Published: (2024) -
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
by: Park, Sangha, et al.
Published: (2025) -
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
by: Oh, Yeongtak, et al.
Published: (2026)