CoCoIns: Consistent Subject Generation via Contrastive Instantiated Concepts
Fuente:
arXiv
Saved in:
| Main Authors: | Hsin-Ying, Lee, Chan, Kelvin C. K., Yang, Ming-Hsuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
by: Chan, Kelvin C. K., et al.
Published: (2024)
by: Chan, Kelvin C. K., et al.
Published: (2024)
CoDi: Subject-Consistent and Pose-Diverse Text-to-Image Generation
by: Gao, Zhanxin, et al.
Published: (2025)
by: Gao, Zhanxin, et al.
Published: (2025)
Exploiting Diffusion Prior for Generalizable Dense Prediction
by: Lee, Hsin-Ying, et al.
Published: (2023)
by: Lee, Hsin-Ying, et al.
Published: (2023)
CoCo4D: Comprehensive and Complex 4D Scene Generation
by: Zhou, Junwei, et al.
Published: (2025)
by: Zhou, Junwei, et al.
Published: (2025)
From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition
by: Lo, Ling, et al.
Published: (2025)
by: Lo, Ling, et al.
Published: (2025)
MotiMotion: Motion-Controlled Video Generation with Visual Reasoning
by: Hsin-Ying, Lee, et al.
Published: (2026)
by: Hsin-Ying, Lee, et al.
Published: (2026)
A Simple Approach to Unifying Diffusion-based Conditional Generation
by: Li, Xirui, et al.
Published: (2024)
by: Li, Xirui, et al.
Published: (2024)
Dual Associated Encoder for Face Restoration
by: Tsai, Yu-Ju, et al.
Published: (2023)
by: Tsai, Yu-Ju, et al.
Published: (2023)
HoliSDiP: Image Super-Resolution via Holistic Semantics and Diffusion Prior
by: Tsao, Li-Yuan, et al.
Published: (2024)
by: Tsao, Li-Yuan, et al.
Published: (2024)
Video-CoM: Interactive Video Reasoning via Chain of Manipulations
by: Rasheed, Hanoona, et al.
Published: (2025)
by: Rasheed, Hanoona, et al.
Published: (2025)
KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities
by: Huang, Hsin-Ping, et al.
Published: (2024)
by: Huang, Hsin-Ping, et al.
Published: (2024)
CoGen: 3D Consistent Video Generation via Adaptive Conditioning for Autonomous Driving
by: Ji, Yishen, et al.
Published: (2025)
by: Ji, Yishen, et al.
Published: (2025)
Effective Adapter for Face Recognition in the Wild
by: Liu, Yunhao, et al.
Published: (2023)
by: Liu, Yunhao, et al.
Published: (2023)
CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation
by: Luo, Xiangyang, et al.
Published: (2026)
by: Luo, Xiangyang, et al.
Published: (2026)
CoBELa: Steering Transparent Generation via Concept Bottlenecks on Energy Landscapes
by: Kim, Sangwon, et al.
Published: (2025)
by: Kim, Sangwon, et al.
Published: (2025)
GeCo: Evaluating Geometric Consistency for Video Generation via Motion and Structure
by: Gu, Leslie, et al.
Published: (2025)
by: Gu, Leslie, et al.
Published: (2025)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
by: Rahman, Tanzila, et al.
Published: (2024)
by: Rahman, Tanzila, et al.
Published: (2024)
Multi-task Image Restoration Guided By Robust DINO Features
by: Lin, Xin, et al.
Published: (2023)
by: Lin, Xin, et al.
Published: (2023)
BindWeave: Subject-Consistent Video Generation via Cross-Modal Integration
by: Li, Zhaoyang, et al.
Published: (2025)
by: Li, Zhaoyang, et al.
Published: (2025)
CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Learning
by: Wu, Yuhui, et al.
Published: (2026)
by: Wu, Yuhui, et al.
Published: (2026)
Concept Guided Co-salient Object Detection
by: Zhu, Jiayi, et al.
Published: (2024)
by: Zhu, Jiayi, et al.
Published: (2024)
Synthesizing Consistent Novel Views via 3D Epipolar Attention without Re-Training
by: Ye, Botao, et al.
Published: (2025)
by: Ye, Botao, et al.
Published: (2025)
CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility
by: Zi, Bojia, et al.
Published: (2024)
by: Zi, Bojia, et al.
Published: (2024)
ReCoSplat: Autoregressive Feed-Forward Gaussian Splatting Using Render-and-Compare
by: Cheng, Freeman, et al.
Published: (2026)
by: Cheng, Freeman, et al.
Published: (2026)
CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection
by: Feng, Huidong, et al.
Published: (2026)
by: Feng, Huidong, et al.
Published: (2026)
CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs
by: Zhang, Daoan, et al.
Published: (2024)
by: Zhang, Daoan, et al.
Published: (2024)
Context Forcing: Consistent Autoregressive Video Generation with Long Context
by: Chen, Shuo, et al.
Published: (2026)
by: Chen, Shuo, et al.
Published: (2026)
OrCo: Towards Better Generalization via Orthogonality and Contrast for Few-Shot Class-Incremental Learning
by: Ahmed, Noor, et al.
Published: (2024)
by: Ahmed, Noor, et al.
Published: (2024)
DC-SAM: In-Context Segment Anything in Images and Videos via Dual Consistency
by: Qi, Mengshi, et al.
Published: (2025)
by: Qi, Mengshi, et al.
Published: (2025)
CoPA: Hierarchical Concept Prompting and Aggregating Network for Explainable Diagnosis
by: Dong, Yiheng, et al.
Published: (2025)
by: Dong, Yiheng, et al.
Published: (2025)
AdaIR: Exploiting Underlying Similarities of Image Restoration Tasks with Adapters
by: Chen, Hao-Wei, et al.
Published: (2024)
by: Chen, Hao-Wei, et al.
Published: (2024)
PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards
by: Wang, Shulei, et al.
Published: (2025)
by: Wang, Shulei, et al.
Published: (2025)
MV-CoLight: Efficient Object Compositing with Consistent Lighting and Shadow Generation
by: Ren, Kerui, et al.
Published: (2025)
by: Ren, Kerui, et al.
Published: (2025)
CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
by: Xu, Dejia, et al.
Published: (2024)
by: Xu, Dejia, et al.
Published: (2024)
CoCoNO: Attention Contrast-and-Complete for Initial Noise Optimization in Text-to-Image Synthesis
by: Sundaram, Aravindan, et al.
Published: (2024)
by: Sundaram, Aravindan, et al.
Published: (2024)
AR-CoPO: Align Autoregressive Video Generation with Contrastive Policy Optimization
by: He, Dailan, et al.
Published: (2026)
by: He, Dailan, et al.
Published: (2026)
HoliGS: Holistic Gaussian Splatting for Embodied View Synthesis
by: Wang, Xiaoyuan, et al.
Published: (2025)
by: Wang, Xiaoyuan, et al.
Published: (2025)
SyCoCa: Symmetrizing Contrastive Captioners with Attentive Masking for Multimodal Alignment
by: Ma, Ziping, et al.
Published: (2024)
by: Ma, Ziping, et al.
Published: (2024)
CoLVR: Enhancing Exploratory Latent Visual Reasoning via Contrastive Optimization
by: Ding, Ziyang, et al.
Published: (2026)
by: Ding, Ziyang, et al.
Published: (2026)
Towards Affordance-Aware Articulation Synthesis for Rigged Objects
by: Yu, Yu-Chu, et al.
Published: (2025)
by: Yu, Yu-Chu, et al.
Published: (2025)
Similar Items
-
Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
by: Chan, Kelvin C. K., et al.
Published: (2024) -
CoDi: Subject-Consistent and Pose-Diverse Text-to-Image Generation
by: Gao, Zhanxin, et al.
Published: (2025) -
Exploiting Diffusion Prior for Generalizable Dense Prediction
by: Lee, Hsin-Ying, et al.
Published: (2023) -
CoCo4D: Comprehensive and Complex 4D Scene Generation
by: Zhou, Junwei, et al.
Published: (2025) -
From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition
by: Lo, Ling, et al.
Published: (2025)