Multi-dimensional Preference Alignment by Conditioning Reward Itself
Fuente:
arXiv
Saved in:
| Main Authors: | Jang, Jiho, Kim, Jinyoung, Baek, Kyungjune, Kwak, Nojun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
by: Park, Seojeong, et al.
Published: (2024)
by: Park, Seojeong, et al.
Published: (2024)
Rethinking Direct Preference Optimization in Diffusion Models
by: Kang, Junyong, et al.
Published: (2025)
by: Kang, Junyong, et al.
Published: (2025)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
by: Baek, Injun, et al.
Published: (2026)
by: Baek, Injun, et al.
Published: (2026)
Conservative Generator, Progressive Discriminator: Coordination of Adversaries in Few-shot Incremental Image Synthesis
by: Kong, Chaerin, et al.
Published: (2022)
by: Kong, Chaerin, et al.
Published: (2022)
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
by: Kim, Yearim, et al.
Published: (2026)
by: Kim, Yearim, et al.
Published: (2026)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
by: Kwon, Minkyung, et al.
Published: (2025)
by: Kwon, Minkyung, et al.
Published: (2025)
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
by: Han, Donghoon, et al.
Published: (2023)
by: Han, Donghoon, et al.
Published: (2023)
DivCon-NeRF: Diverse and Consistent Ray Augmentation for Few-Shot NeRF
by: Lee, Ingyun, et al.
Published: (2025)
by: Lee, Ingyun, et al.
Published: (2025)
Causal Interpretation of Sparse Autoencoder Features in Vision
by: Han, Sangyu, et al.
Published: (2025)
by: Han, Sangyu, et al.
Published: (2025)
Respect the model: Fine-grained and Robust Explanation with Sharing Ratio Decomposition
by: Han, Sangyu, et al.
Published: (2024)
by: Han, Sangyu, et al.
Published: (2024)
Projected Representation Conditioning for High-fidelity Novel View Synthesis
by: Kwak, Min-Seop, et al.
Published: (2026)
by: Kwak, Min-Seop, et al.
Published: (2026)
4DGS360: 360° Gaussian Reconstruction of Dynamic Objects from a Single Video
by: Jang, Jae Won, et al.
Published: (2026)
by: Jang, Jae Won, et al.
Published: (2026)
ReFlex: Text-Guided Editing of Real Images in Rectified Flow via Mid-Step Feature Extraction and Attention Adaptation
by: Kim, Jimyeong, et al.
Published: (2025)
by: Kim, Jimyeong, et al.
Published: (2025)
Decompose the model: Mechanistic interpretability in image models with Generalized Integrated Gradients (GIG)
by: Kim, Yearim, et al.
Published: (2024)
by: Kim, Yearim, et al.
Published: (2024)
Toward Stable World Models: Measuring and Addressing World Instability in Generative Environments
by: Kwon, Soonwoo, et al.
Published: (2025)
by: Kwon, Soonwoo, et al.
Published: (2025)
Unlocking the Potential of Unlabeled Data in Semi-Supervised Domain Generalization
by: Lee, Dongkwan, et al.
Published: (2025)
by: Lee, Dongkwan, et al.
Published: (2025)
VDPP: Video Depth Post-Processing for Speed and Scalability
by: Yoon, Daewon, et al.
Published: (2026)
by: Yoon, Daewon, et al.
Published: (2026)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
by: Song, Yeji, et al.
Published: (2024)
by: Song, Yeji, et al.
Published: (2024)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
by: Song, Jookyung, et al.
Published: (2025)
by: Song, Jookyung, et al.
Published: (2025)
The Role of Teacher Calibration in Knowledge Distillation
by: Kim, Suyoung, et al.
Published: (2025)
by: Kim, Suyoung, et al.
Published: (2025)
Semi-Supervised Domain Adaptation for Wildfire Detection
by: Jang, JooYoung, et al.
Published: (2024)
by: Jang, JooYoung, et al.
Published: (2024)
Coreset Selection for Object Detection
by: Lee, Hojun, et al.
Published: (2024)
by: Lee, Hojun, et al.
Published: (2024)
CSF: Black-box Fingerprinting via Compositional Semantics for Text-to-Image Models
by: Lee, Junhoo, et al.
Published: (2026)
by: Lee, Junhoo, et al.
Published: (2026)
LoGoColor: Local-Global 3D Colorization for 360° Scenes
by: Chang, Yeonjin, et al.
Published: (2025)
by: Chang, Yeonjin, et al.
Published: (2025)
RefReward-SR: LR-Conditioned Reward Modeling for Preference-Aligned Super-Resolution
by: Song, Yushuai, et al.
Published: (2026)
by: Song, Yushuai, et al.
Published: (2026)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
by: Chung, Inseop, et al.
Published: (2024)
by: Chung, Inseop, et al.
Published: (2024)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
by: Kim, Jongha, et al.
Published: (2024)
by: Kim, Jongha, et al.
Published: (2024)
SplatFlow: Multi-View Rectified Flow Model for 3D Gaussian Splatting Synthesis
by: Go, Hyojun, et al.
Published: (2024)
by: Go, Hyojun, et al.
Published: (2024)
ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approximation
by: Kim, Suyoung, et al.
Published: (2026)
by: Kim, Suyoung, et al.
Published: (2026)
ROODI: Reconstructing Occluded Objects with Denoising Inpainters
by: Chang, Yeonjin, et al.
Published: (2025)
by: Chang, Yeonjin, et al.
Published: (2025)
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
by: Song, Yeji, et al.
Published: (2025)
by: Song, Yeji, et al.
Published: (2025)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
by: Cui, Mingxuan, et al.
Published: (2026)
by: Cui, Mingxuan, et al.
Published: (2026)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
by: Kim, Manjin, et al.
Published: (2026)
by: Kim, Manjin, et al.
Published: (2026)
MSG Score: Automated Video Verification for Reliable Multi-Scene Generation
by: Yoon, Daewon, et al.
Published: (2024)
by: Yoon, Daewon, et al.
Published: (2024)
ARC-NeRF: Area Ray Casting for Broader Unseen View Coverage in Few-shot Object Rendering
by: Seo, Seunghyeon, et al.
Published: (2024)
by: Seo, Seunghyeon, et al.
Published: (2024)
TWLV-I: Analysis and Insights from Holistic Evaluation on Video Foundation Models
by: Lee, Hyeongmin, et al.
Published: (2024)
by: Lee, Hyeongmin, et al.
Published: (2024)
VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
by: Lee, Ji Soo, et al.
Published: (2025)
by: Lee, Ji Soo, et al.
Published: (2025)
Rethinking Garment Conditioning in Diffusion-based Virtual Try-On
by: Na, Kihyun, et al.
Published: (2025)
by: Na, Kihyun, et al.
Published: (2025)
Style Composition within Distinct LoRA modules for Traditional Art
by: Lee, Jaehyun, et al.
Published: (2025)
by: Lee, Jaehyun, et al.
Published: (2025)
FREST: Feature RESToration for Semantic Segmentation under Multiple Adverse Conditions
by: Lee, Sohyun, et al.
Published: (2024)
by: Lee, Sohyun, et al.
Published: (2024)
Similar Items
-
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
by: Park, Seojeong, et al.
Published: (2024) -
Rethinking Direct Preference Optimization in Diffusion Models
by: Kang, Junyong, et al.
Published: (2025) -
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
by: Baek, Injun, et al.
Published: (2026) -
Conservative Generator, Progressive Discriminator: Coordination of Adversaries in Few-shot Incremental Image Synthesis
by: Kong, Chaerin, et al.
Published: (2022) -
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
by: Kim, Yearim, et al.
Published: (2026)