Correlation-Weighted Multi-Reward Optimization for Compositional Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wi, Jungmyung, Kim, Hyunsoo, Kim, Donghyun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Training-Free Label Space Alignment for Universal Domain Adaptation
by: Lee, Dujin, et al.
Published: (2025)
by: Lee, Dujin, et al.
Published: (2025)
MahaVar: OOD Detection via Class-wise Mahalanobis Distance Variance under Neural Collapse
by: Kim, Donghwan, et al.
Published: (2026)
by: Kim, Donghwan, et al.
Published: (2026)
Mitigating the Likelihood Paradox in Flow-based OOD Detection via Entropy Manipulation
by: Kim, Donghwan, et al.
Published: (2026)
by: Kim, Donghwan, et al.
Published: (2026)
Human-AI Collaborative Bot Detection in MMORPGs
by: Son, Jaeman, et al.
Published: (2025)
by: Son, Jaeman, et al.
Published: (2025)
Why the Counterintuitive Phenomenon of Likelihood Rarely Appears in Tabular Anomaly Detection with Deep Generative Models?
by: Kim, Donghwan, et al.
Published: (2026)
by: Kim, Donghwan, et al.
Published: (2026)
CAUS: A Dataset for Question Generation based on Human Cognition Leveraging Large Language Models
by: Shin, Minjung, et al.
Published: (2024)
by: Shin, Minjung, et al.
Published: (2024)
MIND: AI Co-Scientist for Material Research
by: Ahn, Geonhee, et al.
Published: (2026)
by: Ahn, Geonhee, et al.
Published: (2026)
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
by: Nguyen, Viet Bac, et al.
Published: (2026)
by: Nguyen, Viet Bac, et al.
Published: (2026)
PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization
by: Kim, Wonjoong, et al.
Published: (2026)
by: Kim, Wonjoong, et al.
Published: (2026)
Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving
by: Kim, Donghyun, et al.
Published: (2026)
by: Kim, Donghyun, et al.
Published: (2026)
Gradient-Free Noise Optimization for Reward Alignment in Generative Models
by: Kim, Jeongsol, et al.
Published: (2026)
by: Kim, Jeongsol, et al.
Published: (2026)
"There Is No Such Thing as a Dumb Question," But There Are Good Ones
by: Shin, Minjung, et al.
Published: (2025)
by: Shin, Minjung, et al.
Published: (2025)
Optimizing Long-Form Clinical Text Generation with Claim-Based Rewards
by: Jhaveri, Samyak, et al.
Published: (2025)
by: Jhaveri, Samyak, et al.
Published: (2025)
Trust the uncertain teacher: distilling dark knowledge via calibrated uncertainty
by: Kim, Jeonghyun, et al.
Published: (2026)
by: Kim, Jeonghyun, et al.
Published: (2026)
Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents
by: Kim, Dong-Hee, et al.
Published: (2026)
by: Kim, Dong-Hee, et al.
Published: (2026)
Difference Inversion: Interpolate and Isolate the Difference with Token Consistency for Image Analogy Generation
by: Kim, Hyunsoo, et al.
Published: (2025)
by: Kim, Hyunsoo, et al.
Published: (2025)
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
by: Kim, Sungyeon, et al.
Published: (2023)
by: Kim, Sungyeon, et al.
Published: (2023)
Adaptive Self-training Framework for Fine-grained Scene Graph Generation
by: Kim, Kibum, et al.
Published: (2024)
by: Kim, Kibum, et al.
Published: (2024)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Is it safe to cross? Interpretable Risk Assessment with GPT-4V for Safety-Aware Street Crossing
by: Hwang, Hochul, et al.
Published: (2024)
by: Hwang, Hochul, et al.
Published: (2024)
MuLMINet: Multi-Layer Multi-Input Transformer Network with Weighted Loss
by: Seong, Minwoo, et al.
Published: (2023)
by: Seong, Minwoo, et al.
Published: (2023)
GOPO: Policy Optimization using Ranked Rewards
by: Choi, Kyuseong, et al.
Published: (2026)
by: Choi, Kyuseong, et al.
Published: (2026)
SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data
by: Kim, Dong-Hee, et al.
Published: (2025)
by: Kim, Dong-Hee, et al.
Published: (2025)
MergeRec: Model Merging for Data-Isolated Cross-Domain Sequential Recommendation
by: Kim, Hyunsoo, et al.
Published: (2026)
by: Kim, Hyunsoo, et al.
Published: (2026)
M3-SLU: Evaluating Speaker-Attributed Reasoning in Multimodal Large Language Models
by: Kwon, Yejin, et al.
Published: (2025)
by: Kwon, Yejin, et al.
Published: (2025)
A Framework for Mining Collectively-Behaving Bots in MMORPGs
by: Kim, Hyunsoo, et al.
Published: (2025)
by: Kim, Hyunsoo, et al.
Published: (2025)
Stage-Wise Reward Shaping for Acrobatic Robots: A Constrained Multi-Objective Reinforcement Learning Approach
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
Bidirectional Multimodal Prompt Learning with Scale-Aware Training for Few-Shot Multi-Class Anomaly Detection
by: Lee, Yujin, et al.
Published: (2024)
by: Lee, Yujin, et al.
Published: (2024)
PAC-BENCH: Evaluating Multi-Agent Collaboration under Privacy Constraints
by: Park, Minjun, et al.
Published: (2026)
by: Park, Minjun, et al.
Published: (2026)
1 Trillion Token (1TT) Platform: A Novel Framework for Efficient Data Sharing and Compensation in Large Language Models
by: Park, Chanjun, et al.
Published: (2024)
by: Park, Chanjun, et al.
Published: (2024)
LP Data Pipeline: Lightweight, Purpose-driven Data Pipeline for Large Language Models
by: Kim, Yungi, et al.
Published: (2024)
by: Kim, Yungi, et al.
Published: (2024)
Rethinking KenLM: Good and Bad Model Ensembles for Efficient Text Quality Filtering in Large Web Corpora
by: Kim, Yungi, et al.
Published: (2024)
by: Kim, Yungi, et al.
Published: (2024)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
by: Kim, Seungwook, et al.
Published: (2026)
by: Kim, Seungwook, et al.
Published: (2026)
PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning
by: Baek, In-Chang, et al.
Published: (2025)
by: Baek, In-Chang, et al.
Published: (2025)
Multi-Level Compositional Reasoning for Interactive Instruction Following
by: Bhambri, Suvaansh, et al.
Published: (2023)
by: Bhambri, Suvaansh, et al.
Published: (2023)
Transformable Gaussian Reward Function for Socially-Aware Navigation with Deep Reinforcement Learning
by: Kim, Jinyeob, et al.
Published: (2024)
by: Kim, Jinyeob, et al.
Published: (2024)
NDST: Neural Driving Style Transfer for Human-Like Vision-Based Autonomous Driving
by: Kim, Donghyun, et al.
Published: (2024)
by: Kim, Donghyun, et al.
Published: (2024)
Large Language Models meet Collaborative Filtering: An Efficient All-round LLM-based Recommender System
by: Kim, Sein, et al.
Published: (2024)
by: Kim, Sein, et al.
Published: (2024)
A Scalable and Transferable Time Series Prediction Framework for Demand Forecasting
by: Park, Young-Jin, et al.
Published: (2024)
by: Park, Young-Jin, et al.
Published: (2024)
WoLF: Wide-scope Large Language Model Framework for CXR Understanding
by: Kang, Seil, et al.
Published: (2024)
by: Kang, Seil, et al.
Published: (2024)
Similar Items
-
Training-Free Label Space Alignment for Universal Domain Adaptation
by: Lee, Dujin, et al.
Published: (2025) -
MahaVar: OOD Detection via Class-wise Mahalanobis Distance Variance under Neural Collapse
by: Kim, Donghwan, et al.
Published: (2026) -
Mitigating the Likelihood Paradox in Flow-based OOD Detection via Entropy Manipulation
by: Kim, Donghwan, et al.
Published: (2026) -
Human-AI Collaborative Bot Detection in MMORPGs
by: Son, Jaeman, et al.
Published: (2025) -
Why the Counterintuitive Phenomenon of Likelihood Rarely Appears in Tabular Anomaly Detection with Deep Generative Models?
by: Kim, Donghwan, et al.
Published: (2026)