Semantic-aware Wasserstein Policy Regularization for Large Language Model Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Na, Byeonghu, Na, Hyungho, Kim, Yeongmin, Jo, Suhyeon, Bae, HeeSun, Kang, Mina, Moon, Il-Chul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reward-based Input Construction for Cross-document Relation Extraction
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
Distillation of Large Language Models via Concrete Score Matching
von: Kim, Yeongmin, et al.
Veröffentlicht: (2025)
von: Kim, Yeongmin, et al.
Veröffentlicht: (2025)
Dirichlet-based Per-Sample Weighting by Transition Matrix for Noisy Label Learning
von: Bae, HeeSun, et al.
Veröffentlicht: (2024)
von: Bae, HeeSun, et al.
Veröffentlicht: (2024)
AMiD: Knowledge Distillation for LLMs with $α$-mixture Assistant Distribution
von: Shin, Donghyeok, et al.
Veröffentlicht: (2025)
von: Shin, Donghyeok, et al.
Veröffentlicht: (2025)
Label-Noise Robust Diffusion Models
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
Preference Optimization by Estimating the Ratio of the Data Distribution
von: Kim, Yeongmin, et al.
Veröffentlicht: (2025)
von: Kim, Yeongmin, et al.
Veröffentlicht: (2025)
Unknown Domain Inconsistency Minimization for Domain Generalization
von: Shin, Seungjae, et al.
Veröffentlicht: (2024)
von: Shin, Seungjae, et al.
Veröffentlicht: (2024)
Diffusion Adaptive Text Embedding for Text-to-Image Diffusion Models
von: Na, Byeonghu, et al.
Veröffentlicht: (2025)
von: Na, Byeonghu, et al.
Veröffentlicht: (2025)
Diffusion Bridge AutoEncoders for Unsupervised Representation Learning
von: Kim, Yeongmin, et al.
Veröffentlicht: (2024)
von: Kim, Yeongmin, et al.
Veröffentlicht: (2024)
Diffusion Rejection Sampling
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models
von: Kim, Yeongmin, et al.
Veröffentlicht: (2026)
von: Kim, Yeongmin, et al.
Veröffentlicht: (2026)
Prompt-Based Safety Guidance Is Ineffective for Unlearned Text-to-Image Diffusion Models
von: Shin, Jiwoo, et al.
Veröffentlicht: (2025)
von: Shin, Jiwoo, et al.
Veröffentlicht: (2025)
Training Unbiased Diffusion Models From Biased Dataset
von: Kim, Yeongmin, et al.
Veröffentlicht: (2024)
von: Kim, Yeongmin, et al.
Veröffentlicht: (2024)
LAGMA: LAtent Goal-guided Multi-Agent Reinforcement Learning
von: Na, Hyungho, et al.
Veröffentlicht: (2024)
von: Na, Hyungho, et al.
Veröffentlicht: (2024)
Distilling Dataset into Neural Field
von: Shin, Donghyeok, et al.
Veröffentlicht: (2025)
von: Shin, Donghyeok, et al.
Veröffentlicht: (2025)
Trajectory-Class-Aware Multi-Agent Reinforcement Learning
von: Na, Hyungho, et al.
Veröffentlicht: (2025)
von: Na, Hyungho, et al.
Veröffentlicht: (2025)
Efficient Episodic Memory Utilization of Cooperative Multi-Agent Reinforcement Learning
von: Na, Hyungho, et al.
Veröffentlicht: (2024)
von: Na, Hyungho, et al.
Veröffentlicht: (2024)
Make Prompts Adaptable: Bayesian Modeling for Vision-Language Prompt Learning with Data-Dependent Prior
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
Missing Pattern Recognized Diffusion Imputation Model for Missing Not At Random
von: Sim, Gyuwon, et al.
Veröffentlicht: (2026)
von: Sim, Gyuwon, et al.
Veröffentlicht: (2026)
Training-Free Safe Text Embedding Guidance for Text-to-Image Diffusion Models
von: Na, Byeonghu, et al.
Veröffentlicht: (2025)
von: Na, Byeonghu, et al.
Veröffentlicht: (2025)
Improving Large Molecular Language Model via Relation-aware Multimodal Collaboration
von: Park, Jinyoung, et al.
Veröffentlicht: (2026)
von: Park, Jinyoung, et al.
Veröffentlicht: (2026)
PromptLoop: Plug-and-Play Prompt Refinement via Latent Feedback for Diffusion Model Alignment
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
Figure 2c from: Na S-M, Shin N-R, Bae Y-S (2026) Two species of Atteva Walker, 1854 (Lepidoptera, Attevidae) new to Laos, with DNA barcodes. Biodiversity Data Journal 14: e189183. https://doi.org/10.3897/BDJ.14.e189183
von: Na, Sol-Moon, et al.
Veröffentlicht: (2026)
von: Na, Sol-Moon, et al.
Veröffentlicht: (2026)
Figure 1c from: Na S-M, Shin N-R, Bae Y-S (2026) Two species of Atteva Walker, 1854 (Lepidoptera, Attevidae) new to Laos, with DNA barcodes. Biodiversity Data Journal 14: e189183. https://doi.org/10.3897/BDJ.14.e189183
von: Na, Sol-Moon, et al.
Veröffentlicht: (2026)
von: Na, Sol-Moon, et al.
Veröffentlicht: (2026)
Electroless Plating on Polymer Surfaces: Comprehensive Review of Mechanism, Process, Analysis, and Future Applications
von: Na Kyoung Kim, et al.
Veröffentlicht: (2025)
von: Na Kyoung Kim, et al.
Veröffentlicht: (2025)
Remote Sensing Large Vision-Language Model: Semantic-augmented Multi-level Alignment and Semantic-aware Expert Modeling
von: Park, Sungjune, et al.
Veröffentlicht: (2025)
von: Park, Sungjune, et al.
Veröffentlicht: (2025)
Decomposed Diffusion Sampler for Accelerating Large-Scale Inverse Problems
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2023)
Feature Structure Distillation with Centered Kernel Alignment in BERT Transferring
von: Jung, Hee-Jun, et al.
Veröffentlicht: (2022)
von: Jung, Hee-Jun, et al.
Veröffentlicht: (2022)
Single-Step Bidirectional Unpaired Image Translation Using Implicit Bridge Consistency Distillation
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
Dynamic range optimization for treatment time reduction in respiratory‐gated proton therapy using RayStation v2025
von: Sungkoo Cho, et al.
Veröffentlicht: (2026)
von: Sungkoo Cho, et al.
Veröffentlicht: (2026)
A Review of Transient Characterization Techniques for Quantum Dot Light‐Emitting Diodes
von: Jaekwon Kim, et al.
Veröffentlicht: (2026)
von: Jaekwon Kim, et al.
Veröffentlicht: (2026)
Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models
von: Moon, Saemi, et al.
Veröffentlicht: (2026)
von: Moon, Saemi, et al.
Veröffentlicht: (2026)
Spontaneous Wrinkle Collapse in Anisotropic Condensed Matter Predicted by Deep Learning
von: Kitae Kim, et al.
Veröffentlicht: (2025)
von: Kitae Kim, et al.
Veröffentlicht: (2025)
Don't Play Favorites: Minority Guidance for Diffusion Models
von: Um, Soobin, et al.
Veröffentlicht: (2023)
von: Um, Soobin, et al.
Veröffentlicht: (2023)
Are Large Vision-Language Models Ready to Guide Blind and Low-Vision Individuals?
von: Kim, Eunki, et al.
Veröffentlicht: (2025)
von: Kim, Eunki, et al.
Veröffentlicht: (2025)
CBT-LLM: A Chinese Large Language Model for Cognitive Behavioral Therapy-based Mental Health Question Answering
von: Na, Hongbin
Veröffentlicht: (2024)
von: Na, Hongbin
Veröffentlicht: (2024)
EEG-Based Speech Decoding: A Novel Approach Using Multi-Kernel Ensemble Diffusion Models
von: Kim, Soowon, et al.
Veröffentlicht: (2024)
von: Kim, Soowon, et al.
Veröffentlicht: (2024)
Enhancing Robustness of Retrieval-Augmented Language Models with In-Context Learning
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
von: Park, Seong-Il, et al.
Veröffentlicht: (2024)
Suppression of spin bath and low-frequency noise for sub-MHz AC magnetometry based on double-dressed spin qubit in diamond
von: Kim, Kihwan, et al.
Veröffentlicht: (2022)
von: Kim, Kihwan, et al.
Veröffentlicht: (2022)
How Blind and Low-Vision Individuals Prefer Large Vision-Language Model-Generated Scene Descriptions
von: An, Na Min, et al.
Veröffentlicht: (2025)
von: An, Na Min, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reward-based Input Construction for Cross-document Relation Extraction
von: Na, Byeonghu, et al.
Veröffentlicht: (2024) -
Distillation of Large Language Models via Concrete Score Matching
von: Kim, Yeongmin, et al.
Veröffentlicht: (2025) -
Dirichlet-based Per-Sample Weighting by Transition Matrix for Noisy Label Learning
von: Bae, HeeSun, et al.
Veröffentlicht: (2024) -
AMiD: Knowledge Distillation for LLMs with $α$-mixture Assistant Distribution
von: Shin, Donghyeok, et al.
Veröffentlicht: (2025) -
Label-Noise Robust Diffusion Models
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)