Reward-Weighted Sampling: Enhancing Non-Autoregressive Characteristics in Masked Diffusion LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Gwak, Daehoon, Jung, Minseo, Park, Junwoo, Park, Minho, Park, ChaeHun, Hyung, Junha, Choo, Jaegul |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Not the Example, but the Process: How Self-Generated Examples Enhance LLM Reasoning
di: Gwak, Daehoon, et al.
Pubblicazione: (2026)
di: Gwak, Daehoon, et al.
Pubblicazione: (2026)
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling
di: Gwak, Daehoon, et al.
Pubblicazione: (2024)
di: Gwak, Daehoon, et al.
Pubblicazione: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
Breaking Chains: Unraveling the Links in Multi-Hop Knowledge Unlearning
di: Choi, Minseok, et al.
Pubblicazione: (2024)
di: Choi, Minseok, et al.
Pubblicazione: (2024)
PairEval: Open-domain Dialogue Evaluation with Pairwise Comparison
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
Can Tool-augmented Large Language Models be Aware of Incomplete Conditions?
di: Yang, Seungbin, et al.
Pubblicazione: (2024)
di: Yang, Seungbin, et al.
Pubblicazione: (2024)
Self-Supervised Contrastive Learning for Long-term Forecasting
di: Park, Junwoo, et al.
Pubblicazione: (2024)
di: Park, Junwoo, et al.
Pubblicazione: (2024)
Revisiting LLMs as Zero-Shot Time-Series Forecasters: Small Noise Can Break Large Models
di: Park, Junwoo, et al.
Pubblicazione: (2025)
di: Park, Junwoo, et al.
Pubblicazione: (2025)
The Comparative Trap: Pairwise Comparisons Amplifies Biased Preferences of LLM Evaluators
di: Jeong, Hawon, et al.
Pubblicazione: (2024)
di: Jeong, Hawon, et al.
Pubblicazione: (2024)
Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
Translation Deserves Better: Analyzing Translation Artifacts in Cross-lingual Visual Question Answering
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
LiveWeb-IE: A Benchmark For Online Web Information Extraction
di: Yang, Seungbin, et al.
Pubblicazione: (2026)
di: Yang, Seungbin, et al.
Pubblicazione: (2026)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
di: Hyung, Junha, et al.
Pubblicazione: (2024)
di: Hyung, Junha, et al.
Pubblicazione: (2024)
EgoX: Egocentric Video Generation from a Single Exocentric Video
di: Kang, Taewoong, et al.
Pubblicazione: (2025)
di: Kang, Taewoong, et al.
Pubblicazione: (2025)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
di: Kim, Kinam, et al.
Pubblicazione: (2025)
di: Kim, Kinam, et al.
Pubblicazione: (2025)
Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
di: Park, Minho, et al.
Pubblicazione: (2024)
di: Park, Minho, et al.
Pubblicazione: (2024)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
MagiCapture: High-Resolution Multi-Concept Portrait Customization
di: Hyung, Junha, et al.
Pubblicazione: (2023)
di: Hyung, Junha, et al.
Pubblicazione: (2023)
ACG: Action Coherence Guidance for Flow-based Vision-Language-Action models
di: Park, Minho, et al.
Pubblicazione: (2025)
di: Park, Minho, et al.
Pubblicazione: (2025)
AlphaFree: Recommendation Free from Users, IDs, and GNNs
di: Jeon, Minseo, et al.
Pubblicazione: (2026)
di: Jeon, Minseo, et al.
Pubblicazione: (2026)
Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
di: Hwang, Sungwon, et al.
Pubblicazione: (2025)
di: Hwang, Sungwon, et al.
Pubblicazione: (2025)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
di: Park, Minho, et al.
Pubblicazione: (2025)
di: Park, Minho, et al.
Pubblicazione: (2025)
Rewarding How Models Think Pedagogically: Integrating Pedagogical Reasoning and Thinking Rewards for LLMs in Education
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
M2S: Multi-turn to Single-turn jailbreak in Red Teaming for LLMs
di: Ha, Junwoo, et al.
Pubblicazione: (2025)
di: Ha, Junwoo, et al.
Pubblicazione: (2025)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
CA-LoRA: Concept-Aware LoRA for Domain-Aligned Segmentation Dataset Generation
di: Park, Minho, et al.
Pubblicazione: (2025)
di: Park, Minho, et al.
Pubblicazione: (2025)
Aligning Reasoning LLMs for Materials Discovery with Physics-aware Rejection Sampling
di: Hyun, Lee, et al.
Pubblicazione: (2025)
di: Hyun, Lee, et al.
Pubblicazione: (2025)
Cross-lingual Collapse: How Language-Centric Foundation Models Shape Reasoning in Large Language Models
di: Park, Cheonbok, et al.
Pubblicazione: (2025)
di: Park, Cheonbok, et al.
Pubblicazione: (2025)
TabReX : Tabular Referenceless eXplainable Evaluation
di: Anvekar, Tejas, et al.
Pubblicazione: (2025)
di: Anvekar, Tejas, et al.
Pubblicazione: (2025)
When Model Meets New Normals: Test-time Adaptation for Unsupervised Time-series Anomaly Detection
di: Kim, Dongmin, et al.
Pubblicazione: (2023)
di: Kim, Dongmin, et al.
Pubblicazione: (2023)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
ObjexMT: Objective Extraction and Metacognitive Calibration for LLM-as-a-Judge under Multi-Turn Jailbreaks
di: Kim, Hyunjun, et al.
Pubblicazione: (2025)
di: Kim, Hyunjun, et al.
Pubblicazione: (2025)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
Effective Rank Analysis and Regularization for Enhanced 3D Gaussian Splatting
di: Hyung, Junha, et al.
Pubblicazione: (2024)
di: Hyung, Junha, et al.
Pubblicazione: (2024)
Cross-Lingual Unlearning of Selective Knowledge in Multilingual Language Models
di: Choi, Minseok, et al.
Pubblicazione: (2024)
di: Choi, Minseok, et al.
Pubblicazione: (2024)
Enhancing Intrinsic Features for Debiasing via Investigating Class-Discerning Common Attributes in Bias-Contrastive Pair
di: Park, Jeonghoon, et al.
Pubblicazione: (2024)
di: Park, Jeonghoon, et al.
Pubblicazione: (2024)
X-Teaming Evolutionary M2S: Automated Discovery of Multi-turn to Single-turn Jailbreak Templates
di: Kim, Hyunjun, et al.
Pubblicazione: (2025)
di: Kim, Hyunjun, et al.
Pubblicazione: (2025)
Quantum-Enhanced Deterministic Inference of $k$-Independent Set Instances on Neutral Atom Arrays
di: Park, Juyoung, et al.
Pubblicazione: (2026)
di: Park, Juyoung, et al.
Pubblicazione: (2026)
Diffusion-Inspired Masked Fine-Tuning for Knowledge Injection in Autoregressive LLMs
di: Pan, Xu, et al.
Pubblicazione: (2025)
di: Pan, Xu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Not the Example, but the Process: How Self-Generated Examples Enhance LLM Reasoning
di: Gwak, Daehoon, et al.
Pubblicazione: (2026) -
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling
di: Gwak, Daehoon, et al.
Pubblicazione: (2024) -
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
di: Park, ChaeHun, et al.
Pubblicazione: (2024) -
Breaking Chains: Unraveling the Links in Multi-Hop Knowledge Unlearning
di: Choi, Minseok, et al.
Pubblicazione: (2024) -
PairEval: Open-domain Dialogue Evaluation with Pairwise Comparison
di: Park, ChaeHun, et al.
Pubblicazione: (2024)