Saved in:
| Main Authors: | Ko, Dayoon, Kim, Jihyuk, Park, Haeju, Kim, Sohyeon, Lee, Dahyun, Jo, Yongrae, Kim, Gunhee, Lee, Moontae, Lee, Kyungjae |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.19113 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Is Enough Not Enough? Illusory Completion in Search Agents
by: Ko, Dayoon, et al.
Published: (2026)
by: Ko, Dayoon, et al.
Published: (2026)
Shifting from Ranking to Set Selection for Retrieval Augmented Generation
by: Lee, Dahyun, et al.
Published: (2025)
by: Lee, Dahyun, et al.
Published: (2025)
DynamicER: Resolving Emerging Mentions to Dynamic Entities for RAG
by: Kim, Jinyoung, et al.
Published: (2024)
by: Kim, Jinyoung, et al.
Published: (2024)
Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models
by: Kim, Jiyeon, et al.
Published: (2026)
by: Kim, Jiyeon, et al.
Published: (2026)
Lost in the Noise: How Reasoning Models Fail with Contextual Distractors
by: Lee, Seongyun, et al.
Published: (2026)
by: Lee, Seongyun, et al.
Published: (2026)
One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL
by: Chae, Hyungjoo, et al.
Published: (2025)
by: Chae, Hyungjoo, et al.
Published: (2025)
Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates
by: Ahn, Jaewoo, et al.
Published: (2025)
by: Ahn, Jaewoo, et al.
Published: (2025)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
by: Lee, Kyungjae, et al.
Published: (2024)
by: Lee, Kyungjae, et al.
Published: (2024)
Can Language Models Laugh at YouTube Short-form Videos?
by: Ko, Dayoon, et al.
Published: (2023)
by: Ko, Dayoon, et al.
Published: (2023)
The CoT Encyclopedia: Analyzing, Predicting, and Controlling how a Reasoning Model will Think
by: Lee, Seongyun, et al.
Published: (2025)
by: Lee, Seongyun, et al.
Published: (2025)
When Should Dense Retrievers Be Updated in Evolving Corpora? Detecting Out-of-Distribution Corpora Using GradNormIR
by: Ko, Dayoon, et al.
Published: (2025)
by: Ko, Dayoon, et al.
Published: (2025)
Self-Corrective Task Planning by Inverse Prompting with Large Language Models
by: Lee, Jiho, et al.
Published: (2025)
by: Lee, Jiho, et al.
Published: (2025)
EXAONE Deep: Reasoning Enhanced Language Models
by: Bae, Kyunghoon, et al.
Published: (2025)
by: Bae, Kyunghoon, et al.
Published: (2025)
Open Ko-LLM Leaderboard2: Bridging Foundational and Practical Evaluation for Korean LLMs
by: Kim, Hyeonwoo, et al.
Published: (2024)
by: Kim, Hyeonwoo, et al.
Published: (2024)
Representing the Under-Represented: Cultural and Core Capability Benchmarks for Developing Thai Large Language Models
by: Kim, Dahyun, et al.
Published: (2024)
by: Kim, Dahyun, et al.
Published: (2024)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
by: Park, Chanjun, et al.
Published: (2024)
by: Park, Chanjun, et al.
Published: (2024)
Byzantine-Robust Decentralized Coordination of LLM Agents
by: Jo, Yongrae, et al.
Published: (2025)
by: Jo, Yongrae, et al.
Published: (2025)
Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning
by: Son, Jaehyeon, et al.
Published: (2025)
by: Son, Jaehyeon, et al.
Published: (2025)
Recasting Continual Learning as Sequence Modeling
by: Lee, Soochan, et al.
Published: (2023)
by: Lee, Soochan, et al.
Published: (2023)
Dataverse: Open-Source ETL (Extract, Transform, Load) Pipeline for Large Language Models
by: Park, Hyunbyung, et al.
Published: (2024)
by: Park, Hyunbyung, et al.
Published: (2024)
Evidentiality-aware Retrieval for Overcoming Abstractiveness in Open-Domain Question Answering
by: Song, Yongho, et al.
Published: (2023)
by: Song, Yongho, et al.
Published: (2023)
EXAONE 3.0 7.8B Instruction Tuned Language Model
by: An, Soyoung, et al.
Published: (2024)
by: An, Soyoung, et al.
Published: (2024)
Compositional Conservatism: A Transductive Approach in Offline Reinforcement Learning
by: Song, Yeda, et al.
Published: (2024)
by: Song, Yeda, et al.
Published: (2024)
Adaptive Retrieval for Reasoning-Intensive Retrieval
by: Kim, Jongho, et al.
Published: (2026)
by: Kim, Jongho, et al.
Published: (2026)
1 Trillion Token (1TT) Platform: A Novel Framework for Efficient Data Sharing and Compensation in Large Language Models
by: Park, Chanjun, et al.
Published: (2024)
by: Park, Chanjun, et al.
Published: (2024)
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR
by: Kim, Hajung, et al.
Published: (2024)
by: Kim, Hajung, et al.
Published: (2024)
Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression
by: Jung, Dahyun, et al.
Published: (2026)
by: Jung, Dahyun, et al.
Published: (2026)
SafeDPO: A Simple Approach to Direct Preference Optimization with Enhanced Safety
by: Kim, Geon-Hyeong, et al.
Published: (2025)
by: Kim, Geon-Hyeong, et al.
Published: (2025)
Think, Verbalize, then Speak: Bridging Complex Thoughts and Comprehensible Speech
by: Woo, Sang Hoon, et al.
Published: (2025)
by: Woo, Sang Hoon, et al.
Published: (2025)
GrowOVER: How Can LLMs Adapt to Growing Real-World Knowledge?
by: Ko, Dayoon, et al.
Published: (2024)
by: Ko, Dayoon, et al.
Published: (2024)
Adaptive Parallel Monte Carlo Tree Search for Efficient Test-time Compute Scaling
by: Kim, Hongbeen, et al.
Published: (2026)
by: Kim, Hongbeen, et al.
Published: (2026)
Learning to Continually Learn with the Bayesian Principle
by: Lee, Soochan, et al.
Published: (2024)
by: Lee, Soochan, et al.
Published: (2024)
Semantic Skill Grounding for Embodied Instruction-Following in Cross-Domain Environments
by: Shin, Sangwoo, et al.
Published: (2024)
by: Shin, Sangwoo, et al.
Published: (2024)
Policy-labeled Preference Learning: Is Preference Enough for RLHF?
by: Cho, Taehyun, et al.
Published: (2025)
by: Cho, Taehyun, et al.
Published: (2025)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
by: Cho, Taehyun, et al.
Published: (2024)
by: Cho, Taehyun, et al.
Published: (2024)
LEGO: Latent-space Exploration for Geometry-aware Optimization of Humanoid Kinematic Design
by: Yoon, Jihwan, et al.
Published: (2026)
by: Yoon, Jihwan, et al.
Published: (2026)
Assessing LLM Reasoning Steps via Principal Knowledge Grounding
by: Hwang, Hyeon, et al.
Published: (2025)
by: Hwang, Hyeon, et al.
Published: (2025)
GRACE: Discriminator-Guided Chain-of-Thought Reasoning
by: Khalifa, Muhammad, et al.
Published: (2023)
by: Khalifa, Muhammad, et al.
Published: (2023)
DiffInject: Revisiting Debias via Synthetic Data Generation using Diffusion-based Style Injection
by: Ko, Donggeun, et al.
Published: (2024)
by: Ko, Donggeun, et al.
Published: (2024)
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
by: Park, Chanjun, et al.
Published: (2024)
by: Park, Chanjun, et al.
Published: (2024)
Similar Items
-
When Is Enough Not Enough? Illusory Completion in Search Agents
by: Ko, Dayoon, et al.
Published: (2026) -
Shifting from Ranking to Set Selection for Retrieval Augmented Generation
by: Lee, Dahyun, et al.
Published: (2025) -
DynamicER: Resolving Emerging Mentions to Dynamic Entities for RAG
by: Kim, Jinyoung, et al.
Published: (2024) -
Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models
by: Kim, Jiyeon, et al.
Published: (2026) -
Lost in the Noise: How Reasoning Models Fail with Contextual Distractors
by: Lee, Seongyun, et al.
Published: (2026)