Efficient LLM Collaboration via Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Byeongchan, Lee, Jonghoon, Kim, Dongyoung, Kim, Jaehyung, Park, Kyungjoon, Lee, Dongjun, Shin, Jinwoo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
by: Kim, Dongyoung, et al.
Published: (2024)
by: Kim, Dongyoung, et al.
Published: (2024)
Debiasing Online Preference Learning via Preference Feature Preservation
by: Kim, Dongyoung, et al.
Published: (2025)
by: Kim, Dongyoung, et al.
Published: (2025)
Training-free LLM Verification via Recycling Few-shot Examples
by: Lee, Dongseok, et al.
Published: (2025)
by: Lee, Dongseok, et al.
Published: (2025)
Robot-R1: Reinforcement Learning for Enhanced Embodied Reasoning in Robotics
by: Kim, Dongyoung, et al.
Published: (2025)
by: Kim, Dongyoung, et al.
Published: (2025)
RoboAlign: Learning Test-Time Reasoning for Language-Action Alignment in Vision-Language-Action Models
by: Kim, Dongyoung, et al.
Published: (2026)
by: Kim, Dongyoung, et al.
Published: (2026)
Learning to Correct for QA Reasoning with Black-box LLMs
by: Kim, Jaehyung, et al.
Published: (2024)
by: Kim, Jaehyung, et al.
Published: (2024)
From Street to Orbit: Training-Free Cross-View Retrieval via Location Semantics and LLM Guidance
by: Min, Jeongho, et al.
Published: (2025)
by: Min, Jeongho, et al.
Published: (2025)
Accelerating Reinforcement Learning with Value-Conditional State Entropy Exploration
by: Kim, Dongyoung, et al.
Published: (2023)
by: Kim, Dongyoung, et al.
Published: (2023)
Verifier-free Test-Time Sampling for Vision Language Action Models
by: Jang, Suhyeok, et al.
Published: (2025)
by: Jang, Suhyeok, et al.
Published: (2025)
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
by: Nam, Jaehyun, et al.
Published: (2024)
by: Nam, Jaehyun, et al.
Published: (2024)
Implicit Contrastive Representation Learning with Guided Stop-gradient
by: Lee, Byeongchan, et al.
Published: (2025)
by: Lee, Byeongchan, et al.
Published: (2025)
Peri-LN: Revisiting Normalization Layer in the Transformer Architecture
by: Kim, Jeonghoon, et al.
Published: (2025)
by: Kim, Jeonghoon, et al.
Published: (2025)
Align to Misalign: Automatic LLM Jailbreak with Meta-Optimized LLM Judges
by: Koo, Hamin, et al.
Published: (2025)
by: Koo, Hamin, et al.
Published: (2025)
SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation
by: Kim, Seoyeon, et al.
Published: (2026)
by: Kim, Seoyeon, et al.
Published: (2026)
RoboCurate: Harnessing Diversity with Action-Verified Neural Trajectory for Robot Learning
by: Kim, Seungku, et al.
Published: (2026)
by: Kim, Seungku, et al.
Published: (2026)
StarFT: Robust Fine-tuning of Zero-shot Models via Spuriosity Alignment
by: Kim, Younghyun, et al.
Published: (2025)
by: Kim, Younghyun, et al.
Published: (2025)
Multi-Step Likelihood-Ratio Correction for Reinforcement Learning with Verifiable Rewards
by: Yoon, Deokgyu, et al.
Published: (2026)
by: Yoon, Deokgyu, et al.
Published: (2026)
Learning to Generate Unit Test via Adversarial Reinforcement Learning
by: Lee, Dongjun, et al.
Published: (2025)
by: Lee, Dongjun, et al.
Published: (2025)
SuRe: Summarizing Retrievals using Answer Candidates for Open-domain QA of LLMs
by: Kim, Jaehyung, et al.
Published: (2024)
by: Kim, Jaehyung, et al.
Published: (2024)
Trust Region Q Adjoint Matching
by: Dong, Yonghoon, et al.
Published: (2026)
by: Dong, Yonghoon, et al.
Published: (2026)
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
by: Kim, Sanghyun, et al.
Published: (2024)
by: Kim, Sanghyun, et al.
Published: (2024)
Visual Representation Learning with Stochastic Frame Prediction
by: Jang, Huiwon, et al.
Published: (2024)
by: Jang, Huiwon, et al.
Published: (2024)
PPMI: Privacy-Preserving LLM Interaction with Socratic Chain-of-Thought Reasoning and Homomorphically Encrypted Vector Databases
by: Bae, Yubeen, et al.
Published: (2025)
by: Bae, Yubeen, et al.
Published: (2025)
InterPol: De-anonymizing LM Arena via Interpolated Preference Learning
by: Cho, Minsung, et al.
Published: (2026)
by: Cho, Minsung, et al.
Published: (2026)
Bridging the Domain Gap: A Simple Domain Matching Method for Reference-based Image Super-Resolution in Remote Sensing
by: Min, Jeongho, et al.
Published: (2024)
by: Min, Jeongho, et al.
Published: (2024)
DiffInject: Revisiting Debias via Synthetic Data Generation using Diffusion-based Style Injection
by: Ko, Donggeun, et al.
Published: (2024)
by: Ko, Donggeun, et al.
Published: (2024)
Reasoning or Fluency? Dissecting Probabilistic Confidence in Best-of-N Selection
by: Kim, Hojin, et al.
Published: (2026)
by: Kim, Hojin, et al.
Published: (2026)
CarPLAN: Context-Adaptive and Robust Planning with Dynamic Scene Awareness for Autonomous Driving
by: Yun, Junyong, et al.
Published: (2026)
by: Yun, Junyong, et al.
Published: (2026)
Real-World Efficient Blind Motion Deblurring via Blur Pixel Discretization
by: Kim, Insoo, et al.
Published: (2024)
by: Kim, Insoo, et al.
Published: (2024)
IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents
by: Choi, Daewon, et al.
Published: (2026)
by: Choi, Daewon, et al.
Published: (2026)
Test-Time Adaptation with Binary Feedback
by: Lee, Taeckyung, et al.
Published: (2025)
by: Lee, Taeckyung, et al.
Published: (2025)
Self-Corrective Task Planning by Inverse Prompting with Large Language Models
by: Lee, Jiho, et al.
Published: (2025)
by: Lee, Jiho, et al.
Published: (2025)
A Comprehensive Survey of Deep Learning for Time Series Forecasting: Architectural Diversity and Open Challenges
by: Kim, Jongseon, et al.
Published: (2024)
by: Kim, Jongseon, et al.
Published: (2024)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
by: Kang, Minjae, et al.
Published: (2026)
by: Kang, Minjae, et al.
Published: (2026)
GuruAgents: Emulating Wise Investors with Prompt-Guided LLM Agents
by: Kim, Yejin, et al.
Published: (2025)
by: Kim, Yejin, et al.
Published: (2025)
Tabular Transfer Learning via Prompting LLMs
by: Nam, Jaehyun, et al.
Published: (2024)
by: Nam, Jaehyun, et al.
Published: (2024)
Debiasing Classifiers by Amplifying Bias with Latent Diffusion and Large Language Models
by: Ko, Donggeun, et al.
Published: (2024)
by: Ko, Donggeun, et al.
Published: (2024)
ARGORA: Orchestrated Argumentation for Causally Grounded LLM Reasoning and Decision Making
by: Jin, Youngjin, et al.
Published: (2026)
by: Jin, Youngjin, et al.
Published: (2026)
Personalized LLM Decoding via Contrasting Personal Preference
by: Bu, Hyungjune, et al.
Published: (2025)
by: Bu, Hyungjune, et al.
Published: (2025)
TiTok: Transfer Token-level Knowledge via Contrastive Excess to Transplant LoRA
by: Jung, Chanjoo, et al.
Published: (2025)
by: Jung, Chanjoo, et al.
Published: (2025)
Similar Items
-
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
by: Kim, Dongyoung, et al.
Published: (2024) -
Debiasing Online Preference Learning via Preference Feature Preservation
by: Kim, Dongyoung, et al.
Published: (2025) -
Training-free LLM Verification via Recycling Few-shot Examples
by: Lee, Dongseok, et al.
Published: (2025) -
Robot-R1: Reinforcement Learning for Enhanced Embodied Reasoning in Robotics
by: Kim, Dongyoung, et al.
Published: (2025) -
RoboAlign: Learning Test-Time Reasoning for Language-Action Alignment in Vision-Language-Action Models
by: Kim, Dongyoung, et al.
Published: (2026)