Saved in:
| Main Authors: | Zhang, Yunxiang, Zhou, Kang, Xu, Zhichao, Ramnath, Kiran, Zhou, Yun, Woo, Sangmin, Ding, Haibo, Cheong, Lin Lee |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.17596 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
Diffusion Language Model Inference with Monte Carlo Tree Search
by: Huang, Zheng, et al.
Published: (2025)
by: Huang, Zheng, et al.
Published: (2025)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
by: Woo, Sangmin, et al.
Published: (2025)
by: Woo, Sangmin, et al.
Published: (2025)
An Empirical Study of Automating Agent Evaluation
by: Zhou, Kang, et al.
Published: (2026)
by: Zhou, Kang, et al.
Published: (2026)
CSPLADE: Learned Sparse Retrieval with Causal Language Models
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
A Systematic Survey of Automatic Prompt Optimization Techniques
by: Ramnath, Kiran, et al.
Published: (2025)
by: Ramnath, Kiran, et al.
Published: (2025)
Reinforcement Learning for LLM Post-Training: A Survey
by: Wang, Zhichao, et al.
Published: (2024)
by: Wang, Zhichao, et al.
Published: (2024)
Synthetic Sandbox for Training Machine Learning Engineering Agents
by: Zhou, Yuhang, et al.
Published: (2026)
by: Zhou, Yuhang, et al.
Published: (2026)
LAMA-UT: Language Agnostic Multilingual ASR through Orthography Unification and Language-Specific Transliteration
by: Lee, Sangmin, et al.
Published: (2024)
by: Lee, Sangmin, et al.
Published: (2024)
PromptPrism: A Linguistically-Inspired Taxonomy for Prompts
by: Jeoung, Sullam, et al.
Published: (2025)
by: Jeoung, Sullam, et al.
Published: (2025)
PAFT: A Parallel Training Paradigm for Effective LLM Fine-Tuning
by: Pentyala, Shiva Kumar, et al.
Published: (2024)
by: Pentyala, Shiva Kumar, et al.
Published: (2024)
MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
by: Chan, Jun Shern, et al.
Published: (2024)
by: Chan, Jun Shern, et al.
Published: (2024)
BayesFlow: A Probability Inference Framework for Meta-Agent Assisted Workflow Generation
by: Yuan, Bo, et al.
Published: (2026)
by: Yuan, Bo, et al.
Published: (2026)
ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering
by: Liu, Zexi, et al.
Published: (2025)
by: Liu, Zexi, et al.
Published: (2025)
Advancing Multi-Agent RAG Systems with Minimalist Reinforcement Learning
by: Wu, Yihong, et al.
Published: (2025)
by: Wu, Yihong, et al.
Published: (2025)
Accelerating Large Language Model Inference via Early-Exiting Algorithms
by: Bae, Sangmin
Published: (2025)
by: Bae, Sangmin
Published: (2025)
SAGE-LD: Towards Scalable and Generalizable End-to-End Language Diarization via Simulated Data Augmentation
by: Lee, Sangmin, et al.
Published: (2025)
by: Lee, Sangmin, et al.
Published: (2025)
LLM-Based Offline Learning for Embodied Agents via Consistency-Guided Reward Ensemble
by: Lee, Yujeong, et al.
Published: (2024)
by: Lee, Yujeong, et al.
Published: (2024)
UniCoM: A Universal Code-Switching Speech Generator
by: Lee, Sangmin, et al.
Published: (2025)
by: Lee, Sangmin, et al.
Published: (2025)
Multi-Agent Visual-Language Reasoning for Comprehensive Highway Scene Understanding
by: Yang, Yunxiang, et al.
Published: (2025)
by: Yang, Yunxiang, et al.
Published: (2025)
Co-Learning: Code Learning for Multi-Agent Reinforcement Collaborative Framework with Conversational Natural Language Interfaces
by: Yu, Jiapeng, et al.
Published: (2024)
by: Yu, Jiapeng, et al.
Published: (2024)
Structured Prompting and Multi-Agent Knowledge Distillation for Traffic Video Interpretation and Risk Inference
by: Yang, Yunxiang, et al.
Published: (2025)
by: Yang, Yunxiang, et al.
Published: (2025)
Scaling Personality Control in LLMs with Big Five Scaler Prompts
by: Cho, Gunhee, et al.
Published: (2025)
by: Cho, Gunhee, et al.
Published: (2025)
FMBench: Adaptive Large Language Model Output Formatting
by: Wang, Yaoting, et al.
Published: (2026)
by: Wang, Yaoting, et al.
Published: (2026)
SLOT: Structuring the Output of Large Language Models
by: Wang, Darren Yow-Bang, et al.
Published: (2025)
by: Wang, Darren Yow-Bang, et al.
Published: (2025)
Symbolic Learning Enables Self-Evolving Agents
by: Zhou, Wangchunshu, et al.
Published: (2024)
by: Zhou, Wangchunshu, et al.
Published: (2024)
Gaming the Judge: Unfaithful Chain-of-Thought Can Undermine Agent Evaluation
by: Khalifa, Muhammad, et al.
Published: (2026)
by: Khalifa, Muhammad, et al.
Published: (2026)
SERM: Self-Evolving Relevance Model with Agent-Driven Learning from Massive Query Streams
by: Wang, Chenglong, et al.
Published: (2026)
by: Wang, Chenglong, et al.
Published: (2026)
ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
by: Wan, Ziyu, et al.
Published: (2025)
by: Wan, Ziyu, et al.
Published: (2025)
Learning to Use Tools via Cooperative and Interactive Agents
by: Shi, Zhengliang, et al.
Published: (2024)
by: Shi, Zhengliang, et al.
Published: (2024)
Learning Video Temporal Dynamics with Cross-Modal Attention for Robust Audio-Visual Speech Recognition
by: Kim, Sungnyun, et al.
Published: (2024)
by: Kim, Sungnyun, et al.
Published: (2024)
GOVERN: Gradient Orientation Vote Ensemble for Multi-Teacher Reinforced Distillation
by: Zhou, Wenjie, et al.
Published: (2024)
by: Zhou, Wenjie, et al.
Published: (2024)
Reinforcement World Model Learning for LLM-based Agents
by: Yu, Xiao, et al.
Published: (2026)
by: Yu, Xiao, et al.
Published: (2026)
Distillation versus Contrastive Learning: How to Train Your Rerankers
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
MMedAgent: Learning to Use Medical Tools with Multi-modal Agent
by: Li, Binxu, et al.
Published: (2024)
by: Li, Binxu, et al.
Published: (2024)
Learning to Retrieve from Agent Trajectories
by: Zhou, Yuqi, et al.
Published: (2026)
by: Zhou, Yuqi, et al.
Published: (2026)
MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval
by: Li, Chunyu, et al.
Published: (2026)
by: Li, Chunyu, et al.
Published: (2026)
Multi-Task Learning for Front-End Text Processing in TTS
by: Kang, Wonjune, et al.
Published: (2024)
by: Kang, Wonjune, et al.
Published: (2024)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
by: Alonso, Silvio
Published: (2025)
by: Alonso, Silvio
Published: (2025)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
by: Alonso, Silvio, et al.
Published: (2025)
by: Alonso, Silvio, et al.
Published: (2025)
Similar Items
-
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
by: Xu, Zhichao, et al.
Published: (2025) -
Diffusion Language Model Inference with Monte Carlo Tree Search
by: Huang, Zheng, et al.
Published: (2025) -
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
by: Woo, Sangmin, et al.
Published: (2025) -
An Empirical Study of Automating Agent Evaluation
by: Zhou, Kang, et al.
Published: (2026) -
CSPLADE: Learned Sparse Retrieval with Causal Language Models
by: Xu, Zhichao, et al.
Published: (2025)