Salvato in:
| Autori principali: | Zhang, Yunxiang, Zhou, Kang, Xu, Zhichao, Ramnath, Kiran, Zhou, Yun, Woo, Sangmin, Ding, Haibo, Cheong, Lin Lee |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2601.17596 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
di: Xu, Zhichao, et al.
Pubblicazione: (2025)
di: Xu, Zhichao, et al.
Pubblicazione: (2025)
Diffusion Language Model Inference with Monte Carlo Tree Search
di: Huang, Zheng, et al.
Pubblicazione: (2025)
di: Huang, Zheng, et al.
Pubblicazione: (2025)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2025)
di: Woo, Sangmin, et al.
Pubblicazione: (2025)
An Empirical Study of Automating Agent Evaluation
di: Zhou, Kang, et al.
Pubblicazione: (2026)
di: Zhou, Kang, et al.
Pubblicazione: (2026)
CSPLADE: Learned Sparse Retrieval with Causal Language Models
di: Xu, Zhichao, et al.
Pubblicazione: (2025)
di: Xu, Zhichao, et al.
Pubblicazione: (2025)
A Systematic Survey of Automatic Prompt Optimization Techniques
di: Ramnath, Kiran, et al.
Pubblicazione: (2025)
di: Ramnath, Kiran, et al.
Pubblicazione: (2025)
Reinforcement Learning for LLM Post-Training: A Survey
di: Wang, Zhichao, et al.
Pubblicazione: (2024)
di: Wang, Zhichao, et al.
Pubblicazione: (2024)
Synthetic Sandbox for Training Machine Learning Engineering Agents
di: Zhou, Yuhang, et al.
Pubblicazione: (2026)
di: Zhou, Yuhang, et al.
Pubblicazione: (2026)
LAMA-UT: Language Agnostic Multilingual ASR through Orthography Unification and Language-Specific Transliteration
di: Lee, Sangmin, et al.
Pubblicazione: (2024)
di: Lee, Sangmin, et al.
Pubblicazione: (2024)
PromptPrism: A Linguistically-Inspired Taxonomy for Prompts
di: Jeoung, Sullam, et al.
Pubblicazione: (2025)
di: Jeoung, Sullam, et al.
Pubblicazione: (2025)
PAFT: A Parallel Training Paradigm for Effective LLM Fine-Tuning
di: Pentyala, Shiva Kumar, et al.
Pubblicazione: (2024)
di: Pentyala, Shiva Kumar, et al.
Pubblicazione: (2024)
MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
di: Chan, Jun Shern, et al.
Pubblicazione: (2024)
di: Chan, Jun Shern, et al.
Pubblicazione: (2024)
BayesFlow: A Probability Inference Framework for Meta-Agent Assisted Workflow Generation
di: Yuan, Bo, et al.
Pubblicazione: (2026)
di: Yuan, Bo, et al.
Pubblicazione: (2026)
ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering
di: Liu, Zexi, et al.
Pubblicazione: (2025)
di: Liu, Zexi, et al.
Pubblicazione: (2025)
Advancing Multi-Agent RAG Systems with Minimalist Reinforcement Learning
di: Wu, Yihong, et al.
Pubblicazione: (2025)
di: Wu, Yihong, et al.
Pubblicazione: (2025)
Accelerating Large Language Model Inference via Early-Exiting Algorithms
di: Bae, Sangmin
Pubblicazione: (2025)
di: Bae, Sangmin
Pubblicazione: (2025)
SAGE-LD: Towards Scalable and Generalizable End-to-End Language Diarization via Simulated Data Augmentation
di: Lee, Sangmin, et al.
Pubblicazione: (2025)
di: Lee, Sangmin, et al.
Pubblicazione: (2025)
LLM-Based Offline Learning for Embodied Agents via Consistency-Guided Reward Ensemble
di: Lee, Yujeong, et al.
Pubblicazione: (2024)
di: Lee, Yujeong, et al.
Pubblicazione: (2024)
UniCoM: A Universal Code-Switching Speech Generator
di: Lee, Sangmin, et al.
Pubblicazione: (2025)
di: Lee, Sangmin, et al.
Pubblicazione: (2025)
Multi-Agent Visual-Language Reasoning for Comprehensive Highway Scene Understanding
di: Yang, Yunxiang, et al.
Pubblicazione: (2025)
di: Yang, Yunxiang, et al.
Pubblicazione: (2025)
Co-Learning: Code Learning for Multi-Agent Reinforcement Collaborative Framework with Conversational Natural Language Interfaces
di: Yu, Jiapeng, et al.
Pubblicazione: (2024)
di: Yu, Jiapeng, et al.
Pubblicazione: (2024)
Structured Prompting and Multi-Agent Knowledge Distillation for Traffic Video Interpretation and Risk Inference
di: Yang, Yunxiang, et al.
Pubblicazione: (2025)
di: Yang, Yunxiang, et al.
Pubblicazione: (2025)
Scaling Personality Control in LLMs with Big Five Scaler Prompts
di: Cho, Gunhee, et al.
Pubblicazione: (2025)
di: Cho, Gunhee, et al.
Pubblicazione: (2025)
FMBench: Adaptive Large Language Model Output Formatting
di: Wang, Yaoting, et al.
Pubblicazione: (2026)
di: Wang, Yaoting, et al.
Pubblicazione: (2026)
SLOT: Structuring the Output of Large Language Models
di: Wang, Darren Yow-Bang, et al.
Pubblicazione: (2025)
di: Wang, Darren Yow-Bang, et al.
Pubblicazione: (2025)
Symbolic Learning Enables Self-Evolving Agents
di: Zhou, Wangchunshu, et al.
Pubblicazione: (2024)
di: Zhou, Wangchunshu, et al.
Pubblicazione: (2024)
Gaming the Judge: Unfaithful Chain-of-Thought Can Undermine Agent Evaluation
di: Khalifa, Muhammad, et al.
Pubblicazione: (2026)
di: Khalifa, Muhammad, et al.
Pubblicazione: (2026)
SERM: Self-Evolving Relevance Model with Agent-Driven Learning from Massive Query Streams
di: Wang, Chenglong, et al.
Pubblicazione: (2026)
di: Wang, Chenglong, et al.
Pubblicazione: (2026)
ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
di: Wan, Ziyu, et al.
Pubblicazione: (2025)
di: Wan, Ziyu, et al.
Pubblicazione: (2025)
Learning to Use Tools via Cooperative and Interactive Agents
di: Shi, Zhengliang, et al.
Pubblicazione: (2024)
di: Shi, Zhengliang, et al.
Pubblicazione: (2024)
Learning Video Temporal Dynamics with Cross-Modal Attention for Robust Audio-Visual Speech Recognition
di: Kim, Sungnyun, et al.
Pubblicazione: (2024)
di: Kim, Sungnyun, et al.
Pubblicazione: (2024)
GOVERN: Gradient Orientation Vote Ensemble for Multi-Teacher Reinforced Distillation
di: Zhou, Wenjie, et al.
Pubblicazione: (2024)
di: Zhou, Wenjie, et al.
Pubblicazione: (2024)
Reinforcement World Model Learning for LLM-based Agents
di: Yu, Xiao, et al.
Pubblicazione: (2026)
di: Yu, Xiao, et al.
Pubblicazione: (2026)
Distillation versus Contrastive Learning: How to Train Your Rerankers
di: Xu, Zhichao, et al.
Pubblicazione: (2025)
di: Xu, Zhichao, et al.
Pubblicazione: (2025)
MMedAgent: Learning to Use Medical Tools with Multi-modal Agent
di: Li, Binxu, et al.
Pubblicazione: (2024)
di: Li, Binxu, et al.
Pubblicazione: (2024)
Learning to Retrieve from Agent Trajectories
di: Zhou, Yuqi, et al.
Pubblicazione: (2026)
di: Zhou, Yuqi, et al.
Pubblicazione: (2026)
MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval
di: Li, Chunyu, et al.
Pubblicazione: (2026)
di: Li, Chunyu, et al.
Pubblicazione: (2026)
Multi-Task Learning for Front-End Text Processing in TTS
di: Kang, Wonjune, et al.
Pubblicazione: (2024)
di: Kang, Wonjune, et al.
Pubblicazione: (2024)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
di: Alonso, Silvio
Pubblicazione: (2025)
di: Alonso, Silvio
Pubblicazione: (2025)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
di: Alonso, Silvio, et al.
Pubblicazione: (2025)
di: Alonso, Silvio, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
di: Xu, Zhichao, et al.
Pubblicazione: (2025) -
Diffusion Language Model Inference with Monte Carlo Tree Search
di: Huang, Zheng, et al.
Pubblicazione: (2025) -
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2025) -
An Empirical Study of Automating Agent Evaluation
di: Zhou, Kang, et al.
Pubblicazione: (2026) -
CSPLADE: Learned Sparse Retrieval with Causal Language Models
di: Xu, Zhichao, et al.
Pubblicazione: (2025)