CoEx -- Co-evolving World-model and Exploration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Minsoo, Hwang, Seung-won |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dual-Scale World Models for LLM Agents Towards Hard-Exploration Problems
von: Kim, Minsoo, et al.
Veröffentlicht: (2025)
von: Kim, Minsoo, et al.
Veröffentlicht: (2025)
Counterfactual-Consistency Prompting for Relative Temporal Understanding in Large Language Models
von: Kim, Jongho, et al.
Veröffentlicht: (2025)
von: Kim, Jongho, et al.
Veröffentlicht: (2025)
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
von: Storaï, Romain, et al.
Veröffentlicht: (2024)
von: Storaï, Romain, et al.
Veröffentlicht: (2024)
SAFE: Stepwise Atomic Feedback for Error correction in Multi-hop Reasoning
von: Kwon, Daeyong, et al.
Veröffentlicht: (2026)
von: Kwon, Daeyong, et al.
Veröffentlicht: (2026)
CREFT: Sequential Multi-Agent LLM for Character Relation Extraction
von: Chun, Ye Eun, et al.
Veröffentlicht: (2025)
von: Chun, Ye Eun, et al.
Veröffentlicht: (2025)
ConvCodeWorld: Benchmarking Conversational Code Generation in Reproducible Feedback Environments
von: Han, Hojae, et al.
Veröffentlicht: (2025)
von: Han, Hojae, et al.
Veröffentlicht: (2025)
Relevance to Utility: Process-Supervised Rewrite for RAG
von: Kim, Jaeyoung, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyoung, et al.
Veröffentlicht: (2025)
Chain of Grounded Objectives: Bridging Process and Goal-oriented Prompting for Code Generation
von: Yeo, Sangyeop, et al.
Veröffentlicht: (2025)
von: Yeo, Sangyeop, et al.
Veröffentlicht: (2025)
Disentangling Questions from Query Generation for Task-Adaptive Retrieval
von: Lee, Yoonsang, et al.
Veröffentlicht: (2024)
von: Lee, Yoonsang, et al.
Veröffentlicht: (2024)
ECoRAG: Evidentiality-guided Compression for Long Context RAG
von: Jeong, Yeonseok, et al.
Veröffentlicht: (2025)
von: Jeong, Yeonseok, et al.
Veröffentlicht: (2025)
Chaining Event Spans for Temporal Relation Grounding
von: Kim, Jongho, et al.
Veröffentlicht: (2025)
von: Kim, Jongho, et al.
Veröffentlicht: (2025)
PERC: Plan-As-Query Example Retrieval for Underrepresented Code Generation
von: Yoo, Jaeseok, et al.
Veröffentlicht: (2024)
von: Yoo, Jaeseok, et al.
Veröffentlicht: (2024)
ArchCode: Incorporating Software Requirements in Code Generation with Large Language Models
von: Han, Hojae, et al.
Veröffentlicht: (2024)
von: Han, Hojae, et al.
Veröffentlicht: (2024)
Ever-Evolving Memory by Blending and Refining the Past
von: Kim, Seo Hyun, et al.
Veröffentlicht: (2024)
von: Kim, Seo Hyun, et al.
Veröffentlicht: (2024)
AcuRank: Uncertainty-Aware Adaptive Computation for Listwise Reranking
von: Yoon, Soyoung, et al.
Veröffentlicht: (2025)
von: Yoon, Soyoung, et al.
Veröffentlicht: (2025)
Agent-as-Judge for Factual Summarization of Long Narratives
von: Jeong, Yeonseok, et al.
Veröffentlicht: (2025)
von: Jeong, Yeonseok, et al.
Veröffentlicht: (2025)
Relation-based Counterfactual Data Augmentation and Contrastive Learning for Robustifying Natural Language Inference Models
von: Yang, Heerin, et al.
Veröffentlicht: (2024)
von: Yang, Heerin, et al.
Veröffentlicht: (2024)
R$^3$-SQL: Ranking Reward and Resampling for Text-to-SQL
von: Han, Hojae, et al.
Veröffentlicht: (2026)
von: Han, Hojae, et al.
Veröffentlicht: (2026)
RoToR: Towards More Reliable Responses for Order-Invariant Inputs
von: Yoon, Soyoung, et al.
Veröffentlicht: (2025)
von: Yoon, Soyoung, et al.
Veröffentlicht: (2025)
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
MARS: Co-evolving Dual-System Deep Research via Multi-Agent Reinforcement Learning
von: Chen, Guoxin, et al.
Veröffentlicht: (2025)
von: Chen, Guoxin, et al.
Veröffentlicht: (2025)
UnIte: Uncertainty-based Iterative Document Sampling for Domain Adaptation in Information Retrieval
von: Kim, Jongyoon, et al.
Veröffentlicht: (2026)
von: Kim, Jongyoon, et al.
Veröffentlicht: (2026)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
The CoT Encyclopedia: Analyzing, Predicting, and Controlling how a Reasoning Model will Think
von: Lee, Seongyun, et al.
Veröffentlicht: (2025)
von: Lee, Seongyun, et al.
Veröffentlicht: (2025)
Semiparametric Token-Sequence Co-Supervision
von: Lee, Hyunji, et al.
Veröffentlicht: (2024)
von: Lee, Hyunji, et al.
Veröffentlicht: (2024)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
von: Park, Eunkyu, et al.
Veröffentlicht: (2025)
von: Park, Eunkyu, et al.
Veröffentlicht: (2025)
KoCoSa: Korean Context-aware Sarcasm Detection Dataset
von: Kim, Yumin, et al.
Veröffentlicht: (2024)
von: Kim, Yumin, et al.
Veröffentlicht: (2024)
Intended Target Identification for Anomia Patients with Gradient-based Selective Augmentation
von: Kim, Jongho, et al.
Veröffentlicht: (2025)
von: Kim, Jongho, et al.
Veröffentlicht: (2025)
Benchmarking Testing in Automated Theorem Proving
von: Kim, Jongyoon, et al.
Veröffentlicht: (2026)
von: Kim, Jongyoon, et al.
Veröffentlicht: (2026)
Learning Co-Speech Gesture for Multimodal Aphasia Type Detection
von: Lee, Daeun, et al.
Veröffentlicht: (2023)
von: Lee, Daeun, et al.
Veröffentlicht: (2023)
Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR
von: Lee, Chanuk, et al.
Veröffentlicht: (2026)
von: Lee, Chanuk, et al.
Veröffentlicht: (2026)
CoCoA: Collaborative Chain-of-Agents for Parametric-Retrieved Knowledge Synergy
von: Jiang, Yi, et al.
Veröffentlicht: (2025)
von: Jiang, Yi, et al.
Veröffentlicht: (2025)
CoMAS: Co-Evolving Multi-Agent Systems via Interaction Rewards
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2025)
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2025)
Interventional Speech Noise Injection for ASR Generalizable Spoken Language Understanding
von: Jung, Yeonjoon, et al.
Veröffentlicht: (2024)
von: Jung, Yeonjoon, et al.
Veröffentlicht: (2024)
Co-EPG: A Framework for Co-Evolution of Planning and Grounding in Autonomous GUI Agents
von: Zhao, Yuan, et al.
Veröffentlicht: (2025)
von: Zhao, Yuan, et al.
Veröffentlicht: (2025)
CoWork-X: Experience-Optimized Co-Evolution for Multi-Agent Collaboration System
von: Lin, Zexin, et al.
Veröffentlicht: (2026)
von: Lin, Zexin, et al.
Veröffentlicht: (2026)
CoCoP: Enhancing Text Classification with LLM through Code Completion Prompt
von: Mohajeri, Mohammad Mahdi, et al.
Veröffentlicht: (2024)
von: Mohajeri, Mohammad Mahdi, et al.
Veröffentlicht: (2024)
CoCoReviewBench: A Completeness- and Correctness-Oriented Benchmark for AI Reviewers
von: Deng, Hexuan, et al.
Veröffentlicht: (2026)
von: Deng, Hexuan, et al.
Veröffentlicht: (2026)
Improving the Reliability of LLMs: Combining CoT, RAG, Self-Consistency, and Self-Verification
von: Kumar, Adarsh, et al.
Veröffentlicht: (2025)
von: Kumar, Adarsh, et al.
Veröffentlicht: (2025)
Beyond Task-Oriented and Chitchat Dialogues: Proactive and Transition-Aware Conversational Agents
von: Yoon, Yejin, et al.
Veröffentlicht: (2025)
von: Yoon, Yejin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Dual-Scale World Models for LLM Agents Towards Hard-Exploration Problems
von: Kim, Minsoo, et al.
Veröffentlicht: (2025) -
Counterfactual-Consistency Prompting for Relative Temporal Understanding in Large Language Models
von: Kim, Jongho, et al.
Veröffentlicht: (2025) -
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
von: Storaï, Romain, et al.
Veröffentlicht: (2024) -
SAFE: Stepwise Atomic Feedback for Error correction in Multi-hop Reasoning
von: Kwon, Daeyong, et al.
Veröffentlicht: (2026) -
CREFT: Sequential Multi-Agent LLM for Character Relation Extraction
von: Chun, Ye Eun, et al.
Veröffentlicht: (2025)