elsciRL: Integrating Language Solutions into Reinforcement Learning Problem Settings
Fuente:
arXiv
Salvato in:
| Autori principali: | Osborne, Philip, Carvalho, Danilo S., Freitas, André |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
di: Wang, Yuchen, et al.
Pubblicazione: (2026)
di: Wang, Yuchen, et al.
Pubblicazione: (2026)
Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce
di: Resh, William G., et al.
Pubblicazione: (2025)
di: Resh, William G., et al.
Pubblicazione: (2025)
Applying Cognitive Design Patterns to General LLM Agents
di: Wray, Robert E., et al.
Pubblicazione: (2025)
di: Wray, Robert E., et al.
Pubblicazione: (2025)
Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay
di: Wang, Xiaohua, et al.
Pubblicazione: (2026)
di: Wang, Xiaohua, et al.
Pubblicazione: (2026)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
di: Chang, Jiale, et al.
Pubblicazione: (2026)
di: Chang, Jiale, et al.
Pubblicazione: (2026)
Post Hoc Extraction of Pareto Fronts for Continuous Control
di: Thakar, Raghav, et al.
Pubblicazione: (2026)
di: Thakar, Raghav, et al.
Pubblicazione: (2026)
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
di: Wang, Zixu, et al.
Pubblicazione: (2026)
di: Wang, Zixu, et al.
Pubblicazione: (2026)
MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents
di: Sidik, Bronislav, et al.
Pubblicazione: (2026)
di: Sidik, Bronislav, et al.
Pubblicazione: (2026)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
di: Li, Bowen, et al.
Pubblicazione: (2026)
di: Li, Bowen, et al.
Pubblicazione: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
di: Costa, Rimom
Pubblicazione: (2025)
di: Costa, Rimom
Pubblicazione: (2025)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
di: Wang, Xiaohua, et al.
Pubblicazione: (2026)
di: Wang, Xiaohua, et al.
Pubblicazione: (2026)
Latent Cache Flow: Model-to-Model Communication Without Text
di: Rossi, Maximillian, et al.
Pubblicazione: (2026)
di: Rossi, Maximillian, et al.
Pubblicazione: (2026)
Automated structural testing of LLM-based agents: methods, framework, and case studies
di: Kohl, Jens, et al.
Pubblicazione: (2026)
di: Kohl, Jens, et al.
Pubblicazione: (2026)
Semantic Commit: Helping Users Update Intent Specifications for AI Memory at Scale
di: Vaithilingam, Priyan, et al.
Pubblicazione: (2025)
di: Vaithilingam, Priyan, et al.
Pubblicazione: (2025)
Who Gets the Kidney? Human-AI Alignment, Indecision, and Moral Values
di: Dickerson, John P., et al.
Pubblicazione: (2025)
di: Dickerson, John P., et al.
Pubblicazione: (2025)
Eliciting Problem Specifications via Large Language Models
di: Wray, Robert E., et al.
Pubblicazione: (2024)
di: Wray, Robert E., et al.
Pubblicazione: (2024)
Contrastive Learning-Enhanced Large Language Models for Monolith-to-Microservice Decomposition
di: Sellami, Khaled, et al.
Pubblicazione: (2025)
di: Sellami, Khaled, et al.
Pubblicazione: (2025)
An Explainable Collaborative Dialogue System using a Theory of Mind
di: Cohen, Philip R., et al.
Pubblicazione: (2023)
di: Cohen, Philip R., et al.
Pubblicazione: (2023)
A Scalable Communication Protocol for Networks of Large Language Models
di: Marro, Samuele, et al.
Pubblicazione: (2024)
di: Marro, Samuele, et al.
Pubblicazione: (2024)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
di: Ray, Aninda
Pubblicazione: (2026)
di: Ray, Aninda
Pubblicazione: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
di: Annapureddy, Sasank
Pubblicazione: (2026)
di: Annapureddy, Sasank
Pubblicazione: (2026)
Advancing Transformer Architecture in Long-Context Large Language Models: A Comprehensive Survey
di: Huang, Yunpeng, et al.
Pubblicazione: (2023)
di: Huang, Yunpeng, et al.
Pubblicazione: (2023)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
di: Zhang, Ke, et al.
Pubblicazione: (2025)
di: Zhang, Ke, et al.
Pubblicazione: (2025)
Towards Resource-Efficient Multimodal Intelligence: Learned Routing among Specialized Expert Models
di: Saini, Mayank, et al.
Pubblicazione: (2025)
di: Saini, Mayank, et al.
Pubblicazione: (2025)
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
di: Qi, Jinhu, et al.
Pubblicazione: (2026)
di: Qi, Jinhu, et al.
Pubblicazione: (2026)
Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models
di: Gokdemir, Ozan, et al.
Pubblicazione: (2025)
di: Gokdemir, Ozan, et al.
Pubblicazione: (2025)
MMiC: Mitigating Modality Incompleteness in Clustered Federated Learning
di: Yang, Lishan, et al.
Pubblicazione: (2025)
di: Yang, Lishan, et al.
Pubblicazione: (2025)
Reinforcement Learning for Scalable and Trustworthy Intelligent Systems
di: Lan, Guangchen
Pubblicazione: (2026)
di: Lan, Guangchen
Pubblicazione: (2026)
VeriPlan: Integrating Formal Verification and LLMs into End-User Planning
di: Lee, Christine, et al.
Pubblicazione: (2025)
di: Lee, Christine, et al.
Pubblicazione: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
di: Wu, Shuai, et al.
Pubblicazione: (2026)
di: Wu, Shuai, et al.
Pubblicazione: (2026)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
di: Figueiredo, Vanessa
Pubblicazione: (2025)
di: Figueiredo, Vanessa
Pubblicazione: (2025)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
di: Jehu-Appiah, Rodney
Pubblicazione: (2026)
di: Jehu-Appiah, Rodney
Pubblicazione: (2026)
RefiningGPT: Specialized language Models for Automated Refinery Unit-level Process Diagram Synthesis
di: Liu, Dongxiao, et al.
Pubblicazione: (2026)
di: Liu, Dongxiao, et al.
Pubblicazione: (2026)
FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints
di: Yang, Lishan, et al.
Pubblicazione: (2025)
di: Yang, Lishan, et al.
Pubblicazione: (2025)
CRAwDAD: Causal Reasoning Augmentation with Dual-Agent Debate
di: Vamosi, Finn G., et al.
Pubblicazione: (2025)
di: Vamosi, Finn G., et al.
Pubblicazione: (2025)
CPEMH: An Agentic Framework for Prompt-Driven Behavior Evaluation and Assurance in Foundation-Model Systems for Mental Health Screening
di: Lorenzoni, Giuliano, et al.
Pubblicazione: (2026)
di: Lorenzoni, Giuliano, et al.
Pubblicazione: (2026)
GSAR: Typed Grounding for Hallucination Detection and Recovery in Multi-Agent LLMs
di: Kamelhar, Federico A.
Pubblicazione: (2026)
di: Kamelhar, Federico A.
Pubblicazione: (2026)
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
di: Kohl, Jens, et al.
Pubblicazione: (2024)
di: Kohl, Jens, et al.
Pubblicazione: (2024)
Agent WARPP: Workflow Adherence via Runtime Parallel Personalization
di: Mazzolenis, Maria Emilia, et al.
Pubblicazione: (2025)
di: Mazzolenis, Maria Emilia, et al.
Pubblicazione: (2025)
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
di: Ueda, Keisuke, et al.
Pubblicazione: (2025)
di: Ueda, Keisuke, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
di: Wang, Yuchen, et al.
Pubblicazione: (2026) -
Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce
di: Resh, William G., et al.
Pubblicazione: (2025) -
Applying Cognitive Design Patterns to General LLM Agents
di: Wray, Robert E., et al.
Pubblicazione: (2025) -
Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay
di: Wang, Xiaohua, et al.
Pubblicazione: (2026) -
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
di: Chang, Jiale, et al.
Pubblicazione: (2026)