Exploring Natural Language-Based Strategies for Efficient Number Learning in Children through Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autor principal: | Mittra, Tirthankar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
por: Sarkar, Bidipta, et al.
Publicado: (2025)
por: Sarkar, Bidipta, et al.
Publicado: (2025)
Multi-Objective Reinforcement Learning for Large Language Model Optimization: Visionary Perspective
por: Kong, Lingxiao, et al.
Publicado: (2025)
por: Kong, Lingxiao, et al.
Publicado: (2025)
ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
por: Wan, Ziyu, et al.
Publicado: (2025)
por: Wan, Ziyu, et al.
Publicado: (2025)
Chain of Uncertain Rewards with Large Language Models for Reinforcement Learning
por: Mo, Shentong
Publicado: (2026)
por: Mo, Shentong
Publicado: (2026)
Communicating Sound Through Natural Language
por: Rossi, Emanuele, et al.
Publicado: (2026)
por: Rossi, Emanuele, et al.
Publicado: (2026)
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality
por: Ren, Baochang, et al.
Publicado: (2025)
por: Ren, Baochang, et al.
Publicado: (2025)
LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions
por: Sun, Chuanneng, et al.
Publicado: (2024)
por: Sun, Chuanneng, et al.
Publicado: (2024)
EPM-RL: Reinforcement Learning for On-Premise Product Mapping in E-Commerce
por: Yu, Minhyeong, et al.
Publicado: (2026)
por: Yu, Minhyeong, et al.
Publicado: (2026)
SPIO: Ensemble and Selective Strategies via LLM-Based Multi-Agent Planning in Automated Data Science
por: Seo, Wonduk, et al.
Publicado: (2025)
por: Seo, Wonduk, et al.
Publicado: (2025)
DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration
por: Nourzad, Narjes, et al.
Publicado: (2025)
por: Nourzad, Narjes, et al.
Publicado: (2025)
Safe Multi-agent Reinforcement Learning with Natural Language Constraints
por: Wang, Ziyan, et al.
Publicado: (2024)
por: Wang, Ziyan, et al.
Publicado: (2024)
Composite Learning Units: Generalized Learning Beyond Parameter Updates to Transform LLMs into Adaptive Reasoners
por: Radha, Santosh Kumar, et al.
Publicado: (2024)
por: Radha, Santosh Kumar, et al.
Publicado: (2024)
MAC: Multi-Agent Constitution Learning
por: Thareja, Rushil, et al.
Publicado: (2026)
por: Thareja, Rushil, et al.
Publicado: (2026)
Memp: Exploring Agent Procedural Memory
por: Fang, Runnan, et al.
Publicado: (2025)
por: Fang, Runnan, et al.
Publicado: (2025)
Can We Predict Before Executing Machine Learning Agents?
por: Zheng, Jingsheng, et al.
Publicado: (2026)
por: Zheng, Jingsheng, et al.
Publicado: (2026)
Towards Emotionally Intelligent and Responsible Reinforcement Learning
por: Keerthana, Garapati, et al.
Publicado: (2025)
por: Keerthana, Garapati, et al.
Publicado: (2025)
MLZero: A Multi-Agent System for End-to-end Machine Learning Automation
por: Fang, Haoyang, et al.
Publicado: (2025)
por: Fang, Haoyang, et al.
Publicado: (2025)
LENS: Learning Ensemble Confidence from Neural States for Multi-LLM Answer Integration
por: Guo, Jizhou
Publicado: (2025)
por: Guo, Jizhou
Publicado: (2025)
Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings
por: Taylor, Russell, et al.
Publicado: (2025)
por: Taylor, Russell, et al.
Publicado: (2025)
MinionsLLM: a Task-adaptive Framework For The Training and Control of Multi-Agent Systems Through Natural Language
por: Rincon, Andres Garcia, et al.
Publicado: (2025)
por: Rincon, Andres Garcia, et al.
Publicado: (2025)
Unleashing Diverse Thinking Modes in LLMs through Multi-Agent Collaboration
por: He, Zhixuan, et al.
Publicado: (2025)
por: He, Zhixuan, et al.
Publicado: (2025)
Language Agents as Optimizable Graphs
por: Zhuge, Mingchen, et al.
Publicado: (2024)
por: Zhuge, Mingchen, et al.
Publicado: (2024)
Representation Learning For Efficient Deep Multi-Agent Reinforcement Learning
por: Huh, Dom, et al.
Publicado: (2024)
por: Huh, Dom, et al.
Publicado: (2024)
Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning
por: Yalcinkaya, Beyazit, et al.
Publicado: (2025)
por: Yalcinkaya, Beyazit, et al.
Publicado: (2025)
ToolOrchestra: Elevating Intelligence via Efficient Model and Tool Orchestration
por: Su, Hongjin, et al.
Publicado: (2025)
por: Su, Hongjin, et al.
Publicado: (2025)
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
por: Shapira, Eilam, et al.
Publicado: (2026)
por: Shapira, Eilam, et al.
Publicado: (2026)
Efficient Multi-agent Reinforcement Learning by Planning
por: Liu, Qihan, et al.
Publicado: (2024)
por: Liu, Qihan, et al.
Publicado: (2024)
In-Context Environments Induce Evaluation-Awareness in Language Models
por: Chaudhary, Maheep
Publicado: (2026)
por: Chaudhary, Maheep
Publicado: (2026)
Breaking the Performance Ceiling in Reinforcement Learning requires Inference Strategies
por: Chalumeau, Felix, et al.
Publicado: (2025)
por: Chalumeau, Felix, et al.
Publicado: (2025)
Language-Guided Tuning: Enhancing Numeric Optimization with Textual Feedback
por: Lu, Yuxing, et al.
Publicado: (2025)
por: Lu, Yuxing, et al.
Publicado: (2025)
Iteration of Thought: Leveraging Inner Dialogue for Autonomous Large Language Model Reasoning
por: Radha, Santosh Kumar, et al.
Publicado: (2024)
por: Radha, Santosh Kumar, et al.
Publicado: (2024)
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
por: Lau, Gregory Kang Ruey, et al.
Publicado: (2024)
por: Lau, Gregory Kang Ruey, et al.
Publicado: (2024)
Doctorina MedBench: End-to-End Evaluation of Agent-Based Medical AI
por: Kozlova, Anna, et al.
Publicado: (2026)
por: Kozlova, Anna, et al.
Publicado: (2026)
Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models
por: Bozdag, Nimet Beyza, et al.
Publicado: (2025)
por: Bozdag, Nimet Beyza, et al.
Publicado: (2025)
Mnemosyne: An Unsupervised, Human-Inspired Long-Term Memory Architecture for Edge-Based LLMs
por: Jonelagadda, Aneesh, et al.
Publicado: (2025)
por: Jonelagadda, Aneesh, et al.
Publicado: (2025)
Learn the Ropes, Then Trust the Wins: Self-imitation with Progressive Exploration for Agentic Reinforcement Learning
por: Qin, Yulei, et al.
Publicado: (2025)
por: Qin, Yulei, et al.
Publicado: (2025)
Imagine, Initialize, and Explore: An Effective Exploration Method in Multi-Agent Reinforcement Learning
por: Liu, Zeyang, et al.
Publicado: (2024)
por: Liu, Zeyang, et al.
Publicado: (2024)
Moira: Language-driven Hierarchical Reinforcement Learning for Pair Trading
por: Giannouris, Polydoros, et al.
Publicado: (2026)
por: Giannouris, Polydoros, et al.
Publicado: (2026)
Large Language Model-Based Reward Design for Deep Reinforcement Learning-Driven Autonomous Cyber Defense
por: Mukherjee, Sayak, et al.
Publicado: (2025)
por: Mukherjee, Sayak, et al.
Publicado: (2025)
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
por: Xu, Zelai, et al.
Publicado: (2023)
por: Xu, Zelai, et al.
Publicado: (2023)
Ejemplares similares
-
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
por: Sarkar, Bidipta, et al.
Publicado: (2025) -
Multi-Objective Reinforcement Learning for Large Language Model Optimization: Visionary Perspective
por: Kong, Lingxiao, et al.
Publicado: (2025) -
ReMA: Learning to Meta-think for LLMs with Multi-Agent Reinforcement Learning
por: Wan, Ziyu, et al.
Publicado: (2025) -
Chain of Uncertain Rewards with Large Language Models for Reinforcement Learning
por: Mo, Shentong
Publicado: (2026) -
Communicating Sound Through Natural Language
por: Rossi, Emanuele, et al.
Publicado: (2026)