Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment
Fuente:
arXiv
Guardado en:
| Autores principales: | Boruch-Gruszecki, Aleksander, Zi, Yangtian, Wu, Zixuan, Oberoi, Tejas, Anderson, Carolyn Jane, Biswas, Joydeep, Guha, Arjun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans
por: Zi, Yangtian, et al.
Publicado: (2025)
por: Zi, Yangtian, et al.
Publicado: (2025)
ReasoningWeekly: A General Knowledge and Verbal Reasoning Challenge for Large Language Models
por: Wu, Zixuan, et al.
Publicado: (2025)
por: Wu, Zixuan, et al.
Publicado: (2025)
"I Would Have Written My Code Differently'': Beginners Struggle to Understand LLM-Generated Code
por: Zi, Yangtian, et al.
Publicado: (2025)
por: Zi, Yangtian, et al.
Publicado: (2025)
More Than a Score: Probing the Impact of Prompt Specificity on LLM Code Generation
por: Zi, Yangtian, et al.
Publicado: (2025)
por: Zi, Yangtian, et al.
Publicado: (2025)
How Beginning Programmers and Code LLMs (Mis)read Each Other
por: Nguyen, Sydney, et al.
Publicado: (2024)
por: Nguyen, Sydney, et al.
Publicado: (2024)
Creating and Repairing Robot Programs in Open-World Domains
por: Schlesinger, Claire, et al.
Publicado: (2024)
por: Schlesinger, Claire, et al.
Publicado: (2024)
Substance Beats Style: Why Beginning Students Fail to Code with LLMs
por: Lucchetti, Francesca, et al.
Publicado: (2024)
por: Lucchetti, Francesca, et al.
Publicado: (2024)
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs
por: Hu, Zichao, et al.
Publicado: (2024)
por: Hu, Zichao, et al.
Publicado: (2024)
Knowledge Transfer from High-Resource to Low-Resource Programming Languages for Code LLMs
por: Cassano, Federico, et al.
Publicado: (2023)
por: Cassano, Federico, et al.
Publicado: (2023)
GlyphPattern: An Abstract Pattern Recognition Benchmark for Vision-Language Models
por: Wu, Zixuan, et al.
Publicado: (2024)
por: Wu, Zixuan, et al.
Publicado: (2024)
Deploying and Evaluating LLMs to Program Service Mobile Robots
por: Hu, Zichao, et al.
Publicado: (2023)
por: Hu, Zichao, et al.
Publicado: (2023)
CLOVER: Context-aware Long-term Object Viewpoint- and Environment- Invariant Representation Learning
por: Lee, Dongmyeong, et al.
Publicado: (2024)
por: Lee, Dongmyeong, et al.
Publicado: (2024)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
por: Banerjee, Arko, et al.
Publicado: (2024)
por: Banerjee, Arko, et al.
Publicado: (2024)
The Role of Environment Access in Agnostic Reinforcement Learning
por: Krishnamurthy, Akshay, et al.
Publicado: (2025)
por: Krishnamurthy, Akshay, et al.
Publicado: (2025)
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
por: Lucchetti, Francesca, et al.
Publicado: (2024)
por: Lucchetti, Francesca, et al.
Publicado: (2024)
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
por: Cassano, Federico, et al.
Publicado: (2023)
por: Cassano, Federico, et al.
Publicado: (2023)
Steering Code LLMs with Activation Directions for Language and Library Control
por: Rahman, Md Mahbubur, et al.
Publicado: (2026)
por: Rahman, Md Mahbubur, et al.
Publicado: (2026)
FLsim: A Modular and Library-Agnostic Simulation Framework for Federated Learning
por: Mukherjee, Arnab, et al.
Publicado: (2025)
por: Mukherjee, Arnab, et al.
Publicado: (2025)
Learning Quantitative Automata Modulo Theories
por: Hsiung, Eric, et al.
Publicado: (2024)
por: Hsiung, Eric, et al.
Publicado: (2024)
Automata Learning from Preference and Equivalence Queries
por: Hsiung, Eric, et al.
Publicado: (2023)
por: Hsiung, Eric, et al.
Publicado: (2023)
Multi-Agent Inverse Reinforcement Learning in Real World Unstructured Pedestrian Crowds
por: Chandra, Rohan, et al.
Publicado: (2024)
por: Chandra, Rohan, et al.
Publicado: (2024)
Magnonics of time-varying media: Giant amplification via phase-transition-driven temporal interfaces
por: Sobucki, Krzysztof, et al.
Publicado: (2025)
por: Sobucki, Krzysztof, et al.
Publicado: (2025)
La colaboración campbell y la Práctica Basada en la evidencia
por: Robert F. Boruch
Publicado: (2002)
por: Robert F. Boruch
Publicado: (2002)
Planning Learning Environments for Library Media Programs: An Introduction.
por: Klasing, Jane P., et al.
Publicado: (1992)
por: Klasing, Jane P., et al.
Publicado: (1992)
Learning Model Agnostic Explanations via Constraint Programming
por: Koriche, Frederic, et al.
Publicado: (2024)
por: Koriche, Frederic, et al.
Publicado: (2024)
Learning Permutation Distributions via Reflected Diffusion on Ranks
por: He, Sizhuang, et al.
Publicado: (2026)
por: He, Sizhuang, et al.
Publicado: (2026)
CoRe-Code: Collaborative Reinforcement Learning for Code Generation
por: Dou, Zhihao, et al.
Publicado: (2026)
por: Dou, Zhihao, et al.
Publicado: (2026)
CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review
por: Kapadnis, Manav Nitin, et al.
Publicado: (2025)
por: Kapadnis, Manav Nitin, et al.
Publicado: (2025)
AnyCalib: On-Manifold Learning for Model-Agnostic Single-View Camera Calibration
por: Tirado-Garín, Javier, et al.
Publicado: (2025)
por: Tirado-Garín, Javier, et al.
Publicado: (2025)
Evaluating Computational Representations of Character: An Austen Character Similarity Benchmark
por: Yang, Funing, et al.
Publicado: (2024)
por: Yang, Funing, et al.
Publicado: (2024)
Safe Continual Reinforcement Learning in Non-stationary Environments
por: Coursey, Austin, et al.
Publicado: (2026)
por: Coursey, Austin, et al.
Publicado: (2026)
Platform-Agnostic Reinforcement Learning Framework for Safe Exploration of Cluttered Environments with Graph Attention
por: Calzolari, Gabriele, et al.
Publicado: (2025)
por: Calzolari, Gabriele, et al.
Publicado: (2025)
Tabular and Deep Reinforcement Learning for Gittins Index
por: Dhankhar, Harshit, et al.
Publicado: (2024)
por: Dhankhar, Harshit, et al.
Publicado: (2024)
Agnostic Reinforcement Learning: Foundations and Algorithms
por: Li, Gene
Publicado: (2025)
por: Li, Gene
Publicado: (2025)
Constrained Meta Agnostic Reinforcement Learning
por: Daaboul, Karam, et al.
Publicado: (2024)
por: Daaboul, Karam, et al.
Publicado: (2024)
Strategic Facility Location with Limited Liars
por: Gruszecki, Yue, et al.
Publicado: (2026)
por: Gruszecki, Yue, et al.
Publicado: (2026)
Library-University Partnerships in Distance Learning.
por: Argentati, Carolyn
Publicado: (1999)
por: Argentati, Carolyn
Publicado: (1999)
GeoContrastNet: Contrastive Key-Value Edge Learning for Language-Agnostic Document Understanding
por: Biescas, Nil, et al.
Publicado: (2024)
por: Biescas, Nil, et al.
Publicado: (2024)
A Theory of Universal Agnostic Learning
por: Hanneke, Steve, et al.
Publicado: (2026)
por: Hanneke, Steve, et al.
Publicado: (2026)
Safe Urban Traffic Control via Uncertainty-Aware Conformal Prediction and World-Model Reinforcement Learning
por: Chandra, Joydeep, et al.
Publicado: (2026)
por: Chandra, Joydeep, et al.
Publicado: (2026)
Ejemplares similares
-
AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans
por: Zi, Yangtian, et al.
Publicado: (2025) -
ReasoningWeekly: A General Knowledge and Verbal Reasoning Challenge for Large Language Models
por: Wu, Zixuan, et al.
Publicado: (2025) -
"I Would Have Written My Code Differently'': Beginners Struggle to Understand LLM-Generated Code
por: Zi, Yangtian, et al.
Publicado: (2025) -
More Than a Score: Probing the Impact of Prompt Specificity on LLM Code Generation
por: Zi, Yangtian, et al.
Publicado: (2025) -
How Beginning Programmers and Code LLMs (Mis)read Each Other
por: Nguyen, Sydney, et al.
Publicado: (2024)