From Reasoning to Code: GRPO Optimization for Underrepresented Languages
Fuente:
arXiv
Saved in:
| Main Authors: | Pennino, Federico, Raimondi, Bianca, Rondelli, Massimo, Gurioli, Andrea, Gabbrielli, Maurizio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoSE: Hierarchical Self-Distillation Enhances Early Layer Embeddings
by: Gurioli, Andrea, et al.
Published: (2025)
by: Gurioli, Andrea, et al.
Published: (2025)
Optimizing the Training Diet: Data Mixture Search for Robust Time Series Forecasting
by: Pennino, Federico, et al.
Published: (2025)
by: Pennino, Federico, et al.
Published: (2025)
Do not copy and paste! Rewriting strategies for code retrieval
by: Gurioli, Andrea, et al.
Published: (2026)
by: Gurioli, Andrea, et al.
Published: (2026)
Exploiting Primacy Effect To Improve Large Language Models
by: Raimondi, Bianca, et al.
Published: (2025)
by: Raimondi, Bianca, et al.
Published: (2025)
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy
by: Raimondi, Bianca, et al.
Published: (2026)
by: Raimondi, Bianca, et al.
Published: (2026)
BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis
by: Rondelli, Massimo, et al.
Published: (2026)
by: Rondelli, Massimo, et al.
Published: (2026)
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets
by: Gurioli, Andrea, et al.
Published: (2026)
by: Gurioli, Andrea, et al.
Published: (2026)
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability
by: Raimondi, Bianca, et al.
Published: (2025)
by: Raimondi, Bianca, et al.
Published: (2025)
Affordably Fine-tuned LLMs Provide Better Answers to Course-specific MCQs
by: Raimondi, Bianca, et al.
Published: (2025)
by: Raimondi, Bianca, et al.
Published: (2025)
From Large to Small: Transferring CUDA Optimization Expertise via Reasoning Graph
by: Gong, Junfeng, et al.
Published: (2025)
by: Gong, Junfeng, et al.
Published: (2025)
Curriculum Learning for Small Code Language Models
by: Naïr, Marwa, et al.
Published: (2024)
by: Naïr, Marwa, et al.
Published: (2024)
Stepwise Guided Policy Optimization: Coloring your Incorrect Reasoning in GRPO
by: Chen, Peter, et al.
Published: (2025)
by: Chen, Peter, et al.
Published: (2025)
SymRTLO: Enhancing RTL Code Optimization with LLMs and Neuron-Inspired Symbolic Reasoning
by: Wang, Yiting, et al.
Published: (2025)
by: Wang, Yiting, et al.
Published: (2025)
Learning Factors in AI-Augmented Education: A Comparative Study of Middle and High School Students
by: Ebli, Gaia, et al.
Published: (2025)
by: Ebli, Gaia, et al.
Published: (2025)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Code Simulation Challenges for Large Language Models
by: La Malfa, Emanuele, et al.
Published: (2024)
by: La Malfa, Emanuele, et al.
Published: (2024)
Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoning
by: Zhang, Xichen, et al.
Published: (2025)
by: Zhang, Xichen, et al.
Published: (2025)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
by: Barone, Antonio Valerio Miceli, et al.
Published: (2026)
by: Barone, Antonio Valerio Miceli, et al.
Published: (2026)
Chain of Execution Supervision Promotes General Reasoning in Large Language Models
by: Chen, Nuo, et al.
Published: (2025)
by: Chen, Nuo, et al.
Published: (2025)
PerfRL: A Small Language Model Framework for Efficient Code Optimization
by: Duan, Shukai, et al.
Published: (2023)
by: Duan, Shukai, et al.
Published: (2023)
ExGRPO: Learning to Reason from Experience
by: Zhan, Runzhe, et al.
Published: (2025)
by: Zhan, Runzhe, et al.
Published: (2025)
LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations
by: Tang, Annabelle Sujun, et al.
Published: (2026)
by: Tang, Annabelle Sujun, et al.
Published: (2026)
ECCO: Evidence-Driven Causal Reasoning for Compiler Optimization
by: Pan, Haolin, et al.
Published: (2026)
by: Pan, Haolin, et al.
Published: (2026)
Bridging the Knowledge Void: Inference-time Acquisition of Unfamiliar Programming Languages for Coding Tasks
by: Shen, Chen, et al.
Published: (2026)
by: Shen, Chen, et al.
Published: (2026)
ProofOptimizer: Training Language Models to Simplify Proofs without Human Demonstrations
by: Gu, Alex, et al.
Published: (2025)
by: Gu, Alex, et al.
Published: (2025)
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
by: Cassano, Federico, et al.
Published: (2023)
by: Cassano, Federico, et al.
Published: (2023)
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
by: Ramesh, Shyam Sundhar, et al.
Published: (2026)
by: Ramesh, Shyam Sundhar, et al.
Published: (2026)
Large Language Models for Code Summarization
by: Szalontai, Balázs, et al.
Published: (2024)
by: Szalontai, Balázs, et al.
Published: (2024)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
by: Kadosh, Tal, et al.
Published: (2023)
by: Kadosh, Tal, et al.
Published: (2023)
Autocomp: A Powerful and Portable Code Optimizer for Tensor Accelerators
by: Hong, Charles, et al.
Published: (2025)
by: Hong, Charles, et al.
Published: (2025)
Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility
by: Maveli, Nickil, et al.
Published: (2026)
by: Maveli, Nickil, et al.
Published: (2026)
GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO
by: Dipta, Shubhashis Roy, et al.
Published: (2026)
by: Dipta, Shubhashis Roy, et al.
Published: (2026)
Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms
by: Zenkner, Janis, et al.
Published: (2025)
by: Zenkner, Janis, et al.
Published: (2025)
LILO: Learning Interpretable Libraries by Compressing and Documenting Code
by: Grand, Gabriel, et al.
Published: (2023)
by: Grand, Gabriel, et al.
Published: (2023)
Program Synthesis using Inductive Logic Programming for the Abstraction and Reasoning Corpus
by: Rocha, Filipe Marinho, et al.
Published: (2024)
by: Rocha, Filipe Marinho, et al.
Published: (2024)
ANCORA: Learning to Question via Manifold-Anchored Self-Play for Verifiable Reasoning
by: Yang, Chengcao
Published: (2026)
by: Yang, Chengcao
Published: (2026)
A Fast, Reliable, and Secure Programming Language for LLM Agents with Code Actions
by: Mell, Stephen, et al.
Published: (2025)
by: Mell, Stephen, et al.
Published: (2025)
A Multi-Expert Large Language Model Architecture for Verilog Code Generation
by: Nadimi, Bardia, et al.
Published: (2024)
by: Nadimi, Bardia, et al.
Published: (2024)
AutoPDL: Automatic Prompt Optimization for LLM Agents
by: Spiess, Claudio, et al.
Published: (2025)
by: Spiess, Claudio, et al.
Published: (2025)
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
by: Pourreza, Mohammadreza, et al.
Published: (2025)
by: Pourreza, Mohammadreza, et al.
Published: (2025)
Similar Items
-
MoSE: Hierarchical Self-Distillation Enhances Early Layer Embeddings
by: Gurioli, Andrea, et al.
Published: (2025) -
Optimizing the Training Diet: Data Mixture Search for Robust Time Series Forecasting
by: Pennino, Federico, et al.
Published: (2025) -
Do not copy and paste! Rewriting strategies for code retrieval
by: Gurioli, Andrea, et al.
Published: (2026) -
Exploiting Primacy Effect To Improve Large Language Models
by: Raimondi, Bianca, et al.
Published: (2025) -
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy
by: Raimondi, Bianca, et al.
Published: (2026)