Dr. Kernel: Reinforcement Learning Done Right for Triton Kernel Generations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Wei, Xu, Jiawei, Li, Yingru, Zheng, Longtao, Li, Tianjian, Liu, Qian, He, Junxian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Geak: Introducing Triton Kernel AI Agent & Evaluation Benchmarks
von: Wang, Jianghui, et al.
Veröffentlicht: (2025)
von: Wang, Jianghui, et al.
Veröffentlicht: (2025)
Liger Kernel: Efficient Triton Kernels for LLM Training
von: Hsu, Pin-Lun, et al.
Veröffentlicht: (2024)
von: Hsu, Pin-Lun, et al.
Veröffentlicht: (2024)
A Note on Hybrid Online Reinforcement and Imitation Learning for LLMs: Formulations and Algorithms
von: Li, Yingru, et al.
Veröffentlicht: (2025)
von: Li, Yingru, et al.
Veröffentlicht: (2025)
AgentKernelArena: Generalization-Aware Benchmarking of GPU Kernel Optimization Agents
von: Younesian, Sharareh, et al.
Veröffentlicht: (2026)
von: Younesian, Sharareh, et al.
Veröffentlicht: (2026)
SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
von: Zeng, Weihao, et al.
Veröffentlicht: (2025)
von: Zeng, Weihao, et al.
Veröffentlicht: (2025)
The Anatomy of a Triton Attention Kernel
von: Ringlein, Burkhard, et al.
Veröffentlicht: (2025)
von: Ringlein, Burkhard, et al.
Veröffentlicht: (2025)
FastKernels: Benchmarking GPU Kernel Generation in Production
von: Oliaro, Gabriele, et al.
Veröffentlicht: (2026)
von: Oliaro, Gabriele, et al.
Veröffentlicht: (2026)
Trust Region Masking for Long-Horizon LLM Reinforcement Learning
von: Li, Yingru, et al.
Veröffentlicht: (2025)
von: Li, Yingru, et al.
Veröffentlicht: (2025)
Dr. MAS: Stable Reinforcement Learning for Multi-Agent LLM Systems
von: Feng, Lang, et al.
Veröffentlicht: (2026)
von: Feng, Lang, et al.
Veröffentlicht: (2026)
True Knowledge Comes from Practice: Aligning LLMs with Embodied Environments via Reinforcement Learning
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
Hybrid Linear Attention Done Right: Efficient Distillation and Effective Architectures for Extremely Long Contexts
von: Chen, Yingfa, et al.
Veröffentlicht: (2026)
von: Chen, Yingfa, et al.
Veröffentlicht: (2026)
DRTriton: Large-Scale Synthetic Data Driven Reinforcement Learning for Triton Kernel Generation
von: Guo, Siqi, et al.
Veröffentlicht: (2026)
von: Guo, Siqi, et al.
Veröffentlicht: (2026)
AutoTriton: Automatic Triton Programming with Reinforcement Learning in LLMs
von: Li, Shangzhan, et al.
Veröffentlicht: (2025)
von: Li, Shangzhan, et al.
Veröffentlicht: (2025)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
von: He, Bingxiang, et al.
Veröffentlicht: (2024)
von: He, Bingxiang, et al.
Veröffentlicht: (2024)
The Optimal Token Baseline: Variance Reduction for Long-Horizon LLM-RL
von: Li, Yingru, et al.
Veröffentlicht: (2026)
von: Li, Yingru, et al.
Veröffentlicht: (2026)
Dr Genre: Reinforcement Learning from Decoupled LLM Feedback for Generic Text Rewriting
von: Li, Yufei, et al.
Veröffentlicht: (2025)
von: Li, Yufei, et al.
Veröffentlicht: (2025)
AutoEval Done Right: Using Synthetic Data for Model Evaluation
von: Boyeau, Pierre, et al.
Veröffentlicht: (2024)
von: Boyeau, Pierre, et al.
Veröffentlicht: (2024)
AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation
von: Du, Weihua, et al.
Veröffentlicht: (2026)
von: Du, Weihua, et al.
Veröffentlicht: (2026)
Beyond Precision: Training-Inference Mismatch is an Optimization Problem and Simple LR Scheduling Fixes It
von: Zhang, Yaxiang, et al.
Veröffentlicht: (2026)
von: Zhang, Yaxiang, et al.
Veröffentlicht: (2026)
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning
von: Liu, Wei, et al.
Veröffentlicht: (2023)
von: Liu, Wei, et al.
Veröffentlicht: (2023)
KITE: Kernelized and Information Theoretic Exemplars for In-Context Learning
von: Singh, Vaibhav, et al.
Veröffentlicht: (2025)
von: Singh, Vaibhav, et al.
Veröffentlicht: (2025)
Kernel Density Bayesian Inverse Reinforcement Learning
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
How Contaminated Is Your Benchmark? Quantifying Dataset Leakage in Large Language Models with Kernel Divergence
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2025)
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2025)
Taming Extreme Tokens: Covariance-Aware GRPO with Gaussian-Kernel Advantage Reweighting
von: Wang, Cheng, et al.
Veröffentlicht: (2026)
von: Wang, Cheng, et al.
Veröffentlicht: (2026)
Learn to Reason Efficiently with Adaptive Length-based Reward Shaping
von: Liu, Wei, et al.
Veröffentlicht: (2025)
von: Liu, Wei, et al.
Veröffentlicht: (2025)
NOVER: Incentive Training for Language Models via Verifier-Free Reinforcement Learning
von: Liu, Wei, et al.
Veröffentlicht: (2025)
von: Liu, Wei, et al.
Veröffentlicht: (2025)
Randomized Antipodal Search Done Right for Data Pareto Improvement of LLM Unlearning
von: Liu, Ziwen, et al.
Veröffentlicht: (2026)
von: Liu, Ziwen, et al.
Veröffentlicht: (2026)
Understanding Emergent In-Context Learning from a Kernel Regression Perspective
von: Han, Chi, et al.
Veröffentlicht: (2023)
von: Han, Chi, et al.
Veröffentlicht: (2023)
Logit Dynamics in Softmax Policy Gradient Methods
von: Li, Yingru
Veröffentlicht: (2025)
von: Li, Yingru
Veröffentlicht: (2025)
Elastic Weight Consolidation Done Right for Continual Learning
von: Liu, Xuan, et al.
Veröffentlicht: (2026)
von: Liu, Xuan, et al.
Veröffentlicht: (2026)
Efficient Estimation of Kernel Surrogate Models for Task Attribution
von: Zhang, Zhenshuo, et al.
Veröffentlicht: (2026)
von: Zhang, Zhenshuo, et al.
Veröffentlicht: (2026)
DISCO Balances the Scales: Adaptive Domain- and Difficulty-Aware Reinforcement Learning on Imbalanced Data
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2025)
Q-Star Meets Scalable Posterior Sampling: Bridging Theory and Practice via HyperAgent
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Feature Structure Distillation with Centered Kernel Alignment in BERT Transferring
von: Jung, Hee-Jun, et al.
Veröffentlicht: (2022)
von: Jung, Hee-Jun, et al.
Veröffentlicht: (2022)
Towards Cold-Start Drafting and Continual Refining: A Value-Driven Memory Approach with Application to NPU Kernel Synthesis
von: Zheng, Yujie, et al.
Veröffentlicht: (2026)
von: Zheng, Yujie, et al.
Veröffentlicht: (2026)
Reinforcement Learning on Pre-Training Data
von: Li, Siheng, et al.
Veröffentlicht: (2025)
von: Li, Siheng, et al.
Veröffentlicht: (2025)
Kernelized Reinforcement Learning with Order Optimal Regret Bounds
von: Vakili, Sattar, et al.
Veröffentlicht: (2023)
von: Vakili, Sattar, et al.
Veröffentlicht: (2023)
Debiasing Kernel-Based Generative Models
von: Qin, Tian, et al.
Veröffentlicht: (2025)
von: Qin, Tian, et al.
Veröffentlicht: (2025)
DPO Kernels: A Semantically-Aware, Kernel-Enhanced, and Divergence-Rich Paradigm for Direct Preference Optimization
von: Das, Amitava, et al.
Veröffentlicht: (2025)
von: Das, Amitava, et al.
Veröffentlicht: (2025)
HISR: Hindsight Information Modulated Segmental Process Rewards For Multi-turn Agentic Reinforcement Learning
von: Lu, Zhicong, et al.
Veröffentlicht: (2026)
von: Lu, Zhicong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Geak: Introducing Triton Kernel AI Agent & Evaluation Benchmarks
von: Wang, Jianghui, et al.
Veröffentlicht: (2025) -
Liger Kernel: Efficient Triton Kernels for LLM Training
von: Hsu, Pin-Lun, et al.
Veröffentlicht: (2024) -
A Note on Hybrid Online Reinforcement and Imitation Learning for LLMs: Formulations and Algorithms
von: Li, Yingru, et al.
Veröffentlicht: (2025) -
AgentKernelArena: Generalization-Aware Benchmarking of GPU Kernel Optimization Agents
von: Younesian, Sharareh, et al.
Veröffentlicht: (2026) -
SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
von: Zeng, Weihao, et al.
Veröffentlicht: (2025)