$\textbf{Only-IF}$:Revealing the Decisive Effect of Instruction Diversity on Generalization
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Dylan, Wang, Justin, Charton, Francois |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Instruction Diversity Drives Generalization To Unseen Tasks
par: Zhang, Dylan, et autres
Publié: (2024)
par: Zhang, Dylan, et autres
Publié: (2024)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
par: Zhang, Dylan, et autres
Publié: (2024)
par: Zhang, Dylan, et autres
Publié: (2024)
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
par: Liu, Zuxin, et autres
Publié: (2024)
par: Liu, Zuxin, et autres
Publié: (2024)
From Symbolic Tasks to Code Generation: Diversification Yields Better Task Performers
par: Zhang, Dylan, et autres
Publié: (2024)
par: Zhang, Dylan, et autres
Publié: (2024)
Learning to Generate Unit Tests for Automated Debugging
par: Prasad, Archiki, et autres
Publié: (2025)
par: Prasad, Archiki, et autres
Publié: (2025)
How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data
par: Wang, Yejie, et autres
Publié: (2024)
par: Wang, Yejie, et autres
Publié: (2024)
Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents
par: Zhang, Kexun, et autres
Publié: (2024)
par: Zhang, Kexun, et autres
Publié: (2024)
XFT: Unlocking the Power of Code Instruction Tuning by Simply Merging Upcycled Mixture-of-Experts
par: Ding, Yifeng, et autres
Publié: (2024)
par: Ding, Yifeng, et autres
Publié: (2024)
The Art of Repair: Optimizing Iterative Program Repair with Instruction-Tuned Models
par: Ruiz, Fernando Vallecillos, et autres
Publié: (2025)
par: Ruiz, Fernando Vallecillos, et autres
Publié: (2025)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
par: Ni, Xinyi, et autres
Publié: (2025)
par: Ni, Xinyi, et autres
Publié: (2025)
Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation
par: Yu, Zhuohao, et autres
Publié: (2024)
par: Yu, Zhuohao, et autres
Publié: (2024)
Selective Prompt Anchoring for Code Generation
par: Tian, Yuan, et autres
Publié: (2024)
par: Tian, Yuan, et autres
Publié: (2024)
A Survey on Code Generation with LLM-based Agents
par: Dong, Yihong, et autres
Publié: (2025)
par: Dong, Yihong, et autres
Publié: (2025)
Automata-Based Steering of Large Language Models for Diverse Structured Generation
par: Luan, Xiaokun, et autres
Publié: (2025)
par: Luan, Xiaokun, et autres
Publié: (2025)
AFlow: Automating Agentic Workflow Generation
par: Zhang, Jiayi, et autres
Publié: (2024)
par: Zhang, Jiayi, et autres
Publié: (2024)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
par: Fan, Lishui, et autres
Publié: (2025)
par: Fan, Lishui, et autres
Publié: (2025)
Solution-oriented Agent-based Models Generation with Verifier-assisted Iterative In-context Learning
par: Niu, Tong, et autres
Publié: (2024)
par: Niu, Tong, et autres
Publié: (2024)
DeepCRCEval: Revisiting the Evaluation of Code Review Comment Generation
par: Lu, Junyi, et autres
Publié: (2024)
par: Lu, Junyi, et autres
Publié: (2024)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
par: Wang, Xinchen, et autres
Publié: (2025)
par: Wang, Xinchen, et autres
Publié: (2025)
AlgoTune: Can Language Models Speed Up General-Purpose Numerical Programs?
par: Press, Ori, et autres
Publié: (2025)
par: Press, Ori, et autres
Publié: (2025)
CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion
par: Ren, Qibing, et autres
Publié: (2024)
par: Ren, Qibing, et autres
Publié: (2024)
SWE-Bench++: A Framework for the Scalable Generation of Software Engineering Benchmarks from Open-Source Repositories
par: Wang, Lilin, et autres
Publié: (2025)
par: Wang, Lilin, et autres
Publié: (2025)
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
par: Zhuo, Terry Yue, et autres
Publié: (2024)
par: Zhuo, Terry Yue, et autres
Publié: (2024)
Rethinking Repetition Problems of LLMs in Code Generation
par: Dong, Yihong, et autres
Publié: (2025)
par: Dong, Yihong, et autres
Publié: (2025)
Improving Code Generation by Training with Natural Language Feedback
par: Chen, Angelica, et autres
Publié: (2023)
par: Chen, Angelica, et autres
Publié: (2023)
An Approach for Auto Generation of Labeling Functions for Software Engineering Chatbots
par: Alor, Ebube, et autres
Publié: (2024)
par: Alor, Ebube, et autres
Publié: (2024)
The Larger the Better? Improved LLM Code-Generation via Budget Reallocation
par: Hassid, Michael, et autres
Publié: (2024)
par: Hassid, Michael, et autres
Publié: (2024)
DDPT: Diffusion-Driven Prompt Tuning for Large Language Model Code Generation
par: Li, Jinyang, et autres
Publié: (2025)
par: Li, Jinyang, et autres
Publié: (2025)
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
par: Yang, Dayu, et autres
Publié: (2025)
par: Yang, Dayu, et autres
Publié: (2025)
Bridging Online and Offline RL: Contextual Bandit Learning for Multi-Turn Code Generation
par: Chen, Ziru, et autres
Publié: (2026)
par: Chen, Ziru, et autres
Publié: (2026)
Automated Unity Game Template Generation from GDDs via NLP and Multi-Modal LLMs
par: Hassan, Amna
Publié: (2025)
par: Hassan, Amna
Publié: (2025)
mcdok at SemEval-2026 Task 13: Finetuning LLMs for Detection of Machine-Generated Code
par: Skurla, Adam, et autres
Publié: (2026)
par: Skurla, Adam, et autres
Publié: (2026)
Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?
par: Xia, Chunqiu Steven, et autres
Publié: (2025)
par: Xia, Chunqiu Steven, et autres
Publié: (2025)
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
par: Wei, Yuxiang, et autres
Publié: (2025)
par: Wei, Yuxiang, et autres
Publié: (2025)
Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code
par: Song, Zhenghan, et autres
Publié: (2026)
par: Song, Zhenghan, et autres
Publié: (2026)
Let the Code LLM Edit Itself When You Edit the Code
par: He, Zhenyu, et autres
Publié: (2024)
par: He, Zhenyu, et autres
Publié: (2024)
KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?
par: Jiang, Xue, et autres
Publié: (2026)
par: Jiang, Xue, et autres
Publié: (2026)
From I/O to Code with Discovery Agent
par: Dong, Yihong, et autres
Publié: (2026)
par: Dong, Yihong, et autres
Publié: (2026)
CODEMENV: Benchmarking Large Language Models on Code Migration
par: Cheng, Keyuan, et autres
Publié: (2025)
par: Cheng, Keyuan, et autres
Publié: (2025)
Confucius Code Agent: Scalable Agent Scaffolding for Real-World Codebases
par: Wong, Sherman, et autres
Publié: (2025)
par: Wong, Sherman, et autres
Publié: (2025)
Documents similaires
-
Instruction Diversity Drives Generalization To Unseen Tasks
par: Zhang, Dylan, et autres
Publié: (2024) -
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
par: Zhang, Dylan, et autres
Publié: (2024) -
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
par: Liu, Zuxin, et autres
Publié: (2024) -
From Symbolic Tasks to Code Generation: Diversification Yields Better Task Performers
par: Zhang, Dylan, et autres
Publié: (2024) -
Learning to Generate Unit Tests for Automated Debugging
par: Prasad, Archiki, et autres
Publié: (2025)