Code-Optimise: Self-Generated Preference Data for Correctness and Efficiency
Fuente:
arXiv
Saved in:
| Main Authors: | Gee, Leonidas, Gritta, Milan, Lampouras, Gerasimos, Iacobacci, Ignacio |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
by: Gritta, Milan, et al.
Published: (2024)
by: Gritta, Milan, et al.
Published: (2024)
DReSD: Dense Retrieval for Speculative Decoding
by: Gritta, Milan, et al.
Published: (2025)
by: Gritta, Milan, et al.
Published: (2025)
DRIFT: Decompose, Retrieve, Illustrate, then Formalize Theorems
by: Zhang, Meiru, et al.
Published: (2025)
by: Zhang, Meiru, et al.
Published: (2025)
Mixture of Attentions For Speculative Decoding
by: Zimmer, Matthieu, et al.
Published: (2024)
by: Zimmer, Matthieu, et al.
Published: (2024)
Text-to-Code Generation with Modality-relative Pre-training
by: Christopoulou, Fenia, et al.
Published: (2024)
by: Christopoulou, Fenia, et al.
Published: (2024)
Findings of the First Workshop on Simulating Conversational Intelligence in Chat
by: Graham, Yvette, et al.
Published: (2024)
by: Graham, Yvette, et al.
Published: (2024)
TopoAlign: A Framework for Aligning Code to Math via Topological Decomposition
by: Li, Yupei, et al.
Published: (2025)
by: Li, Yupei, et al.
Published: (2025)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
by: Christopoulou, Fenia, et al.
Published: (2024)
by: Christopoulou, Fenia, et al.
Published: (2024)
Conjecturing: An Overlooked Step in Formal Mathematical Reasoning
by: Sivakumar, Jasivan Alex, et al.
Published: (2025)
by: Sivakumar, Jasivan Alex, et al.
Published: (2025)
Correct and Optimal: the Regular Expression Inference Challenge
by: Valizadeh, Mojtaba, et al.
Published: (2023)
by: Valizadeh, Mojtaba, et al.
Published: (2023)
MULAN: A Multi Layer Annotated Dataset for Controllable Text-to-Image Generation
by: Tudosiu, Petru-Daniel, et al.
Published: (2024)
by: Tudosiu, Petru-Daniel, et al.
Published: (2024)
Encoder-Decoder Framework for Interactive Free Verses with Generation with Controllable High-Quality Rhyming
by: Pasini, Tommaso, et al.
Published: (2024)
by: Pasini, Tommaso, et al.
Published: (2024)
Are Compressed Language Models Less Subgroup Robust?
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
Increasing the Thinking Budget is Not All You Need
by: Iacobacci, Ignacio, et al.
Published: (2025)
by: Iacobacci, Ignacio, et al.
Published: (2025)
Self-Correcting Code Generation Using Small Language Models
by: Cho, Jeonghun, et al.
Published: (2025)
by: Cho, Jeonghun, et al.
Published: (2025)
Multi-word Tokenization for Sequence Compression
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
Human-inspired Episodic Memory for Infinite Context LLMs
by: Fountas, Zafeirios, et al.
Published: (2024)
by: Fountas, Zafeirios, et al.
Published: (2024)
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
EffiLearner: Enhancing Efficiency of Generated Code via Self-Optimization
by: Huang, Dong, et al.
Published: (2024)
by: Huang, Dong, et al.
Published: (2024)
A Benchmark for Deep Information Synthesis
by: Paul, Debjit, et al.
Published: (2026)
by: Paul, Debjit, et al.
Published: (2026)
Fast Vocabulary Transfer for Language Model Compression
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
CodeLutra: Boosting LLM Code Generation via Preference-Guided Refinement
by: Tao, Leitian, et al.
Published: (2024)
by: Tao, Leitian, et al.
Published: (2024)
ECCO: Can We Improve Model-Generated Code Efficiency Without Sacrificing Functional Correctness?
by: Waghjale, Siddhant, et al.
Published: (2024)
by: Waghjale, Siddhant, et al.
Published: (2024)
It Helps to Take a Second Opinion: Teaching Smaller LLMs to Deliberate Mutually via Selective Rationale Optimisation
by: Patnaik, Sohan, et al.
Published: (2025)
by: Patnaik, Sohan, et al.
Published: (2025)
Dutch CrowS-Pairs: Adapting a Challenge Dataset for Measuring Social Biases in Language Models for Dutch
by: Strazda, Elza, et al.
Published: (2025)
by: Strazda, Elza, et al.
Published: (2025)
Self-Infilling Code Generation
by: Zheng, Lin, et al.
Published: (2023)
by: Zheng, Lin, et al.
Published: (2023)
Visualising Policy-Reward Interplay to Inform Zeroth-Order Preference Optimisation of Large Language Models
by: Galatolo, Alessio, et al.
Published: (2025)
by: Galatolo, Alessio, et al.
Published: (2025)
SGPO: Self-Generated Preference Optimization based on Self-Improver
by: Lee, Hyeonji, et al.
Published: (2025)
by: Lee, Hyeonji, et al.
Published: (2025)
Dynamic Noise Preference Optimization: Self-Improvement of Large Language Models with Self-Synthetic Data
by: Yang, Haoyan, et al.
Published: (2025)
by: Yang, Haoyan, et al.
Published: (2025)
ALMol: Aligned Language-Molecule Translation LLMs through Offline Preference Contrastive Optimisation
by: Gkoumas, Dimitris
Published: (2024)
by: Gkoumas, Dimitris
Published: (2024)
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
by: Pal, Arka, et al.
Published: (2024)
by: Pal, Arka, et al.
Published: (2024)
Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL
by: Zheng, Kunhao, et al.
Published: (2026)
by: Zheng, Kunhao, et al.
Published: (2026)
Who Would Chatbots Vote For? Political Preferences of ChatGPT and Gemini in the 2024 European Union Elections
by: Haman, Michael, et al.
Published: (2024)
by: Haman, Michael, et al.
Published: (2024)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
by: Shi, Yuling, et al.
Published: (2024)
by: Shi, Yuling, et al.
Published: (2024)
From FusHa to Folk: Exploring Cross-Lingual Transfer in Arabic Language Models
by: Khalak, Abdulmuizz, et al.
Published: (2026)
by: Khalak, Abdulmuizz, et al.
Published: (2026)
EffiBench: Benchmarking the Efficiency of Automatically Generated Code
by: Huang, Dong, et al.
Published: (2024)
by: Huang, Dong, et al.
Published: (2024)
Evaluating and Aligning CodeLLMs on Human Preference
by: Yang, Jian, et al.
Published: (2024)
by: Yang, Jian, et al.
Published: (2024)
SelfCodeAlign: Self-Alignment for Code Generation
by: Wei, Yuxiang, et al.
Published: (2024)
by: Wei, Yuxiang, et al.
Published: (2024)
Aligning Teacher with Student Preferences for Tailored Training Data Generation
by: Liu, Yantao, et al.
Published: (2024)
by: Liu, Yantao, et al.
Published: (2024)
MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data
by: Saxena, Vageesh, et al.
Published: (2024)
by: Saxena, Vageesh, et al.
Published: (2024)
Similar Items
-
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
by: Gritta, Milan, et al.
Published: (2024) -
DReSD: Dense Retrieval for Speculative Decoding
by: Gritta, Milan, et al.
Published: (2025) -
DRIFT: Decompose, Retrieve, Illustrate, then Formalize Theorems
by: Zhang, Meiru, et al.
Published: (2025) -
Mixture of Attentions For Speculative Decoding
by: Zimmer, Matthieu, et al.
Published: (2024) -
Text-to-Code Generation with Modality-relative Pre-training
by: Christopoulou, Fenia, et al.
Published: (2024)