ECCO: Can We Improve Model-Generated Code Efficiency Without Sacrificing Functional Correctness?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Waghjale, Siddhant, Veerendranath, Vishruth, Wang, Zora Zhiruo, Fried, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pre-Calc: Learning to Use the Calculator Improves Numeracy in Language Models
von: Veerendranath, Vishruth, et al.
Veröffentlicht: (2024)
von: Veerendranath, Vishruth, et al.
Veröffentlicht: (2024)
API-Assisted Code Generation for Question Answering on Varied Table Structures
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
What Are Tools Anyway? A Survey from the Language Model Perspective
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
cAST: Enhancing Code Retrieval-Augmented Generation with Structural Chunking via Abstract Syntax Tree
von: Zhang, Yilin, et al.
Veröffentlicht: (2025)
von: Zhang, Yilin, et al.
Veröffentlicht: (2025)
CodeRAG-Bench: Can Retrieval Augment Code Generation?
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
Inducing Programmatic Skills for Agentic Tasks
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
Agent Workflow Memory
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation
von: Huq, Faria, et al.
Veröffentlicht: (2025)
von: Huq, Faria, et al.
Veröffentlicht: (2025)
Reasoning Models Can Be Effective Without Thinking
von: Ma, Wenjie, et al.
Veröffentlicht: (2025)
von: Ma, Wenjie, et al.
Veröffentlicht: (2025)
CRScore: Grounding Automated Evaluation of Code Review Comments in Code Claims and Smells
von: Naik, Atharva, et al.
Veröffentlicht: (2024)
von: Naik, Atharva, et al.
Veröffentlicht: (2024)
SkillWeaver: Web Agents can Self-Improve by Discovering and Honing Skills
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
Dual-objective Language Models: Training Efficiency Without Overfitting
von: Samuel, David, et al.
Veröffentlicht: (2025)
von: Samuel, David, et al.
Veröffentlicht: (2025)
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
Do Cognitively Interpretable Reasoning Traces Improve LLM Performance?
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
Is the Pope Catholic? Yes, the Pope is Catholic. Generative Evaluation of Non-Literal Intent Resolution in LLMs
von: Yerukola, Akhila, et al.
Veröffentlicht: (2024)
von: Yerukola, Akhila, et al.
Veröffentlicht: (2024)
Can We Trust LLM Detectors?
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
Can We Verify Step by Step for Incorrect Answer Detection?
von: Xu, Xin, et al.
Veröffentlicht: (2024)
von: Xu, Xin, et al.
Veröffentlicht: (2024)
Can We Locate and Prevent Stereotypes in LLMs?
von: D'Souza, Alex
Veröffentlicht: (2026)
von: D'Souza, Alex
Veröffentlicht: (2026)
Can Large Language Models Understand, Reason About, and Generate Code-Switched Text?
von: Winata, Genta Indra, et al.
Veröffentlicht: (2026)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2026)
How Far Are We from Optimal Reasoning Efficiency?
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2025)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
von: Shi, Yuling, et al.
Veröffentlicht: (2024)
von: Shi, Yuling, et al.
Veröffentlicht: (2024)
LLMs Can Plan Only If We Tell Them
von: Sel, Bilgehan, et al.
Veröffentlicht: (2025)
von: Sel, Bilgehan, et al.
Veröffentlicht: (2025)
SAIL: Self-Improving Efficient Online Alignment of Large Language Models
von: Ding, Mucong, et al.
Veröffentlicht: (2024)
von: Ding, Mucong, et al.
Veröffentlicht: (2024)
Improving the Efficiency of Visually Augmented Language Models
von: Ontalvilla, Paula, et al.
Veröffentlicht: (2024)
von: Ontalvilla, Paula, et al.
Veröffentlicht: (2024)
DreamPRM-Code: Function-as-Step Process Reward Model with Label Correction for LLM Coding
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
Improving Training Efficiency and Reducing Maintenance Costs via Language Specific Model Merging
von: Dmonte, Alphaeus, et al.
Veröffentlicht: (2026)
von: Dmonte, Alphaeus, et al.
Veröffentlicht: (2026)
Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL
von: Zheng, Kunhao, et al.
Veröffentlicht: (2026)
von: Zheng, Kunhao, et al.
Veröffentlicht: (2026)
LLM-REVal: Can We Trust LLM Reviewers Yet?
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Can We Edit LLMs for Long-Tail Biomedical Knowledge?
von: Yi, Xinhao, et al.
Veröffentlicht: (2025)
von: Yi, Xinhao, et al.
Veröffentlicht: (2025)
We Can't Understand AI Using our Existing Vocabulary
von: Hewitt, John, et al.
Veröffentlicht: (2025)
von: Hewitt, John, et al.
Veröffentlicht: (2025)
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
Can Hallucination Correction Improve Video-Language Alignment?
von: Zhao, Lingjun, et al.
Veröffentlicht: (2025)
von: Zhao, Lingjun, et al.
Veröffentlicht: (2025)
On the Brittle Foundations of ReAct Prompting for Agentic Large Language Models
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
ECCO: Evidence-Driven Causal Reasoning for Compiler Optimization
von: Pan, Haolin, et al.
Veröffentlicht: (2026)
von: Pan, Haolin, et al.
Veröffentlicht: (2026)
"In Dialogues We Learn": Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning
von: Cheng, Chuanqi, et al.
Veröffentlicht: (2024)
von: Cheng, Chuanqi, et al.
Veröffentlicht: (2024)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
Can Language Model Moderators Improve the Health of Online Discourse?
von: Cho, Hyundong, et al.
Veröffentlicht: (2023)
von: Cho, Hyundong, et al.
Veröffentlicht: (2023)
CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
von: Gou, Zhibin, et al.
Veröffentlicht: (2023)
von: Gou, Zhibin, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Pre-Calc: Learning to Use the Calculator Improves Numeracy in Language Models
von: Veerendranath, Vishruth, et al.
Veröffentlicht: (2024) -
API-Assisted Code Generation for Question Answering on Varied Table Structures
von: Cao, Yihan, et al.
Veröffentlicht: (2023) -
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025) -
What Are Tools Anyway? A Survey from the Language Model Perspective
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024) -
cAST: Enhancing Code Retrieval-Augmented Generation with Structural Chunking via Abstract Syntax Tree
von: Zhang, Yilin, et al.
Veröffentlicht: (2025)