Smart but Costly? Benchmarking LLMs on Functional Accuracy and Energy Efficiency
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mehditabar, Mohammadjavad, Rajput, Saurabhsingh, Mastropaolo, Antonio, Sharma, Tushar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Validated Taxonomy on Software Energy Smells
von: Mehditabar, Mohammadjavad, et al.
Veröffentlicht: (2026)
von: Mehditabar, Mohammadjavad, et al.
Veröffentlicht: (2026)
Energy Flow Graph: Modeling Software Energy Consumption
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
CodeGreen: Towards Improving Precision and Portability in Software Energy Measurement
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
Tu(r)ning AI Green: Exploring Energy Efficiency Cascading with Orthogonal Optimizations
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2025)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2025)
FlipFlop: A Static Analysis-based Energy Optimization Framework for GPU Kernels
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)
COMET: Generating Commit Messages using Delta Graph Context Representation
von: Mandli, Abhinav Reddy, et al.
Veröffentlicht: (2024)
von: Mandli, Abhinav Reddy, et al.
Veröffentlicht: (2024)
Enhancing Energy-Awareness in Deep Learning through Fine-Grained Energy Measurement
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2023)
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2023)
Not All Tokens Matter: Data-Centric Optimization for Efficient Code Summarization
von: Afrin, Saima, et al.
Veröffentlicht: (2026)
von: Afrin, Saima, et al.
Veröffentlicht: (2026)
A Path Less Traveled: Reimagining Software Engineering Automation via a Neurosymbolic Paradigm
von: Mastropaolo, Antonio, et al.
Veröffentlicht: (2025)
von: Mastropaolo, Antonio, et al.
Veröffentlicht: (2025)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
von: Cheng, Zaiyu, et al.
Veröffentlicht: (2026)
von: Cheng, Zaiyu, et al.
Veröffentlicht: (2026)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
Is Quantization a Deal-breaker? Empirical Insights from Large Code Models
von: Afrin, Saima, et al.
Veröffentlicht: (2025)
von: Afrin, Saima, et al.
Veröffentlicht: (2025)
Towards Comprehensive Benchmarking Infrastructure for LLMs In Software Engineering
von: Rodriguez-Cardenas, Daniel, et al.
Veröffentlicht: (2026)
von: Rodriguez-Cardenas, Daniel, et al.
Veröffentlicht: (2026)
A Systematic Literature Review of Parameter-Efficient Fine-Tuning for Large Code Models
von: Afrin, Saima, et al.
Veröffentlicht: (2025)
von: Afrin, Saima, et al.
Veröffentlicht: (2025)
Parameter-Efficient Multi-Task Fine-Tuning in Code-Related Tasks
von: Haque, Md Zahidul, et al.
Veröffentlicht: (2026)
von: Haque, Md Zahidul, et al.
Veröffentlicht: (2026)
SmartLLMs Scheduler: A Framework for Cost-Effective LLMs Utilization
von: Liu, Yueyue, et al.
Veröffentlicht: (2025)
von: Liu, Yueyue, et al.
Veröffentlicht: (2025)
Why Attention Fails: A Taxonomy of Faults in Attention-Based Neural Networks
von: Jahan, Sigma, et al.
Veröffentlicht: (2025)
von: Jahan, Sigma, et al.
Veröffentlicht: (2025)
Prompt-Driven Code Summarization: A Systematic Literature Review
von: Farjana, Afia, et al.
Veröffentlicht: (2026)
von: Farjana, Afia, et al.
Veröffentlicht: (2026)
Generating refactored code accurately using reinforcement learning
von: Palit, Indranil, et al.
Veröffentlicht: (2024)
von: Palit, Indranil, et al.
Veröffentlicht: (2024)
How the Training Procedure Impacts the Performance of Deep Learning-based Vulnerability Patching
von: Mastropaolo, Antonio, et al.
Veröffentlicht: (2024)
von: Mastropaolo, Antonio, et al.
Veröffentlicht: (2024)
Optimizing Datasets for Code Summarization: Is Code-Comment Coherence Enough?
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
Code Review Automation: Strengths and Weaknesses of the State of the Art
von: Tufano, Rosalia, et al.
Veröffentlicht: (2024)
von: Tufano, Rosalia, et al.
Veröffentlicht: (2024)
Beyond Code Similarity: Benchmarking the Plausibility, Efficiency, and Complexity of LLM-Generated Smart Contracts
von: Salzano, Francesco, et al.
Veröffentlicht: (2025)
von: Salzano, Francesco, et al.
Veröffentlicht: (2025)
Resource-Efficient & Effective Code Summarization
von: Afrin, Saima, et al.
Veröffentlicht: (2025)
von: Afrin, Saima, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Post-Training Quantization on Large Language Models for Code Generation
von: Giagnorio, Alessandro, et al.
Veröffentlicht: (2025)
von: Giagnorio, Alessandro, et al.
Veröffentlicht: (2025)
Fine-grained Multi-Document Extraction and Generation of Code Change Rationale
von: Sun, Mehedi, et al.
Veröffentlicht: (2026)
von: Sun, Mehedi, et al.
Veröffentlicht: (2026)
A Taxonomy of Self-Admitted Technical Debt in Deep Learning Systems
von: Pepe, Federica, et al.
Veröffentlicht: (2024)
von: Pepe, Federica, et al.
Veröffentlicht: (2024)
On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization
von: Crupi, Giuseppe, et al.
Veröffentlicht: (2025)
von: Crupi, Giuseppe, et al.
Veröffentlicht: (2025)
Towards Summarizing Code Snippets Using Pre-Trained Transformers
von: Mastropaolo, Antonio, et al.
Veröffentlicht: (2024)
von: Mastropaolo, Antonio, et al.
Veröffentlicht: (2024)
CONCORD: Towards a DSL for Configurable Graph Code Representation
von: Saad, Mootez, et al.
Veröffentlicht: (2024)
von: Saad, Mootez, et al.
Veröffentlicht: (2024)
Unveiling ChatGPT's Usage in Open Source Projects: A Mining-based Study
von: Tufano, Rosalia, et al.
Veröffentlicht: (2024)
von: Tufano, Rosalia, et al.
Veröffentlicht: (2024)
Evaluating the Energy-Efficiency of the Code Generated by LLMs
von: Islam, Md Arman, et al.
Veröffentlicht: (2025)
von: Islam, Md Arman, et al.
Veröffentlicht: (2025)
TS-Detector : Detecting Feature Toggle Usage Patterns
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2025)
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2025)
Broken Windows: Exploring the Applicability of a Controversial Theory on Code Quality
von: Spinellis, Diomidis, et al.
Veröffentlicht: (2024)
von: Spinellis, Diomidis, et al.
Veröffentlicht: (2024)
Hierarchical Evaluation of Software Design Capabilities of Large Language Models of Code
von: Saad, Mootez, et al.
Veröffentlicht: (2025)
von: Saad, Mootez, et al.
Veröffentlicht: (2025)
ALPINE: An adaptive language-agnostic pruning method for language models for code
von: Saad, Mootez, et al.
Veröffentlicht: (2024)
von: Saad, Mootez, et al.
Veröffentlicht: (2024)
On Inter-dataset Code Duplication and Data Leakage in Large Language Models
von: López, José Antonio Hernández, et al.
Veröffentlicht: (2024)
von: López, José Antonio Hernández, et al.
Veröffentlicht: (2024)
Developers and Generative AI: A Study of Self-Admitted Usage in Open Source Projects
von: Tufano, Rosalia, et al.
Veröffentlicht: (2026)
von: Tufano, Rosalia, et al.
Veröffentlicht: (2026)
On the Effect of Token Merging on Pre-trained Models for Code
von: Saad, Mootez, et al.
Veröffentlicht: (2025)
von: Saad, Mootez, et al.
Veröffentlicht: (2025)
Toward Neurosymbolic Program Comprehension
von: Velasco, Alejandro, et al.
Veröffentlicht: (2025)
von: Velasco, Alejandro, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Validated Taxonomy on Software Energy Smells
von: Mehditabar, Mohammadjavad, et al.
Veröffentlicht: (2026) -
Energy Flow Graph: Modeling Software Energy Consumption
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026) -
CodeGreen: Towards Improving Precision and Portability in Software Energy Measurement
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026) -
Tu(r)ning AI Green: Exploring Energy Efficiency Cascading with Orthogonal Optimizations
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2025) -
FlipFlop: A Static Analysis-based Energy Optimization Framework for GPU Kernels
von: Rajput, Saurabhsingh, et al.
Veröffentlicht: (2026)