FairCoder: Evaluating Social Bias of LLMs in Code Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Du, Yongkang, Huang, Jen-tse, Zhao, Jieyu, Lin, Lu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
di: Dou, Shihan, et al.
Pubblicazione: (2024)
di: Dou, Shihan, et al.
Pubblicazione: (2024)
MaintainCoder: Maintainable Code Generation Under Dynamic Requirements
di: Wang, Zhengren, et al.
Pubblicazione: (2025)
di: Wang, Zhengren, et al.
Pubblicazione: (2025)
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct
di: Wu, Yutong, et al.
Pubblicazione: (2024)
di: Wu, Yutong, et al.
Pubblicazione: (2024)
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
di: Ni, Yuansheng, et al.
Pubblicazione: (2025)
di: Ni, Yuansheng, et al.
Pubblicazione: (2025)
Evaluation of Code LLMs on Geospatial Code Generation
di: Gramacki, Piotr, et al.
Pubblicazione: (2024)
di: Gramacki, Piotr, et al.
Pubblicazione: (2024)
Seed-Coder: Let the Code Model Curate Data for Itself
di: Seed, ByteDance, et al.
Pubblicazione: (2025)
di: Seed, ByteDance, et al.
Pubblicazione: (2025)
CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation
di: Doan, Duy Tung, et al.
Pubblicazione: (2026)
di: Doan, Duy Tung, et al.
Pubblicazione: (2026)
InstructCoder: Instruction Tuning Large Language Models for Code Editing
di: Li, Kaixin, et al.
Pubblicazione: (2023)
di: Li, Kaixin, et al.
Pubblicazione: (2023)
EffiCoder: Enhancing Code Generation in Large Language Models through Efficiency-Aware Fine-tuning
di: Huang, Dong, et al.
Pubblicazione: (2024)
di: Huang, Dong, et al.
Pubblicazione: (2024)
New Job, New Gender? Measuring the Social Bias in Image Generation Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
UnitCoder: Scalable Iterative Code Synthesis with Unit Test Guidance
di: Ma, Yichuan, et al.
Pubblicazione: (2025)
di: Ma, Yichuan, et al.
Pubblicazione: (2025)
ShortCoder: Knowledge-Augmented Syntax Optimization for Token-Efficient Code Generation
di: Liu, Sicong, et al.
Pubblicazione: (2026)
di: Liu, Sicong, et al.
Pubblicazione: (2026)
AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions
di: Wang, Zihan, et al.
Pubblicazione: (2025)
di: Wang, Zihan, et al.
Pubblicazione: (2025)
WaveCoder: Widespread And Versatile Enhancement For Code Large Language Models By Instruction Tuning
di: Yu, Zhaojian, et al.
Pubblicazione: (2023)
di: Yu, Zhaojian, et al.
Pubblicazione: (2023)
R2C2-Coder: Enhancing and Benchmarking Real-world Repository-level Code Completion Abilities of Code Large Language Models
di: Deng, Ken, et al.
Pubblicazione: (2024)
di: Deng, Ken, et al.
Pubblicazione: (2024)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey
di: Wang, Junqiao, et al.
Pubblicazione: (2024)
di: Wang, Junqiao, et al.
Pubblicazione: (2024)
SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades
di: Lam, Man Ho, et al.
Pubblicazione: (2026)
di: Lam, Man Ho, et al.
Pubblicazione: (2026)
Code-Vision: Evaluating Multimodal LLMs Logic Understanding and Code Generation Capabilities
di: Wang, Hanbin, et al.
Pubblicazione: (2025)
di: Wang, Hanbin, et al.
Pubblicazione: (2025)
From If-Statements to ML Pipelines: Revisiting Bias in Code-Generation
di: Bui, Minh Duc, et al.
Pubblicazione: (2026)
di: Bui, Minh Duc, et al.
Pubblicazione: (2026)
From Output to Evaluation: Does Raw Instruction-Tuned Code LLMs Output Suffice for Fill-in-the-Middle Code Generation?
di: Ahmad, Wasi Uddin, et al.
Pubblicazione: (2025)
di: Ahmad, Wasi Uddin, et al.
Pubblicazione: (2025)
CodeScope: An Execution-based Multilingual Multitask Multidimensional Benchmark for Evaluating LLMs on Code Understanding and Generation
di: Yan, Weixiang, et al.
Pubblicazione: (2023)
di: Yan, Weixiang, et al.
Pubblicazione: (2023)
Measuring the Influence of Incorrect Code on Test Generation
di: Huang, Dong, et al.
Pubblicazione: (2024)
di: Huang, Dong, et al.
Pubblicazione: (2024)
Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval
di: Wang, Jiexin, et al.
Pubblicazione: (2024)
di: Wang, Jiexin, et al.
Pubblicazione: (2024)
SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning
di: Ding, Yangruibo, et al.
Pubblicazione: (2024)
di: Ding, Yangruibo, et al.
Pubblicazione: (2024)
State-of-the-art Small Language Coder Model: Mify-Coder
di: Parmar, Abhinav, et al.
Pubblicazione: (2025)
di: Parmar, Abhinav, et al.
Pubblicazione: (2025)
VisCoder2: Building Multi-Language Visualization Coding Agents
di: Ni, Yuansheng, et al.
Pubblicazione: (2025)
di: Ni, Yuansheng, et al.
Pubblicazione: (2025)
ComboBench: Can LLMs Manipulate Physical Devices to Play Virtual Reality Games?
di: Li, Shuqing, et al.
Pubblicazione: (2025)
di: Li, Shuqing, et al.
Pubblicazione: (2025)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
JumpCoder: Go Beyond Autoregressive Coder via Online Modification
di: Chen, Mouxiang, et al.
Pubblicazione: (2024)
di: Chen, Mouxiang, et al.
Pubblicazione: (2024)
ProjectEval: A Benchmark for Programming Agents Automated Evaluation on Project-Level Code Generation
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
CONCUR: Benchmarking LLMs for Concurrent Code Generation
di: Huang, Jue, et al.
Pubblicazione: (2026)
di: Huang, Jue, et al.
Pubblicazione: (2026)
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
di: Tian, Yuchen, et al.
Pubblicazione: (2024)
di: Tian, Yuchen, et al.
Pubblicazione: (2024)
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
di: Jiang, Juyong, et al.
Pubblicazione: (2026)
di: Jiang, Juyong, et al.
Pubblicazione: (2026)
CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation
di: Chen, Zaoyu, et al.
Pubblicazione: (2026)
di: Chen, Zaoyu, et al.
Pubblicazione: (2026)
OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs
di: Ahmad, Wasi Uddin, et al.
Pubblicazione: (2025)
di: Ahmad, Wasi Uddin, et al.
Pubblicazione: (2025)
ObscuraCoder: Powering Efficient Code LM Pre-Training Via Obfuscation Grounding
di: Paul, Indraneil, et al.
Pubblicazione: (2025)
di: Paul, Indraneil, et al.
Pubblicazione: (2025)
IQuest-Coder-V1 Technical Report
di: Yang, Jian, et al.
Pubblicazione: (2026)
di: Yang, Jian, et al.
Pubblicazione: (2026)
Benchmarking LLMs for Unit Test Generation from Real-World Functions
di: Huang, Dong, et al.
Pubblicazione: (2025)
di: Huang, Dong, et al.
Pubblicazione: (2025)
Evaluating and Achieving Controllable Code Completion in Code LLM
di: Zhang, Jiajun, et al.
Pubblicazione: (2026)
di: Zhang, Jiajun, et al.
Pubblicazione: (2026)
Documenti analoghi
-
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
di: Dou, Shihan, et al.
Pubblicazione: (2024) -
MaintainCoder: Maintainable Code Generation Under Dynamic Requirements
di: Wang, Zhengren, et al.
Pubblicazione: (2025) -
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct
di: Wu, Yutong, et al.
Pubblicazione: (2024) -
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
di: Ni, Yuansheng, et al.
Pubblicazione: (2025) -
Evaluation of Code LLMs on Geospatial Code Generation
di: Gramacki, Piotr, et al.
Pubblicazione: (2024)