Large Language Models are overconfident and amplify human bias
Fuente:
arXiv
Salvato in:
| Autori principali: | Sun, Fengfei, Li, Ningke, Wang, Kailong, Goette, Lorenz |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
di: Li, Ningke, et al.
Pubblicazione: (2024)
di: Li, Ningke, et al.
Pubblicazione: (2024)
Towards adaptive trajectories for mixed autonomous and human-operated ships
di: Pianini, Danilo, et al.
Pubblicazione: (2024)
di: Pianini, Danilo, et al.
Pubblicazione: (2024)
The Systems Engineering Approach in Times of Large Language Models
di: Cabrera, Christian, et al.
Pubblicazione: (2024)
di: Cabrera, Christian, et al.
Pubblicazione: (2024)
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
di: Zheng, Xinyi, et al.
Pubblicazione: (2025)
di: Zheng, Xinyi, et al.
Pubblicazione: (2025)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
di: Li, Yuxi, et al.
Pubblicazione: (2024)
di: Li, Yuxi, et al.
Pubblicazione: (2024)
Thoughts on Learning Human and Programming Languages
di: Katz, Daniel S., et al.
Pubblicazione: (2024)
di: Katz, Daniel S., et al.
Pubblicazione: (2024)
FastFixer: An Efficient and Effective Approach for Repairing Programming Assignments
di: Liu, Fang, et al.
Pubblicazione: (2024)
di: Liu, Fang, et al.
Pubblicazione: (2024)
The Potential of Citizen Platforms for Requirements Engineering of Large Socio-Technical Software Systems
di: Ruohonen, Jukka, et al.
Pubblicazione: (2024)
di: Ruohonen, Jukka, et al.
Pubblicazione: (2024)
Regulatory Requirements Engineering in Large Enterprises: An Interview Study on the European Accessibility Act
di: Kosenkov, Oleksandr, et al.
Pubblicazione: (2024)
di: Kosenkov, Oleksandr, et al.
Pubblicazione: (2024)
"Write in English, Nobody Understands Your Language Here": A Study of Non-English Trends in Open-Source Repositories
di: Bhuiyan, Masudul Hasan Masud, et al.
Pubblicazione: (2026)
di: Bhuiyan, Masudul Hasan Masud, et al.
Pubblicazione: (2026)
Enhancing Semantic Understanding in Pointer Analysis using Large Language Models
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
A Conceptual Model and Methodology for Sustainability-aware, IoT-enhanced Business Processes
di: Bosch, Victoria Torres, et al.
Pubblicazione: (2025)
di: Bosch, Victoria Torres, et al.
Pubblicazione: (2025)
Model-based Elaboration of a Requirements and Design Pattern Catalogue for Sustainable Systems
di: Ponsard, Christophe
Pubblicazione: (2025)
di: Ponsard, Christophe
Pubblicazione: (2025)
Early Results from Teaching Modelling for Software Comprehension in New-Hire Onboarding
di: Kumar, Mrityunjay, et al.
Pubblicazione: (2025)
di: Kumar, Mrityunjay, et al.
Pubblicazione: (2025)
Diversity in Software Engineering Education: Exploring Motivations, Influences, and Role Models Among Undergraduate Students
di: Santos, Ronnie de Souza, et al.
Pubblicazione: (2024)
di: Santos, Ronnie de Souza, et al.
Pubblicazione: (2024)
From Generation to Adaptation: Comparing AI-Assisted Strategies in High School Programming Education
di: Hu, Tong, et al.
Pubblicazione: (2025)
di: Hu, Tong, et al.
Pubblicazione: (2025)
Discovering Ideologies of the Open Source Software Movement
di: Yue, Yang, et al.
Pubblicazione: (2025)
di: Yue, Yang, et al.
Pubblicazione: (2025)
Can Large-Language Models Help us Better Understand and Teach the Development of Energy-Efficient Software?
di: Hasler, Ryan, et al.
Pubblicazione: (2024)
di: Hasler, Ryan, et al.
Pubblicazione: (2024)
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT
di: Qi, Yi, et al.
Pubblicazione: (2023)
di: Qi, Yi, et al.
Pubblicazione: (2023)
Software Engineering Through Community-Engaged Learning and an Inclusive Network
di: Arony, Nowshin Nawar, et al.
Pubblicazione: (2023)
di: Arony, Nowshin Nawar, et al.
Pubblicazione: (2023)
Ideology in Open Source Development
di: Yue, Yang, et al.
Pubblicazione: (2021)
di: Yue, Yang, et al.
Pubblicazione: (2021)
Understanding the Effectiveness of Coverage Criteria for Large Language Models: A Special Angle from Jailbreak Attacks
di: Zhou, Shide, et al.
Pubblicazione: (2024)
di: Zhou, Shide, et al.
Pubblicazione: (2024)
Semantic-Enhanced Indirect Call Analysis with Large Language Models
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
GraphCue for SDN Configuration Code Synthesis
di: Qi, Haomin, et al.
Pubblicazione: (2025)
di: Qi, Haomin, et al.
Pubblicazione: (2025)
The Fusion of Large Language Models and Formal Methods for Trustworthy AI Agents: A Roadmap
di: Zhang, Yedi, et al.
Pubblicazione: (2024)
di: Zhang, Yedi, et al.
Pubblicazione: (2024)
A Scenario Analysis of Ethical Issues in Dark Patterns and Their Research
di: Ruohonen, Jukka, et al.
Pubblicazione: (2025)
di: Ruohonen, Jukka, et al.
Pubblicazione: (2025)
Towards a Knowledge Base of Common Sustainability Weaknesses in Green Software Development
di: Pathania, Priyavanshi, et al.
Pubblicazione: (2025)
di: Pathania, Priyavanshi, et al.
Pubblicazione: (2025)
Overcoming Obstacles: Challenges of Gender Inequality in Undergraduate ICT Programs
di: Souza, Angelica Pereira, et al.
Pubblicazione: (2025)
di: Souza, Angelica Pereira, et al.
Pubblicazione: (2025)
Robustness tests for biomedical foundation models should tailor to specifications
di: Xian, R. Patrick, et al.
Pubblicazione: (2025)
di: Xian, R. Patrick, et al.
Pubblicazione: (2025)
How frontier AI companies could implement an internal audit function
di: Gomez, Francesca, et al.
Pubblicazione: (2025)
di: Gomez, Francesca, et al.
Pubblicazione: (2025)
From Pre-labeling to Production: Engineering Lessons from a Machine Learning Pipeline in the Public Sector
di: Ferreira, Ronivaldo, et al.
Pubblicazione: (2025)
di: Ferreira, Ronivaldo, et al.
Pubblicazione: (2025)
A Systematic Mapping on Software Fairness: Focus, Trends and Industrial Context
di: Nepomuceno, Kessia, et al.
Pubblicazione: (2025)
di: Nepomuceno, Kessia, et al.
Pubblicazione: (2025)
Assessing Simulation Knowledge and Proficiency Among Undergraduate Computing Students in Brazil: Insights and Results from a Survey Research
di: Rodrigues, Fernando Brito, et al.
Pubblicazione: (2025)
di: Rodrigues, Fernando Brito, et al.
Pubblicazione: (2025)
A Bot-based Approach to Manage Codes of Conduct in Open-Source Projects
di: Cobos, Sergio, et al.
Pubblicazione: (2025)
di: Cobos, Sergio, et al.
Pubblicazione: (2025)
Calculating Software's Energy Use and Carbon Emissions: A Survey of the State of Art, Challenges, and the Way Ahead
di: Pathania, Priyavanshi, et al.
Pubblicazione: (2025)
di: Pathania, Priyavanshi, et al.
Pubblicazione: (2025)
Measuring What Matters: A Framework for Evaluating Safety Risks in Real-World LLM Applications
di: Goh, Jia Yi, et al.
Pubblicazione: (2025)
di: Goh, Jia Yi, et al.
Pubblicazione: (2025)
The Evaluation of Open Source Software Innovativeness
di: Benkeltoum, Nordine
Pubblicazione: (2025)
di: Benkeltoum, Nordine
Pubblicazione: (2025)
Software is infrastructure: failures, successes, costs, and the case for formal verification
di: Bernardi, Giovanni, et al.
Pubblicazione: (2025)
di: Bernardi, Giovanni, et al.
Pubblicazione: (2025)
Mining a Decade of Event Impacts on Contributor Dynamics in Ethereum: A Longitudinal Study
di: Vaccargiu, Matteo, et al.
Pubblicazione: (2025)
di: Vaccargiu, Matteo, et al.
Pubblicazione: (2025)
The Impact of Team Diversity in Agile Development Education
di: Torchiano, Marco, et al.
Pubblicazione: (2025)
di: Torchiano, Marco, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
di: Li, Ningke, et al.
Pubblicazione: (2024) -
Towards adaptive trajectories for mixed autonomous and human-operated ships
di: Pianini, Danilo, et al.
Pubblicazione: (2024) -
The Systems Engineering Approach in Times of Large Language Models
di: Cabrera, Christian, et al.
Pubblicazione: (2024) -
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
di: Zheng, Xinyi, et al.
Pubblicazione: (2025) -
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
di: Li, Yuxi, et al.
Pubblicazione: (2024)