The LLM Has Left The Chat: Evidence of Bail Preferences in Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ensign, Danielle, Sleight, Henry, Fish, Kyle |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
When the Domain Expert Has No Time and the LLM Developer Has No Clinical Expertise: Real-World Lessons from LLM Co-Design in a Safety-Net Hospital
par: Kothari, Avni, et autres
Publié: (2025)
par: Kothari, Avni, et autres
Publié: (2025)
Survey on Plagiarism Detection in Large Language Models: The Impact of ChatGPT and Gemini on Academic Integrity
par: Pudasaini, Shushanta, et autres
Publié: (2024)
par: Pudasaini, Shushanta, et autres
Publié: (2024)
All Code, No Thought: Current Language Models Struggle to Reason in Ciphered Language
par: Guo, Shiyuan, et autres
Publié: (2025)
par: Guo, Shiyuan, et autres
Publié: (2025)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
par: Zhou, Han, et autres
Publié: (2024)
par: Zhou, Han, et autres
Publié: (2024)
Large Language Models as Urban Residents: An LLM Agent Framework for Personal Mobility Generation
par: Wang, Jiawei, et autres
Publié: (2024)
par: Wang, Jiawei, et autres
Publié: (2024)
Toward Preference-aligned Large Language Models via Residual-based Model Steering
par: La Cava, Lucio, et autres
Publié: (2025)
par: La Cava, Lucio, et autres
Publié: (2025)
Defining and Evaluating Physical Safety for Large Language Models
par: Tang, Yung-Chen, et autres
Publié: (2024)
par: Tang, Yung-Chen, et autres
Publié: (2024)
Evaluating Large Language Models for Fair and Reliable Organ Allocation
par: Kim, Brian Hyeongseok, et autres
Publié: (2025)
par: Kim, Brian Hyeongseok, et autres
Publié: (2025)
Existential Conversations with Large Language Models: Content, Community, and Culture
par: Shanahan, Murray, et autres
Publié: (2024)
par: Shanahan, Murray, et autres
Publié: (2024)
Leveraging Large Language Models for Tacit Knowledge Discovery in Organizational Contexts
par: Zuin, Gianlucca, et autres
Publié: (2025)
par: Zuin, Gianlucca, et autres
Publié: (2025)
Participatory Assessment of Large Language Model Applications in an Academic Medical Center
par: Carra, Giorgia, et autres
Publié: (2024)
par: Carra, Giorgia, et autres
Publié: (2024)
I Prefer not to Say: Protecting User Consent in Models with Optional Personal Data
par: Leemann, Tobias, et autres
Publié: (2022)
par: Leemann, Tobias, et autres
Publié: (2022)
Fairness of ChatGPT
par: Li, Yunqi, et autres
Publié: (2023)
par: Li, Yunqi, et autres
Publié: (2023)
Mechanistic Interpretability with SAEs: Probing Religion, Violence, and Geography in Large Language Models
par: Simbeck, Katharina, et autres
Publié: (2025)
par: Simbeck, Katharina, et autres
Publié: (2025)
Correlated Errors in Large Language Models
par: Kim, Elliot, et autres
Publié: (2025)
par: Kim, Elliot, et autres
Publié: (2025)
Harnessing Large Language Models for Mental Health: Opportunities, Challenges, and Ethical Considerations
par: Pandey, Hari Mohan
Publié: (2024)
par: Pandey, Hari Mohan
Publié: (2024)
Hypothesis Generation with Large Language Models
par: Zhou, Yangqiaoyu, et autres
Publié: (2024)
par: Zhou, Yangqiaoyu, et autres
Publié: (2024)
Large Language Models are Geographically Biased
par: Manvi, Rohin, et autres
Publié: (2024)
par: Manvi, Rohin, et autres
Publié: (2024)
No Culture Left Behind: ArtELingo-28, a Benchmark of WikiArt with Captions in 28 Languages
par: Mohamed, Youssef, et autres
Publié: (2024)
par: Mohamed, Youssef, et autres
Publié: (2024)
Machine Learners Should Acknowledge the Legal Implications of Large Language Models as Personal Data
par: Nolte, Henrik, et autres
Publié: (2025)
par: Nolte, Henrik, et autres
Publié: (2025)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
par: Urchs, Stefanie, et autres
Publié: (2023)
par: Urchs, Stefanie, et autres
Publié: (2023)
Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas
par: Chiu, Yu Ying, et autres
Publié: (2025)
par: Chiu, Yu Ying, et autres
Publié: (2025)
Psychological Counseling Ability of Large Language Models
par: Peng, Fangyu, et autres
Publié: (2025)
par: Peng, Fangyu, et autres
Publié: (2025)
Assessing Large Language Models on Climate Information
par: Bulian, Jannis, et autres
Publié: (2023)
par: Bulian, Jannis, et autres
Publié: (2023)
Judging by Appearances? Auditing and Intervening Vision-Language Models for Bail Prediction
par: Basu, Sagnik, et autres
Publié: (2025)
par: Basu, Sagnik, et autres
Publié: (2025)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
par: Sehwag, Udari Madhushani, et autres
Publié: (2025)
par: Sehwag, Udari Madhushani, et autres
Publié: (2025)
Evaluating Retrieval-Augmented Generation Strategies for Large Language Models in Travel Mode Choice Prediction
par: Xu, Yiming, et autres
Publié: (2025)
par: Xu, Yiming, et autres
Publié: (2025)
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans?
par: Zakazov, Ivan, et autres
Publié: (2024)
par: Zakazov, Ivan, et autres
Publié: (2024)
Learning the Value Systems of Societies from Preferences
par: Holgado-Sánchez, Andrés, et autres
Publié: (2025)
par: Holgado-Sánchez, Andrés, et autres
Publié: (2025)
Investigating Thematic Patterns and User Preferences in LLM Interactions using BERTopic
par: Bhandarkar, Abhay, et autres
Publié: (2025)
par: Bhandarkar, Abhay, et autres
Publié: (2025)
Critical Foreign Policy Decisions (CFPD)-Benchmark: Measuring Diplomatic Preferences in Large Language Models
par: Jensen, Benjamin, et autres
Publié: (2025)
par: Jensen, Benjamin, et autres
Publié: (2025)
A Taxonomy of Stereotype Content in Large Language Models
par: Nicolas, Gandalf, et autres
Publié: (2024)
par: Nicolas, Gandalf, et autres
Publié: (2024)
Bias and Fairness in Large Language Models: A Survey
par: Gallegos, Isabel O., et autres
Publié: (2023)
par: Gallegos, Isabel O., et autres
Publié: (2023)
Transforming Agency. On the mode of existence of Large Language Models
par: Barandiaran, Xabier E., et autres
Publié: (2024)
par: Barandiaran, Xabier E., et autres
Publié: (2024)
LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
par: Faiz, Ahmad, et autres
Publié: (2023)
par: Faiz, Ahmad, et autres
Publié: (2023)
The AI Companion in Education: Analyzing the Pedagogical Potential of ChatGPT in Computer Science and Engineering
par: He, Zhangying, et autres
Publié: (2024)
par: He, Zhangying, et autres
Publié: (2024)
ResumeAtlas: Revisiting Resume Classification with Large-Scale Datasets and Large Language Models
par: Heakl, Ahmed, et autres
Publié: (2024)
par: Heakl, Ahmed, et autres
Publié: (2024)
Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey
par: Deng, Chengyuan, et autres
Publié: (2024)
par: Deng, Chengyuan, et autres
Publié: (2024)
Foundational Challenges in Assuring Alignment and Safety of Large Language Models
par: Anwar, Usman, et autres
Publié: (2024)
par: Anwar, Usman, et autres
Publié: (2024)
Exploring Accuracy-Fairness Trade-off in Large Language Models
par: Zhang, Qingquan, et autres
Publié: (2024)
par: Zhang, Qingquan, et autres
Publié: (2024)
Documents similaires
-
When the Domain Expert Has No Time and the LLM Developer Has No Clinical Expertise: Real-World Lessons from LLM Co-Design in a Safety-Net Hospital
par: Kothari, Avni, et autres
Publié: (2025) -
Survey on Plagiarism Detection in Large Language Models: The Impact of ChatGPT and Gemini on Academic Integrity
par: Pudasaini, Shushanta, et autres
Publié: (2024) -
All Code, No Thought: Current Language Models Struggle to Reason in Ciphered Language
par: Guo, Shiyuan, et autres
Publié: (2025) -
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
par: Zhou, Han, et autres
Publié: (2024) -
Large Language Models as Urban Residents: An LLM Agent Framework for Personal Mobility Generation
par: Wang, Jiawei, et autres
Publié: (2024)