Guardado en:
| Autores principales: | Bignotti, Camilla, Camassa, Carolina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2407.19760 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Prompting for Policy: Forecasting Macroeconomic Scenarios with Synthetic LLM Personas
por: Iadisernia, Giulia, et al.
Publicado: (2025)
por: Iadisernia, Giulia, et al.
Publicado: (2025)
Chat Bankman-Fried: an Exploration of LLM Alignment in Finance
por: Biancotti, Claudia, et al.
Publicado: (2024)
por: Biancotti, Claudia, et al.
Publicado: (2024)
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
por: Camassa, Carolina, et al.
Publicado: (2026)
por: Camassa, Carolina, et al.
Publicado: (2026)
LLMs Provide Unstable Answers to Legal Questions
por: Blair-Stanek, Andrew, et al.
Publicado: (2025)
por: Blair-Stanek, Andrew, et al.
Publicado: (2025)
Towards Next-Generation Medical Agent: How o1 is Reshaping Decision-Making in Medical Scenarios
por: Xu, Shaochen, et al.
Publicado: (2024)
por: Xu, Shaochen, et al.
Publicado: (2024)
Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms
por: Ashkinaze, Joshua, et al.
Publicado: (2024)
por: Ashkinaze, Joshua, et al.
Publicado: (2024)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
por: Arita, Takaya, et al.
Publicado: (2025)
por: Arita, Takaya, et al.
Publicado: (2025)
Are LLMs Court-Ready? Evaluating Frontier Models on Indian Legal Reasoning
por: Juvekar, Kush, et al.
Publicado: (2025)
por: Juvekar, Kush, et al.
Publicado: (2025)
RTP-LX: Can LLMs Evaluate Toxicity in Multilingual Scenarios?
por: de Wynter, Adrian, et al.
Publicado: (2024)
por: de Wynter, Adrian, et al.
Publicado: (2024)
Leveraging Large Language Models (LLMs) for Traffic Management at Urban Intersections: The Case of Mixed Traffic Scenarios
por: Masri, Sari, et al.
Publicado: (2024)
por: Masri, Sari, et al.
Publicado: (2024)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
por: Anzenberg, Eitan, et al.
Publicado: (2025)
por: Anzenberg, Eitan, et al.
Publicado: (2025)
RAGAT-Mind: A Multi-Granular Modeling Approach for Rumor Detection Based on MindSpore
por: Qin, Zhenkai, et al.
Publicado: (2025)
por: Qin, Zhenkai, et al.
Publicado: (2025)
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
por: AlDahoul, Nouar, et al.
Publicado: (2025)
por: AlDahoul, Nouar, et al.
Publicado: (2025)
PLawBench: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal Practice
por: Shi, Yuzhen, et al.
Publicado: (2026)
por: Shi, Yuzhen, et al.
Publicado: (2026)
Does Claude's Constitution Have a Culture?
por: Pourdavood, Parham
Publicado: (2026)
por: Pourdavood, Parham
Publicado: (2026)
Epistemic Constitutionalism Or: how to avoid coherence bias
por: Loi, Michele
Publicado: (2026)
por: Loi, Michele
Publicado: (2026)
Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction
por: Liu, Junkai, et al.
Publicado: (2024)
por: Liu, Junkai, et al.
Publicado: (2024)
Are Models Trained on Indian Legal Data Fair?
por: Girhepuje, Sahil, et al.
Publicado: (2023)
por: Girhepuje, Sahil, et al.
Publicado: (2023)
Towards Grammatical Tagging for the Legal Language of Cybersecurity
por: Castiglione, Gianpietro, et al.
Publicado: (2023)
por: Castiglione, Gianpietro, et al.
Publicado: (2023)
Mining Legal Arguments to Study Judicial Formalism
por: Koref, Tomáš, et al.
Publicado: (2025)
por: Koref, Tomáš, et al.
Publicado: (2025)
Toward Robust Legal Text Formalization into Defeasible Deontic Logic using LLMs
por: Horner, Elias, et al.
Publicado: (2025)
por: Horner, Elias, et al.
Publicado: (2025)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
por: Shin, Jisu, et al.
Publicado: (2025)
por: Shin, Jisu, et al.
Publicado: (2025)
Gender Bias in LLMs: Preliminary Evidence from Shared Parenting Scenario in Czech Family Law
por: Harasta, Jakub, et al.
Publicado: (2026)
por: Harasta, Jakub, et al.
Publicado: (2026)
Artificial Intelligence and Civil Discourse: How LLMs Moderate Climate Change Conversations
por: Fan, Wenlu, et al.
Publicado: (2025)
por: Fan, Wenlu, et al.
Publicado: (2025)
Minding the Politeness Gap in Cross-cultural Communication
por: Machino, Yuka, et al.
Publicado: (2025)
por: Machino, Yuka, et al.
Publicado: (2025)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
por: Dahl, Matthew, et al.
Publicado: (2024)
por: Dahl, Matthew, et al.
Publicado: (2024)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
por: Gajewska, Ewelina, et al.
Publicado: (2025)
por: Gajewska, Ewelina, et al.
Publicado: (2025)
Caveat Lector: Large Language Models in Legal Practice
por: Mik, Eliza
Publicado: (2024)
por: Mik, Eliza
Publicado: (2024)
Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensions
por: Patil, Parth, et al.
Publicado: (2026)
por: Patil, Parth, et al.
Publicado: (2026)
How Large Language Models (LLMs) Extrapolate: From Guided Missiles to Guided Prompts
por: Cao, Xuenan
Publicado: (2024)
por: Cao, Xuenan
Publicado: (2024)
Bridging Legal Interpretation and Formal Logic: Faithfulness, Assumption, and the Future of AI Legal Reasoning
por: Wang, Olivia Peiyu, et al.
Publicado: (2026)
por: Wang, Olivia Peiyu, et al.
Publicado: (2026)
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
por: Shankar, Hari, et al.
Publicado: (2026)
por: Shankar, Hari, et al.
Publicado: (2026)
From Perceptions to Decisions: Wildfire Evacuation Decision Prediction with Behavioral Theory-informed LLMs
por: Chen, Ruxiao, et al.
Publicado: (2025)
por: Chen, Ruxiao, et al.
Publicado: (2025)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
por: Magesh, Varun, et al.
Publicado: (2024)
por: Magesh, Varun, et al.
Publicado: (2024)
How Far Are LLMs from Believable AI? A Benchmark for Evaluating the Believability of Human Behavior Simulation
por: Xiao, Yang, et al.
Publicado: (2023)
por: Xiao, Yang, et al.
Publicado: (2023)
Red Lines and Grey Zones in the Fog of War: Benchmarking Legal Risk, Moral Harm, and Regional Bias in Large Language Model Military Decision-Making
por: Drinkall, Toby
Publicado: (2025)
por: Drinkall, Toby
Publicado: (2025)
Persuadability and LLMs as Legal Decision Tools
por: Suttle, Oisin, et al.
Publicado: (2026)
por: Suttle, Oisin, et al.
Publicado: (2026)
Few-shot Hate Speech Detection Based on the MindSpore Framework
por: Qin, Zhenkai, et al.
Publicado: (2025)
por: Qin, Zhenkai, et al.
Publicado: (2025)
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models
por: Hijazi, Faris, et al.
Publicado: (2024)
por: Hijazi, Faris, et al.
Publicado: (2024)
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
por: Hou, Abe Bohan, et al.
Publicado: (2024)
por: Hou, Abe Bohan, et al.
Publicado: (2024)
Ejemplares similares
-
Prompting for Policy: Forecasting Macroeconomic Scenarios with Synthetic LLM Personas
por: Iadisernia, Giulia, et al.
Publicado: (2025) -
Chat Bankman-Fried: an Exploration of LLM Alignment in Finance
por: Biancotti, Claudia, et al.
Publicado: (2024) -
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
por: Camassa, Carolina, et al.
Publicado: (2026) -
LLMs Provide Unstable Answers to Legal Questions
por: Blair-Stanek, Andrew, et al.
Publicado: (2025) -
Towards Next-Generation Medical Agent: How o1 is Reshaping Decision-Making in Medical Scenarios
por: Xu, Shaochen, et al.
Publicado: (2024)