Salvato in:
| Autori principali: | Bignotti, Camilla, Camassa, Carolina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.19760 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Prompting for Policy: Forecasting Macroeconomic Scenarios with Synthetic LLM Personas
di: Iadisernia, Giulia, et al.
Pubblicazione: (2025)
di: Iadisernia, Giulia, et al.
Pubblicazione: (2025)
Chat Bankman-Fried: an Exploration of LLM Alignment in Finance
di: Biancotti, Claudia, et al.
Pubblicazione: (2024)
di: Biancotti, Claudia, et al.
Pubblicazione: (2024)
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
di: Camassa, Carolina, et al.
Pubblicazione: (2026)
di: Camassa, Carolina, et al.
Pubblicazione: (2026)
LLMs Provide Unstable Answers to Legal Questions
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
Towards Next-Generation Medical Agent: How o1 is Reshaping Decision-Making in Medical Scenarios
di: Xu, Shaochen, et al.
Pubblicazione: (2024)
di: Xu, Shaochen, et al.
Pubblicazione: (2024)
Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms
di: Ashkinaze, Joshua, et al.
Pubblicazione: (2024)
di: Ashkinaze, Joshua, et al.
Pubblicazione: (2024)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
di: Arita, Takaya, et al.
Pubblicazione: (2025)
di: Arita, Takaya, et al.
Pubblicazione: (2025)
Are LLMs Court-Ready? Evaluating Frontier Models on Indian Legal Reasoning
di: Juvekar, Kush, et al.
Pubblicazione: (2025)
di: Juvekar, Kush, et al.
Pubblicazione: (2025)
RTP-LX: Can LLMs Evaluate Toxicity in Multilingual Scenarios?
di: de Wynter, Adrian, et al.
Pubblicazione: (2024)
di: de Wynter, Adrian, et al.
Pubblicazione: (2024)
Leveraging Large Language Models (LLMs) for Traffic Management at Urban Intersections: The Case of Mixed Traffic Scenarios
di: Masri, Sari, et al.
Pubblicazione: (2024)
di: Masri, Sari, et al.
Pubblicazione: (2024)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
di: Anzenberg, Eitan, et al.
Pubblicazione: (2025)
di: Anzenberg, Eitan, et al.
Pubblicazione: (2025)
RAGAT-Mind: A Multi-Granular Modeling Approach for Rumor Detection Based on MindSpore
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
di: AlDahoul, Nouar, et al.
Pubblicazione: (2025)
PLawBench: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal Practice
di: Shi, Yuzhen, et al.
Pubblicazione: (2026)
di: Shi, Yuzhen, et al.
Pubblicazione: (2026)
Does Claude's Constitution Have a Culture?
di: Pourdavood, Parham
Pubblicazione: (2026)
di: Pourdavood, Parham
Pubblicazione: (2026)
Epistemic Constitutionalism Or: how to avoid coherence bias
di: Loi, Michele
Pubblicazione: (2026)
di: Loi, Michele
Pubblicazione: (2026)
Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction
di: Liu, Junkai, et al.
Pubblicazione: (2024)
di: Liu, Junkai, et al.
Pubblicazione: (2024)
Are Models Trained on Indian Legal Data Fair?
di: Girhepuje, Sahil, et al.
Pubblicazione: (2023)
di: Girhepuje, Sahil, et al.
Pubblicazione: (2023)
Towards Grammatical Tagging for the Legal Language of Cybersecurity
di: Castiglione, Gianpietro, et al.
Pubblicazione: (2023)
di: Castiglione, Gianpietro, et al.
Pubblicazione: (2023)
Mining Legal Arguments to Study Judicial Formalism
di: Koref, Tomáš, et al.
Pubblicazione: (2025)
di: Koref, Tomáš, et al.
Pubblicazione: (2025)
Toward Robust Legal Text Formalization into Defeasible Deontic Logic using LLMs
di: Horner, Elias, et al.
Pubblicazione: (2025)
di: Horner, Elias, et al.
Pubblicazione: (2025)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
di: Shin, Jisu, et al.
Pubblicazione: (2025)
di: Shin, Jisu, et al.
Pubblicazione: (2025)
Gender Bias in LLMs: Preliminary Evidence from Shared Parenting Scenario in Czech Family Law
di: Harasta, Jakub, et al.
Pubblicazione: (2026)
di: Harasta, Jakub, et al.
Pubblicazione: (2026)
Artificial Intelligence and Civil Discourse: How LLMs Moderate Climate Change Conversations
di: Fan, Wenlu, et al.
Pubblicazione: (2025)
di: Fan, Wenlu, et al.
Pubblicazione: (2025)
Minding the Politeness Gap in Cross-cultural Communication
di: Machino, Yuka, et al.
Pubblicazione: (2025)
di: Machino, Yuka, et al.
Pubblicazione: (2025)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
di: Dahl, Matthew, et al.
Pubblicazione: (2024)
di: Dahl, Matthew, et al.
Pubblicazione: (2024)
Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
di: Gajewska, Ewelina, et al.
Pubblicazione: (2025)
Caveat Lector: Large Language Models in Legal Practice
di: Mik, Eliza
Pubblicazione: (2024)
di: Mik, Eliza
Pubblicazione: (2024)
Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensions
di: Patil, Parth, et al.
Pubblicazione: (2026)
di: Patil, Parth, et al.
Pubblicazione: (2026)
How Large Language Models (LLMs) Extrapolate: From Guided Missiles to Guided Prompts
di: Cao, Xuenan
Pubblicazione: (2024)
di: Cao, Xuenan
Pubblicazione: (2024)
Bridging Legal Interpretation and Formal Logic: Faithfulness, Assumption, and the Future of AI Legal Reasoning
di: Wang, Olivia Peiyu, et al.
Pubblicazione: (2026)
di: Wang, Olivia Peiyu, et al.
Pubblicazione: (2026)
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
di: Shankar, Hari, et al.
Pubblicazione: (2026)
di: Shankar, Hari, et al.
Pubblicazione: (2026)
From Perceptions to Decisions: Wildfire Evacuation Decision Prediction with Behavioral Theory-informed LLMs
di: Chen, Ruxiao, et al.
Pubblicazione: (2025)
di: Chen, Ruxiao, et al.
Pubblicazione: (2025)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
di: Magesh, Varun, et al.
Pubblicazione: (2024)
di: Magesh, Varun, et al.
Pubblicazione: (2024)
How Far Are LLMs from Believable AI? A Benchmark for Evaluating the Believability of Human Behavior Simulation
di: Xiao, Yang, et al.
Pubblicazione: (2023)
di: Xiao, Yang, et al.
Pubblicazione: (2023)
Red Lines and Grey Zones in the Fog of War: Benchmarking Legal Risk, Moral Harm, and Regional Bias in Large Language Model Military Decision-Making
di: Drinkall, Toby
Pubblicazione: (2025)
di: Drinkall, Toby
Pubblicazione: (2025)
Persuadability and LLMs as Legal Decision Tools
di: Suttle, Oisin, et al.
Pubblicazione: (2026)
di: Suttle, Oisin, et al.
Pubblicazione: (2026)
Few-shot Hate Speech Detection Based on the MindSpore Framework
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
di: Qin, Zhenkai, et al.
Pubblicazione: (2025)
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models
di: Hijazi, Faris, et al.
Pubblicazione: (2024)
di: Hijazi, Faris, et al.
Pubblicazione: (2024)
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
di: Hou, Abe Bohan, et al.
Pubblicazione: (2024)
di: Hou, Abe Bohan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Prompting for Policy: Forecasting Macroeconomic Scenarios with Synthetic LLM Personas
di: Iadisernia, Giulia, et al.
Pubblicazione: (2025) -
Chat Bankman-Fried: an Exploration of LLM Alignment in Finance
di: Biancotti, Claudia, et al.
Pubblicazione: (2024) -
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
di: Camassa, Carolina, et al.
Pubblicazione: (2026) -
LLMs Provide Unstable Answers to Legal Questions
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025) -
Towards Next-Generation Medical Agent: How o1 is Reshaping Decision-Making in Medical Scenarios
di: Xu, Shaochen, et al.
Pubblicazione: (2024)