Can Large Language Models Grasp Legal Theories? Enhance Legal Reasoning with Insights from Multi-Agent Collaboration
Fuente:
arXiv
Guardado en:
| Autores principales: | Yuan, Weikang, Cao, Junjie, Jiang, Zhuoren, Kang, Yangyang, Lin, Jun, Song, Kaisong, lin, tianqianjin, Yan, Pengwei, Sun, Changlong, Liu, Xiaozhong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation
por: Yuan, Weikang, et al.
Publicado: (2025)
por: Yuan, Weikang, et al.
Publicado: (2025)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
por: Ovcharov, Volodymyr
Publicado: (2026)
por: Ovcharov, Volodymyr
Publicado: (2026)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
por: Bouchekif, Abdessalam, et al.
Publicado: (2026)
por: Bouchekif, Abdessalam, et al.
Publicado: (2026)
BLT: Can Large Language Models Handle Basic Legal Text?
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
por: Wang, Yongjie, et al.
Publicado: (2025)
por: Wang, Yongjie, et al.
Publicado: (2025)
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
por: van der Meer, Virgill, et al.
Publicado: (2026)
por: van der Meer, Virgill, et al.
Publicado: (2026)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
por: Zhang, Li, et al.
Publicado: (2025)
por: Zhang, Li, et al.
Publicado: (2025)
Assessing Large Language Models on Islamic Legal Reasoning: Evidence from Inheritance Law Evaluation
por: Bouchekif, Abdessalam, et al.
Publicado: (2025)
por: Bouchekif, Abdessalam, et al.
Publicado: (2025)
LLMs for Legal Subsumption in German Employment Contracts
por: Wardas, Oliver, et al.
Publicado: (2025)
por: Wardas, Oliver, et al.
Publicado: (2025)
From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text
por: Le, Van-Truong
Publicado: (2026)
por: Le, Van-Truong
Publicado: (2026)
LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification
por: Neto, Pedro Barbosa de Carvalho
Publicado: (2026)
por: Neto, Pedro Barbosa de Carvalho
Publicado: (2026)
Unlocking Legal Knowledge with Multi-Layered Embedding-Based Retrieval
por: Lima, João Alberto de Oliveira
Publicado: (2024)
por: Lima, João Alberto de Oliveira
Publicado: (2024)
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
por: Elganayni, Mohamed Hesham, et al.
Publicado: (2026)
por: Elganayni, Mohamed Hesham, et al.
Publicado: (2026)
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
por: Zhang, Yiqing, et al.
Publicado: (2026)
por: Zhang, Yiqing, et al.
Publicado: (2026)
Can LLM Graph Reasoning Generalize beyond Pattern Memorization?
por: Zhang, Yizhuo, et al.
Publicado: (2024)
por: Zhang, Yizhuo, et al.
Publicado: (2024)
Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs
por: Ovcharov, Volodymyr
Publicado: (2026)
por: Ovcharov, Volodymyr
Publicado: (2026)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
por: Zhang, Li, et al.
Publicado: (2026)
por: Zhang, Li, et al.
Publicado: (2026)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
por: Fernandes, Rean, et al.
Publicado: (2025)
por: Fernandes, Rean, et al.
Publicado: (2025)
Swiss-Bench SBP-002: A Frontier Model Comparison on Swiss Legal and Regulatory Tasks
por: Uenal, Fatih
Publicado: (2026)
por: Uenal, Fatih
Publicado: (2026)
MapAgent: A Hierarchical Agent for Geospatial Reasoning with Dynamic Map Tool Integration
por: Hasan, Md Hasebul, et al.
Publicado: (2025)
por: Hasan, Md Hasebul, et al.
Publicado: (2025)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
por: Souza, Débora, et al.
Publicado: (2026)
por: Souza, Débora, et al.
Publicado: (2026)
LangGFM: A Large Language Model Alone Can be a Powerful Graph Foundation Model
por: Lin, Tianqianjin, et al.
Publicado: (2024)
por: Lin, Tianqianjin, et al.
Publicado: (2024)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
por: Pradhan, Anu, et al.
Publicado: (2025)
por: Pradhan, Anu, et al.
Publicado: (2025)
Maat: The Agentic Legal Research Assistant for Competition Protection
por: Mounir, Basant, et al.
Publicado: (2026)
por: Mounir, Basant, et al.
Publicado: (2026)
The Judge Variable: Challenging Judge-Agnostic Legal Judgment Prediction
por: Zambrano, Guillaume
Publicado: (2025)
por: Zambrano, Guillaume
Publicado: (2025)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
Can LLMs Compute with Reasons?
por: Sandilya, Harshit, et al.
Publicado: (2024)
por: Sandilya, Harshit, et al.
Publicado: (2024)
Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning
por: Vlachos, Angelos, et al.
Publicado: (2025)
por: Vlachos, Angelos, et al.
Publicado: (2025)
From Extraction to Synthesis: Entangled Heuristics for Agent-Augmented Strategic Reasoning
por: Ghisellini, Renato, et al.
Publicado: (2025)
por: Ghisellini, Renato, et al.
Publicado: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
LegalGuardian: A Privacy-Preserving Framework for Secure Integration of Large Language Models in Legal Practice
por: Demir, M. Mikail, et al.
Publicado: (2025)
por: Demir, M. Mikail, et al.
Publicado: (2025)
RAVR: Reference-Answer-guided Variational Reasoning for Large Language Models
por: Lin, Tianqianjin, et al.
Publicado: (2025)
por: Lin, Tianqianjin, et al.
Publicado: (2025)
The Unified Cognitive Consciousness Theory for Language Models: Anchoring Semantics, Thresholds of Activation, and Emergent Reasoning
por: Chang, Edward Y., et al.
Publicado: (2025)
por: Chang, Edward Y., et al.
Publicado: (2025)
Efficient Reasoning via Thought-Training and Thought-Free Inference
por: Wu, Canhui, et al.
Publicado: (2025)
por: Wu, Canhui, et al.
Publicado: (2025)
AI Can Learn Scientific Taste
por: Tong, Jingqi, et al.
Publicado: (2026)
por: Tong, Jingqi, et al.
Publicado: (2026)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
por: Han, Xudong, et al.
Publicado: (2025)
por: Han, Xudong, et al.
Publicado: (2025)
CoE: Collaborative Entropy for Uncertainty Quantification in Agentic Multi-LLM Systems
por: Sun, Kangkang, et al.
Publicado: (2026)
por: Sun, Kangkang, et al.
Publicado: (2026)
ContractBench: Can LLM Agents Preserve Observation Contracts?
por: Wang, Jicheng, et al.
Publicado: (2026)
por: Wang, Jicheng, et al.
Publicado: (2026)
Auditing Meta-Cognitive Hallucinations in Reasoning Large Language Models
por: Lu, Haolang, et al.
Publicado: (2025)
por: Lu, Haolang, et al.
Publicado: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
por: Saji, Alan, et al.
Publicado: (2025)
por: Saji, Alan, et al.
Publicado: (2025)
Ejemplares similares
-
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation
por: Yuan, Weikang, et al.
Publicado: (2025) -
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
por: Ovcharov, Volodymyr
Publicado: (2026) -
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
por: Bouchekif, Abdessalam, et al.
Publicado: (2026) -
BLT: Can Large Language Models Handle Basic Legal Text?
por: Blair-Stanek, Andrew, et al.
Publicado: (2023) -
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
por: Wang, Yongjie, et al.
Publicado: (2025)