Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Li, Grabmair, Matthias, Gray, Morgan, Ashley, Kevin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
di: Zhang, Li, et al.
Pubblicazione: (2026)
di: Zhang, Li, et al.
Pubblicazione: (2026)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
di: Lei, Xiang, et al.
Pubblicazione: (2025)
di: Lei, Xiang, et al.
Pubblicazione: (2025)
Do LLMs Truly Understand When a Precedent Is Overruled?
di: Zhang, Li, et al.
Pubblicazione: (2025)
di: Zhang, Li, et al.
Pubblicazione: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
di: Haque, Md. Asraful, et al.
Pubblicazione: (2026)
di: Haque, Md. Asraful, et al.
Pubblicazione: (2026)
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
di: Vernie, Julius, et al.
Pubblicazione: (2026)
di: Vernie, Julius, et al.
Pubblicazione: (2026)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
di: Da, Longchao, et al.
Pubblicazione: (2025)
di: Da, Longchao, et al.
Pubblicazione: (2025)
BMAM: Brain-inspired Multi-Agent Memory Framework
di: Li, Yang, et al.
Pubblicazione: (2026)
di: Li, Yang, et al.
Pubblicazione: (2026)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
di: Amamou, Hazem, et al.
Pubblicazione: (2026)
di: Amamou, Hazem, et al.
Pubblicazione: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
di: Da, Longchao, et al.
Pubblicazione: (2025)
di: Da, Longchao, et al.
Pubblicazione: (2025)
Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration
di: Lin, Shuhang, et al.
Pubblicazione: (2026)
di: Lin, Shuhang, et al.
Pubblicazione: (2026)
AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs
di: Perera, Manoj Madushanka, et al.
Pubblicazione: (2026)
di: Perera, Manoj Madushanka, et al.
Pubblicazione: (2026)
Hopscotch: Discovering and Skipping Redundancies in Language Models
di: Eyceoz, Mustafa, et al.
Pubblicazione: (2025)
di: Eyceoz, Mustafa, et al.
Pubblicazione: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
di: Cao, Deyu, et al.
Pubblicazione: (2025)
di: Cao, Deyu, et al.
Pubblicazione: (2025)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
di: Kaiser, Daniel, et al.
Pubblicazione: (2025)
di: Kaiser, Daniel, et al.
Pubblicazione: (2025)
DPDisc: From Factoid Questions to Data Product Requests for Open-World Data Product Discovery over Tables and Text
di: Zhang, Liangliang, et al.
Pubblicazione: (2025)
di: Zhang, Liangliang, et al.
Pubblicazione: (2025)
CAPE: Corrective Actions from Precondition Errors using Large Language Models
di: Raman, Shreyas Sundara, et al.
Pubblicazione: (2022)
di: Raman, Shreyas Sundara, et al.
Pubblicazione: (2022)
LLMs for Legal Subsumption in German Employment Contracts
di: Wardas, Oliver, et al.
Pubblicazione: (2025)
di: Wardas, Oliver, et al.
Pubblicazione: (2025)
Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data
di: Borisov, Vadim
Pubblicazione: (2026)
di: Borisov, Vadim
Pubblicazione: (2026)
Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor
di: Nwokocha, Caleb Princewill
Pubblicazione: (2022)
di: Nwokocha, Caleb Princewill
Pubblicazione: (2022)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
di: Chen, Tiejin, et al.
Pubblicazione: (2026)
di: Chen, Tiejin, et al.
Pubblicazione: (2026)
From Capabilities to Performance: Evaluating Key Functional Properties of LLM Architectures in Penetration Testing
di: Huang, Lanxiao, et al.
Pubblicazione: (2025)
di: Huang, Lanxiao, et al.
Pubblicazione: (2025)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
di: Liu, Xiaoou, et al.
Pubblicazione: (2026)
di: Liu, Xiaoou, et al.
Pubblicazione: (2026)
Exploring State Tracking Capabilities of Large Language Models
di: Rezaee, Kiamehr, et al.
Pubblicazione: (2025)
di: Rezaee, Kiamehr, et al.
Pubblicazione: (2025)
Fast Quiet-STaR: Thinking Without Thought Tokens
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
A Computational Approach to Modeling Conversational Systems: Analyzing Large-Scale Quasi-Patterned Dialogue Flows
di: Ammar, Mohamed Achref Ben, et al.
Pubblicazione: (2025)
di: Ammar, Mohamed Achref Ben, et al.
Pubblicazione: (2025)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
di: Guo, Willis, et al.
Pubblicazione: (2024)
di: Guo, Willis, et al.
Pubblicazione: (2024)
Judgment2vec: Apply Graph Analytics to Searching and Recommendation of Similar Judgments
di: Shao, Hsuan-Lei
Pubblicazione: (2024)
di: Shao, Hsuan-Lei
Pubblicazione: (2024)
Ontology Learning with LLMs: A Benchmark Study on Axiom Identification
di: Bakker, Roos M., et al.
Pubblicazione: (2025)
di: Bakker, Roos M., et al.
Pubblicazione: (2025)
Sliced-Wasserstein Distribution Alignment Loss Improves the Ultra-Low-Bit Quantization of Large Language Models
di: Cao, Deyu, et al.
Pubblicazione: (2026)
di: Cao, Deyu, et al.
Pubblicazione: (2026)
Can LLMs Compute with Reasons?
di: Sandilya, Harshit, et al.
Pubblicazione: (2024)
di: Sandilya, Harshit, et al.
Pubblicazione: (2024)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
di: Tu, Songjun, et al.
Pubblicazione: (2026)
di: Tu, Songjun, et al.
Pubblicazione: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
di: Lasbordes, Maxence, et al.
Pubblicazione: (2026)
di: Lasbordes, Maxence, et al.
Pubblicazione: (2026)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
di: Nieth, Björn, et al.
Pubblicazione: (2026)
di: Nieth, Björn, et al.
Pubblicazione: (2026)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
di: Sun, Jingyi, et al.
Pubblicazione: (2024)
di: Sun, Jingyi, et al.
Pubblicazione: (2024)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
di: Wang, Zhilin, et al.
Pubblicazione: (2026)
di: Wang, Zhilin, et al.
Pubblicazione: (2026)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2023)
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2023)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
di: Pranida, Salsabila Zahirah, et al.
Pubblicazione: (2025)
di: Pranida, Salsabila Zahirah, et al.
Pubblicazione: (2025)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
di: Johnson, Warren
Pubblicazione: (2026)
di: Johnson, Warren
Pubblicazione: (2026)
Documenti analoghi
-
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
di: Zhang, Li, et al.
Pubblicazione: (2026) -
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
di: Lei, Xiang, et al.
Pubblicazione: (2025) -
Do LLMs Truly Understand When a Precedent Is Overruled?
di: Zhang, Li, et al.
Pubblicazione: (2025) -
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
di: Haque, Md. Asraful, et al.
Pubblicazione: (2026) -
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
di: Vernie, Julius, et al.
Pubblicazione: (2026)