Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Li, Grabmair, Matthias, Gray, Morgan, Ashley, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
Do LLMs Truly Understand When a Precedent Is Overruled?
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
von: Vernie, Julius, et al.
Veröffentlicht: (2026)
von: Vernie, Julius, et al.
Veröffentlicht: (2026)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
BMAM: Brain-inspired Multi-Agent Memory Framework
von: Li, Yang, et al.
Veröffentlicht: (2026)
von: Li, Yang, et al.
Veröffentlicht: (2026)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
von: Amamou, Hazem, et al.
Veröffentlicht: (2026)
von: Amamou, Hazem, et al.
Veröffentlicht: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration
von: Lin, Shuhang, et al.
Veröffentlicht: (2026)
von: Lin, Shuhang, et al.
Veröffentlicht: (2026)
AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2026)
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2026)
Hopscotch: Discovering and Skipping Redundancies in Language Models
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
von: Cao, Deyu, et al.
Veröffentlicht: (2025)
von: Cao, Deyu, et al.
Veröffentlicht: (2025)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
DPDisc: From Factoid Questions to Data Product Requests for Open-World Data Product Discovery over Tables and Text
von: Zhang, Liangliang, et al.
Veröffentlicht: (2025)
von: Zhang, Liangliang, et al.
Veröffentlicht: (2025)
CAPE: Corrective Actions from Precondition Errors using Large Language Models
von: Raman, Shreyas Sundara, et al.
Veröffentlicht: (2022)
von: Raman, Shreyas Sundara, et al.
Veröffentlicht: (2022)
LLMs for Legal Subsumption in German Employment Contracts
von: Wardas, Oliver, et al.
Veröffentlicht: (2025)
von: Wardas, Oliver, et al.
Veröffentlicht: (2025)
Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data
von: Borisov, Vadim
Veröffentlicht: (2026)
von: Borisov, Vadim
Veröffentlicht: (2026)
Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor
von: Nwokocha, Caleb Princewill
Veröffentlicht: (2022)
von: Nwokocha, Caleb Princewill
Veröffentlicht: (2022)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
From Capabilities to Performance: Evaluating Key Functional Properties of LLM Architectures in Penetration Testing
von: Huang, Lanxiao, et al.
Veröffentlicht: (2025)
von: Huang, Lanxiao, et al.
Veröffentlicht: (2025)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
von: Liu, Xiaoou, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoou, et al.
Veröffentlicht: (2026)
Exploring State Tracking Capabilities of Large Language Models
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
Fast Quiet-STaR: Thinking Without Thought Tokens
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
A Computational Approach to Modeling Conversational Systems: Analyzing Large-Scale Quasi-Patterned Dialogue Flows
von: Ammar, Mohamed Achref Ben, et al.
Veröffentlicht: (2025)
von: Ammar, Mohamed Achref Ben, et al.
Veröffentlicht: (2025)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
von: Guo, Willis, et al.
Veröffentlicht: (2024)
von: Guo, Willis, et al.
Veröffentlicht: (2024)
Judgment2vec: Apply Graph Analytics to Searching and Recommendation of Similar Judgments
von: Shao, Hsuan-Lei
Veröffentlicht: (2024)
von: Shao, Hsuan-Lei
Veröffentlicht: (2024)
Ontology Learning with LLMs: A Benchmark Study on Axiom Identification
von: Bakker, Roos M., et al.
Veröffentlicht: (2025)
von: Bakker, Roos M., et al.
Veröffentlicht: (2025)
Sliced-Wasserstein Distribution Alignment Loss Improves the Ultra-Low-Bit Quantization of Large Language Models
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
Can LLMs Compute with Reasons?
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2026)
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2026)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
Ähnliche Einträge
-
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026) -
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
von: Lei, Xiang, et al.
Veröffentlicht: (2025) -
Do LLMs Truly Understand When a Precedent Is Overruled?
von: Zhang, Li, et al.
Veröffentlicht: (2025) -
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026) -
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
von: Vernie, Julius, et al.
Veröffentlicht: (2026)