Do LLMs Truly Understand When a Precedent Is Overruled?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Li, Savelka, Jaromir, Ashley, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
von: Lei, Xiang, et al.
Veröffentlicht: (2025)
Hopscotch: Discovering and Skipping Redundancies in Language Models
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
Ontology Learning with LLMs: A Benchmark Study on Axiom Identification
von: Bakker, Roos M., et al.
Veröffentlicht: (2025)
von: Bakker, Roos M., et al.
Veröffentlicht: (2025)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
CAPE: Corrective Actions from Precondition Errors using Large Language Models
von: Raman, Shreyas Sundara, et al.
Veröffentlicht: (2022)
von: Raman, Shreyas Sundara, et al.
Veröffentlicht: (2022)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
von: Liu, Xiaoou, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoou, et al.
Veröffentlicht: (2026)
Pun Unintended: LLMs and the Illusion of Humor Understanding
von: Zangari, Alessandro, et al.
Veröffentlicht: (2025)
von: Zangari, Alessandro, et al.
Veröffentlicht: (2025)
BMAM: Brain-inspired Multi-Agent Memory Framework
von: Li, Yang, et al.
Veröffentlicht: (2026)
von: Li, Yang, et al.
Veröffentlicht: (2026)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
von: Amamou, Hazem, et al.
Veröffentlicht: (2026)
von: Amamou, Hazem, et al.
Veröffentlicht: (2026)
Large Language Models Can Better Understand Knowledge Graphs Than We Thought
von: Dai, Xinbang, et al.
Veröffentlicht: (2024)
von: Dai, Xinbang, et al.
Veröffentlicht: (2024)
Do LLMs have a Gender (Entropy) Bias?
von: Prabhune, Sonal, et al.
Veröffentlicht: (2025)
von: Prabhune, Sonal, et al.
Veröffentlicht: (2025)
Sliced-Wasserstein Distribution Alignment Loss Improves the Ultra-Low-Bit Quantization of Large Language Models
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
Learning When to Think: Shaping Adaptive Reasoning in R1-Style Models via Multi-Stage RL
von: Tu, Songjun, et al.
Veröffentlicht: (2025)
von: Tu, Songjun, et al.
Veröffentlicht: (2025)
Breaking Free Transformer Models: Task-specific Context Attribution Promises Improved Generalizability Without Fine-tuning Pre-trained LLMs
von: Tytarenko, Stepan, et al.
Veröffentlicht: (2024)
von: Tytarenko, Stepan, et al.
Veröffentlicht: (2024)
AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2026)
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2026)
Judgment2vec: Apply Graph Analytics to Searching and Recommendation of Similar Judgments
von: Shao, Hsuan-Lei
Veröffentlicht: (2024)
von: Shao, Hsuan-Lei
Veröffentlicht: (2024)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
Dynamic Demonstration Retrieval and Cognitive Understanding for Emotional Support Conversation
von: Xu, Zhe, et al.
Veröffentlicht: (2024)
von: Xu, Zhe, et al.
Veröffentlicht: (2024)
Causally Grounded Mechanistic Interpretability for LLMs with Faithful Natural-Language Explanations
von: Mahale, Ajay Pravin
Veröffentlicht: (2026)
von: Mahale, Ajay Pravin
Veröffentlicht: (2026)
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
von: Fang, Xi, et al.
Veröffentlicht: (2025)
von: Fang, Xi, et al.
Veröffentlicht: (2025)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English
von: Anderson, Bryce, et al.
Veröffentlicht: (2025)
von: Anderson, Bryce, et al.
Veröffentlicht: (2025)
The CLEF-2025 CheckThat! Lab: Subjectivity, Fact-Checking, Claim Normalization, and Retrieval
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
Textual Data Bias Detection and Mitigation -- An Extensible Pipeline with Experimental Evaluation
von: Görge, Rebekka, et al.
Veröffentlicht: (2025)
von: Görge, Rebekka, et al.
Veröffentlicht: (2025)
Understanding the Effects of RLHF on the Quality and Detectability of LLM-Generated Texts
von: Xu, Beining, et al.
Veröffentlicht: (2025)
von: Xu, Beining, et al.
Veröffentlicht: (2025)
SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
Task Complexity Matters: An Empirical Study of Reasoning in LLMs for Sentiment Analysis
von: Huang, Donghao, et al.
Veröffentlicht: (2026)
von: Huang, Donghao, et al.
Veröffentlicht: (2026)
MORABLES: A Benchmark for Assessing Abstract Moral Reasoning in LLMs with Fables
von: Marcuzzo, Matteo, et al.
Veröffentlicht: (2025)
von: Marcuzzo, Matteo, et al.
Veröffentlicht: (2025)
Question Answering Over Spatio-Temporal Knowledge Graph
von: Dai, Xinbang, et al.
Veröffentlicht: (2024)
von: Dai, Xinbang, et al.
Veröffentlicht: (2024)
A Pluggable Common Sense-Enhanced Framework for Knowledge Graph Completion
von: Niu, Guanglin, et al.
Veröffentlicht: (2024)
von: Niu, Guanglin, et al.
Veröffentlicht: (2024)
Cognitive Workspace: Active Memory Management for LLMs -- An Empirical Study of Functional Infinite Context
von: An, Tao
Veröffentlicht: (2025)
von: An, Tao
Veröffentlicht: (2025)
UnifiedCrawl: Aggregated Common Crawl for Affordable Adaptation of LLMs on Low-Resource Languages
von: Tessema, Bethel Melesse, et al.
Veröffentlicht: (2024)
von: Tessema, Bethel Melesse, et al.
Veröffentlicht: (2024)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
von: Wang, Yongjie, et al.
Veröffentlicht: (2025)
von: Wang, Yongjie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026) -
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
von: Zhang, Li, et al.
Veröffentlicht: (2025) -
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
von: Haque, Md. Asraful, et al.
Veröffentlicht: (2026) -
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
von: Lei, Xiang, et al.
Veröffentlicht: (2025) -
Hopscotch: Discovering and Skipping Redundancies in Language Models
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)