LexGenius: An Expert-Level Benchmark for Large Language Models in Legal General Intelligence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Wenjin, Luo, Haoran, Feng, Xin, Ji, Xiang, Zhou, Lijuan, Mao, Rui, Wang, Jiapu, Pan, Shirui, Cambria, Erik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning
von: Liu, Wenjin, et al.
Veröffentlicht: (2025)
von: Liu, Wenjin, et al.
Veröffentlicht: (2025)
A Systematic Analysis of Biases in Large Language Models
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
von: Zhang, Xulang, et al.
Veröffentlicht: (2025)
Unifying Large Language Models and Knowledge Graphs: A Roadmap
von: Pan, Shirui, et al.
Veröffentlicht: (2023)
von: Pan, Shirui, et al.
Veröffentlicht: (2023)
FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
von: Zhang, Mingda, et al.
Veröffentlicht: (2026)
von: Zhang, Mingda, et al.
Veröffentlicht: (2026)
A Survey of Large Language Models for Healthcare: from Data, Technology, and Applications to Accountability and Ethics
von: He, Kai, et al.
Veröffentlicht: (2023)
von: He, Kai, et al.
Veröffentlicht: (2023)
Deriving Strategic Market Insights with Large Language Models: A Benchmark for Forward Counterfactual Generation
von: Ong, Keane, et al.
Veröffentlicht: (2025)
von: Ong, Keane, et al.
Veröffentlicht: (2025)
LexEval: A Comprehensive Chinese Legal Benchmark for Evaluating Large Language Models
von: Li, Haitao, et al.
Veröffentlicht: (2024)
von: Li, Haitao, et al.
Veröffentlicht: (2024)
LexSumm and LexT5: Benchmarking and Modeling Legal Summarization Tasks in English
von: Santosh, T. Y. S. S., et al.
Veröffentlicht: (2024)
von: Santosh, T. Y. S. S., et al.
Veröffentlicht: (2024)
LLMdoctor: Token-Level Flow-Guided Preference Optimization for Efficient Test-Time Alignment of Large Language Models
von: Shen, Tiesunlong, et al.
Veröffentlicht: (2026)
von: Shen, Tiesunlong, et al.
Veröffentlicht: (2026)
SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration
von: Zhang, Mingda, et al.
Veröffentlicht: (2026)
von: Zhang, Mingda, et al.
Veröffentlicht: (2026)
LexTime: A Benchmark for Temporal Ordering of Legal Events
von: Barale, Claire, et al.
Veröffentlicht: (2025)
von: Barale, Claire, et al.
Veröffentlicht: (2025)
Large Language Models-guided Dynamic Adaptation for Temporal Knowledge Graph Reasoning
von: Wang, Jiapu, et al.
Veröffentlicht: (2024)
von: Wang, Jiapu, et al.
Veröffentlicht: (2024)
Individualized Cognitive Simulation in Large Language Models: Evaluating Different Cognitive Representation Methods
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
Logical Reasoning over Natural Language as Knowledge Representation: A Survey
von: Yang, Zonglin, et al.
Veröffentlicht: (2023)
von: Yang, Zonglin, et al.
Veröffentlicht: (2023)
Effective Instruction Parsing Plugin for Complex Logical Query Answering on Knowledge Graphs
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2024)
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2024)
LexRel: Benchmarking Legal Relation Extraction for Chinese Civil Cases
von: Cai, Yida, et al.
Veröffentlicht: (2025)
von: Cai, Yida, et al.
Veröffentlicht: (2025)
ClimaEmpact: Domain-Aligned Small Language Models and Datasets for Extreme Weather Analytics
von: Varshney, Deeksha, et al.
Veröffentlicht: (2025)
von: Varshney, Deeksha, et al.
Veröffentlicht: (2025)
Knowledge Reasoning Language Model: Unifying Knowledge and Language for Inductive Knowledge Graph Reasoning
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2025)
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2025)
Enhancing Language Models for Robust Greenwashing Detection
von: Braun, Neil Heinrich, et al.
Veröffentlicht: (2026)
von: Braun, Neil Heinrich, et al.
Veröffentlicht: (2026)
GLARE: Guided LexRank for Advanced Retrieval in Legal Analysis
von: Gregório, Fabio, et al.
Veröffentlicht: (2024)
von: Gregório, Fabio, et al.
Veröffentlicht: (2024)
Towards Faithful Natural Language Explanations: A Study Using Activation Patching in Large Language Models
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2024)
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2024)
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)
UniER: A Unified Benchmark for Item-level and Path-level Exercise Recommendation
von: Cheng, Xinghe, et al.
Veröffentlicht: (2026)
von: Cheng, Xinghe, et al.
Veröffentlicht: (2026)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
von: Li, Haitao, et al.
Veröffentlicht: (2025)
von: Li, Haitao, et al.
Veröffentlicht: (2025)
BenchBench: Benchmarking Automated Benchmark Generation
von: Zheng, Yandan, et al.
Veröffentlicht: (2026)
von: Zheng, Yandan, et al.
Veröffentlicht: (2026)
Cumulative Path-Level Semantic Reasoning for Inductive Knowledge Graph Completion
von: Wang, Jiapu, et al.
Veröffentlicht: (2026)
von: Wang, Jiapu, et al.
Veröffentlicht: (2026)
Bridging Minds and Machines: Toward an Integration of AI and Cognitive Science
von: Mao, Rui, et al.
Veröffentlicht: (2025)
von: Mao, Rui, et al.
Veröffentlicht: (2025)
Towards Robust ESG Analysis Against Greenwashing Risks: Aspect-Action Analysis with Cross-Category Generalization
von: Ong, Keane, et al.
Veröffentlicht: (2025)
von: Ong, Keane, et al.
Veröffentlicht: (2025)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
von: Mao, Rui, et al.
Veröffentlicht: (2023)
von: Mao, Rui, et al.
Veröffentlicht: (2023)
Large Language Models for Few-Shot Named Entity Recognition
von: Zhao, Yufei, et al.
Veröffentlicht: (2018)
von: Zhao, Yufei, et al.
Veröffentlicht: (2018)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling
von: Wang, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Wang, Jiaxuan, et al.
Veröffentlicht: (2026)
Unifying Unsupervised Graph-Level Anomaly Detection and Out-of-Distribution Detection: A Benchmark
von: Wang, Yili, et al.
Veröffentlicht: (2024)
von: Wang, Yili, et al.
Veröffentlicht: (2024)
ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence
von: Ma, Menghe, et al.
Veröffentlicht: (2026)
von: Ma, Menghe, et al.
Veröffentlicht: (2026)
GraphRAG-Induced Dual Knowledge Structure Graphs for Personalized Learning Path Recommendation
von: Cheng, Xinghe, et al.
Veröffentlicht: (2025)
von: Cheng, Xinghe, et al.
Veröffentlicht: (2025)
LexAbSumm: Aspect-based Summarization of Legal Decisions
von: Santosh, T. Y. S. S, et al.
Veröffentlicht: (2024)
von: Santosh, T. Y. S. S, et al.
Veröffentlicht: (2024)
Molecular dynamics and optimization studies of horse prion protein wild type and its S167D mutant
von: Zhang, Jiapu
Veröffentlicht: (2021)
von: Zhang, Jiapu
Veröffentlicht: (2021)
Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
IME: Integrating Multi-curvature Shared and Specific Embedding for Temporal Knowledge Graph Completion
von: Wang, Jiapu, et al.
Veröffentlicht: (2024)
von: Wang, Jiapu, et al.
Veröffentlicht: (2024)
Explainable Natural Language Processing for Corporate Sustainability Analysis
von: Ong, Keane, et al.
Veröffentlicht: (2024)
von: Ong, Keane, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning
von: Liu, Wenjin, et al.
Veröffentlicht: (2025) -
A Systematic Analysis of Biases in Large Language Models
von: Zhang, Xulang, et al.
Veröffentlicht: (2025) -
Unifying Large Language Models and Knowledge Graphs: A Roadmap
von: Pan, Shirui, et al.
Veröffentlicht: (2023) -
FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
von: Zhang, Mingda, et al.
Veröffentlicht: (2026) -
A Survey of Large Language Models for Healthcare: from Data, Technology, and Applications to Accountability and Ethics
von: He, Kai, et al.
Veröffentlicht: (2023)