Comparison of Unsupervised Metrics for Evaluating Judicial Decision Extraction
Fuente:
arXiv
Saved in:
| Main Authors: | Litvak, Ivan Leonidovich, Kostin, Anton, Lashkin, Fedor, Maksiyan, Tatiana, Lagutin, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
by: Yang, Zhenye, et al.
Published: (2025)
by: Yang, Zhenye, et al.
Published: (2025)
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
by: Du, Yuchen, et al.
Published: (2026)
by: Du, Yuchen, et al.
Published: (2026)
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
HiFi-RAG: Hierarchical Content Filtering and Two-Pass Generation for Open-Domain RAG
by: Nuengsigkapian, Cattalyya
Published: (2025)
by: Nuengsigkapian, Cattalyya
Published: (2025)
Retrieval Augmented Thought Process for Private Data Handling in Healthcare
by: Pouplin, Thomas, et al.
Published: (2024)
by: Pouplin, Thomas, et al.
Published: (2024)
BERTopic for Topic Modeling of Hindi Short Texts: A Comparative Study
by: Mutsaddi, Atharva, et al.
Published: (2025)
by: Mutsaddi, Atharva, et al.
Published: (2025)
HySemRAG: A Hybrid Semantic Retrieval-Augmented Generation Framework for Automated Literature Synthesis and Methodological Gap Analysis
by: Godinez, Alejandro
Published: (2025)
by: Godinez, Alejandro
Published: (2025)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
by: Pawar, Tejas, et al.
Published: (2025)
by: Pawar, Tejas, et al.
Published: (2025)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
by: Jiang, Yuning, et al.
Published: (2025)
by: Jiang, Yuning, et al.
Published: (2025)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
by: Shi, Kainan, et al.
Published: (2025)
by: Shi, Kainan, et al.
Published: (2025)
Logging Policy Design for Off-Policy Evaluation
by: Douglas, Connor, et al.
Published: (2026)
by: Douglas, Connor, et al.
Published: (2026)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
by: Egami, Shusaku, et al.
Published: (2026)
by: Egami, Shusaku, et al.
Published: (2026)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
by: Iana, Andreea, et al.
Published: (2023)
by: Iana, Andreea, et al.
Published: (2023)
Session Context Embedding for Intent Understanding in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Scaling Multilingual Semantic Search in Uber Eats Delivery
by: Ling, Bo, et al.
Published: (2026)
by: Ling, Bo, et al.
Published: (2026)
Does UMBRELA Work on Other LLMs?
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
by: Oruesagasti, Julen
Published: (2026)
by: Oruesagasti, Julen
Published: (2026)
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
by: Koutsiaris, Christos
Published: (2026)
by: Koutsiaris, Christos
Published: (2026)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
by: Che, Jiarui, et al.
Published: (2026)
by: Che, Jiarui, et al.
Published: (2026)
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025)
by: Cahoon, Joyce, et al.
Published: (2025)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026)
by: Liu, Jianan, et al.
Published: (2026)
Method for Aggregating Unstructured Data Using Large Language Models
by: Lazebnyi, Vsevolod, et al.
Published: (2026)
by: Lazebnyi, Vsevolod, et al.
Published: (2026)
Enhancing Document AI Data Generation Through Graph-Based Synthetic Layouts
by: Agarwal, Amit, et al.
Published: (2024)
by: Agarwal, Amit, et al.
Published: (2024)
Biomedical systems biology workflow orchestration and execution with PoSyMed
by: Süwer, Simon, et al.
Published: (2026)
by: Süwer, Simon, et al.
Published: (2026)
Conversational No-code, Multi-agentic Disease Module Identification and Drug Repurposing Prediction with ChatDRex
by: Süwer, Simon, et al.
Published: (2025)
by: Süwer, Simon, et al.
Published: (2025)
Tulip Agent -- Enabling LLM-Based Agents to Solve Tasks Using Large Tool Libraries
by: Ocker, Felix, et al.
Published: (2024)
by: Ocker, Felix, et al.
Published: (2024)
From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
by: Gill, Gurbinder, et al.
Published: (2025)
by: Gill, Gurbinder, et al.
Published: (2025)
Uncovering the Limitations of Query Performance Prediction: Failures, Insights, and Implications for Selective Query Processing
by: Chifu, Adrian-Gabriel, et al.
Published: (2025)
by: Chifu, Adrian-Gabriel, et al.
Published: (2025)
Evaluating Perspectival Biases in Cross-Modal Retrieval
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
Smart ETL and LLM-based contents classification: the European Smart Tourism Tools Observatory experience
by: Cosme, Diogo, et al.
Published: (2024)
by: Cosme, Diogo, et al.
Published: (2024)
LLM Reasoning for Cold-Start Item Recommendation
by: Li, Shijun, et al.
Published: (2025)
by: Li, Shijun, et al.
Published: (2025)
The Reasoning-Creativity Trade-off: Toward Creativity-Driven Problem Solving
by: Luyten, Max Ruiz, et al.
Published: (2026)
by: Luyten, Max Ruiz, et al.
Published: (2026)
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
by: Song, Seyoung
Published: (2025)
by: Song, Seyoung
Published: (2025)
DySK-Attn: A Framework for Efficient, Real-Time Knowledge Updating in Large Language Models via Dynamic Sparse Knowledge Attention
by: Khan, Kabir, et al.
Published: (2025)
by: Khan, Kabir, et al.
Published: (2025)
Task Memory Engine: Spatial Memory for Robust Multi-Step LLM Agents
by: Ye, Ye
Published: (2025)
by: Ye, Ye
Published: (2025)
Similar Items
-
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
by: Bose, Joy
Published: (2026) -
STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
by: Yang, Zhenye, et al.
Published: (2025) -
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
by: Du, Yuchen, et al.
Published: (2026) -
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
by: Chen, Hao, et al.
Published: (2025) -
HiFi-RAG: Hierarchical Content Filtering and Two-Pass Generation for Open-Domain RAG
by: Nuengsigkapian, Cattalyya
Published: (2025)