RDF-Based Structured Quality Assessment Representation of Multilingual LLM Evaluations
Fuente:
arXiv
Salvato in:
| Autori principali: | Gwozdz, Jonas, Both, Andreas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Text-to-SPARQL Goes Beyond English: Multilingual Question Answering Over Knowledge Graphs through Human-Inspired Reasoning
di: Perevalov, Aleksandr, et al.
Pubblicazione: (2025)
di: Perevalov, Aleksandr, et al.
Pubblicazione: (2025)
SPARQL Query Generation with LLMs: Measuring the Impact of Training Data Memorization and Knowledge Injection
di: Gashkov, Aleksandr, et al.
Pubblicazione: (2025)
di: Gashkov, Aleksandr, et al.
Pubblicazione: (2025)
Auto-ARGUE: LLM-Based Report Generation Evaluation
di: Walden, William, et al.
Pubblicazione: (2025)
di: Walden, William, et al.
Pubblicazione: (2025)
Multilingual Information Retrieval with a Monolingual Knowledge Base
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)
Analysis and Detection of Multilingual Hate Speech Using Transformer Based Deep Learning
di: Das, Arijit, et al.
Pubblicazione: (2024)
di: Das, Arijit, et al.
Pubblicazione: (2024)
CLEF HIPE-2026: Evaluating Accurate and Efficient Person-Place Relation Extraction from Multilingual Historical Texts
di: Opitz, Juri, et al.
Pubblicazione: (2026)
di: Opitz, Juri, et al.
Pubblicazione: (2026)
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
di: Zhang, Ming, et al.
Pubblicazione: (2026)
di: Zhang, Ming, et al.
Pubblicazione: (2026)
Are LLMs Truly Multilingual? Exploring Zero-Shot Multilingual Capability of LLMs for Information Retrieval: An Italian Healthcare Use Case
di: Kembu, Vignesh Kumar, et al.
Pubblicazione: (2025)
di: Kembu, Vignesh Kumar, et al.
Pubblicazione: (2025)
No Free Lunch in Active Learning: LLM Embedding Quality Dictates Query Strategy Success
di: Rauch, Lukas, et al.
Pubblicazione: (2025)
di: Rauch, Lukas, et al.
Pubblicazione: (2025)
Prompt Compression in the Wild: Measuring Latency, Rate Adherence, and Quality for Faster LLM Inference
di: Kummer, Cornelius, et al.
Pubblicazione: (2026)
di: Kummer, Cornelius, et al.
Pubblicazione: (2026)
Enhancing LLM Medical Coding with Structured External Knowledge
di: Gan, Yidong, et al.
Pubblicazione: (2026)
di: Gan, Yidong, et al.
Pubblicazione: (2026)
MMTEB: Massive Multilingual Text Embedding Benchmark
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2025)
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2025)
KG-LLM-Bench: A Scalable Benchmark for Evaluating LLM Reasoning on Textualized Knowledge Graphs
di: Markowitz, Elan, et al.
Pubblicazione: (2025)
di: Markowitz, Elan, et al.
Pubblicazione: (2025)
Structure-R1: Dynamically Leveraging Structural Knowledge in LLM Reasoning through Reinforcement Learning
di: Wu, Junlin, et al.
Pubblicazione: (2025)
di: Wu, Junlin, et al.
Pubblicazione: (2025)
The 2021 Tokyo Olympics Multilingual News Article Dataset
di: Novak, Erik, et al.
Pubblicazione: (2025)
di: Novak, Erik, et al.
Pubblicazione: (2025)
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
di: Islamaj, Rezarta, et al.
Pubblicazione: (2026)
di: Islamaj, Rezarta, et al.
Pubblicazione: (2026)
Within-Document Event Coreference with BERT-Based Contextualized Representations
di: Ahmed, Shafiuddin Rehan, et al.
Pubblicazione: (2021)
di: Ahmed, Shafiuddin Rehan, et al.
Pubblicazione: (2021)
MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval
di: Khanghah, Kiarash Naghavi, et al.
Pubblicazione: (2026)
di: Khanghah, Kiarash Naghavi, et al.
Pubblicazione: (2026)
The Effect of Document Summarization on LLM-Based Relevance Judgments
di: Mohtadi, Samaneh, et al.
Pubblicazione: (2025)
di: Mohtadi, Samaneh, et al.
Pubblicazione: (2025)
Generative Query Expansion with Multilingual LLMs for Cross-Lingual Information Retrieval
di: Macmillan-Scott, Olivia, et al.
Pubblicazione: (2025)
di: Macmillan-Scott, Olivia, et al.
Pubblicazione: (2025)
Enhancing Multilingual Embeddings via Multi-Way Parallel Text Alignment
di: Fazili, Barah, et al.
Pubblicazione: (2026)
di: Fazili, Barah, et al.
Pubblicazione: (2026)
Enhancing LLM Generation with Knowledge Hypergraph for Evidence-Based Medicine
di: Dou, Chengfeng, et al.
Pubblicazione: (2025)
di: Dou, Chengfeng, et al.
Pubblicazione: (2025)
Scientific Paper Retrieval with LLM-Guided Semantic-Based Ranking
di: Zhang, Yunyi, et al.
Pubblicazione: (2025)
di: Zhang, Yunyi, et al.
Pubblicazione: (2025)
A Survey of LLM-based Deep Search Agents: Paradigm, Optimization, Evaluation, and Challenges
di: Xi, Yunjia, et al.
Pubblicazione: (2025)
di: Xi, Yunjia, et al.
Pubblicazione: (2025)
Support Evaluation for the TREC 2024 RAG Track: Comparing Human versus LLM Judges
di: Thakur, Nandan, et al.
Pubblicazione: (2025)
di: Thakur, Nandan, et al.
Pubblicazione: (2025)
Bridging Language Gaps: Advances in Cross-Lingual Information Retrieval with Multilingual LLMs
di: Goworek, Roksana, et al.
Pubblicazione: (2025)
di: Goworek, Roksana, et al.
Pubblicazione: (2025)
What Drives Cross-lingual Ranking? Retrieval Approaches with Multilingual Language Models
di: Goworek, Roksana, et al.
Pubblicazione: (2025)
di: Goworek, Roksana, et al.
Pubblicazione: (2025)
Faux Polyglot: A Study on Information Disparity in Multilingual Large Language Models
di: Sharma, Nikhil, et al.
Pubblicazione: (2024)
di: Sharma, Nikhil, et al.
Pubblicazione: (2024)
Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLP
di: Tan, Zhiyin, et al.
Pubblicazione: (2026)
di: Tan, Zhiyin, et al.
Pubblicazione: (2026)
WebFAQ: A Multilingual Collection of Natural Q&A Datasets for Dense Retrieval
di: Dinzinger, Michael, et al.
Pubblicazione: (2025)
di: Dinzinger, Michael, et al.
Pubblicazione: (2025)
Leveraging LLMs for Synthesizing Training Data Across Many Languages in Multilingual Dense Retrieval
di: Thakur, Nandan, et al.
Pubblicazione: (2023)
di: Thakur, Nandan, et al.
Pubblicazione: (2023)
LEMUR: A Corpus for Robust Fine-Tuning of Multilingual Law Embedding Models for Retrieval
di: Ahmadi, Narges Baba, et al.
Pubblicazione: (2026)
di: Ahmadi, Narges Baba, et al.
Pubblicazione: (2026)
Leveraging the Power of LLMs: A Fine-Tuning Approach for High-Quality Aspect-Based Summarization
di: Mullick, Ankan, et al.
Pubblicazione: (2024)
di: Mullick, Ankan, et al.
Pubblicazione: (2024)
Evaluating Structured Decoding for Text-to-Table Generation: Evidence from Three Datasets
di: Oestreich, Julian, et al.
Pubblicazione: (2025)
di: Oestreich, Julian, et al.
Pubblicazione: (2025)
SwasthLLM: a Unified Cross-Lingual, Multi-Task, and Meta-Learning Zero-Shot Framework for Medical Diagnosis Using Contrastive Representations
di: Sar, Ayan, et al.
Pubblicazione: (2025)
di: Sar, Ayan, et al.
Pubblicazione: (2025)
SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems
di: Hao, Haochang, et al.
Pubblicazione: (2026)
di: Hao, Haochang, et al.
Pubblicazione: (2026)
Efficient and Versatile Model for Multilingual Information Retrieval of Islamic Text: Development and Deployment in Real-World Scenarios
di: Pavlova, Vera, et al.
Pubblicazione: (2025)
di: Pavlova, Vera, et al.
Pubblicazione: (2025)
WebFAQ 2.0: A Multilingual QA Dataset with Mined Hard Negatives for Dense Retrieval
di: Dinzinger, Michael, et al.
Pubblicazione: (2026)
di: Dinzinger, Michael, et al.
Pubblicazione: (2026)
Table Meets LLM: Can Large Language Models Understand Structured Table Data? A Benchmark and Empirical Study
di: Sui, Yuan, et al.
Pubblicazione: (2023)
di: Sui, Yuan, et al.
Pubblicazione: (2023)
fact check AI at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-checked Claim Retrieval
di: Rastogi, Pranshu
Pubblicazione: (2025)
di: Rastogi, Pranshu
Pubblicazione: (2025)
Documenti analoghi
-
Text-to-SPARQL Goes Beyond English: Multilingual Question Answering Over Knowledge Graphs through Human-Inspired Reasoning
di: Perevalov, Aleksandr, et al.
Pubblicazione: (2025) -
SPARQL Query Generation with LLMs: Measuring the Impact of Training Data Memorization and Knowledge Injection
di: Gashkov, Aleksandr, et al.
Pubblicazione: (2025) -
Auto-ARGUE: LLM-Based Report Generation Evaluation
di: Walden, William, et al.
Pubblicazione: (2025) -
Multilingual Information Retrieval with a Monolingual Knowledge Base
di: Zhuang, Yingying, et al.
Pubblicazione: (2025) -
Analysis and Detection of Multilingual Hate Speech Using Transformer Based Deep Learning
di: Das, Arijit, et al.
Pubblicazione: (2024)