Assessing Web Search Credibility and Response Groundedness in Chat Assistants
Fuente:
arXiv
Saved in:
| Main Authors: | Vykopal, Ivan, Pikuliak, Matúš, Ostermann, Simon, Šimko, Marián |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Large Language Models in Automated Fact-Checking: A Survey
by: Vykopal, Ivan, et al.
Published: (2024)
by: Vykopal, Ivan, et al.
Published: (2024)
Large Language Models for Multilingual Previously Fact-Checked Claim Detection
by: Vykopal, Ivan, et al.
Published: (2025)
by: Vykopal, Ivan, et al.
Published: (2025)
Soft Language Prompts for Language Transfer
by: Vykopal, Ivan, et al.
Published: (2024)
by: Vykopal, Ivan, et al.
Published: (2024)
Women Are Beautiful, Men Are Leaders: Gender Stereotypes in Machine Translation and Language Modeling
by: Pikuliak, Matúš, et al.
Published: (2023)
by: Pikuliak, Matúš, et al.
Published: (2023)
Disinformation Capabilities of Large Language Models
by: Vykopal, Ivan, et al.
Published: (2023)
by: Vykopal, Ivan, et al.
Published: (2023)
GenderBench: Evaluation Suite for Gender Biases in LLMs
by: Pikuliak, Matúš
Published: (2025)
by: Pikuliak, Matúš
Published: (2025)
Multilingual Previously Fact-Checked Claim Retrieval
by: Pikuliak, Matúš, et al.
Published: (2023)
by: Pikuliak, Matúš, et al.
Published: (2023)
Small Models, Big Impact: Efficient Corpus and Graph-Based Adaptation of Small Multilingual Language Models for Low-Resource Languages
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
Investigating Language and Retrieval Bias in Multilingual Previously Fact-Checked Claim Detection
by: Vykopal, Ivan, et al.
Published: (2025)
by: Vykopal, Ivan, et al.
Published: (2025)
A Generative-AI-Driven Claim Retrieval System Capable of Detecting and Retrieving Claims from Social Media Platforms in Multiple Languages
by: Vykopal, Ivan, et al.
Published: (2025)
by: Vykopal, Ivan, et al.
Published: (2025)
SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
by: Peng, Qiwei, et al.
Published: (2025)
by: Peng, Qiwei, et al.
Published: (2025)
A Rigorous Evaluation of LLM Data Generation Strategies for Low-Resource Languages
by: Anikina, Tatiana, et al.
Published: (2025)
by: Anikina, Tatiana, et al.
Published: (2025)
Multilingual Political Views of Large Language Models: Identification and Steering
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
Measuring the Groundedness of Legal Question-Answering Systems
by: Trautmann, Dietrich, et al.
Published: (2024)
by: Trautmann, Dietrich, et al.
Published: (2024)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
by: Macko, Dominik, et al.
Published: (2023)
by: Macko, Dominik, et al.
Published: (2023)
Groundedness in Retrieval-augmented Long-form Generation: An Empirical Study
by: Stolfo, Alessandro
Published: (2024)
by: Stolfo, Alessandro
Published: (2024)
Small Encoders Can Rival Large Decoders in Detecting Groundedness
by: Abbes, Istabrak, et al.
Published: (2025)
by: Abbes, Istabrak, et al.
Published: (2025)
Evaluating AI for Finance: Is AI Credible at Assessing Investment Risk?
by: Chawla, Divij, et al.
Published: (2025)
by: Chawla, Divij, et al.
Published: (2025)
MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs
by: Gao, Yufei, et al.
Published: (2025)
by: Gao, Yufei, et al.
Published: (2025)
Task Prompt Vectors: Effective Initialization through Multi-Task Soft-Prompt Transfer
by: Belanec, Robert, et al.
Published: (2024)
by: Belanec, Robert, et al.
Published: (2024)
RoSE: Round-robin Synthetic Data Evaluation for Selecting LLM Generators without Human Test Sets
by: Cegin, Jan, et al.
Published: (2025)
by: Cegin, Jan, et al.
Published: (2025)
Variation between Credible and Non-Credible News Across Topics
by: Francis, Emilie
Published: (2024)
by: Francis, Emilie
Published: (2024)
The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks
by: Pomerenke, David, et al.
Published: (2025)
by: Pomerenke, David, et al.
Published: (2025)
GrEmLIn: A Repository of Green Baseline Embeddings for 87 Low-Resource Languages Injected with Multilingual Graph Knowledge
by: Gurgurov, Daniil, et al.
Published: (2024)
by: Gurgurov, Daniil, et al.
Published: (2024)
A Survey on Automatic Credibility Assessment Using Textual Credibility Signals in the Era of Large Language Models
by: Srba, Ivan, et al.
Published: (2024)
by: Srba, Ivan, et al.
Published: (2024)
OpenFActScore: Open-Source Atomic Evaluation of Factuality in Text Generation
by: Lage, Lucas Fonseca, et al.
Published: (2025)
by: Lage, Lucas Fonseca, et al.
Published: (2025)
Beyond Image-Text Matching: Verb Understanding in Multimodal Transformers Using Guided Masking
by: Beňová, Ivana, et al.
Published: (2024)
by: Beňová, Ivana, et al.
Published: (2024)
LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
Political Leaning and Politicalness Classification of Texts
by: Volf, Matous, et al.
Published: (2025)
by: Volf, Matous, et al.
Published: (2025)
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
by: Hyben, Martin, et al.
Published: (2023)
by: Hyben, Martin, et al.
Published: (2023)
The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs
by: Baeumel, Tanja, et al.
Published: (2025)
by: Baeumel, Tanja, et al.
Published: (2025)
Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms
by: Baeumel, Tanja, et al.
Published: (2026)
by: Baeumel, Tanja, et al.
Published: (2026)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
by: Vijayakumar, Soniya, et al.
Published: (2024)
by: Vijayakumar, Soniya, et al.
Published: (2024)
Automatic Fact-checking in English and Telugu
by: Chikkala, Ravi Kiran, et al.
Published: (2025)
by: Chikkala, Ravi Kiran, et al.
Published: (2025)
ExU: AI Models for Examining Multilingual Disinformation Narratives and Understanding their Spread
by: Vasilakes, Jake, et al.
Published: (2024)
by: Vasilakes, Jake, et al.
Published: (2024)
Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters
by: Gurgurov, Daniil, et al.
Published: (2024)
by: Gurgurov, Daniil, et al.
Published: (2024)
LLMs vs Established Text Augmentation Techniques for Classification: When do the Benefits Outweight the Costs?
by: Cegin, Jan, et al.
Published: (2024)
by: Cegin, Jan, et al.
Published: (2024)
Cross-Prompt Encoder for Low-Performing Languages
by: Mikaberidze, Beso, et al.
Published: (2025)
by: Mikaberidze, Beso, et al.
Published: (2025)
Conversational Assistants to support Heart Failure Patients: comparing a Neurosymbolic Architecture with ChatGPT
by: Tayal, Anuja, et al.
Published: (2025)
by: Tayal, Anuja, et al.
Published: (2025)
Multi-User Chat Assistant (MUCA): a Framework Using LLMs to Facilitate Group Conversations
by: Mao, Manqing, et al.
Published: (2024)
by: Mao, Manqing, et al.
Published: (2024)
Similar Items
-
Generative Large Language Models in Automated Fact-Checking: A Survey
by: Vykopal, Ivan, et al.
Published: (2024) -
Large Language Models for Multilingual Previously Fact-Checked Claim Detection
by: Vykopal, Ivan, et al.
Published: (2025) -
Soft Language Prompts for Language Transfer
by: Vykopal, Ivan, et al.
Published: (2024) -
Women Are Beautiful, Men Are Leaders: Gender Stereotypes in Machine Translation and Language Modeling
by: Pikuliak, Matúš, et al.
Published: (2023) -
Disinformation Capabilities of Large Language Models
by: Vykopal, Ivan, et al.
Published: (2023)