The Chronicles of RiDiC: Generating Datasets with Controlled Popularity Distribution for Long-form Factuality Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Braslavski, Pavel, Iarosh, Dmitrii, Sushko, Nikita, Sakhovskiy, Andrey, Konovalov, Vasily, Tutubalina, Elena, Panchenko, Alexander |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
by: Sakhovskiy, Andrey, et al.
Published: (2025)
by: Sakhovskiy, Andrey, et al.
Published: (2025)
SynthDetoxM: Modern LLMs are Few-Shot Parallel Detoxification Data Annotators
by: Moskovskiy, Daniil, et al.
Published: (2025)
by: Moskovskiy, Daniil, et al.
Published: (2025)
How Much Knowledge Can You Pack into a LoRA Adapter without Harming LLM?
by: Pletenev, Sergey, et al.
Published: (2025)
by: Pletenev, Sergey, et al.
Published: (2025)
Beyond Detection: Rethinking Education in the Age of AI-writing
by: Marina, Maria, et al.
Published: (2026)
by: Marina, Maria, et al.
Published: (2026)
Konstruktor: A Strong Baseline for Simple Knowledge Graph Question Answering
by: Lysyuk, Maria, et al.
Published: (2024)
by: Lysyuk, Maria, et al.
Published: (2024)
Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning
by: Borisiuk, Anna, et al.
Published: (2026)
by: Borisiuk, Anna, et al.
Published: (2026)
Through the Looking Glass: Common Sense Consistency Evaluation of Weird Images
by: Rykov, Elisei, et al.
Published: (2025)
by: Rykov, Elisei, et al.
Published: (2025)
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
by: Rykov, Elisei, et al.
Published: (2025)
by: Rykov, Elisei, et al.
Published: (2025)
Eigen-Factors a Bilevel Optimization for Plane SLAM of 3D Point Clouds
by: Ferrer, Gonzalo, et al.
Published: (2023)
by: Ferrer, Gonzalo, et al.
Published: (2023)
KoWit-24: A Richly Annotated Dataset of Wordplay in News Headlines
by: Baranov, Alexander, et al.
Published: (2025)
by: Baranov, Alexander, et al.
Published: (2025)
CoRoVA: Compressed Representations for Vector-Augmented Code Completion
by: Cherniuk, Daria, et al.
Published: (2025)
by: Cherniuk, Daria, et al.
Published: (2025)
Leveraging LLM Parametric Knowledge for Fact Checking without Retrieval
by: Vazhentsev, Artem, et al.
Published: (2026)
by: Vazhentsev, Artem, et al.
Published: (2026)
RuCCoD: Towards Automated ICD Coding in Russian
by: Nesterov, Aleksandr, et al.
Published: (2025)
by: Nesterov, Aleksandr, et al.
Published: (2025)
Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures
by: Somov, Oleg, et al.
Published: (2026)
by: Somov, Oleg, et al.
Published: (2026)
KazQAD: Kazakh Open-Domain Question Answering Dataset
by: Yeshpanov, Rustem, et al.
Published: (2024)
by: Yeshpanov, Rustem, et al.
Published: (2024)
LLM-Independent Adaptive RAG: Let the Question Speak for Itself
by: Marina, Maria, et al.
Published: (2025)
by: Marina, Maria, et al.
Published: (2025)
Will It Still Be True Tomorrow? Multilingual Evergreen Question Classification to Improve Trustworthy QA
by: Pletenev, Sergey, et al.
Published: (2025)
by: Pletenev, Sergey, et al.
Published: (2025)
OCC-RAG: Optimal Cognitive Core for Faithful Question Answering
by: Savkin, Maksim, et al.
Published: (2026)
by: Savkin, Maksim, et al.
Published: (2026)
Beyond Factual Accuracy: Evaluating Coverage of Diverse Factual Information in Long-form Text Generation
by: Samarinas, Chris, et al.
Published: (2025)
by: Samarinas, Chris, et al.
Published: (2025)
The benefits of query-based KGQA systems for complex and temporal questions in LLM era
by: Alekseev, Artem, et al.
Published: (2025)
by: Alekseev, Artem, et al.
Published: (2025)
When Punctuation Matters: A Large-Scale Comparison of Prompt Robustness Methods for LLMs
by: Seleznyov, Mikhail, et al.
Published: (2025)
by: Seleznyov, Mikhail, et al.
Published: (2025)
Evolutionary Search for Automated Design of Uncertainty Quantification Methods
by: Seleznyov, Mikhail, et al.
Published: (2026)
by: Seleznyov, Mikhail, et al.
Published: (2026)
DeepPavlov at SemEval-2024 Task 8: Leveraging Transfer Learning for Detecting Boundaries of Machine-Generated Texts
by: Voznyuk, Anastasia, et al.
Published: (2024)
by: Voznyuk, Anastasia, et al.
Published: (2024)
Adaptive Retrieval Without Self-Knowledge? Bringing Uncertainty Back Home
by: Moskvoretskii, Viktor, et al.
Published: (2025)
by: Moskvoretskii, Viktor, et al.
Published: (2025)
Demographic Prediction Based on User Reviews about Medications
by: Elena Tutubalina
Published: (2017)
by: Elena Tutubalina
Published: (2017)
One Task Vector is not Enough: A Large-Scale Study for In-Context Learning
by: Tikhonov, Pavel, et al.
Published: (2025)
by: Tikhonov, Pavel, et al.
Published: (2025)
Score-based deterministic density sampling
by: Ilin, Vasily, et al.
Published: (2025)
by: Ilin, Vasily, et al.
Published: (2025)
Geopolitical biases in LLMs: what are the "good" and the "bad" countries according to contemporary language models
by: Salnikov, Mikhail, et al.
Published: (2025)
by: Salnikov, Mikhail, et al.
Published: (2025)
Harnessing non-adversarial robustness in large language models
by: Zhou, Qinghua, et al.
Published: (2026)
by: Zhou, Qinghua, et al.
Published: (2026)
When Models Lie, We Learn: Multilingual Span-Level Hallucination Detection with PsiloQA
by: Rykov, Elisei, et al.
Published: (2025)
by: Rykov, Elisei, et al.
Published: (2025)
Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs
by: Afonin, Nikita, et al.
Published: (2025)
by: Afonin, Nikita, et al.
Published: (2025)
Team Anotheroption at SemEval-2025 Task 8: Bridging the Gap Between Open-Source and Proprietary LLMs in Table QA
by: Evkarpidi, Nikolas, et al.
Published: (2025)
by: Evkarpidi, Nikolas, et al.
Published: (2025)
Confidence Estimation for Error Detection in Text-to-SQL Systems
by: Somov, Oleg, et al.
Published: (2025)
by: Somov, Oleg, et al.
Published: (2025)
Data on the bat fauna of the Northern Black Sea Region based on results of the work of bat contact centres
by: Panchenko, Pavel, et al.
Published: (2018)
by: Panchenko, Pavel, et al.
Published: (2018)
Aligning Distributionally Robust Optimization with Practical Deep Learning Needs
by: Feoktistov, Dmitrii, et al.
Published: (2025)
by: Feoktistov, Dmitrii, et al.
Published: (2025)
RiWeCIS – Supporting Datasets
by: Panozzo, Silvia, et al.
Published: (2026)
by: Panozzo, Silvia, et al.
Published: (2026)
WildHallucinations: Evaluating Long-form Factuality in LLMs with Real-World Entity Queries
by: Zhao, Wenting, et al.
Published: (2024)
by: Zhao, Wenting, et al.
Published: (2024)
Bring the Apple, Not the Sofa: Impact of Irrelevant Context in Embodied AI Commands on VLA Models
by: Pugacheva, Daria, et al.
Published: (2025)
by: Pugacheva, Daria, et al.
Published: (2025)
OLAPH: Improving Factuality in Biomedical Long-form Question Answering
by: Jeong, Minbyul, et al.
Published: (2024)
by: Jeong, Minbyul, et al.
Published: (2024)
Distributed Chronicles to Faults Recognition
by: Juan Vizcarrondo
Published: (2015)
by: Juan Vizcarrondo
Published: (2015)
Similar Items
-
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
by: Sakhovskiy, Andrey, et al.
Published: (2025) -
SynthDetoxM: Modern LLMs are Few-Shot Parallel Detoxification Data Annotators
by: Moskovskiy, Daniil, et al.
Published: (2025) -
How Much Knowledge Can You Pack into a LoRA Adapter without Harming LLM?
by: Pletenev, Sergey, et al.
Published: (2025) -
Beyond Detection: Rethinking Education in the Age of AI-writing
by: Marina, Maria, et al.
Published: (2026) -
Konstruktor: A Strong Baseline for Simple Knowledge Graph Question Answering
by: Lysyuk, Maria, et al.
Published: (2024)