FictionalQA: A Dataset for Studying Memorization and Knowledge Acquisition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kirchenbauer, John, Mongkolsupawan, Janny, Wen, Yuxin, Goldstein, Tom, Ippolito, Daphne |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LMD3: Language Model Data Density Dependence
von: Kirchenbauer, John, et al.
Veröffentlicht: (2024)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2024)
A Watermark for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
Multi-Token Prediction via Self-Distillation
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026)
GenQA: Generating Millions of Instructions from a Handful of Prompts
von: Chen, Jiuhai, et al.
Veröffentlicht: (2024)
von: Chen, Jiuhai, et al.
Veröffentlicht: (2024)
On the Reliability of Watermarks for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024)
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024)
OPTune: Efficient Online Preference Tuning
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
Forcing Diffuse Distributions out of Language Models
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
Antidistillation Fingerprinting
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026)
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026)
Measuring Non-Adversarial Reproduction of Training Data in Large Language Models
von: Aerni, Michael, et al.
Veröffentlicht: (2024)
von: Aerni, Michael, et al.
Veröffentlicht: (2024)
Quantifying Cross-Modality Memorization in Vision-Language Models
von: Wen, Yuxin, et al.
Veröffentlicht: (2025)
von: Wen, Yuxin, et al.
Veröffentlicht: (2025)
FoQA: A Faroese Question-Answering Dataset
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
Coercing LLMs to do and reveal (almost) anything
von: Geiping, Jonas, et al.
Veröffentlicht: (2024)
von: Geiping, Jonas, et al.
Veröffentlicht: (2024)
Exploring the Limits of Model Compression in LLMs: A Knowledge Distillation Study on QA Tasks
von: Datta, Joyeeta, et al.
Veröffentlicht: (2025)
von: Datta, Joyeeta, et al.
Veröffentlicht: (2025)
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
von: McLeish, Sean, et al.
Veröffentlicht: (2025)
von: McLeish, Sean, et al.
Veröffentlicht: (2025)
GATES: Self-Distillation under Privileged Context with Consensus Gating
von: Stein, Alex, et al.
Veröffentlicht: (2026)
von: Stein, Alex, et al.
Veröffentlicht: (2026)
Memorization vs. Reasoning: Updating LLMs with New Knowledge
von: Li, Aochong Oliver, et al.
Veröffentlicht: (2025)
von: Li, Aochong Oliver, et al.
Veröffentlicht: (2025)
Provable Knowledge Acquisition and Extraction in One-Layer Transformers
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
von: Xu, Ruichen, et al.
Veröffentlicht: (2025)
Hubble: a Model Suite to Advance the Study of LLM Memorization
von: Wei, Johnny Tian-Zheng, et al.
Veröffentlicht: (2025)
von: Wei, Johnny Tian-Zheng, et al.
Veröffentlicht: (2025)
P-RAG: Prompt-Enhanced Parametric RAG with LoRA and Selective CoT for Biomedical and Multi-Hop QA
von: Lyu, Xingda, et al.
Veröffentlicht: (2026)
von: Lyu, Xingda, et al.
Veröffentlicht: (2026)
ViQA-COVID: COVID-19 Machine Reading Comprehension Dataset for Vietnamese
von: Nguyen-Phung, Hai-Chung, et al.
Veröffentlicht: (2025)
von: Nguyen-Phung, Hai-Chung, et al.
Veröffentlicht: (2025)
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text
von: Kandpal, Nikhil, et al.
Veröffentlicht: (2025)
von: Kandpal, Nikhil, et al.
Veröffentlicht: (2025)
SyllabusQA: A Course Logistics Question Answering Dataset
von: Fernandez, Nigel, et al.
Veröffentlicht: (2024)
von: Fernandez, Nigel, et al.
Veröffentlicht: (2024)
RoBiologyDataChoiceQA: A Romanian Dataset for improving Biology understanding of Large Language Models
von: Ghinea, Dragos-Dumitru, et al.
Veröffentlicht: (2025)
von: Ghinea, Dragos-Dumitru, et al.
Veröffentlicht: (2025)
Reason to Rote: Rethinking Memorization in Reasoning
von: Du, Yupei, et al.
Veröffentlicht: (2025)
von: Du, Yupei, et al.
Veröffentlicht: (2025)
A Lightweight Method to Disrupt Memorized Sequences in LLM
von: Prashant, Parjanya Prajakta, et al.
Veröffentlicht: (2025)
von: Prashant, Parjanya Prajakta, et al.
Veröffentlicht: (2025)
The Detection and Understanding of Fictional Discourse
von: Piper, Andrew, et al.
Veröffentlicht: (2024)
von: Piper, Andrew, et al.
Veröffentlicht: (2024)
MedConceptsQA: Open Source Medical Concepts QA Benchmark
von: Shoham, Ofir Ben, et al.
Veröffentlicht: (2024)
von: Shoham, Ofir Ben, et al.
Veröffentlicht: (2024)
From Memorization to Reasoning in the Spectrum of Loss Curvature
von: Merullo, Jack, et al.
Veröffentlicht: (2025)
von: Merullo, Jack, et al.
Veröffentlicht: (2025)
Demystifying Verbatim Memorization in Large Language Models
von: Huang, Jing, et al.
Veröffentlicht: (2024)
von: Huang, Jing, et al.
Veröffentlicht: (2024)
PeruMedQA: Benchmarking Large Language Models (LLMs) on Peruvian Medical Exams -- Dataset Construction and Evaluation
von: Carrillo-Larco, Rodrigo M., et al.
Veröffentlicht: (2025)
von: Carrillo-Larco, Rodrigo M., et al.
Veröffentlicht: (2025)
Benchmarking ChatGPT on Algorithmic Reasoning
von: McLeish, Sean, et al.
Veröffentlicht: (2024)
von: McLeish, Sean, et al.
Veröffentlicht: (2024)
Rethinking LLM Memorization through the Lens of Adversarial Compression
von: Schwarzschild, Avi, et al.
Veröffentlicht: (2024)
von: Schwarzschild, Avi, et al.
Veröffentlicht: (2024)
Memorization in In-Context Learning
von: Golchin, Shahriar, et al.
Veröffentlicht: (2024)
von: Golchin, Shahriar, et al.
Veröffentlicht: (2024)
Think Before You Lie: How Reasoning Leads to Honesty
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
Memorization: A Close Look at Books
von: Ma, Iris, et al.
Veröffentlicht: (2025)
von: Ma, Iris, et al.
Veröffentlicht: (2025)
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation
von: Kundu, Debarshi
Veröffentlicht: (2024)
von: Kundu, Debarshi
Veröffentlicht: (2024)
Fact or Fiction? Improving Fact Verification with Knowledge Graphs through Simplified Subgraph Retrievals
von: Opsahl, Tobias A.
Veröffentlicht: (2024)
von: Opsahl, Tobias A.
Veröffentlicht: (2024)
Data Mixing Can Induce Phase Transitions in Knowledge Acquisition
von: Gu, Xinran, et al.
Veröffentlicht: (2025)
von: Gu, Xinran, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LMD3: Language Model Data Density Dependence
von: Kirchenbauer, John, et al.
Veröffentlicht: (2024) -
A Watermark for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023) -
Multi-Token Prediction via Self-Distillation
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026) -
GenQA: Generating Millions of Instructions from a Handful of Prompts
von: Chen, Jiuhai, et al.
Veröffentlicht: (2024) -
On the Reliability of Watermarks for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)