Overview of the Plagiarism Detection Task at PAN 2025
Fuente:
arXiv
Saved in:
| Main Authors: | Greiner-Petter, André, Fröbe, Maik, Wahle, Jan Philip, Ruas, Terry, Gipp, Bela, Aizawa, Akiko, Potthast, Martin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Taxonomy of Mathematical Plagiarism
by: Satpute, Ankit, et al.
Published: (2024)
by: Satpute, Ankit, et al.
Published: (2024)
Can LLMs Master Math? Investigating Large Language Models on Math Stack Exchange
by: Satpute, Ankit, et al.
Published: (2024)
by: Satpute, Ankit, et al.
Published: (2024)
Aspect-Aware Content-Based Recommendations for Mathematical Research Papers
by: Satpute, Ankit, et al.
Published: (2026)
by: Satpute, Ankit, et al.
Published: (2026)
How Large Language Models are Transforming Machine-Paraphrased Plagiarism
by: Wahle, Jan Philip, et al.
Published: (2022)
by: Wahle, Jan Philip, et al.
Published: (2022)
Paraphrase Types for Generation and Detection
by: Wahle, Jan Philip, et al.
Published: (2023)
by: Wahle, Jan Philip, et al.
Published: (2023)
MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions
by: Horych, Tomáš, et al.
Published: (2024)
by: Horych, Tomáš, et al.
Published: (2024)
The Promises and Pitfalls of LLM Annotations in Dataset Labeling: a Case Study on Media Bias Detection
by: Horych, Tomas, et al.
Published: (2024)
by: Horych, Tomas, et al.
Published: (2024)
Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges
by: Becker, Jonas, et al.
Published: (2024)
by: Becker, Jonas, et al.
Published: (2024)
Analyzing Adversarial Attacks on Sequence-to-Sequence Relevance Models
by: Parry, Andrew, et al.
Published: (2024)
by: Parry, Andrew, et al.
Published: (2024)
Overview of PAN 2026: Voight-Kampff Generative AI Detection, Text Watermarking, Multi-Author Writing Style Analysis, Generative Plagiarism Detection, and Reasoning Trajectory Detection
by: Bevendorff, Janek, et al.
Published: (2026)
by: Bevendorff, Janek, et al.
Published: (2026)
Simplified Longitudinal Retrieval Experiments: A Case Study on Query Expansion and Document Boosting
by: Keller, Jüri, et al.
Published: (2025)
by: Keller, Jüri, et al.
Published: (2025)
Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains
by: Schmidt, Finn, et al.
Published: (2026)
by: Schmidt, Finn, et al.
Published: (2026)
Paraphrase Types Elicit Prompt Engineering Capabilities
by: Wahle, Jan Philip, et al.
Published: (2024)
by: Wahle, Jan Philip, et al.
Published: (2024)
Piecing Together Cross-Document Coreference Resolution Datasets: Systematic Dataset Analysis and Unification
by: Zhukova, Anastasia, et al.
Published: (2026)
by: Zhukova, Anastasia, et al.
Published: (2026)
An Analysis of Datasets, Metrics and Models in Keyphrase Generation
by: Boudin, Florian, et al.
Published: (2025)
by: Boudin, Florian, et al.
Published: (2025)
What's under the hood: Investigating Automatic Metrics on Meeting Summarization
by: Kirstein, Frederic, et al.
Published: (2024)
by: Kirstein, Frederic, et al.
Published: (2024)
CADS: A Systematic Literature Review on the Challenges of Abstractive Dialogue Summarization
by: Kirstein, Frederic, et al.
Published: (2024)
by: Kirstein, Frederic, et al.
Published: (2024)
Towards Human Understanding of Paraphrase Types in Large Language Models
by: Meier, Dominik, et al.
Published: (2024)
by: Meier, Dominik, et al.
Published: (2024)
The Viability of Crowdsourcing for RAG Evaluation
by: Gienapp, Lukas, et al.
Published: (2025)
by: Gienapp, Lukas, et al.
Published: (2025)
Counterfactual Query Rewriting to Use Historical Relevance Feedback
by: Keller, Jüri, et al.
Published: (2025)
by: Keller, Jüri, et al.
Published: (2025)
Big Tech-Funded AI Papers Have Higher Citation Impact, Greater Insularity, and Larger Recency Bias
by: Gnewuch, Max Martin, et al.
Published: (2025)
by: Gnewuch, Max Martin, et al.
Published: (2025)
Evaluating Generative Ad Hoc Information Retrieval
by: Gienapp, Lukas, et al.
Published: (2023)
by: Gienapp, Lukas, et al.
Published: (2023)
Link Prediction for Event Logs in the Process Industry
by: Zhukova, Anastasia, et al.
Published: (2025)
by: Zhukova, Anastasia, et al.
Published: (2025)
Contrastive Learning Using Graph Embeddings for Domain Adaptation of Language Models in the Process Industry
by: Zhukova, Anastasia, et al.
Published: (2025)
by: Zhukova, Anastasia, et al.
Published: (2025)
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
by: Kaesberg, Lars Benedikt, et al.
Published: (2024)
by: Kaesberg, Lars Benedikt, et al.
Published: (2024)
D3: A Massive Dataset of Scholarly Metadata for Analyzing the State of Computer Science Research
by: Wahle, Jan Philip, et al.
Published: (2022)
by: Wahle, Jan Philip, et al.
Published: (2022)
SPaRC: A Spatial Pathfinding Reasoning Challenge
by: Kaesberg, Lars Benedikt, et al.
Published: (2025)
by: Kaesberg, Lars Benedikt, et al.
Published: (2025)
Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation
by: He, Xuhong, et al.
Published: (2026)
by: He, Xuhong, et al.
Published: (2026)
Systematic Evaluation of Neural Retrieval Models on the Touché 2020 Argument Retrieval Subset of BEIR
by: Thakur, Nandan, et al.
Published: (2024)
by: Thakur, Nandan, et al.
Published: (2024)
Lightning IR: Straightforward Fine-tuning and Inference of Transformer-based Language Models for Information Retrieval
by: Schlatt, Ferdinand, et al.
Published: (2024)
by: Schlatt, Ferdinand, et al.
Published: (2024)
Investigating the Effects of Sparse Attention on Cross-Encoders
by: Schlatt, Ferdinand, et al.
Published: (2023)
by: Schlatt, Ferdinand, et al.
Published: (2023)
Beyond Chains: Bridging Large Language Models and Knowledge Bases in Complex Question Answering
by: Zhu, Yihua, et al.
Published: (2025)
by: Zhu, Yihua, et al.
Published: (2025)
Self-Compositional Data Augmentation for Scientific Keyphrase Generation
by: Houbre, Mael, et al.
Published: (2024)
by: Houbre, Mael, et al.
Published: (2024)
You need to MIMIC to get FAME: Solving Meeting Transcript Scarcity with a Multi-Agent Conversations
by: Kirstein, Frederic, et al.
Published: (2025)
by: Kirstein, Frederic, et al.
Published: (2025)
TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
by: Meier, Dominik, et al.
Published: (2025)
by: Meier, Dominik, et al.
Published: (2025)
Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking
by: Schlatt, Ferdinand, et al.
Published: (2024)
by: Schlatt, Ferdinand, et al.
Published: (2024)
Set-Encoder: Permutation-Invariant Inter-Passage Attention for Listwise Passage Re-Ranking with Cross-Encoders
by: Schlatt, Ferdinand, et al.
Published: (2024)
by: Schlatt, Ferdinand, et al.
Published: (2024)
Detecting Generated Native Ads in Conversational Search
by: Schmidt, Sebastian, et al.
Published: (2024)
by: Schmidt, Sebastian, et al.
Published: (2024)
Med-CoDE: Medical Critique based Disagreement Evaluation Framework
by: Gupta, Mohit, et al.
Published: (2025)
by: Gupta, Mohit, et al.
Published: (2025)
Testing the Generalization of Neural Language Models for COVID-19 Misinformation Detection
by: Wahle, Jan Philip, et al.
Published: (2021)
by: Wahle, Jan Philip, et al.
Published: (2021)
Similar Items
-
Taxonomy of Mathematical Plagiarism
by: Satpute, Ankit, et al.
Published: (2024) -
Can LLMs Master Math? Investigating Large Language Models on Math Stack Exchange
by: Satpute, Ankit, et al.
Published: (2024) -
Aspect-Aware Content-Based Recommendations for Mathematical Research Papers
by: Satpute, Ankit, et al.
Published: (2026) -
How Large Language Models are Transforming Machine-Paraphrased Plagiarism
by: Wahle, Jan Philip, et al.
Published: (2022) -
Paraphrase Types for Generation and Detection
by: Wahle, Jan Philip, et al.
Published: (2023)