Wikipedia is Not a Dictionary, Delete! Text Classification as a Proxy for Analysing Wiki Deletion Discussions
Fuente:
arXiv
Saved in:
| Main Authors: | Borkakoty, Hsuvas, Espinosa-Anke, Luis |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WiDe-analysis: Enabling One-click Content Moderation Analysis on Wikipedia's Articles for Deletion
by: Borkakoty, Hsuvas, et al.
Published: (2024)
by: Borkakoty, Hsuvas, et al.
Published: (2024)
CHEW: A Dataset of CHanging Events in Wikipedia
by: Borkakoty, Hsuvas, et al.
Published: (2024)
by: Borkakoty, Hsuvas, et al.
Published: (2024)
Hoaxpedia: A Unified Wikipedia Hoax Articles Dataset
by: Borkakoty, Hsuvas, et al.
Published: (2024)
by: Borkakoty, Hsuvas, et al.
Published: (2024)
Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor Discussions
by: Kaffee, Lucie-Aimée, et al.
Published: (2023)
by: Kaffee, Lucie-Aimée, et al.
Published: (2023)
Automatic Extraction of Metaphoric Analogies from Literary Texts: Task Formulation, Dataset Construction, and Evaluation
by: Boisson, Joanne, et al.
Published: (2024)
by: Boisson, Joanne, et al.
Published: (2024)
Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis
by: Fu, Rong, et al.
Published: (2026)
by: Fu, Rong, et al.
Published: (2026)
SliceGPT: Compress Large Language Models by Deleting Rows and Columns
by: Ashkboos, Saleh, et al.
Published: (2024)
by: Ashkboos, Saleh, et al.
Published: (2024)
Where Knowledge Collides: A Mechanistic Study of Intra-Memory Knowledge Conflict in Language Models
by: Pham, Minh Vu, et al.
Published: (2026)
by: Pham, Minh Vu, et al.
Published: (2026)
MultiDelete for Multimodal Machine Unlearning
by: Cheng, Jiali, et al.
Published: (2023)
by: Cheng, Jiali, et al.
Published: (2023)
WikiContradict: A Benchmark for Evaluating LLMs on Real-World Knowledge Conflicts from Wikipedia
by: Hou, Yufang, et al.
Published: (2024)
by: Hou, Yufang, et al.
Published: (2024)
AdaptDel: Adaptable Deletion Rate Randomized Smoothing for Certified Robustness
by: Huang, Zhuoqun, et al.
Published: (2025)
by: Huang, Zhuoqun, et al.
Published: (2025)
SpaceByte: Towards Deleting Tokenization from Large Language Modeling
by: Slagle, Kevin
Published: (2024)
by: Slagle, Kevin
Published: (2024)
Lightweight Conceptual Dictionary Learning for Text Classification Using Information Compression
by: Wan, Li, et al.
Published: (2024)
by: Wan, Li, et al.
Published: (2024)
Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs
by: Xu, Xiaoyu, et al.
Published: (2025)
by: Xu, Xiaoyu, et al.
Published: (2025)
Beyond Masks: Efficient, Flexible Diffusion Language Models via Deletion-Insertion Processes
by: Ding, Fangyu, et al.
Published: (2026)
by: Ding, Fangyu, et al.
Published: (2026)
Delete My Account: Impact of Data Deletion on Machine Learning Classifiers
by: Dam, Tobias, et al.
Published: (2023)
by: Dam, Tobias, et al.
Published: (2023)
Rare but Severe Neural Machine Translation Errors Induced by Minimal Deletion: An Empirical Study on Chinese and English
by: Shi, Ruikang, et al.
Published: (2022)
by: Shi, Ruikang, et al.
Published: (2022)
DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain
by: Borkakoty, Hsuvas, et al.
Published: (2026)
by: Borkakoty, Hsuvas, et al.
Published: (2026)
Simulation, Modelling and Classification of Wiki Contributors: Spotting The Good, The Bad, and The Ugly
by: Méndez, Silvia García, et al.
Published: (2024)
by: Méndez, Silvia García, et al.
Published: (2024)
Detecting Sockpuppetry on Wikipedia Using Meta-Learning
by: Raszewski, Luc, et al.
Published: (2025)
by: Raszewski, Luc, et al.
Published: (2025)
WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMs
by: Thede, Lukas, et al.
Published: (2025)
by: Thede, Lukas, et al.
Published: (2025)
Delete and Retain: Efficient Unlearning for Document Classification
by: Goel, Aadya, et al.
Published: (2025)
by: Goel, Aadya, et al.
Published: (2025)
Wiki-Quantities and Wiki-Measurements: Datasets of Quantities and their Measurement Context from Wikipedia
by: Göpfert, Jan, et al.
Published: (2025)
by: Göpfert, Jan, et al.
Published: (2025)
Proxy Compression for Language Modeling
by: Zheng, Lin, et al.
Published: (2026)
by: Zheng, Lin, et al.
Published: (2026)
GEAR: A Simple GENERATE, EMBED, AVERAGE AND RANK Approach for Unsupervised Reverse Dictionary
by: Almeman, Fatemah, et al.
Published: (2024)
by: Almeman, Fatemah, et al.
Published: (2024)
DICTDIS: Dictionary Constrained Disambiguation for Improved NMT
by: Maheshwari, Ayush, et al.
Published: (2022)
by: Maheshwari, Ayush, et al.
Published: (2022)
PhantomWiki: On-Demand Datasets for Reasoning and Retrieval Evaluation
by: Gong, Albert, et al.
Published: (2025)
by: Gong, Albert, et al.
Published: (2025)
Shifting Perspectives: Steering Vectors for Robust Bias Mitigation in LLMs
by: Siddique, Zara, et al.
Published: (2025)
by: Siddique, Zara, et al.
Published: (2025)
On the Fragility of Active Learners for Text Classification
by: Ghose, Abhishek, et al.
Published: (2024)
by: Ghose, Abhishek, et al.
Published: (2024)
Universal Cross-Lingual Text Classification
by: Savant, Riya, et al.
Published: (2024)
by: Savant, Riya, et al.
Published: (2024)
Improving Rare Word Translation With Dictionaries and Attention Masking
by: Sible, Kenneth J., et al.
Published: (2024)
by: Sible, Kenneth J., et al.
Published: (2024)
Forecasting Downstream Performance of LLMs With Proxy Metrics
by: Patel, Arkil, et al.
Published: (2026)
by: Patel, Arkil, et al.
Published: (2026)
TextAge: A Curated and Diverse Text Dataset for Age Classification
by: Cheekati, Shravan, et al.
Published: (2024)
by: Cheekati, Shravan, et al.
Published: (2024)
Fighting Spurious Correlations in Text Classification via a Causal Learning Perspective
by: Zhou, Yuqing, et al.
Published: (2024)
by: Zhou, Yuqing, et al.
Published: (2024)
Wikipedia in the Era of LLMs: Evolution and Risks
by: Huang, Siming, et al.
Published: (2025)
by: Huang, Siming, et al.
Published: (2025)
Proxy-RLHF: Decoupling Generation and Alignment in Large Language Model with Proxy
by: Zhu, Yu, et al.
Published: (2024)
by: Zhu, Yu, et al.
Published: (2024)
Fair Text Classification via Transferable Representations
by: Leteno, Thibaud, et al.
Published: (2025)
by: Leteno, Thibaud, et al.
Published: (2025)
Revisiting Hierarchical Text Classification: Inference and Metrics
by: Plaud, Roman, et al.
Published: (2024)
by: Plaud, Roman, et al.
Published: (2024)
Ensembling Finetuned Language Models for Text Classification
by: Arango, Sebastian Pineda, et al.
Published: (2024)
by: Arango, Sebastian Pineda, et al.
Published: (2024)
AQuA -- Combining Experts' and Non-Experts' Views To Assess Deliberation Quality in Online Discussions Using LLMs
by: Behrendt, Maike, et al.
Published: (2024)
by: Behrendt, Maike, et al.
Published: (2024)
Similar Items
-
WiDe-analysis: Enabling One-click Content Moderation Analysis on Wikipedia's Articles for Deletion
by: Borkakoty, Hsuvas, et al.
Published: (2024) -
CHEW: A Dataset of CHanging Events in Wikipedia
by: Borkakoty, Hsuvas, et al.
Published: (2024) -
Hoaxpedia: A Unified Wikipedia Hoax Articles Dataset
by: Borkakoty, Hsuvas, et al.
Published: (2024) -
Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor Discussions
by: Kaffee, Lucie-Aimée, et al.
Published: (2023) -
Automatic Extraction of Metaphoric Analogies from Literary Texts: Task Formulation, Dataset Construction, and Evaluation
by: Boisson, Joanne, et al.
Published: (2024)