WinoWhat: A Parallel Corpus of Paraphrased WinoGrande Sentences with Common Sense Categorization
Fuente:
arXiv
Saved in:
| Main Authors: | Gevers, Ine, De Marez, Victor, De Bruyne, Luna, Daelemans, Walter |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Agentic AI Framework for Training General Practitioner Student Skills
by: De Marez, Victor, et al.
Published: (2025)
by: De Marez, Victor, et al.
Published: (2025)
Estonian WinoGrande Dataset: Comparative Analysis of LLM Performance on Human and Machine Translation
by: Ojastu, Marii, et al.
Published: (2025)
by: Ojastu, Marii, et al.
Published: (2025)
Bag of Lies: Robustness in Continuous Pre-training BERT
by: Gevers, Ine, et al.
Published: (2024)
by: Gevers, Ine, et al.
Published: (2024)
Do You Get the Hint? Benchmarking LLMs on the Board Game Concept
by: Gevers, Ine, et al.
Published: (2025)
by: Gevers, Ine, et al.
Published: (2025)
One Size Does Not Fit All: Exploring Variable Thresholds for Distance-Based Multi-Label Text Classification
by: Van Nooten, Jens, et al.
Published: (2025)
by: Van Nooten, Jens, et al.
Published: (2025)
WinoViz: Probing Visual Properties of Objects Under Different States
by: Jin, Woojeong, et al.
Published: (2024)
by: Jin, Woojeong, et al.
Published: (2024)
WinoPron: Revisiting English Winogender Schemas for Consistency, Coverage, and Grammatical Case
by: Gautam, Vagrant, et al.
Published: (2024)
by: Gautam, Vagrant, et al.
Published: (2024)
Quantifying Geospatial in the Common Crawl Corpus
by: Ilyankou, Ilya, et al.
Published: (2024)
by: Ilyankou, Ilya, et al.
Published: (2024)
WinoQueer: A Community-in-the-Loop Benchmark for Anti-LGBTQ+ Bias in Large Language Models
by: Felkner, Virginia K., et al.
Published: (2023)
by: Felkner, Virginia K., et al.
Published: (2023)
AliMark: Enhancing Robustness of Sentence-Level Watermarking Against Text Paraphrasing
by: Li, Yuexin, et al.
Published: (2026)
by: Li, Yuexin, et al.
Published: (2026)
AdParaphrase: Paraphrase Dataset for Analyzing Linguistic Features toward Generating Attractive Ad Texts
by: Murakami, Soichiro, et al.
Published: (2025)
by: Murakami, Soichiro, et al.
Published: (2025)
Predicate-Argument Structure Divergences in Chinese and English Parallel Sentences and their Impact on Language Transfer
by: Tripodi, Rocco, et al.
Published: (2025)
by: Tripodi, Rocco, et al.
Published: (2025)
GlotCC: An Open Broad-Coverage CommonCrawl Corpus and Pipeline for Minority Languages
by: Kargaran, Amir Hossein, et al.
Published: (2024)
by: Kargaran, Amir Hossein, et al.
Published: (2024)
Action Controlled Paraphrasing
by: Shi, Ning, et al.
Published: (2024)
by: Shi, Ning, et al.
Published: (2024)
BiST: A Gold Standard Bangla-English Bilingual Corpus for Sentence Structure and Tense Classification with Inter-Annotator Agreement
by: Shafi, Abdullah Al, et al.
Published: (2026)
by: Shafi, Abdullah Al, et al.
Published: (2026)
Language Modeling and Understanding Through Paraphrase Generation and Detection
by: Wahle, Jan Philip
Published: (2026)
by: Wahle, Jan Philip
Published: (2026)
OpenWHO: A Document-Level Parallel Corpus for Health Translation in Low-Resource Languages
by: Merx, Raphaël, et al.
Published: (2025)
by: Merx, Raphaël, et al.
Published: (2025)
Can LLMs Detect Intrinsic Hallucinations in Paraphrasing and Machine Translation?
by: Gogoulou, Evangelia, et al.
Published: (2025)
by: Gogoulou, Evangelia, et al.
Published: (2025)
Temperature Matters: Enhancing Watermark Robustness Against Paraphrasing Attacks
by: Idrissi, Badr Youbi, et al.
Published: (2025)
by: Idrissi, Badr Youbi, et al.
Published: (2025)
Spotting AI's Touch: Identifying LLM-Paraphrased Spans in Text
by: Li, Yafu, et al.
Published: (2024)
by: Li, Yafu, et al.
Published: (2024)
How Large Language Models are Transforming Machine-Paraphrased Plagiarism
by: Wahle, Jan Philip, et al.
Published: (2022)
by: Wahle, Jan Philip, et al.
Published: (2022)
Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities
by: Liu, Zhichen, et al.
Published: (2026)
by: Liu, Zhichen, et al.
Published: (2026)
BanglaNirTox: A Large-scale Parallel Corpus for Explainable AI in Bengali Text Detoxification
by: Mohsin, Ayesha Afroza, et al.
Published: (2025)
by: Mohsin, Ayesha Afroza, et al.
Published: (2025)
Latent Paraphrasing: Perturbation on Layers Improves Knowledge Injection in Language Models
by: Kang, Minki, et al.
Published: (2024)
by: Kang, Minki, et al.
Published: (2024)
On the Relationship between Sentence Analogy Identification and Sentence Structure Encoding in Large Language Models
by: Wijesiriwardene, Thilini, et al.
Published: (2023)
by: Wijesiriwardene, Thilini, et al.
Published: (2023)
PANER: A Paraphrase-Augmented Framework for Low-Resource Named Entity Recognition
by: Rengarajan, Nanda Kumar, et al.
Published: (2025)
by: Rengarajan, Nanda Kumar, et al.
Published: (2025)
ParaGuide: Guided Diffusion Paraphrasers for Plug-and-Play Textual Style Transfer
by: Horvitz, Zachary, et al.
Published: (2023)
by: Horvitz, Zachary, et al.
Published: (2023)
Target-Side Paraphrase Augmentation for Sign Language Translation with Large Language Models
by: Bianco, Pedro Dal, et al.
Published: (2026)
by: Bianco, Pedro Dal, et al.
Published: (2026)
How to Understand Named Entities: Using Common Sense for News Captioning
by: Xu, Ning, et al.
Published: (2024)
by: Xu, Ning, et al.
Published: (2024)
Mechanistic Decomposition of Sentence Representations
by: Tehenan, Matthieu, et al.
Published: (2025)
by: Tehenan, Matthieu, et al.
Published: (2025)
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
by: Kaneko, Masahiro
Published: (2026)
by: Kaneko, Masahiro
Published: (2026)
Meta4XNLI: A Crosslingual Parallel Corpus for Metaphor Detection and Interpretation
by: Sanchez-Bayona, Elisa, et al.
Published: (2024)
by: Sanchez-Bayona, Elisa, et al.
Published: (2024)
Demonstrations Are All You Need: Advancing Offensive Content Paraphrasing using In-Context Learning
by: Som, Anirudh, et al.
Published: (2023)
by: Som, Anirudh, et al.
Published: (2023)
PADBen: A Comprehensive Benchmark for Evaluating AI Text Detectors Against Paraphrase Attacks
by: Zha, Yiwei, et al.
Published: (2025)
by: Zha, Yiwei, et al.
Published: (2025)
Training With "Paraphrasing the Original Text" Teaches LLM to Better Retrieve in Long-context Tasks
by: Yu, Yijiong, et al.
Published: (2023)
by: Yu, Yijiong, et al.
Published: (2023)
Common Sense vs. Morality: The Curious Case of Narrative Focus Bias in LLMs
by: Purkayastha, Saugata, et al.
Published: (2026)
by: Purkayastha, Saugata, et al.
Published: (2026)
Navigating Semantic Relations: Challenges for Language Models in Abstract Common-Sense Reasoning
by: Gawin, Cole, et al.
Published: (2025)
by: Gawin, Cole, et al.
Published: (2025)
Edit-Constrained Decoding for Sentence Simplification
by: Zetsu, Tatsuya, et al.
Published: (2024)
by: Zetsu, Tatsuya, et al.
Published: (2024)
THInC: A Theory-Driven Framework for Computational Humor Detection
by: De Marez, Victor, et al.
Published: (2024)
by: De Marez, Victor, et al.
Published: (2024)
NepTam: A Nepali-Tamang Parallel Corpus and Baseline Machine Translation Experiments
by: Ghimire, Rupak Raj, et al.
Published: (2026)
by: Ghimire, Rupak Raj, et al.
Published: (2026)
Similar Items
-
An Agentic AI Framework for Training General Practitioner Student Skills
by: De Marez, Victor, et al.
Published: (2025) -
Estonian WinoGrande Dataset: Comparative Analysis of LLM Performance on Human and Machine Translation
by: Ojastu, Marii, et al.
Published: (2025) -
Bag of Lies: Robustness in Continuous Pre-training BERT
by: Gevers, Ine, et al.
Published: (2024) -
Do You Get the Hint? Benchmarking LLMs on the Board Game Concept
by: Gevers, Ine, et al.
Published: (2025) -
One Size Does Not Fit All: Exploring Variable Thresholds for Distance-Based Multi-Label Text Classification
by: Van Nooten, Jens, et al.
Published: (2025)