Thai Winograd Schemas: A Benchmark for Thai Commonsense Reasoning
Fuente:
arXiv
Saved in:
| Main Author: | Artkaew, Phakphum |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ThaiOCRBench: A Task-Diverse Benchmark for Vision-Language Understanding in Thai
by: Nonesung, Surapon, et al.
Published: (2025)
by: Nonesung, Surapon, et al.
Published: (2025)
ThaiCoref: Thai Coreference Resolution Dataset
by: Trakuekul, Pontakorn, et al.
Published: (2024)
by: Trakuekul, Pontakorn, et al.
Published: (2024)
OpenThaiGPT 1.6 and R1: Thai-Centric Open Source and Reasoning Large Language Models
by: Yuenyong, Sumeth, et al.
Published: (2025)
by: Yuenyong, Sumeth, et al.
Published: (2025)
ThaiSafetyBench: Assessing Language Model Safety in Thai Cultural Contexts
by: Ukarapol, Trapoom, et al.
Published: (2026)
by: Ukarapol, Trapoom, et al.
Published: (2026)
Typhoon T1: An Open Thai Reasoning Model
by: Taveekitworachai, Pittawat, et al.
Published: (2025)
by: Taveekitworachai, Pittawat, et al.
Published: (2025)
PyThaiNLP: Thai Natural Language Processing in Python
by: Phatthiyaphaibun, Wannaphong, et al.
Published: (2023)
by: Phatthiyaphaibun, Wannaphong, et al.
Published: (2023)
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries
by: Sun, Jing Han, et al.
Published: (2024)
by: Sun, Jing Han, et al.
Published: (2024)
Picturing Ambiguity: A Visual Twist on the Winograd Schema Challenge
by: Park, Brendan, et al.
Published: (2024)
by: Park, Brendan, et al.
Published: (2024)
Thai Universal Dependency Treebank
by: Sriwirote, Panyut, et al.
Published: (2024)
by: Sriwirote, Panyut, et al.
Published: (2024)
OpenThaiGPT 1.5: A Thai-Centric Open Source Large Language Model
by: Yuenyong, Sumeth, et al.
Published: (2024)
by: Yuenyong, Sumeth, et al.
Published: (2024)
THaLLE-ThaiLLM: Domain-Specialized Small LLMs for Finance and Thai -- Technical Report
by: Labs, KBTG, et al.
Published: (2026)
by: Labs, KBTG, et al.
Published: (2026)
WangchanThaiInstruct: An instruction-following Dataset for Culture-Aware, Multitask, and Multi-domain Evaluation in Thai
by: Limkonchotiwat, Peerat, et al.
Published: (2025)
by: Limkonchotiwat, Peerat, et al.
Published: (2025)
PhayaThaiBERT: Enhancing a Pretrained Thai Language Model with Unassimilated Loanwords
by: Sriwirote, Panyut, et al.
Published: (2023)
by: Sriwirote, Panyut, et al.
Published: (2023)
WSC+: Enhancing The Winograd Schema Challenge Using Tree-of-Experts
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
Assessing Thai Dialect Performance in LLMs with Automatic Benchmarks and Human Evaluation
by: Limkonchotiwat, Peerat, et al.
Published: (2025)
by: Limkonchotiwat, Peerat, et al.
Published: (2025)
Concept-Reversed Winograd Schema Challenge: Evaluating and Improving Robust Reasoning in Large Language Models via Abstraction
by: Han, Kaiqiao, et al.
Published: (2024)
by: Han, Kaiqiao, et al.
Published: (2024)
Eir: Thai Medical Large Language Models
by: Thiprak, Yutthakorn, et al.
Published: (2024)
by: Thiprak, Yutthakorn, et al.
Published: (2024)
JaiTTS: A Thai Voice Cloning Model
by: Karnjanaekarin, Jullajak, et al.
Published: (2026)
by: Karnjanaekarin, Jullajak, et al.
Published: (2026)
Solving the Challenge Set without Solving the Task: On Winograd Schemas as a Test of Pronominal Coreference Resolution
by: Porada, Ian, et al.
Published: (2024)
by: Porada, Ian, et al.
Published: (2024)
Can Group Relative Policy Optimization Improve Thai Legal Reasoning and Question Answering?
by: Akarajaradwong, Pawitsapak, et al.
Published: (2025)
by: Akarajaradwong, Pawitsapak, et al.
Published: (2025)
Headline-Guided Extractive Summarization for Thai News Articles
by: Kositcharoensuk, Pimpitchaya, et al.
Published: (2024)
by: Kositcharoensuk, Pimpitchaya, et al.
Published: (2024)
Mangosteen: An Open Thai Corpus for Language Model Pretraining
by: Phatthiyaphaibun, Wannaphong, et al.
Published: (2025)
by: Phatthiyaphaibun, Wannaphong, et al.
Published: (2025)
JAI-1: A Thai-Centric Large Language Model
by: Rutherford, Attapol T., et al.
Published: (2025)
by: Rutherford, Attapol T., et al.
Published: (2025)
AyutthayaAlpha: A Thai-Latin Script Transliteration Transformer
by: Lauc, Davor, et al.
Published: (2024)
by: Lauc, Davor, et al.
Published: (2024)
Representing the Under-Represented: Cultural and Core Capability Benchmarks for Developing Thai Large Language Models
by: Kim, Dahyun, et al.
Published: (2024)
by: Kim, Dahyun, et al.
Published: (2024)
Typhoon OCR: Open Vision-Language Model For Thai Document Extraction
by: Nonesung, Surapon, et al.
Published: (2026)
by: Nonesung, Surapon, et al.
Published: (2026)
LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning
by: Junias, Obed, et al.
Published: (2026)
by: Junias, Obed, et al.
Published: (2026)
Thai Financial Domain Adaptation of THaLLE -- Technical Report
by: Labs, KBTG, et al.
Published: (2024)
by: Labs, KBTG, et al.
Published: (2024)
SCoRE: Benchmarking Long-Chain Reasoning in Commonsense Scenarios
by: Zhan, Weidong, et al.
Published: (2025)
by: Zhan, Weidong, et al.
Published: (2025)
LOTUSDIS: A Thai far-field meeting corpus for robust conversational ASR
by: Tipaksorn, Pattara, et al.
Published: (2025)
by: Tipaksorn, Pattara, et al.
Published: (2025)
SiamGPT: Quality-First Fine-Tuning for Stable Thai Text Generation
by: Pairatsuppawat, Thittipat, et al.
Published: (2025)
by: Pairatsuppawat, Thittipat, et al.
Published: (2025)
Benchmarking Chinese Commonsense Reasoning with a Multi-hop Reasoning Perspective
by: You, Wangjie, et al.
Published: (2025)
by: You, Wangjie, et al.
Published: (2025)
Thai Semantic End-of-Turn Detection for Real-Time Voice Agents
by: Popit, Thanapol, et al.
Published: (2025)
by: Popit, Thanapol, et al.
Published: (2025)
The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning
by: Cui, Shaobo, et al.
Published: (2024)
by: Cui, Shaobo, et al.
Published: (2024)
Benchmarking Chinese Commonsense Reasoning of LLMs: From Chinese-Specifics to Reasoning-Memorization Correlations
by: Sun, Jiaxing, et al.
Published: (2024)
by: Sun, Jiaxing, et al.
Published: (2024)
NitiBench: A Comprehensive Study of LLM Framework Capabilities for Thai Legal Question Answering
by: Akarajaradwong, Pawitsapak, et al.
Published: (2025)
by: Akarajaradwong, Pawitsapak, et al.
Published: (2025)
Typhoon ASR Real-time: FastConformer-Transducer for Thai Automatic Speech Recognition
by: Sirichotedumrong, Warit, et al.
Published: (2026)
by: Sirichotedumrong, Warit, et al.
Published: (2026)
Aspect-Level Obfuscated Sentiment in Thai Financial Disclosures and Its Impact on Abnormal Returns
by: Rutherford, Attapol T., et al.
Published: (2025)
by: Rutherford, Attapol T., et al.
Published: (2025)
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning
by: Tang, Zhisheng, et al.
Published: (2024)
by: Tang, Zhisheng, et al.
Published: (2024)
Plausibly Problematic Questions in Multiple-Choice Benchmarks for Commonsense Reasoning
by: Palta, Shramay, et al.
Published: (2024)
by: Palta, Shramay, et al.
Published: (2024)
Similar Items
-
ThaiOCRBench: A Task-Diverse Benchmark for Vision-Language Understanding in Thai
by: Nonesung, Surapon, et al.
Published: (2025) -
ThaiCoref: Thai Coreference Resolution Dataset
by: Trakuekul, Pontakorn, et al.
Published: (2024) -
OpenThaiGPT 1.6 and R1: Thai-Centric Open Source and Reasoning Large Language Models
by: Yuenyong, Sumeth, et al.
Published: (2025) -
ThaiSafetyBench: Assessing Language Model Safety in Thai Cultural Contexts
by: Ukarapol, Trapoom, et al.
Published: (2026) -
Typhoon T1: An Open Thai Reasoning Model
by: Taveekitworachai, Pittawat, et al.
Published: (2025)