A Large-Scale Dataset and Citation Intent Classification in Turkish with LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Karaca, Kemal Sami, Eravcı, Bahaeddin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal Stock Price Prediction
von: Karadaş, Furkan, et al.
Veröffentlicht: (2025)
von: Karadaş, Furkan, et al.
Veröffentlicht: (2025)
REIC: RAG-Enhanced Intent Classification at Scale
von: Zhang, Ziji, et al.
Veröffentlicht: (2025)
von: Zhang, Ziji, et al.
Veröffentlicht: (2025)
WolBanking77: Wolof Banking Speech Intent Classification Dataset
von: Kandji, Abdou Karim, et al.
Veröffentlicht: (2025)
von: Kandji, Abdou Karim, et al.
Veröffentlicht: (2025)
ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models
von: Düzkar, Kemal
Veröffentlicht: (2026)
von: Düzkar, Kemal
Veröffentlicht: (2026)
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)
From Intents to Conversations: Generating Intent-Driven Dialogues with Contrastive Learning for Multi-Turn Classification
von: Liu, Junhua, et al.
Veröffentlicht: (2024)
von: Liu, Junhua, et al.
Veröffentlicht: (2024)
Ellipsoid-Based Decision Boundaries for Open Intent Classification
von: Zou, Yuetian, et al.
Veröffentlicht: (2025)
von: Zou, Yuetian, et al.
Veröffentlicht: (2025)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
BIRDTurk: Adaptation of the BIRD Text-to-SQL Dataset to Turkish
von: Aktaş, Burak, et al.
Veröffentlicht: (2026)
von: Aktaş, Burak, et al.
Veröffentlicht: (2026)
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
Organic Data-Driven Approach for Turkish Grammatical Error Correction and LLMs
von: Ersoy, Asım, et al.
Veröffentlicht: (2024)
von: Ersoy, Asım, et al.
Veröffentlicht: (2024)
Evaluating the Quality of Benchmark Datasets for Low-Resource Languages: A Case Study on Turkish
von: Cengiz, Ayşe Aysu, et al.
Veröffentlicht: (2025)
von: Cengiz, Ayşe Aysu, et al.
Veröffentlicht: (2025)
Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues
von: Bayrak, Ahmet Tuğrul, et al.
Veröffentlicht: (2026)
von: Bayrak, Ahmet Tuğrul, et al.
Veröffentlicht: (2026)
BengaliSent140: A Large-Scale Bengali Binary Sentiment Dataset for Hate and Non-Hate Speech Classification
von: Islam, Akif, et al.
Veröffentlicht: (2026)
von: Islam, Akif, et al.
Veröffentlicht: (2026)
TurkBench: A Benchmark for Evaluating Turkish Large Language Models
von: Toraman, Çağrı, et al.
Veröffentlicht: (2026)
von: Toraman, Çağrı, et al.
Veröffentlicht: (2026)
TR-EduVSum: A Turkish-Focused Dataset and Consensus Framework for Educational Video Summarization
von: Eğin, Figen, et al.
Veröffentlicht: (2026)
von: Eğin, Figen, et al.
Veröffentlicht: (2026)
ResumeAtlas: Revisiting Resume Classification with Large-Scale Datasets and Large Language Models
von: Heakl, Ahmed, et al.
Veröffentlicht: (2024)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2024)
Training LLMs with Reinforcement Learning for Intent-Aware Personalized Question Answering
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2026)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2026)
Intent Mismatch Causes LLMs to Get Lost in Multi-Turn Conversation
von: Liu, Geng, et al.
Veröffentlicht: (2026)
von: Liu, Geng, et al.
Veröffentlicht: (2026)
Citekit: A Modular Toolkit for Large Language Model Citation Generation
von: Shen, Jiajun, et al.
Veröffentlicht: (2024)
von: Shen, Jiajun, et al.
Veröffentlicht: (2024)
Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation
von: Karakaş, Sercan, et al.
Veröffentlicht: (2026)
von: Karakaş, Sercan, et al.
Veröffentlicht: (2026)
Non-Halting Queries: Exploiting Fixed Points in LLMs
von: Hammouri, Ghaith, et al.
Veröffentlicht: (2024)
von: Hammouri, Ghaith, et al.
Veröffentlicht: (2024)
Aligning Large Language Model Behavior with Human Citation Preferences
von: Ando, Kenichiro, et al.
Veröffentlicht: (2026)
von: Ando, Kenichiro, et al.
Veröffentlicht: (2026)
Impact of Stickers on Multimodal Sentiment and Intent in Social Media: A New Task, Dataset and Baseline
von: Shi, Yuanchen, et al.
Veröffentlicht: (2024)
von: Shi, Yuanchen, et al.
Veröffentlicht: (2024)
PromptTailor: Multi-turn Intent-Aligned Prompt Synthesis for Lightweight LLMs
von: Xu, Yizhou, et al.
Veröffentlicht: (2025)
von: Xu, Yizhou, et al.
Veröffentlicht: (2025)
MessIRve: A Large-Scale Spanish Information Retrieval Dataset
von: Valentini, Francisco, et al.
Veröffentlicht: (2024)
von: Valentini, Francisco, et al.
Veröffentlicht: (2024)
Where Fake Citations Are Made: Tracing Field-Level Hallucination to Specific Neurons in LLMs
von: Chen, Yuefei, et al.
Veröffentlicht: (2026)
von: Chen, Yuefei, et al.
Veröffentlicht: (2026)
Optimal Turkish Subword Strategies at Scale: Systematic Evaluation of Data, Vocabulary, Morphology Interplay
von: Altinok, Duygu
Veröffentlicht: (2026)
von: Altinok, Duygu
Veröffentlicht: (2026)
Beyond Facts: Evaluating Intent Hallucination in Large Language Models
von: Hao, Yijie, et al.
Veröffentlicht: (2025)
von: Hao, Yijie, et al.
Veröffentlicht: (2025)
Hierarchical Memorization in Large Language Models: Evidence from Citation Generation
von: Niimi, Junichiro
Veröffentlicht: (2025)
von: Niimi, Junichiro
Veröffentlicht: (2025)
Learning Fine-Grained Grounded Citations for Attributed Large Language Models
von: Huang, Lei, et al.
Veröffentlicht: (2024)
von: Huang, Lei, et al.
Veröffentlicht: (2024)
BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
von: Kumar, Divyanshu, et al.
Veröffentlicht: (2024)
von: Kumar, Divyanshu, et al.
Veröffentlicht: (2024)
Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs
von: Gody, Reem, et al.
Veröffentlicht: (2025)
von: Gody, Reem, et al.
Veröffentlicht: (2025)
Is the Pope Catholic? Yes, the Pope is Catholic. Generative Evaluation of Non-Literal Intent Resolution in LLMs
von: Yerukola, Akhila, et al.
Veröffentlicht: (2024)
von: Yerukola, Akhila, et al.
Veröffentlicht: (2024)
VietLyrics: A Large-Scale Dataset and Models for Vietnamese Automatic Lyrics Transcription
von: Nguyen, Quoc Anh, et al.
Veröffentlicht: (2025)
von: Nguyen, Quoc Anh, et al.
Veröffentlicht: (2025)
SlovKE: A Large-Scale Dataset and LLM Evaluation for Slovak Keyphrase Extraction
von: Števaňák, David, et al.
Veröffentlicht: (2026)
von: Števaňák, David, et al.
Veröffentlicht: (2026)
OCRTurk: A Comprehensive OCR Benchmark for Turkish
von: Yılmaz, Deniz, et al.
Veröffentlicht: (2026)
von: Yılmaz, Deniz, et al.
Veröffentlicht: (2026)
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction
von: Saoud, Abdulkader, et al.
Veröffentlicht: (2024)
von: Saoud, Abdulkader, et al.
Veröffentlicht: (2024)
Large-Scale Constraint Generation -- Can LLMs Parse Hundreds of Constraints?
von: Boffa, Matteo, et al.
Veröffentlicht: (2025)
von: Boffa, Matteo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multimodal Stock Price Prediction
von: Karadaş, Furkan, et al.
Veröffentlicht: (2025) -
REIC: RAG-Enhanced Intent Classification at Scale
von: Zhang, Ziji, et al.
Veröffentlicht: (2025) -
WolBanking77: Wolof Banking Speech Intent Classification Dataset
von: Kandji, Abdou Karim, et al.
Veröffentlicht: (2025) -
ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models
von: Düzkar, Kemal
Veröffentlicht: (2026) -
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
von: Wu, Chengwei, et al.
Veröffentlicht: (2025)