LingVarBench: Benchmarking LLMs on Entity Recognitions and Linguistic Verbalization Patterns in Phone-Call Transcripts
Fuente:
arXiv
Saved in:
| Main Authors: | Mohammadi, Seyedali, Paldhe, Manas, Chhabra, Amit, Son, Youngseo, Seshagiri, Vishal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph Integrated Language Transformers for Next Action Prediction in Complex Phone Calls
by: Marani, Amin Hosseiny, et al.
Published: (2024)
by: Marani, Amin Hosseiny, et al.
Published: (2024)
Auto Review: Second Stage Error Detection for Highly Accurate Information Extraction from Phone Conversations
by: Qamar, Ayesha, et al.
Published: (2025)
by: Qamar, Ayesha, et al.
Published: (2025)
Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models
by: Mohammadi, Seyedali, et al.
Published: (2026)
by: Mohammadi, Seyedali, et al.
Published: (2026)
LingBench++: A Linguistically-Informed Benchmark and Reasoning Framework for Multi-Step and Cross-Cultural Inference with LLMs
by: Lian, Da-Chen, et al.
Published: (2025)
by: Lian, Da-Chen, et al.
Published: (2025)
SAM2LoRA: Composite Loss-Guided, Parameter-Efficient Finetuning of SAM2 for Retinal Fundus Segmentation
by: Mandal, Sayan, et al.
Published: (2025)
by: Mandal, Sayan, et al.
Published: (2025)
IoT-Based Preventive Mental Health Using Knowledge Graphs and Standards for Better Well-Being
by: Gyrard, Amelie, et al.
Published: (2024)
by: Gyrard, Amelie, et al.
Published: (2024)
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
by: Zbeeb, Mohammad, et al.
Published: (2025)
by: Zbeeb, Mohammad, et al.
Published: (2025)
LingGym: How Far Are LLMs from Thinking Like Field Linguists?
by: Yang, Changbing, et al.
Published: (2025)
by: Yang, Changbing, et al.
Published: (2025)
Do LLMs Adhere to Label Definitions? Examining Their Receptivity to External Label Definitions
by: Mohammadi, Seyedali, et al.
Published: (2025)
by: Mohammadi, Seyedali, et al.
Published: (2025)
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions
by: Mohammadi, Seyedali, et al.
Published: (2024)
by: Mohammadi, Seyedali, et al.
Published: (2024)
LIEDER: Linguistically-Informed Evaluation for Discourse Entity Recognition
by: Zhu, Xiaomeng, et al.
Published: (2024)
by: Zhu, Xiaomeng, et al.
Published: (2024)
VarBench: Robust Language Model Benchmarking Through Dynamic Variable Perturbation
by: Qian, Kun, et al.
Published: (2024)
by: Qian, Kun, et al.
Published: (2024)
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation
by: Mohseni, Seyedreza, et al.
Published: (2024)
by: Mohseni, Seyedreza, et al.
Published: (2024)
TailNLG: A Multilingual Benchmark Addressing Verbalization of Long-Tail Entities
by: Draetta, Lia, et al.
Published: (2026)
by: Draetta, Lia, et al.
Published: (2026)
NoiseBench: Benchmarking the Impact of Real Label Noise on Named Entity Recognition
by: Merdjanovska, Elena, et al.
Published: (2024)
by: Merdjanovska, Elena, et al.
Published: (2024)
LingML: Linguistic-Informed Machine Learning for Enhanced Fake News Detection
by: Singh, Jasraj, et al.
Published: (2024)
by: Singh, Jasraj, et al.
Published: (2024)
modeLing: A Novel Dataset for Testing Linguistic Reasoning in Language Models
by: Chi, Nathan A., et al.
Published: (2024)
by: Chi, Nathan A., et al.
Published: (2024)
LingGen: Scalable Multi-Attribute Linguistic Control via Power-Law Masking
by: Elgaar, Mohamed, et al.
Published: (2024)
by: Elgaar, Mohamed, et al.
Published: (2024)
/UnmuteAll: Modeling Verbal Communication Patterns of Collaborative Contexts in MOBA Games
by: Son, Yongchan, et al.
Published: (2025)
by: Son, Yongchan, et al.
Published: (2025)
Multimodal Contextual Dialogue Breakdown Detection for Conversational AI Models
by: Miah, Md Messal Monem, et al.
Published: (2024)
by: Miah, Md Messal Monem, et al.
Published: (2024)
Generative Voice Bursts during Phone Call
by: Ranjan, Paritosh, et al.
Published: (2025)
by: Ranjan, Paritosh, et al.
Published: (2025)
Named Clinical Entity Recognition Benchmark
by: Abdul, Wadood M, et al.
Published: (2024)
by: Abdul, Wadood M, et al.
Published: (2024)
LLMs as Bridges: Reformulating Grounded Multimodal Named Entity Recognition
by: Li, Jinyuan, et al.
Published: (2024)
by: Li, Jinyuan, et al.
Published: (2024)
Attribution in Scientific Literature: New Benchmark and Methods
by: Saxena, Yash, et al.
Published: (2024)
by: Saxena, Yash, et al.
Published: (2024)
IOLBENCH: Benchmarking LLMs on Linguistic Reasoning
by: Goyal, Satyam, et al.
Published: (2025)
by: Goyal, Satyam, et al.
Published: (2025)
BenchHub: A Unified Benchmark Suite for Holistic and Customizable LLM Evaluation
by: Kim, Eunsu, et al.
Published: (2025)
by: Kim, Eunsu, et al.
Published: (2025)
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
by: Lajčinová, Bibiána, et al.
Published: (2024)
by: Lajčinová, Bibiána, et al.
Published: (2024)
LLMs in Biomedicine: A study on clinical Named Entity Recognition
by: Monajatipoor, Masoud, et al.
Published: (2024)
by: Monajatipoor, Masoud, et al.
Published: (2024)
On Verbalized Confidence Scores for LLMs
by: Yang, Daniel, et al.
Published: (2024)
by: Yang, Daniel, et al.
Published: (2024)
DarkBench: Benchmarking Dark Patterns in Large Language Models
by: Kran, Esben, et al.
Published: (2025)
by: Kran, Esben, et al.
Published: (2025)
Understanding Emotion in Discourse: Recognition Insights and Linguistic Patterns for Generation
by: Jeong, Cheonkam, et al.
Published: (2026)
by: Jeong, Cheonkam, et al.
Published: (2026)
Extracting Biomedical Entities from Noisy Audio Transcripts
by: Ebadi, Nima, et al.
Published: (2024)
by: Ebadi, Nima, et al.
Published: (2024)
Linguistic Loops and Geometric Invariants as a Way to Pre-Verbal Thought?
by: Corradetti, Daniele, et al.
Published: (2025)
by: Corradetti, Daniele, et al.
Published: (2025)
GlyphPattern: An Abstract Pattern Recognition Benchmark for Vision-Language Models
by: Wu, Zixuan, et al.
Published: (2024)
by: Wu, Zixuan, et al.
Published: (2024)
Verbalizing LLMs' assumptions to explain and control sycophancy
by: Cheng, Myra, et al.
Published: (2026)
by: Cheng, Myra, et al.
Published: (2026)
Cross-Linguistic Transcription and Phonological Representation in the Huìtóngguǎnxì Huáyíyìyǔ
by: Kim, Ji-eun
Published: (2026)
by: Kim, Ji-eun
Published: (2026)
An Empirical Recipe for Universal Phone Recognition
by: Bharadwaj, Shikhar, et al.
Published: (2026)
by: Bharadwaj, Shikhar, et al.
Published: (2026)
PRiSM: Benchmarking Phone Realization in Speech Models
by: Bharadwaj, Shikhar, et al.
Published: (2026)
by: Bharadwaj, Shikhar, et al.
Published: (2026)
Medical Spoken Named Entity Recognition
by: Le-Duc, Khai, et al.
Published: (2024)
by: Le-Duc, Khai, et al.
Published: (2024)
BioUNER: A Benchmark Dataset for Clinical Urdu Named Entity Recognition
by: Ali, Wazir, et al.
Published: (2026)
by: Ali, Wazir, et al.
Published: (2026)
Similar Items
-
Graph Integrated Language Transformers for Next Action Prediction in Complex Phone Calls
by: Marani, Amin Hosseiny, et al.
Published: (2024) -
Auto Review: Second Stage Error Detection for Highly Accurate Information Extraction from Phone Conversations
by: Qamar, Ayesha, et al.
Published: (2025) -
Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models
by: Mohammadi, Seyedali, et al.
Published: (2026) -
LingBench++: A Linguistically-Informed Benchmark and Reasoning Framework for Multi-Step and Cross-Cultural Inference with LLMs
by: Lian, Da-Chen, et al.
Published: (2025) -
SAM2LoRA: Composite Loss-Guided, Parameter-Efficient Finetuning of SAM2 for Retinal Fundus Segmentation
by: Mandal, Sayan, et al.
Published: (2025)