Splits! Flexible Sociocultural Linguistic Investigation at Scale
Fuente:
arXiv
Saved in:
| Main Authors: | Caplan, Eylon, Chakraborty, Tania, Goldwasser, Dan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ConceptCarve: Dynamic Realization of Evidence
by: Caplan, Eylon, et al.
Published: (2025)
by: Caplan, Eylon, et al.
Published: (2025)
VIBE: Can a VLM Read the Room?
by: Chakraborty, Tania, et al.
Published: (2025)
by: Chakraborty, Tania, et al.
Published: (2025)
TAIGR: Towards Modeling Influencer Content on Social Media via Structured, Pragmatic Inference
by: Nakshatri, Nishanth Sridhar, et al.
Published: (2026)
by: Nakshatri, Nishanth Sridhar, et al.
Published: (2026)
LLM-Human Pipeline for Cultural Context Grounding of Conversations
by: Pujari, Rajkumar, et al.
Published: (2024)
by: Pujari, Rajkumar, et al.
Published: (2024)
CoLa: Learning to Interactively Collaborate with Large Language Models
by: Sharma, Abhishek, et al.
Published: (2025)
by: Sharma, Abhishek, et al.
Published: (2025)
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
by: Seoh, Ronald, et al.
Published: (2025)
by: Seoh, Ronald, et al.
Published: (2025)
Post-hoc Study of Climate Microtargeting on Social Media Ads with LLMs: Thematic Insights and Fairness Evaluation
by: Islam, Tunazzina, et al.
Published: (2024)
by: Islam, Tunazzina, et al.
Published: (2024)
Discovering Latent Themes in Social Media Messaging: A Machine-in-the-Loop Approach Integrating LLMs
by: Islam, Tunazzina, et al.
Published: (2024)
by: Islam, Tunazzina, et al.
Published: (2024)
Uncovering Latent Arguments in Social Media Messaging by Employing LLMs-in-the-Loop Strategy
by: Islam, Tunazzina, et al.
Published: (2024)
by: Islam, Tunazzina, et al.
Published: (2024)
Can LLMs Assist Annotators in Identifying Morality Frames? -- Case Study on Vaccination Debate on Social Media
by: Islam, Tunazzina, et al.
Published: (2025)
by: Islam, Tunazzina, et al.
Published: (2025)
An Investigation of Linguistic Biases in LLM-Based Recommendations
by: Venkateswaran, Nitin, et al.
Published: (2026)
by: Venkateswaran, Nitin, et al.
Published: (2026)
Large Linguistic Models: Investigating LLMs' metalinguistic abilities
by: Beguš, Gašper, et al.
Published: (2023)
by: Beguš, Gašper, et al.
Published: (2023)
Investigating Large Language Models' Linguistic Abilities for Text Preprocessing
by: Braga, Marco, et al.
Published: (2025)
by: Braga, Marco, et al.
Published: (2025)
Linguistic and Argument Diversity in Synthetic Data for Function-Calling Agents
by: Greenstein, Dan, et al.
Published: (2026)
by: Greenstein, Dan, et al.
Published: (2026)
Integrating Linguistics and AI: Morphological Analysis and Corpus development of Endangered Toto Language of West Bengal
by: Guha, Ambalika, et al.
Published: (2025)
by: Guha, Ambalika, et al.
Published: (2025)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
by: Cheng, Myra, et al.
Published: (2024)
by: Cheng, Myra, et al.
Published: (2024)
Unsupervised Translation of Emergent Communication
by: Levy, Ido, et al.
Published: (2025)
by: Levy, Ido, et al.
Published: (2025)
Parallel Scaling Law: Unveiling Reasoning Generalization through A Cross-Linguistic Perspective
by: Yang, Wen, et al.
Published: (2025)
by: Yang, Wen, et al.
Published: (2025)
LLM NL2SQL Robustness: Surface Noise vs. Linguistic Variation in Traditional and Agentic Settings
by: Tu, Lifu, et al.
Published: (2026)
by: Tu, Lifu, et al.
Published: (2026)
A Linguistic Analysis of Spontaneous Thoughts: Investigating Experiences of Déjà Vu, Unexpected Thoughts, and Involuntary Autobiographical Memories
by: Venkatesha, Videep, et al.
Published: (2025)
by: Venkatesha, Videep, et al.
Published: (2025)
LLaMAX: Scaling Linguistic Horizons of LLM by Enhancing Translation Capabilities Beyond 100 Languages
by: Lu, Yinquan, et al.
Published: (2024)
by: Lu, Yinquan, et al.
Published: (2024)
Affine-Scaled Attention: Towards Flexible and Stable Transformer Attention
by: Bae, Jeongin, et al.
Published: (2026)
by: Bae, Jeongin, et al.
Published: (2026)
Hallucination Detection in Foundation Models for Decision-Making: A Flexible Definition and Review of the State of the Art
by: Chakraborty, Neeloy, et al.
Published: (2024)
by: Chakraborty, Neeloy, et al.
Published: (2024)
Improving Bangla Linguistics: Advanced LSTM, Bi-LSTM, and Seq2Seq Models for Translating Sylheti to Modern Bangla
by: Das, Sourav Kumar, et al.
Published: (2025)
by: Das, Sourav Kumar, et al.
Published: (2025)
Leveraging Wikidata for Geographically Informed Sociocultural Bias Dataset Creation: Application to Latin America
by: Karmim, Yannis, et al.
Published: (2026)
by: Karmim, Yannis, et al.
Published: (2026)
MELA: Multilingual Evaluation of Linguistic Acceptability
by: Zhang, Ziyin, et al.
Published: (2023)
by: Zhang, Ziyin, et al.
Published: (2023)
Natural Language Processing RELIES on Linguistics
by: Opitz, Juri, et al.
Published: (2024)
by: Opitz, Juri, et al.
Published: (2024)
Linguistically Conditioned Semantic Textual Similarity
by: Tu, Jingxuan, et al.
Published: (2024)
by: Tu, Jingxuan, et al.
Published: (2024)
Efficiency at Scale: Investigating the Performance of Diminutive Language Models in Clinical Tasks
by: Taylor, Niall, et al.
Published: (2024)
by: Taylor, Niall, et al.
Published: (2024)
Say It Differently: Linguistic Styles as Jailbreak Vectors
by: Panda, Srikant, et al.
Published: (2025)
by: Panda, Srikant, et al.
Published: (2025)
Linguistic traces of stochastic empathy in language models
by: Kleinberg, Bennett, et al.
Published: (2024)
by: Kleinberg, Bennett, et al.
Published: (2024)
Inductive Linguistic Reasoning with Large Language Models
by: Ramji, Raghav, et al.
Published: (2024)
by: Ramji, Raghav, et al.
Published: (2024)
Linguistic Structure Induction from Language Models
by: Momen, Omar
Published: (2024)
by: Momen, Omar
Published: (2024)
Linguistic Profiling of a Neural Language Model
by: Miaschi, Alessio, et al.
Published: (2020)
by: Miaschi, Alessio, et al.
Published: (2020)
Linguistic Blind Spots in Clinical Decision Extraction
by: Elgaar, Mohamed, et al.
Published: (2026)
by: Elgaar, Mohamed, et al.
Published: (2026)
Linguistics-Aware Non-Distortionary LLM Watermarking
by: Park, Shinwoo, et al.
Published: (2026)
by: Park, Shinwoo, et al.
Published: (2026)
Scaling, Simplification, and Adaptation: Lessons from Pretraining on Machine-Translated Text
by: Velasco, Dan John, et al.
Published: (2025)
by: Velasco, Dan John, et al.
Published: (2025)
Geometric Monomial (GEM): a family of rational 2N-differentiable activation functions
by: Krause, Eylon E.
Published: (2026)
by: Krause, Eylon E.
Published: (2026)
Mapping Clinical Doubt: Locating Linguistic Uncertainty in LLMs
by: Sridhar, Srivarshinee, et al.
Published: (2025)
by: Sridhar, Srivarshinee, et al.
Published: (2025)
Are Large Language Models the future crowd workers of Linguistics?
by: Ferrazzo, Iris
Published: (2025)
by: Ferrazzo, Iris
Published: (2025)
Similar Items
-
ConceptCarve: Dynamic Realization of Evidence
by: Caplan, Eylon, et al.
Published: (2025) -
VIBE: Can a VLM Read the Room?
by: Chakraborty, Tania, et al.
Published: (2025) -
TAIGR: Towards Modeling Influencer Content on Social Media via Structured, Pragmatic Inference
by: Nakshatri, Nishanth Sridhar, et al.
Published: (2026) -
LLM-Human Pipeline for Cultural Context Grounding of Conversations
by: Pujari, Rajkumar, et al.
Published: (2024) -
CoLa: Learning to Interactively Collaborate with Large Language Models
by: Sharma, Abhishek, et al.
Published: (2025)