Strategies for political-statement segmentation and labelling in unstructured text
Fuente:
arXiv
Saved in:
| Main Authors: | Nikolaev, Dmitry, Papay, Sean |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Language Modeling for Low-Resource Settings with Hybrid RNN-Transformer Architectures
by: Lindenmaier, Gabriel, et al.
Published: (2025)
by: Lindenmaier, Gabriel, et al.
Published: (2025)
Actor Identification in Discourse: A Challenge for LLMs?
by: Barić, Ana, et al.
Published: (2024)
by: Barić, Ana, et al.
Published: (2024)
Are Humans as Brittle as Large Language Models?
by: Li, Jiahui, et al.
Published: (2025)
by: Li, Jiahui, et al.
Published: (2025)
Classifier identification in Ancient Egyptian as a low-resource sequence-labelling task
by: Nikolaev, Dmitry, et al.
Published: (2024)
by: Nikolaev, Dmitry, et al.
Published: (2024)
Regular-pattern-sensitive CRFs for Distant Label Interactions
by: Papay, Sean, et al.
Published: (2024)
by: Papay, Sean, et al.
Published: (2024)
Beyond prompt brittleness: Evaluating the reliability and consistency of political worldviews in LLMs
by: Ceron, Tanise, et al.
Published: (2024)
by: Ceron, Tanise, et al.
Published: (2024)
iPOE: Interpretable Prompt Optimization via Explanations
by: Li, Jiahui, et al.
Published: (2026)
by: Li, Jiahui, et al.
Published: (2026)
Identifying attributions of causality in political text
by: Garcia-Corral, Paulina
Published: (2025)
by: Garcia-Corral, Paulina
Published: (2025)
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts
by: Gokceoglu, Gokcen, et al.
Published: (2024)
by: Gokceoglu, Gokcen, et al.
Published: (2024)
Approximate Attributions for Off-the-Shelf Siamese Transformers
by: Möller, Lucas, et al.
Published: (2024)
by: Möller, Lucas, et al.
Published: (2024)
KeNet:Knowledge-enhanced Doc-Label Attention Network for Multi-label text classification
by: Li, Bo, et al.
Published: (2024)
by: Li, Bo, et al.
Published: (2024)
A large-scale image-text dataset benchmark for farmland segmentation
by: Tao, Chao, et al.
Published: (2025)
by: Tao, Chao, et al.
Published: (2025)
The study of short texts in digital politics: Document aggregation for topic modeling
by: Nakka, Nitheesha, et al.
Published: (2025)
by: Nakka, Nitheesha, et al.
Published: (2025)
What Is The Political Content in LLMs' Pre- and Post-Training Data?
by: Ceron, Tanise, et al.
Published: (2025)
by: Ceron, Tanise, et al.
Published: (2025)
AlleNoise: large-scale text classification benchmark dataset with real-world label noise
by: Rączkowska, Alicja, et al.
Published: (2024)
by: Rączkowska, Alicja, et al.
Published: (2024)
AI-generated text boundary detection with RoFT
by: Kushnareva, Laida, et al.
Published: (2023)
by: Kushnareva, Laida, et al.
Published: (2023)
Cropping outperforms dropout as an augmentation strategy for self-supervised training of text embeddings
by: González-Márquez, Rita, et al.
Published: (2025)
by: González-Márquez, Rita, et al.
Published: (2025)
LLMs left, right, and center: Assessing GPT's capabilities to label political bias from web domains
by: Hernandes, Raphael, et al.
Published: (2024)
by: Hernandes, Raphael, et al.
Published: (2024)
Text to model via SysML: Automated generation of dynamical system computational models from unstructured natural language text via enhanced System Modeling Language diagrams
by: Hendricks, Matthew Anderson, et al.
Published: (2025)
by: Hendricks, Matthew Anderson, et al.
Published: (2025)
A suite of LMs comprehend puzzle statements as well as humans
by: Goldberg, Adele E, et al.
Published: (2025)
by: Goldberg, Adele E, et al.
Published: (2025)
Synthetically generated text for supervised text analysis
by: Halterman, Andrew
Published: (2023)
by: Halterman, Andrew
Published: (2023)
Automating Venture Capital: Founder assessment using LLM-powered segmentation, feature engineering and automated labeling techniques
by: Ozince, Ekin, et al.
Published: (2024)
by: Ozince, Ekin, et al.
Published: (2024)
Donate or Create? Comparing Data Collection Strategies for Emotion-labeled Multimodal Social Media Posts
by: Bagdon, Christopher, et al.
Published: (2025)
by: Bagdon, Christopher, et al.
Published: (2025)
Which Demographics do LLMs Default to During Annotation?
by: Schäfer, Johannes, et al.
Published: (2024)
by: Schäfer, Johannes, et al.
Published: (2024)
Efficient slot labelling
by: Vlasov, Vladimir
Published: (2024)
by: Vlasov, Vladimir
Published: (2024)
Prompt-based Pseudo-labeling Strategy for Sample-Efficient Semi-Supervised Extractive Summarization
by: Sahu, Gaurav, et al.
Published: (2023)
by: Sahu, Gaurav, et al.
Published: (2023)
BP-Seg: A graphical model approach to unsupervised and non-contiguous text segmentation using belief propagation
by: Li, Fengyi, et al.
Published: (2025)
by: Li, Fengyi, et al.
Published: (2025)
Measuring short-form factuality in large language models
by: Wei, Jason, et al.
Published: (2024)
by: Wei, Jason, et al.
Published: (2024)
Can large language models interpret unstructured chat data on dynamic group decision-making processes? Evidence on joint destination choice
by: Lim, Sung-Yoo, et al.
Published: (2026)
by: Lim, Sung-Yoo, et al.
Published: (2026)
How do we measure privacy in text? A survey of text anonymization metrics
by: Ren, Yaxuan, et al.
Published: (2025)
by: Ren, Yaxuan, et al.
Published: (2025)
Identifying social isolation themes in NVDRS text narratives using topic modeling and text-classification methods
by: Walker, Drew, et al.
Published: (2025)
by: Walker, Drew, et al.
Published: (2025)
Qwen it detect machine-generated text?
by: Marchitan, Teodor-George, et al.
Published: (2025)
by: Marchitan, Teodor-George, et al.
Published: (2025)
Few-shot text-based emotion detection
by: Marchitan, Teodor-George, et al.
Published: (2025)
by: Marchitan, Teodor-George, et al.
Published: (2025)
Transferable text data distillation by trajectory matching
by: Yao, Rong, et al.
Published: (2025)
by: Yao, Rong, et al.
Published: (2025)
LUQ: Long-text Uncertainty Quantification for LLMs
by: Zhang, Caiqi, et al.
Published: (2024)
by: Zhang, Caiqi, et al.
Published: (2024)
Leveraging the power of transformers for guilt detection in text
by: Meque, Abdul Gafar Manuel, et al.
Published: (2024)
by: Meque, Abdul Gafar Manuel, et al.
Published: (2024)
Analyzing Finnish Inflectional Classes through Discriminative Lexicon and Deep Learning Models
by: Nikolaev, Alexandre, et al.
Published: (2025)
by: Nikolaev, Alexandre, et al.
Published: (2025)
QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies
by: Khoroshilov, Alexey, et al.
Published: (2026)
by: Khoroshilov, Alexey, et al.
Published: (2026)
BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents
by: Wei, Jason, et al.
Published: (2025)
by: Wei, Jason, et al.
Published: (2025)
MatchXML: An Efficient Text-label Matching Framework for Extreme Multi-label Text Classification
by: Ye, Hui, et al.
Published: (2023)
by: Ye, Hui, et al.
Published: (2023)
Similar Items
-
Efficient Language Modeling for Low-Resource Settings with Hybrid RNN-Transformer Architectures
by: Lindenmaier, Gabriel, et al.
Published: (2025) -
Actor Identification in Discourse: A Challenge for LLMs?
by: Barić, Ana, et al.
Published: (2024) -
Are Humans as Brittle as Large Language Models?
by: Li, Jiahui, et al.
Published: (2025) -
Classifier identification in Ancient Egyptian as a low-resource sequence-labelling task
by: Nikolaev, Dmitry, et al.
Published: (2024) -
Regular-pattern-sensitive CRFs for Distant Label Interactions
by: Papay, Sean, et al.
Published: (2024)