Recent Trends in Linear Text Segmentation: a Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Ghinassi, Iacopo, Wang, Lin, Newell, Chris, Purver, Matthew |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly
by: Hosseini, Peyman, et al.
Published: (2024)
by: Hosseini, Peyman, et al.
Published: (2024)
Efficient Aspect-Based Summarization of Climate Change Reports with Small Language Models
by: Ghinassi, Iacopo, et al.
Published: (2024)
by: Ghinassi, Iacopo, et al.
Published: (2024)
Lon-ea at SemEval-2023 Task 11: A Comparison of Activation Functions for Soft and Hard Label Prediction
by: Hosseini, Peyman, et al.
Published: (2023)
by: Hosseini, Peyman, et al.
Published: (2023)
Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
by: Ma, Longxuan, et al.
Published: (2024)
by: Ma, Longxuan, et al.
Published: (2024)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors
by: Dugan, Liam, et al.
Published: (2024)
by: Dugan, Liam, et al.
Published: (2024)
Controllable Text Generation for Large Language Models: A Survey
by: Liang, Xun, et al.
Published: (2024)
by: Liang, Xun, et al.
Published: (2024)
CG-TTRL: Context-Guided Test-Time Reinforcement Learning for On-Device Large Language Models
by: Hosseini, Peyman, et al.
Published: (2025)
by: Hosseini, Peyman, et al.
Published: (2025)
Clinical Document Corpora -- Real Ones, Translated and Synthetic Substitutes, and Assorted Domain Proxies: A Survey of Diversity in Corpus Design, with Focus on German Text Data
by: Hahn, Udo
Published: (2024)
by: Hahn, Udo
Published: (2024)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
by: CH-Wang, Sky, et al.
Published: (2025)
by: CH-Wang, Sky, et al.
Published: (2025)
Text Summarization With Graph Attention Networks
by: Ardestani, Mohammadreza, et al.
Published: (2026)
by: Ardestani, Mohammadreza, et al.
Published: (2026)
Investigating the Impact of Text Summarization on Topic Modeling
by: Khandelwal, Trishia
Published: (2024)
by: Khandelwal, Trishia
Published: (2024)
Active Few-Shot Learning for Text Classification
by: Ahmadnia, Saeed, et al.
Published: (2025)
by: Ahmadnia, Saeed, et al.
Published: (2025)
Normalization of Lithuanian Text Using Regular Expressions
by: Kasparaitis, Pijus
Published: (2023)
by: Kasparaitis, Pijus
Published: (2023)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
VertAttack: Taking advantage of Text Classifiers' horizontal vision
by: Rusert, Jonathan
Published: (2024)
by: Rusert, Jonathan
Published: (2024)
Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation
by: Dejl, Adam, et al.
Published: (2025)
by: Dejl, Adam, et al.
Published: (2025)
UM_FHS at the CLEF 2025 SimpleText Track: Comparing No-Context and Fine-Tune Approaches for GPT-4.1 Models in Sentence and Document-Level Text Simplification
by: Kocbek, Primoz, et al.
Published: (2025)
by: Kocbek, Primoz, et al.
Published: (2025)
Universal-2-TF: Robust All-Neural Text Formatting for ASR
by: Khare, Yash, et al.
Published: (2025)
by: Khare, Yash, et al.
Published: (2025)
SegNSP: Revisiting Next Sentence Prediction for Linear Text Segmentation
by: Isidro, José, et al.
Published: (2026)
by: Isidro, José, et al.
Published: (2026)
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data
by: Klöser, Lars, et al.
Published: (2024)
by: Klöser, Lars, et al.
Published: (2024)
A RoBERTa-Based Functional Syntax Annotation Model for Chinese Texts
by: Xiaohui, Han, et al.
Published: (2025)
by: Xiaohui, Han, et al.
Published: (2025)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
by: Wanjawa, Barack Wamkaya, et al.
Published: (2025)
by: Wanjawa, Barack Wamkaya, et al.
Published: (2025)
Vocabulary Transfer for Biomedical Texts: Add Tokens if You Can Not Add Data
by: Singh, Priyanka, et al.
Published: (2022)
by: Singh, Priyanka, et al.
Published: (2022)
The Effect of Data Partitioning Strategy on Model Generalizability: A Case Study of Morphological Segmentation
by: Liu, Zoey, et al.
Published: (2024)
by: Liu, Zoey, et al.
Published: (2024)
A Survey of Text Watermarking in the Era of Large Language Models
by: Liu, Aiwei, et al.
Published: (2023)
by: Liu, Aiwei, et al.
Published: (2023)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
by: Hashemi, Helia, et al.
Published: (2024)
by: Hashemi, Helia, et al.
Published: (2024)
Comparing Complex Concepts with Transformers: Matching Patent Claims Against Natural Language Text
by: Blume, Matthias, et al.
Published: (2024)
by: Blume, Matthias, et al.
Published: (2024)
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
by: Lajčinová, Bibiána, et al.
Published: (2024)
by: Lajčinová, Bibiána, et al.
Published: (2024)
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
by: Levine, Lauren, et al.
Published: (2024)
by: Levine, Lauren, et al.
Published: (2024)
RuOpinionNE-2024: Extraction of Opinion Tuples from Russian News Texts
by: Loukachevitch, Natalia, et al.
Published: (2025)
by: Loukachevitch, Natalia, et al.
Published: (2025)
SEED: Enhancing Text-to-SQL Performance and Practical Usability Through Automatic Evidence Generation
by: Yun, Janghyeon, et al.
Published: (2025)
by: Yun, Janghyeon, et al.
Published: (2025)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
by: Sauter, Adrian, et al.
Published: (2025)
by: Sauter, Adrian, et al.
Published: (2025)
On the Influence of Discourse Relations in Persuasive Texts
by: Turk, Nawar, et al.
Published: (2025)
by: Turk, Nawar, et al.
Published: (2025)
Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension and Text Generation
by: Costa-jussà, Marta R., et al.
Published: (2024)
by: Costa-jussà, Marta R., et al.
Published: (2024)
Beyond Demographics: Fine-tuning Large Language Models to Predict Individuals' Subjective Text Perceptions
by: Orlikowski, Matthias, et al.
Published: (2025)
by: Orlikowski, Matthias, et al.
Published: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
by: Saji, Alan, et al.
Published: (2025)
by: Saji, Alan, et al.
Published: (2025)
Controllable Text Summarization: Unraveling Challenges, Approaches, and Prospects -- A Survey
by: Urlana, Ashok, et al.
Published: (2023)
by: Urlana, Ashok, et al.
Published: (2023)
A Novel Word Pair-based Gaussian Sentence Similarity Algorithm For Bengali Extractive Text Summarization
by: Morshed, Fahim, et al.
Published: (2024)
by: Morshed, Fahim, et al.
Published: (2024)
Similar Items
-
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly
by: Hosseini, Peyman, et al.
Published: (2024) -
Efficient Aspect-Based Summarization of Climate Change Reports with Small Language Models
by: Ghinassi, Iacopo, et al.
Published: (2024) -
Lon-ea at SemEval-2023 Task 11: A Comparison of Activation Functions for Soft and Hard Label Prediction
by: Hosseini, Peyman, et al.
Published: (2023) -
Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
by: Ma, Longxuan, et al.
Published: (2024) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)