Are most sentences unique? An empirical examination of Chomskyan claims
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Ring, Hiram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The taggedPBC: Annotating a massive parallel corpus for crosslinguistic investigations
von: Ring, Hiram
Veröffentlicht: (2025)
von: Ring, Hiram
Veröffentlicht: (2025)
Extending dependencies to the taggedPBC: Word order in transitive clauses
von: Ring, Hiram
Veröffentlicht: (2025)
von: Ring, Hiram
Veröffentlicht: (2025)
Word length predicts word order: "Min-max"-ing drives language evolution
von: Ring, Hiram
Veröffentlicht: (2025)
von: Ring, Hiram
Veröffentlicht: (2025)
Chomskyan (R)evolutions
Veröffentlicht: (2020)
Veröffentlicht: (2020)
Variation of sentence length across time and genre
von: Rudnicka, Karolina
Veröffentlicht: (2025)
von: Rudnicka, Karolina
Veröffentlicht: (2025)
Escaping the sentence-level paradigm in machine translation
von: Post, Matt, et al.
Veröffentlicht: (2023)
von: Post, Matt, et al.
Veröffentlicht: (2023)
POS-tagging to highlight the skeletal structure of sentences
von: Churakov, Grigorii
Veröffentlicht: (2024)
von: Churakov, Grigorii
Veröffentlicht: (2024)
Recovering document annotations for sentence-level bitext
von: Wicks, Rachel, et al.
Veröffentlicht: (2024)
von: Wicks, Rachel, et al.
Veröffentlicht: (2024)
Neural paraphrasing by automatically crawled and aligned sentence pairs
von: Globo, Achille, et al.
Veröffentlicht: (2024)
von: Globo, Achille, et al.
Veröffentlicht: (2024)
Linguistic features for sentence difficulty prediction in ABSA
von: Chifu, Adrian-Gabriel, et al.
Veröffentlicht: (2024)
von: Chifu, Adrian-Gabriel, et al.
Veröffentlicht: (2024)
PEACH: A sentence-aligned Parallel English-Arabic Corpus for Healthcare
von: Al-Sabbagh, Rania
Veröffentlicht: (2025)
von: Al-Sabbagh, Rania
Veröffentlicht: (2025)
Can LLMs capture stable human-generated sentence entropy measures?
von: Pivel-Villanueva, Estrella, et al.
Veröffentlicht: (2026)
von: Pivel-Villanueva, Estrella, et al.
Veröffentlicht: (2026)
Video sentence grounding with temporally global textual knowledge
von: Chen, Cai, et al.
Veröffentlicht: (2024)
von: Chen, Cai, et al.
Veröffentlicht: (2024)
MEXMA: Token-level objectives improve sentence representations
von: Janeiro, João Maria, et al.
Veröffentlicht: (2024)
von: Janeiro, João Maria, et al.
Veröffentlicht: (2024)
NMT-Obfuscator Attack: Ignore a sentence in translation with only one word
von: Sadrizadeh, Sahar, et al.
Veröffentlicht: (2024)
von: Sadrizadeh, Sahar, et al.
Veröffentlicht: (2024)
Detection and Positive Reconstruction of Cognitive Distortion sentences: Mandarin Dataset and Evaluation
von: Lin, Shuya, et al.
Veröffentlicht: (2024)
von: Lin, Shuya, et al.
Veröffentlicht: (2024)
FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings
von: Kesiraju, Santosh, et al.
Veröffentlicht: (2026)
von: Kesiraju, Santosh, et al.
Veröffentlicht: (2026)
Hyperbolic sentence representations for solving Textual Entailment
von: Petrovski, Igor
Veröffentlicht: (2024)
von: Petrovski, Igor
Veröffentlicht: (2024)
Attention-aware semantic relevance predicting Chinese sentence reading
von: Sun, Kun
Veröffentlicht: (2024)
von: Sun, Kun
Veröffentlicht: (2024)
Power in Numbers: Robust reading comprehension by finetuning with four adversarial sentences per example
von: Marcus, Ariel
Veröffentlicht: (2024)
von: Marcus, Ariel
Veröffentlicht: (2024)
Deconstructing sentence disambiguation by joint latent modeling of reading paradigms: LLM surprisal is not enough
von: Paape, Dario, et al.
Veröffentlicht: (2026)
von: Paape, Dario, et al.
Veröffentlicht: (2026)
Are there identifiable structural parts in the sentence embedding whole?
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
Generating bilingual example sentences with large language models as lexicography assistants
von: Merx, Raphael, et al.
Veröffentlicht: (2024)
von: Merx, Raphael, et al.
Veröffentlicht: (2024)
The role of inhibitory control in garden-path sentence processing: A Chinese-English bilingual perspective
von: Rao, Xiaohui, et al.
Veröffentlicht: (2024)
von: Rao, Xiaohui, et al.
Veröffentlicht: (2024)
Readers make targeted regressions to plausible errors in reanalysis of "noisy-channel garden-path" sentences
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
von: Clark, Thomas Hikaru, et al.
Veröffentlicht: (2026)
DefSent+: Improving sentence embeddings of language models by projecting definition sentences into a quasi-isotropic or isotropic vector space of unlimited dictionary entries
von: Liu, Xiaodong
Veröffentlicht: (2024)
von: Liu, Xiaodong
Veröffentlicht: (2024)
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts
von: Lamsal, Rabindra, et al.
Veröffentlicht: (2023)
von: Lamsal, Rabindra, et al.
Veröffentlicht: (2023)
Investigating large language models for their competence in extracting grammatically sound sentences from transcribed noisy utterances
von: Wróblewska, Alina
Veröffentlicht: (2024)
von: Wróblewska, Alina
Veröffentlicht: (2024)
Jailbreak Instruction-Tuned LLMs via end-of-sentence MLP Re-weighting
von: Luo, Yifan, et al.
Veröffentlicht: (2024)
von: Luo, Yifan, et al.
Veröffentlicht: (2024)
MultiCaption: Detecting disinformation using multilingual visual claims
von: Frade, Rafael Martins, et al.
Veröffentlicht: (2026)
von: Frade, Rafael Martins, et al.
Veröffentlicht: (2026)
Exploring Italian sentence embeddings properties through multi-tasking
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
VERISCORE: Evaluating the factuality of verifiable claims in long-form text generation
von: Song, Yixiao, et al.
Veröffentlicht: (2024)
von: Song, Yixiao, et al.
Veröffentlicht: (2024)
Modeling the language cortex with form-independent and enriched representations of sentence meaning reveals remarkable semantic abstractness
von: Saha, Shreya, et al.
Veröffentlicht: (2025)
von: Saha, Shreya, et al.
Veröffentlicht: (2025)
A new approach for fine-tuning sentence transformers for intent classification and out-of-scope detection tasks
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
Hope Speech Detection in Social Media English Corpora: Performance of Traditional and Transformer Models
von: Ramos, Luis, et al.
Veröffentlicht: (2025)
von: Ramos, Luis, et al.
Veröffentlicht: (2025)
Tracking linguistic information in transformer-based sentence embeddings through targeted sparsification
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
Exploring syntactic information in sentence embeddings through multilingual subject-verb agreement
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
Open Conversational LLMs do not know most Spanish words
von: Conde, Javier, et al.
Veröffentlicht: (2024)
von: Conde, Javier, et al.
Veröffentlicht: (2024)
SentenceVAE: Enable Next-sentence Prediction for Large Language Models with Faster Speed, Higher Accuracy and Longer Context
von: An, Hongjun, et al.
Veröffentlicht: (2024)
von: An, Hongjun, et al.
Veröffentlicht: (2024)
Testing the assumptions about the geometry of sentence embedding spaces: the cosine measure need not apply
von: Nastase, Vivi, et al.
Veröffentlicht: (2025)
von: Nastase, Vivi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The taggedPBC: Annotating a massive parallel corpus for crosslinguistic investigations
von: Ring, Hiram
Veröffentlicht: (2025) -
Extending dependencies to the taggedPBC: Word order in transitive clauses
von: Ring, Hiram
Veröffentlicht: (2025) -
Word length predicts word order: "Min-max"-ing drives language evolution
von: Ring, Hiram
Veröffentlicht: (2025) -
Chomskyan (R)evolutions
Veröffentlicht: (2020) -
Variation of sentence length across time and genre
von: Rudnicka, Karolina
Veröffentlicht: (2025)