Gespeichert in:
| Hauptverfasser: | Lochter, Johannes V., Silva, Renato M., Almeida, Tiago A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2020
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2007.07318 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
OmniDraft: A Cross-vocabulary, Online Adaptive Drafter for On-device Speculative Decoding
von: Ramakrishnan, Ramchalam Kinattinkara, et al.
Veröffentlicht: (2025)
von: Ramakrishnan, Ramchalam Kinattinkara, et al.
Veröffentlicht: (2025)
Multi-word Tokenization for Sequence Compression
von: Gee, Leonidas, et al.
Veröffentlicht: (2024)
von: Gee, Leonidas, et al.
Veröffentlicht: (2024)
Transformers represent belief state geometry in their residual stream
von: Shai, Adam S., et al.
Veröffentlicht: (2024)
von: Shai, Adam S., et al.
Veröffentlicht: (2024)
Vocabulary shapes cross-lingual variation of word-order learnability in language models
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)
Critical biblical studies via word frequency analysis: unveiling text authorship
von: Faigenbaum-Golovin, Shira, et al.
Veröffentlicht: (2024)
von: Faigenbaum-Golovin, Shira, et al.
Veröffentlicht: (2024)
Forcing Diffuse Distributions out of Language Models
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
AIDetx: a compression-based method for identification of machine-learning generated text
von: Almeida, Leonardo, et al.
Veröffentlicht: (2024)
von: Almeida, Leonardo, et al.
Veröffentlicht: (2024)
From communities to interpretable network and word embedding: an unified approach
von: Prouteau, Thibault, et al.
Veröffentlicht: (2024)
von: Prouteau, Thibault, et al.
Veröffentlicht: (2024)
Effects of term weighting approach with and without stop words removing on Arabic text classification
von: Alhenawi, Esra'a, et al.
Veröffentlicht: (2024)
von: Alhenawi, Esra'a, et al.
Veröffentlicht: (2024)
Detecting out-of-distribution text using topological features of transformer-based language models
von: Pollano, Andres, et al.
Veröffentlicht: (2023)
von: Pollano, Andres, et al.
Veröffentlicht: (2023)
Convergence and Divergence of Language Models under Different Random Seeds
von: Fehlauer, Finlay, et al.
Veröffentlicht: (2025)
von: Fehlauer, Finlay, et al.
Veröffentlicht: (2025)
A light-weight and efficient punctuation and word casing prediction model for on-device streaming ASR
von: You, Jian, et al.
Veröffentlicht: (2024)
von: You, Jian, et al.
Veröffentlicht: (2024)
Deep learning and abstractive summarisation for radiological reports: an empirical study for adapting the PEGASUS models' family with scarce data
von: Benzoni, Claudio, et al.
Veröffentlicht: (2025)
von: Benzoni, Claudio, et al.
Veröffentlicht: (2025)
Enhancing ASD detection accuracy: a combined approach of machine learning and deep learning models with natural language processing
von: Rubio-Martín, Sergio, et al.
Veröffentlicht: (2024)
von: Rubio-Martín, Sergio, et al.
Veröffentlicht: (2024)
How do language models learn facts? Dynamics, curricula and hallucinations
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2025)
von: Zucchet, Nicolas, et al.
Veröffentlicht: (2025)
Forecasting Events in Soccer Matches Through Language
von: Mendes-Neves, Tiago, et al.
Veröffentlicht: (2024)
von: Mendes-Neves, Tiago, et al.
Veröffentlicht: (2024)
Deep literature reviews: an application of fine-tuned language models to migration research
von: Iacus, Stefano M., et al.
Veröffentlicht: (2025)
von: Iacus, Stefano M., et al.
Veröffentlicht: (2025)
Prompt reinforcing for long-term planning of large language models
von: Lin, Hsien-Chin, et al.
Veröffentlicht: (2025)
von: Lin, Hsien-Chin, et al.
Veröffentlicht: (2025)
The representation landscape of few-shot learning and fine-tuning in large language models
von: Doimo, Diego, et al.
Veröffentlicht: (2024)
von: Doimo, Diego, et al.
Veröffentlicht: (2024)
Topic Modeling with Fine-tuning LLMs and Bag of Sentences
von: Schneider, Johannes
Veröffentlicht: (2024)
von: Schneider, Johannes
Veröffentlicht: (2024)
Improving Next Tokens via Second-to-Last Predictions with Generate and Refine
von: Schneider, Johannes
Veröffentlicht: (2024)
von: Schneider, Johannes
Veröffentlicht: (2024)
Efficient and Flexible Topic Modeling using Pretrained Embeddings and Bag of Sentences
von: Schneider, Johannes
Veröffentlicht: (2023)
von: Schneider, Johannes
Veröffentlicht: (2023)
Tokenisation via Convex Relaxations
von: Tempus, Jan, et al.
Veröffentlicht: (2026)
von: Tempus, Jan, et al.
Veröffentlicht: (2026)
Empowering machine learning models with contextual knowledge for enhancing the detection of eating disorders in social media posts
von: Benítez-Andrades, José Alberto, et al.
Veröffentlicht: (2024)
von: Benítez-Andrades, José Alberto, et al.
Veröffentlicht: (2024)
Leveraging large language models for structured information extraction from pathology reports
von: Balasubramanian, Jeya Balaji, et al.
Veröffentlicht: (2025)
von: Balasubramanian, Jeya Balaji, et al.
Veröffentlicht: (2025)
Do Generalisation Results Generalise?
von: Boglioni, Matteo, et al.
Veröffentlicht: (2025)
von: Boglioni, Matteo, et al.
Veröffentlicht: (2025)
Playing with words: Comparing the vocabulary and lexical diversity of ChatGPT and humans
von: Reviriego, Pedro, et al.
Veröffentlicht: (2023)
von: Reviriego, Pedro, et al.
Veröffentlicht: (2023)
A meta-analysis on the performance of machine-learning based language models for sentiment analysis
von: Rohde, Elena, et al.
Veröffentlicht: (2025)
von: Rohde, Elena, et al.
Veröffentlicht: (2025)
Negation Neglect: When models fail to learn negations in training
von: Mayne, Harry, et al.
Veröffentlicht: (2026)
von: Mayne, Harry, et al.
Veröffentlicht: (2026)
Latent learning: episodic memory complements parametric learning by enabling flexible reuse of experiences
von: Lampinen, Andrew Kyle, et al.
Veröffentlicht: (2025)
von: Lampinen, Andrew Kyle, et al.
Veröffentlicht: (2025)
Predicting ATP binding sites in protein sequences using Deep Learning and Natural Language Processing
von: V, Shreyas, et al.
Veröffentlicht: (2024)
von: V, Shreyas, et al.
Veröffentlicht: (2024)
The broader spectrum of in-context learning
von: Lampinen, Andrew Kyle, et al.
Veröffentlicht: (2024)
von: Lampinen, Andrew Kyle, et al.
Veröffentlicht: (2024)
Where is the signal in tokenization space?
von: Geh, Renato Lui, et al.
Veröffentlicht: (2024)
von: Geh, Renato Lui, et al.
Veröffentlicht: (2024)
Large language models reorganize representational geometry during in-context learning
von: Xiong, Hua-Dong, et al.
Veröffentlicht: (2026)
von: Xiong, Hua-Dong, et al.
Veröffentlicht: (2026)
CausalLM is not optimal for in-context learning
von: Ding, Nan, et al.
Veröffentlicht: (2023)
von: Ding, Nan, et al.
Veröffentlicht: (2023)
More than words: Advancements and challenges in speech recognition for singing
von: Kruspe, Anna
Veröffentlicht: (2024)
von: Kruspe, Anna
Veröffentlicht: (2024)
JMI at SemEval 2024 Task 3: Two-step approach for multimodal ECAC using in-context learning with GPT and instruction-tuned Llama models
von: Arefa, et al.
Veröffentlicht: (2024)
von: Arefa, et al.
Veröffentlicht: (2024)
Deep sequence models tend to memorize geometrically; it is unclear why
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025)
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025)
Rethinking Reinforcement Fine-Tuning in LVLM: Convergence, Reward Decomposition, and Generalization
von: Adams, Carter, et al.
Veröffentlicht: (2026)
von: Adams, Carter, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026) -
OmniDraft: A Cross-vocabulary, Online Adaptive Drafter for On-device Speculative Decoding
von: Ramakrishnan, Ramchalam Kinattinkara, et al.
Veröffentlicht: (2025) -
Multi-word Tokenization for Sequence Compression
von: Gee, Leonidas, et al.
Veröffentlicht: (2024) -
Transformers represent belief state geometry in their residual stream
von: Shai, Adam S., et al.
Veröffentlicht: (2024) -
Vocabulary shapes cross-lingual variation of word-order learnability in language models
von: Martins, Jonas Mayer, et al.
Veröffentlicht: (2026)