Enregistré dans:
| Auteurs principaux: | Berman, David S., Stapleton, Alexander G. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2601.03368 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Infusing clinical knowledge into tokenisers for language models
par: Hasan, Abul, et autres
Publié: (2024)
par: Hasan, Abul, et autres
Publié: (2024)
Mixture-of-Depths: Dynamically allocating compute in transformer-based language models
par: Raposo, David, et autres
Publié: (2024)
par: Raposo, David, et autres
Publié: (2024)
Zipf Distributions from Two-Stage Symbolic Processes: Stability Under Stochastic Lexical Filtering
par: Berman, Vladimir
Publié: (2025)
par: Berman, Vladimir
Publié: (2025)
InsurTech innovation using natural language processing
par: Dong, Panyi, et autres
Publié: (2025)
par: Dong, Panyi, et autres
Publié: (2025)
Applications of natural language processing in aviation safety: A review and qualitative analysis
par: Nanyonga, Aziida, et autres
Publié: (2025)
par: Nanyonga, Aziida, et autres
Publié: (2025)
On the evolution of research in hypersonics: application of natural language processing and machine learning
par: Ebadi, Ashkan, et autres
Publié: (2022)
par: Ebadi, Ashkan, et autres
Publié: (2022)
Random Text, Zipf's Law, Critical Length,and Implications for Large Language Models
par: Berman, Vladimir
Publié: (2025)
par: Berman, Vladimir
Publié: (2025)
Comparison of different Unique hard attention transformer models by the formal languages they can recognize
par: Ryvkin, Leonid
Publié: (2025)
par: Ryvkin, Leonid
Publié: (2025)
MedicalBERT: enhancing biomedical natural language processing using pretrained BERT-based model
par: Reddy, K. Sahit, et autres
Publié: (2025)
par: Reddy, K. Sahit, et autres
Publié: (2025)
Benchmarking large language models for biomedical natural language processing applications and recommendations
par: Chen, Qingyu, et autres
Publié: (2023)
par: Chen, Qingyu, et autres
Publié: (2023)
EEG-CLIP : Learning EEG representations from natural language descriptions
par: Ndir, Tidiane Camaret, et autres
Publié: (2025)
par: Ndir, Tidiane Camaret, et autres
Publié: (2025)
A comparison of pipelines for the translation of a low resource language based on transformers
par: Bonfanti, Chiara, et autres
Publié: (2025)
par: Bonfanti, Chiara, et autres
Publié: (2025)
Detecting out-of-distribution text using topological features of transformer-based language models
par: Pollano, Andres, et autres
Publié: (2023)
par: Pollano, Andres, et autres
Publié: (2023)
Physical models realizing the transformer architecture of large language models
par: Chen, Zeqian
Publié: (2025)
par: Chen, Zeqian
Publié: (2025)
SSCAE -- Semantic, Syntactic, and Context-aware natural language Adversarial Examples generator
par: Asl, Javad Rafiei, et autres
Publié: (2024)
par: Asl, Javad Rafiei, et autres
Publié: (2024)
Enhancing ASD detection accuracy: a combined approach of machine learning and deep learning models with natural language processing
par: Rubio-Martín, Sergio, et autres
Publié: (2024)
par: Rubio-Martín, Sergio, et autres
Publié: (2024)
Block removal for large language models through constrained binary optimization
par: Jansen, David, et autres
Publié: (2026)
par: Jansen, David, et autres
Publié: (2026)
Predicting potentially abusive clauses in Chilean terms of services with natural language processing
par: Loeffler, Christoffer, et autres
Publié: (2025)
par: Loeffler, Christoffer, et autres
Publié: (2025)
Innovative tokenisation of structured data for LLM training
par: Karim, Kayvan, et autres
Publié: (2025)
par: Karim, Kayvan, et autres
Publié: (2025)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
par: Byers, Neil, et autres
Publié: (2025)
par: Byers, Neil, et autres
Publié: (2025)
BRAIn: Bayesian Reward-conditioned Amortized Inference for natural language generation from feedback
par: Pandey, Gaurav, et autres
Publié: (2024)
par: Pandey, Gaurav, et autres
Publié: (2024)
Optimization Strategies for Enhancing Resource Efficiency in Transformers & Large Language Models
par: Wallace, Tom, et autres
Publié: (2025)
par: Wallace, Tom, et autres
Publié: (2025)
Transformers need glasses! Information over-squashing in language tasks
par: Barbero, Federico, et autres
Publié: (2024)
par: Barbero, Federico, et autres
Publié: (2024)
A mean teacher algorithm for unlearning of language models
par: Klochkov, Yegor
Publié: (2025)
par: Klochkov, Yegor
Publié: (2025)
Aligning language models with human preferences
par: Korbak, Tomasz
Publié: (2024)
par: Korbak, Tomasz
Publié: (2024)
Evaluating language models as risk scores
par: Cruz, André F., et autres
Publié: (2024)
par: Cruz, André F., et autres
Publié: (2024)
DevBench: A multimodal developmental benchmark for language learning
par: Tan, Alvin Wei Ming, et autres
Publié: (2024)
par: Tan, Alvin Wei Ming, et autres
Publié: (2024)
Amortizing intractable inference in large language models
par: Hu, Edward J., et autres
Publié: (2023)
par: Hu, Edward J., et autres
Publié: (2023)
Attribution analysis of legal language as used by LLM
par: Belew, Richard K.
Publié: (2025)
par: Belew, Richard K.
Publié: (2025)
Human-interpretable clustering of short-text using large language models
par: Miller, Justin K., et autres
Publié: (2024)
par: Miller, Justin K., et autres
Publié: (2024)
Learning from flowsheets: A generative transformer model for autocompletion of flowsheets
par: Vogel, Gabriel, et autres
Publié: (2022)
par: Vogel, Gabriel, et autres
Publié: (2022)
Dynamic layer selection in decoder-only transformers
par: Glavas, Theodore, et autres
Publié: (2024)
par: Glavas, Theodore, et autres
Publié: (2024)
The SMeL Test: A simple benchmark for media literacy in language models
par: Ahdritz, Gustaf, et autres
Publié: (2025)
par: Ahdritz, Gustaf, et autres
Publié: (2025)
Perturbed examples reveal invariances shared by language models
par: Rawal, Ruchit, et autres
Publié: (2023)
par: Rawal, Ruchit, et autres
Publié: (2023)
Do language models plan ahead for future tokens?
par: Wu, Wilson, et autres
Publié: (2024)
par: Wu, Wilson, et autres
Publié: (2024)
Visualizing token importance for black-box language models
par: Rauba, Paulius, et autres
Publié: (2025)
par: Rauba, Paulius, et autres
Publié: (2025)
DataComp-LM: In search of the next generation of training sets for language models
par: Li, Jeffrey, et autres
Publié: (2024)
par: Li, Jeffrey, et autres
Publié: (2024)
Prompt reinforcing for long-term planning of large language models
par: Lin, Hsien-Chin, et autres
Publié: (2025)
par: Lin, Hsien-Chin, et autres
Publié: (2025)
Machine-generated text detection prevents language model collapse
par: Drayson, George, et autres
Publié: (2025)
par: Drayson, George, et autres
Publié: (2025)
Alignment faking in large language models
par: Greenblatt, Ryan, et autres
Publié: (2024)
par: Greenblatt, Ryan, et autres
Publié: (2024)
Documents similaires
-
Infusing clinical knowledge into tokenisers for language models
par: Hasan, Abul, et autres
Publié: (2024) -
Mixture-of-Depths: Dynamically allocating compute in transformer-based language models
par: Raposo, David, et autres
Publié: (2024) -
Zipf Distributions from Two-Stage Symbolic Processes: Stability Under Stochastic Lexical Filtering
par: Berman, Vladimir
Publié: (2025) -
InsurTech innovation using natural language processing
par: Dong, Panyi, et autres
Publié: (2025) -
Applications of natural language processing in aviation safety: A review and qualitative analysis
par: Nanyonga, Aziida, et autres
Publié: (2025)