Saved in:
| Main Authors: | Cohen, Ido, Gidron, Jacob, Pinto, Idan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.02417 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MosaicBERT: A Bidirectional Encoder Optimized for Fast Pretraining
by: Portes, Jacob, et al.
Published: (2023)
by: Portes, Jacob, et al.
Published: (2023)
An investigation of structures responsible for gender bias in BERT and DistilBERT
by: Leteno, Thibaud, et al.
Published: (2024)
by: Leteno, Thibaud, et al.
Published: (2024)
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
by: Damani, Mehul, et al.
Published: (2024)
by: Damani, Mehul, et al.
Published: (2024)
LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users
by: Hilel, Almog, et al.
Published: (2025)
by: Hilel, Almog, et al.
Published: (2025)
DE$^3$-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks
by: He, Jianing, et al.
Published: (2024)
by: He, Jianing, et al.
Published: (2024)
MaxPoolBERT: Enhancing BERT Classification via Layer- and Token-Wise Aggregation
by: Behrendt, Maike, et al.
Published: (2025)
by: Behrendt, Maike, et al.
Published: (2025)
NeoDictaBERT: Pushing the Frontier of BERT models for Hebrew
by: Shmidman, Shaltiel, et al.
Published: (2025)
by: Shmidman, Shaltiel, et al.
Published: (2025)
Empowering Interdisciplinary Research with BERT-Based Models: An Approach Through SciBERT-CNN with Topic Modeling
by: Likhareva, Darya, et al.
Published: (2024)
by: Likhareva, Darya, et al.
Published: (2024)
Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models
by: Puri, Isha, et al.
Published: (2026)
by: Puri, Isha, et al.
Published: (2026)
Energy and Carbon Considerations of Fine-Tuning BERT
by: Wang, Xiaorong, et al.
Published: (2023)
by: Wang, Xiaorong, et al.
Published: (2023)
Never Train from Scratch: Fair Comparison of Long-Sequence Models Requires Data-Driven Priors
by: Amos, Ido, et al.
Published: (2023)
by: Amos, Ido, et al.
Published: (2023)
Ensemble Self-Training for Unsupervised Machine Translation
by: Aharon, Ido, et al.
Published: (2026)
by: Aharon, Ido, et al.
Published: (2026)
A Contextualized BERT model for Knowledge Graph Completion
by: Gul, Haji, et al.
Published: (2024)
by: Gul, Haji, et al.
Published: (2024)
CEEBERT: Cross-Domain Inference in Early Exit BERT
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
Weight-Inherited Distillation for Task-Agnostic BERT Compression
by: Wu, Taiqiang, et al.
Published: (2023)
by: Wu, Taiqiang, et al.
Published: (2023)
GottBERT: a pure German Language Model
by: Scheible, Raphael, et al.
Published: (2020)
by: Scheible, Raphael, et al.
Published: (2020)
Sarcasm Detection in Tweets with BERT and GloVe Embeddings
by: Khatri, Akshay, et al.
Published: (2020)
by: Khatri, Akshay, et al.
Published: (2020)
From BERT to Qwen: Hate Detection across architectures
by: Mon, Ariadna, et al.
Published: (2025)
by: Mon, Ariadna, et al.
Published: (2025)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
by: Damani, Mehul, et al.
Published: (2025)
by: Damani, Mehul, et al.
Published: (2025)
Structured RAG for Answering Aggregative Questions
by: Koshorek, Omri, et al.
Published: (2025)
by: Koshorek, Omri, et al.
Published: (2025)
MedicalBERT: enhancing biomedical natural language processing using pretrained BERT-based model
by: Reddy, K. Sahit, et al.
Published: (2025)
by: Reddy, K. Sahit, et al.
Published: (2025)
Non-Contextual BERT or FastText? A Comparative Analysis
by: Shanbhag, Abhay, et al.
Published: (2024)
by: Shanbhag, Abhay, et al.
Published: (2024)
Breaking MLPerf Training: A Case Study on Optimizing BERT
by: Kim, Yongdeok, et al.
Published: (2024)
by: Kim, Yongdeok, et al.
Published: (2024)
Improving Sampling Methods for Fine-tuning SentenceBERT in Text Streams
by: Garcia, Cristiano Mesquita, et al.
Published: (2024)
by: Garcia, Cristiano Mesquita, et al.
Published: (2024)
Research on the Application of Deep Learning-based BERT Model in Sentiment Analysis
by: Wu, Yichao, et al.
Published: (2024)
by: Wu, Yichao, et al.
Published: (2024)
Towards Building Efficient Sentence BERT Models using Layer Pruning
by: Shelke, Anushka, et al.
Published: (2024)
by: Shelke, Anushka, et al.
Published: (2024)
Enhancing BERT Fine-Tuning for Sentiment Analysis in Lower-Resourced Languages
by: Kubík, Jozef, et al.
Published: (2025)
by: Kubík, Jozef, et al.
Published: (2025)
Ensemble BERT for Medication Event Classification on Electronic Health Records (EHRs)
by: Sarker, Shouvon, et al.
Published: (2025)
by: Sarker, Shouvon, et al.
Published: (2025)
Breaking BERT: Gradient Attack on Twitter Sentiment Analysis for Targeted Misclassification
by: Subedi, Akil Raj, et al.
Published: (2025)
by: Subedi, Akil Raj, et al.
Published: (2025)
On Importance of Layer Pruning for Smaller BERT Models and Low Resource Languages
by: Shirke, Mayur, et al.
Published: (2025)
by: Shirke, Mayur, et al.
Published: (2025)
MelHuBERT: A simplified HuBERT on Mel spectrograms
by: Lin, Tzu-Quan, et al.
Published: (2022)
by: Lin, Tzu-Quan, et al.
Published: (2022)
BERT Learns (and Teaches) Chemistry
by: Payne, Josh, et al.
Published: (2020)
by: Payne, Josh, et al.
Published: (2020)
Enhancing TinyBERT for Financial Sentiment Analysis Using GPT-Augmented FinBERT Distillation
by: Thomas, Graison Jos
Published: (2024)
by: Thomas, Graison Jos
Published: (2024)
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models
by: Zhang, Ying, et al.
Published: (2024)
by: Zhang, Ying, et al.
Published: (2024)
LATMiX: Learnable Affine Transformations for Microscaling Quantization of LLMs
by: Gordon, Ofir, et al.
Published: (2026)
by: Gordon, Ofir, et al.
Published: (2026)
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language
by: Parankusham, Kanishka, et al.
Published: (2025)
by: Parankusham, Kanishka, et al.
Published: (2025)
Benchmarking BERT-based Models for Sentence-level Topic Classification in Nepali Language
by: Karki, Nischal, et al.
Published: (2026)
by: Karki, Nischal, et al.
Published: (2026)
A BERT-based Hierarchical Classification Model with Applications in Chinese Commodity Classification
by: Liu, Kun, et al.
Published: (2025)
by: Liu, Kun, et al.
Published: (2025)
MahaParaphrase: A Marathi Paraphrase Detection Corpus and BERT-based Models
by: Jadhav, Suramya, et al.
Published: (2025)
by: Jadhav, Suramya, et al.
Published: (2025)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
Similar Items
-
MosaicBERT: A Bidirectional Encoder Optimized for Fast Pretraining
by: Portes, Jacob, et al.
Published: (2023) -
An investigation of structures responsible for gender bias in BERT and DistilBERT
by: Leteno, Thibaud, et al.
Published: (2024) -
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
by: Damani, Mehul, et al.
Published: (2024) -
LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users
by: Hilel, Almog, et al.
Published: (2025) -
DE$^3$-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks
by: He, Jianing, et al.
Published: (2024)