Isotropy, Clusters, and Classifiers
Fuente:
arXiv
Saved in:
| Main Authors: | Mickus, Timothee, Grönroos, Stig-Arne, Attieh, Joseph |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Life Cycle-Aware Evaluation of Knowledge Distillation for Machine Translation: Environmental Impact and Translation Quality Trade-offs
by: Attieh, Joseph, et al.
Published: (2026)
by: Attieh, Joseph, et al.
Published: (2026)
I Have an Attention Bridge to Sell You: Generalization Capabilities of Modular Translation Architectures
by: Mickus, Timothee, et al.
Published: (2024)
by: Mickus, Timothee, et al.
Published: (2024)
MAMMOTH: Massively Multilingual Modular Open Translation @ Helsinki
by: Mickus, Timothee, et al.
Published: (2024)
by: Mickus, Timothee, et al.
Published: (2024)
KD4MT: A Survey of Knowledge Distillation for Machine Translation
by: de Gibert, Ona, et al.
Published: (2026)
by: de Gibert, Ona, et al.
Published: (2026)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025)
by: Kadlčík, Marek, et al.
Published: (2025)
Adapting Definition Modeling for New Languages: A Case Study on Belarusian
by: Kazakouskaya, Daniela, et al.
Published: (2025)
by: Kazakouskaya, Daniela, et al.
Published: (2025)
Your Model is Overconfident, and Other Lies We Tell Ourselves
by: Mickus, Timothee, et al.
Published: (2025)
by: Mickus, Timothee, et al.
Published: (2025)
Language Models Learn Universal Representations of Numbers and Here's Why You Should Care
by: Štefánik, Michal, et al.
Published: (2025)
by: Štefánik, Michal, et al.
Published: (2025)
Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation
by: Bhardwaj, Dhrupad, et al.
Published: (2025)
by: Bhardwaj, Dhrupad, et al.
Published: (2025)
Continual Learning Under Language Shift
by: Gogoulou, Evangelia, et al.
Published: (2023)
by: Gogoulou, Evangelia, et al.
Published: (2023)
Can Machine Translation Bridge Multilingual Pretraining and Cross-lingual Transfer Learning?
by: Ji, Shaoxiong, et al.
Published: (2024)
by: Ji, Shaoxiong, et al.
Published: (2024)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
by: Kimera, Richard, et al.
Published: (2024)
by: Kimera, Richard, et al.
Published: (2024)
Transparent Neighborhood Approximation for Text Classifier Explanation
by: Cai, Yi, et al.
Published: (2024)
by: Cai, Yi, et al.
Published: (2024)
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
by: Hughes, Anthony, et al.
Published: (2026)
by: Hughes, Anthony, et al.
Published: (2026)
DISCERN: Decoding Systematic Errors in Natural Language for Text Classifiers
by: Menon, Rakesh R., et al.
Published: (2024)
by: Menon, Rakesh R., et al.
Published: (2024)
DALLMi: Domain Adaption for LLM-based Multi-label Classifier
by: Beţianu, Miruna, et al.
Published: (2024)
by: Beţianu, Miruna, et al.
Published: (2024)
Classifying and Addressing the Diversity of Errors in Retrieval-Augmented Generation Systems
by: Leung, Kin Kwan, et al.
Published: (2025)
by: Leung, Kin Kwan, et al.
Published: (2025)
Domain-specific or Uncertainty-aware models: Does it really make a difference for biomedical text classification?
by: Sinha, Aman, et al.
Published: (2024)
by: Sinha, Aman, et al.
Published: (2024)
A Comparison of Language Modeling and Translation as Multilingual Pretraining Objectives
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
The Data-Quality Illusion: Rethinking Classifier-Based Quality Filtering for LLM Pretraining
by: Saada, Thiziri Nait, et al.
Published: (2025)
by: Saada, Thiziri Nait, et al.
Published: (2025)
Improving Preference Extraction In LLMs By Identifying Latent Knowledge Through Classifying Probes
by: Maiya, Sharan, et al.
Published: (2025)
by: Maiya, Sharan, et al.
Published: (2025)
ylmmcl at Multilingual Text Detoxification 2025: Lexicon-Guided Detoxification and Classifier-Gated Rewriting
by: Lai-Lopez, Nicole, et al.
Published: (2025)
by: Lai-Lopez, Nicole, et al.
Published: (2025)
I Can't Believe It's Not Robust: Catastrophic Collapse of Safety Classifiers under Embedding Drift
by: Sahoo, Subramanyam, et al.
Published: (2026)
by: Sahoo, Subramanyam, et al.
Published: (2026)
NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated Data
by: Bogdanov, Sergei, et al.
Published: (2024)
by: Bogdanov, Sergei, et al.
Published: (2024)
Natural Language Processing and Deep Learning Models to Classify Phase of Flight in Aviation Safety Occurrences
by: Nanyonga, Aziida, et al.
Published: (2025)
by: Nanyonga, Aziida, et al.
Published: (2025)
Lexical Hints of Accuracy in LLM Reasoning Chains
by: Vanhoyweghen, Arne, et al.
Published: (2025)
by: Vanhoyweghen, Arne, et al.
Published: (2025)
Explaining Text Classifiers with Counterfactual Representations
by: Lemberger, Pirmin, et al.
Published: (2024)
by: Lemberger, Pirmin, et al.
Published: (2024)
Machines Do See Color: A Guideline to Classify Different Forms of Racist Discourse in Large Corpora
by: Gordillo, Diana Davila, et al.
Published: (2024)
by: Gordillo, Diana Davila, et al.
Published: (2024)
Aviation Safety Enhancement via NLP & Deep Learning: Classifying Flight Phases in ATSB Safety Reports
by: Nanyonga, Aziida, et al.
Published: (2025)
by: Nanyonga, Aziida, et al.
Published: (2025)
TaCo: Targeted Concept Erasure Prevents Non-Linear Classifiers From Detecting Protected Attributes
by: Jourdan, Fanny, et al.
Published: (2023)
by: Jourdan, Fanny, et al.
Published: (2023)
Characterizing and Classifying Developer Forum Posts with their Intentions
by: Wu, Xingfang, et al.
Published: (2023)
by: Wu, Xingfang, et al.
Published: (2023)
SDEC: Semantic Deep Embedded Clustering
by: Rahman, Mohammad Wali Ur, et al.
Published: (2025)
by: Rahman, Mohammad Wali Ur, et al.
Published: (2025)
Summaries as Centroids for Interpretable and Scalable Text Clustering
by: Diaz-Rodriguez, Jairo
Published: (2025)
by: Diaz-Rodriguez, Jairo
Published: (2025)
Reversible Jump Attack to Textual Classifiers with Modification Reduction
by: Ni, Mingze, et al.
Published: (2024)
by: Ni, Mingze, et al.
Published: (2024)
Binary Classifier Optimization for Large Language Model Alignment
by: Jung, Seungjae, et al.
Published: (2024)
by: Jung, Seungjae, et al.
Published: (2024)
CHAI: Clustered Head Attention for Efficient LLM Inference
by: Agarwal, Saurabh, et al.
Published: (2024)
by: Agarwal, Saurabh, et al.
Published: (2024)
Voices in a Crowd: Searching for Clusters of Unique Perspectives
by: Vitsakis, Nikolas, et al.
Published: (2024)
by: Vitsakis, Nikolas, et al.
Published: (2024)
The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious
by: Chua, James, et al.
Published: (2026)
by: Chua, James, et al.
Published: (2026)
Normalisation of SWIFT Message Counterparties with Feature Extraction and Clustering
by: Schoinas, Thanasis, et al.
Published: (2025)
by: Schoinas, Thanasis, et al.
Published: (2025)
Similar Items
-
Life Cycle-Aware Evaluation of Knowledge Distillation for Machine Translation: Environmental Impact and Translation Quality Trade-offs
by: Attieh, Joseph, et al.
Published: (2026) -
I Have an Attention Bridge to Sell You: Generalization Capabilities of Modular Translation Architectures
by: Mickus, Timothee, et al.
Published: (2024) -
MAMMOTH: Massively Multilingual Modular Open Translation @ Helsinki
by: Mickus, Timothee, et al.
Published: (2024) -
KD4MT: A Survey of Knowledge Distillation for Machine Translation
by: de Gibert, Ona, et al.
Published: (2026) -
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025)