Isotropy, Clusters, and Classifiers
Fuente:
arXiv
Salvato in:
| Autori principali: | Mickus, Timothee, Grönroos, Stig-Arne, Attieh, Joseph |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Life Cycle-Aware Evaluation of Knowledge Distillation for Machine Translation: Environmental Impact and Translation Quality Trade-offs
di: Attieh, Joseph, et al.
Pubblicazione: (2026)
di: Attieh, Joseph, et al.
Pubblicazione: (2026)
I Have an Attention Bridge to Sell You: Generalization Capabilities of Modular Translation Architectures
di: Mickus, Timothee, et al.
Pubblicazione: (2024)
di: Mickus, Timothee, et al.
Pubblicazione: (2024)
MAMMOTH: Massively Multilingual Modular Open Translation @ Helsinki
di: Mickus, Timothee, et al.
Pubblicazione: (2024)
di: Mickus, Timothee, et al.
Pubblicazione: (2024)
KD4MT: A Survey of Knowledge Distillation for Machine Translation
di: de Gibert, Ona, et al.
Pubblicazione: (2026)
di: de Gibert, Ona, et al.
Pubblicazione: (2026)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
di: Kadlčík, Marek, et al.
Pubblicazione: (2025)
di: Kadlčík, Marek, et al.
Pubblicazione: (2025)
Adapting Definition Modeling for New Languages: A Case Study on Belarusian
di: Kazakouskaya, Daniela, et al.
Pubblicazione: (2025)
di: Kazakouskaya, Daniela, et al.
Pubblicazione: (2025)
Your Model is Overconfident, and Other Lies We Tell Ourselves
di: Mickus, Timothee, et al.
Pubblicazione: (2025)
di: Mickus, Timothee, et al.
Pubblicazione: (2025)
Language Models Learn Universal Representations of Numbers and Here's Why You Should Care
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation
di: Bhardwaj, Dhrupad, et al.
Pubblicazione: (2025)
di: Bhardwaj, Dhrupad, et al.
Pubblicazione: (2025)
Continual Learning Under Language Shift
di: Gogoulou, Evangelia, et al.
Pubblicazione: (2023)
di: Gogoulou, Evangelia, et al.
Pubblicazione: (2023)
Can Machine Translation Bridge Multilingual Pretraining and Cross-lingual Transfer Learning?
di: Ji, Shaoxiong, et al.
Pubblicazione: (2024)
di: Ji, Shaoxiong, et al.
Pubblicazione: (2024)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
di: Kimera, Richard, et al.
Pubblicazione: (2024)
di: Kimera, Richard, et al.
Pubblicazione: (2024)
Transparent Neighborhood Approximation for Text Classifier Explanation
di: Cai, Yi, et al.
Pubblicazione: (2024)
di: Cai, Yi, et al.
Pubblicazione: (2024)
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
di: Neshaei, Seyed Parsa, et al.
Pubblicazione: (2024)
di: Neshaei, Seyed Parsa, et al.
Pubblicazione: (2024)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
di: Hughes, Anthony, et al.
Pubblicazione: (2026)
di: Hughes, Anthony, et al.
Pubblicazione: (2026)
DISCERN: Decoding Systematic Errors in Natural Language for Text Classifiers
di: Menon, Rakesh R., et al.
Pubblicazione: (2024)
di: Menon, Rakesh R., et al.
Pubblicazione: (2024)
DALLMi: Domain Adaption for LLM-based Multi-label Classifier
di: Beţianu, Miruna, et al.
Pubblicazione: (2024)
di: Beţianu, Miruna, et al.
Pubblicazione: (2024)
Classifying and Addressing the Diversity of Errors in Retrieval-Augmented Generation Systems
di: Leung, Kin Kwan, et al.
Pubblicazione: (2025)
di: Leung, Kin Kwan, et al.
Pubblicazione: (2025)
Domain-specific or Uncertainty-aware models: Does it really make a difference for biomedical text classification?
di: Sinha, Aman, et al.
Pubblicazione: (2024)
di: Sinha, Aman, et al.
Pubblicazione: (2024)
A Comparison of Language Modeling and Translation as Multilingual Pretraining Objectives
di: Li, Zihao, et al.
Pubblicazione: (2024)
di: Li, Zihao, et al.
Pubblicazione: (2024)
The Data-Quality Illusion: Rethinking Classifier-Based Quality Filtering for LLM Pretraining
di: Saada, Thiziri Nait, et al.
Pubblicazione: (2025)
di: Saada, Thiziri Nait, et al.
Pubblicazione: (2025)
Improving Preference Extraction In LLMs By Identifying Latent Knowledge Through Classifying Probes
di: Maiya, Sharan, et al.
Pubblicazione: (2025)
di: Maiya, Sharan, et al.
Pubblicazione: (2025)
ylmmcl at Multilingual Text Detoxification 2025: Lexicon-Guided Detoxification and Classifier-Gated Rewriting
di: Lai-Lopez, Nicole, et al.
Pubblicazione: (2025)
di: Lai-Lopez, Nicole, et al.
Pubblicazione: (2025)
I Can't Believe It's Not Robust: Catastrophic Collapse of Safety Classifiers under Embedding Drift
di: Sahoo, Subramanyam, et al.
Pubblicazione: (2026)
di: Sahoo, Subramanyam, et al.
Pubblicazione: (2026)
NuNER: Entity Recognition Encoder Pre-training via LLM-Annotated Data
di: Bogdanov, Sergei, et al.
Pubblicazione: (2024)
di: Bogdanov, Sergei, et al.
Pubblicazione: (2024)
Natural Language Processing and Deep Learning Models to Classify Phase of Flight in Aviation Safety Occurrences
di: Nanyonga, Aziida, et al.
Pubblicazione: (2025)
di: Nanyonga, Aziida, et al.
Pubblicazione: (2025)
Lexical Hints of Accuracy in LLM Reasoning Chains
di: Vanhoyweghen, Arne, et al.
Pubblicazione: (2025)
di: Vanhoyweghen, Arne, et al.
Pubblicazione: (2025)
Explaining Text Classifiers with Counterfactual Representations
di: Lemberger, Pirmin, et al.
Pubblicazione: (2024)
di: Lemberger, Pirmin, et al.
Pubblicazione: (2024)
Machines Do See Color: A Guideline to Classify Different Forms of Racist Discourse in Large Corpora
di: Gordillo, Diana Davila, et al.
Pubblicazione: (2024)
di: Gordillo, Diana Davila, et al.
Pubblicazione: (2024)
Aviation Safety Enhancement via NLP & Deep Learning: Classifying Flight Phases in ATSB Safety Reports
di: Nanyonga, Aziida, et al.
Pubblicazione: (2025)
di: Nanyonga, Aziida, et al.
Pubblicazione: (2025)
TaCo: Targeted Concept Erasure Prevents Non-Linear Classifiers From Detecting Protected Attributes
di: Jourdan, Fanny, et al.
Pubblicazione: (2023)
di: Jourdan, Fanny, et al.
Pubblicazione: (2023)
Characterizing and Classifying Developer Forum Posts with their Intentions
di: Wu, Xingfang, et al.
Pubblicazione: (2023)
di: Wu, Xingfang, et al.
Pubblicazione: (2023)
SDEC: Semantic Deep Embedded Clustering
di: Rahman, Mohammad Wali Ur, et al.
Pubblicazione: (2025)
di: Rahman, Mohammad Wali Ur, et al.
Pubblicazione: (2025)
Summaries as Centroids for Interpretable and Scalable Text Clustering
di: Diaz-Rodriguez, Jairo
Pubblicazione: (2025)
di: Diaz-Rodriguez, Jairo
Pubblicazione: (2025)
Reversible Jump Attack to Textual Classifiers with Modification Reduction
di: Ni, Mingze, et al.
Pubblicazione: (2024)
di: Ni, Mingze, et al.
Pubblicazione: (2024)
Binary Classifier Optimization for Large Language Model Alignment
di: Jung, Seungjae, et al.
Pubblicazione: (2024)
di: Jung, Seungjae, et al.
Pubblicazione: (2024)
CHAI: Clustered Head Attention for Efficient LLM Inference
di: Agarwal, Saurabh, et al.
Pubblicazione: (2024)
di: Agarwal, Saurabh, et al.
Pubblicazione: (2024)
Voices in a Crowd: Searching for Clusters of Unique Perspectives
di: Vitsakis, Nikolas, et al.
Pubblicazione: (2024)
di: Vitsakis, Nikolas, et al.
Pubblicazione: (2024)
The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious
di: Chua, James, et al.
Pubblicazione: (2026)
di: Chua, James, et al.
Pubblicazione: (2026)
Normalisation of SWIFT Message Counterparties with Feature Extraction and Clustering
di: Schoinas, Thanasis, et al.
Pubblicazione: (2025)
di: Schoinas, Thanasis, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Life Cycle-Aware Evaluation of Knowledge Distillation for Machine Translation: Environmental Impact and Translation Quality Trade-offs
di: Attieh, Joseph, et al.
Pubblicazione: (2026) -
I Have an Attention Bridge to Sell You: Generalization Capabilities of Modular Translation Architectures
di: Mickus, Timothee, et al.
Pubblicazione: (2024) -
MAMMOTH: Massively Multilingual Modular Open Translation @ Helsinki
di: Mickus, Timothee, et al.
Pubblicazione: (2024) -
KD4MT: A Survey of Knowledge Distillation for Machine Translation
di: de Gibert, Ona, et al.
Pubblicazione: (2026) -
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
di: Kadlčík, Marek, et al.
Pubblicazione: (2025)