Logistic Regression makes small LLMs strong and explainable "tens-of-shot" classifiers
Fuente:
arXiv
Saved in:
| Main Authors: | Buckmann, Marcus, Hill, Edward |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
by: Wang, Huanqian, et al.
Published: (2024)
by: Wang, Huanqian, et al.
Published: (2024)
A Survey on Collaborating Small and Large Language Models for Performance, Cost-effectiveness, Cloud-edge Privacy, and Trustworthiness
by: Wang, Fali, et al.
Published: (2025)
by: Wang, Fali, et al.
Published: (2025)
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
by: Wang, Fali, et al.
Published: (2024)
by: Wang, Fali, et al.
Published: (2024)
Do Reasoning Models Enhance Embedding Models?
by: Chan, Wun Yu, et al.
Published: (2026)
by: Chan, Wun Yu, et al.
Published: (2026)
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
by: Kuz, Mykola, et al.
Published: (2025)
by: Kuz, Mykola, et al.
Published: (2025)
Linguistic Collapse: Neural Collapse in (Large) Language Models
by: Wu, Robert, et al.
Published: (2024)
by: Wu, Robert, et al.
Published: (2024)
Atyaephyra at SemEval-2025 Task 4: Low-Rank Negative Preference Optimization
by: Bronec, Jan, et al.
Published: (2025)
by: Bronec, Jan, et al.
Published: (2025)
Pairwise Comparison for Bias Identification and Quantification
by: Haak, Fabian, et al.
Published: (2025)
by: Haak, Fabian, et al.
Published: (2025)
Make Literature-Based Discovery Great Again through Reproducible Pipelines
by: Cestnik, Bojan, et al.
Published: (2025)
by: Cestnik, Bojan, et al.
Published: (2025)
Uncovering Uncertainty in Transformer Inference
by: Brothers, Greyson, et al.
Published: (2024)
by: Brothers, Greyson, et al.
Published: (2024)
Drama Engine: A Framework for Narrative Agents
by: Pichlmair, Martin, et al.
Published: (2024)
by: Pichlmair, Martin, et al.
Published: (2024)
The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care
by: Faust, Douglas K., et al.
Published: (2026)
by: Faust, Douglas K., et al.
Published: (2026)
Intelligence Without Integrity: Why Capable LLMs May Undermine Reliability
by: Allen, Ryan, et al.
Published: (2026)
by: Allen, Ryan, et al.
Published: (2026)
Semantic Retention and Extreme Compression in LLMs: Can We Have Both?
by: Laborde, Stanislas, et al.
Published: (2025)
by: Laborde, Stanislas, et al.
Published: (2025)
Whisper-LM: Improving ASR Models with Language Models for Low-Resource Languages
by: de Zuazo, Xabier, et al.
Published: (2025)
by: de Zuazo, Xabier, et al.
Published: (2025)
ProactBench: Beyond What The User Asked For
by: Harfi, Sepehr, et al.
Published: (2026)
by: Harfi, Sepehr, et al.
Published: (2026)
Quantum NLP models on Natural Language Inference
by: Sun, Ling, et al.
Published: (2025)
by: Sun, Ling, et al.
Published: (2025)
Generative AI Models: Opportunities and Risks for Industry and Authorities
by: Alt, Tobias, et al.
Published: (2024)
by: Alt, Tobias, et al.
Published: (2024)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
by: Yim, Wen-wai, et al.
Published: (2025)
by: Yim, Wen-wai, et al.
Published: (2025)
DanceHA: A Multi-Agent Framework for Document-Level Aspect-Based Sentiment Analysis
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
Symphonym: Universal Phonetic Embeddings for Cross-Script Name Matching
by: Gadd, Stephen
Published: (2026)
by: Gadd, Stephen
Published: (2026)
From Black Box to Glass Box: Cross-Model ASR Disagreement to Prioto Review in Ambient AI Scribe Documentation
by: Karbalaie, Abdolamir, et al.
Published: (2026)
by: Karbalaie, Abdolamir, et al.
Published: (2026)
AI in Investment Analysis: LLMs for Equity Stock Ratings
by: Papasotiriou, Kassiani, et al.
Published: (2024)
by: Papasotiriou, Kassiani, et al.
Published: (2024)
Text Clustering with Large Language Model Embeddings
by: Petukhova, Alina, et al.
Published: (2024)
by: Petukhova, Alina, et al.
Published: (2024)
Neural Tucker Convolutional Network for Water Quality Analysis
by: Si, Hongnan, et al.
Published: (2025)
by: Si, Hongnan, et al.
Published: (2025)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
by: Moradbeiki, Pardis, et al.
Published: (2024)
by: Moradbeiki, Pardis, et al.
Published: (2024)
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer
by: Zhang, Tony, et al.
Published: (2025)
by: Zhang, Tony, et al.
Published: (2025)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
by: Breneur, Oleksandr Marchenko, et al.
Published: (2026)
by: Breneur, Oleksandr Marchenko, et al.
Published: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
by: Lasbordes, Maxence, et al.
Published: (2026)
by: Lasbordes, Maxence, et al.
Published: (2026)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
by: Kaiser, Daniel, et al.
Published: (2025)
by: Kaiser, Daniel, et al.
Published: (2025)
TRAWL: Tensor Reduced and Approximated Weights for Large Language Models
by: Luo, Yiran, et al.
Published: (2024)
by: Luo, Yiran, et al.
Published: (2024)
TableMoE: Neuro-Symbolic Routing for Structured Expert Reasoning in Multimodal Table Understanding
by: Zhang, Junwen, et al.
Published: (2025)
by: Zhang, Junwen, et al.
Published: (2025)
Increasing the Difficulty of Automatically Generated Questions via Reinforcement Learning with Synthetic Preference
by: Thorne, William, et al.
Published: (2024)
by: Thorne, William, et al.
Published: (2024)
GraphSkill: Documentation-Guided Hierarchical Retrieval-Augmented Coding for Complex Graph Reasoning
by: Wang, Fali, et al.
Published: (2026)
by: Wang, Fali, et al.
Published: (2026)
Semi-automated extraction of research topics and trends from NCI funding in radiological sciences from 2000-2020
by: Nguyen, Mark, et al.
Published: (2023)
by: Nguyen, Mark, et al.
Published: (2023)
Claim Automation using Large Language Model
by: Mo, Zhengda, et al.
Published: (2026)
by: Mo, Zhengda, et al.
Published: (2026)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
by: Miliani, Martina, et al.
Published: (2025)
by: Miliani, Martina, et al.
Published: (2025)
Context Aware Lemmatization and Morphological Tagging Method in Turkish
by: Sayallar, Cagri
Published: (2025)
by: Sayallar, Cagri
Published: (2025)
MetaCheckGPT -- A Multi-task Hallucination Detector Using LLM Uncertainty and Meta-models
by: Mehta, Rahul, et al.
Published: (2024)
by: Mehta, Rahul, et al.
Published: (2024)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
by: Fu, Tianyu, et al.
Published: (2025)
by: Fu, Tianyu, et al.
Published: (2025)
Similar Items
-
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
by: Wang, Huanqian, et al.
Published: (2024) -
A Survey on Collaborating Small and Large Language Models for Performance, Cost-effectiveness, Cloud-edge Privacy, and Trustworthiness
by: Wang, Fali, et al.
Published: (2025) -
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
by: Wang, Fali, et al.
Published: (2024) -
Do Reasoning Models Enhance Embedding Models?
by: Chan, Wun Yu, et al.
Published: (2026) -
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
by: Kuz, Mykola, et al.
Published: (2025)