Large Language Models for Patent Classification: Strengths, Trade-offs, and the Long Tail Effect
Fuente:
arXiv
Guardado en:
| Autores principales: | Emer, Lorenzo, Lippi, Marco, Mina, Andrea, Vandin, Andrea |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Advanced Stock Market Prediction Using Long Short-Term Memory Networks: A Comprehensive Deep Learning Framework
por: Chaudhary, Rajneesh
Publicado: (2025)
por: Chaudhary, Rajneesh
Publicado: (2025)
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
por: Nauen, Tobias Christian, et al.
Publicado: (2024)
por: Nauen, Tobias Christian, et al.
Publicado: (2024)
Graph Connectionist Temporal Classification for Phoneme Recognition
por: Grafé, Henry, et al.
Publicado: (2025)
por: Grafé, Henry, et al.
Publicado: (2025)
GradES: Significantly Faster Training in Transformers with Gradient-Based Early Stopping
por: Wen, Qifu, et al.
Publicado: (2025)
por: Wen, Qifu, et al.
Publicado: (2025)
Semantic Retention and Extreme Compression in LLMs: Can We Have Both?
por: Laborde, Stanislas, et al.
Publicado: (2025)
por: Laborde, Stanislas, et al.
Publicado: (2025)
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer
por: Zhang, Tony, et al.
Publicado: (2025)
por: Zhang, Tony, et al.
Publicado: (2025)
Symphonym: Universal Phonetic Embeddings for Cross-Script Name Matching
por: Gadd, Stephen
Publicado: (2026)
por: Gadd, Stephen
Publicado: (2026)
Complex-Valued Phase-Coherent Transformer
por: Hioki, Leona
Publicado: (2026)
por: Hioki, Leona
Publicado: (2026)
Optimized Gradient Clipping for Noisy Label Learning
por: Ye, Xichen, et al.
Publicado: (2024)
por: Ye, Xichen, et al.
Publicado: (2024)
Transparent but Powerful: Explainability, Accuracy, and Generalizability in ADHD Detection from Social Media Data
por: Wiechmann, D., et al.
Publicado: (2024)
por: Wiechmann, D., et al.
Publicado: (2024)
Leveraging large multimodal models for audio-video deepfake detection: a pilot study
por: Cao, Songjun, et al.
Publicado: (2026)
por: Cao, Songjun, et al.
Publicado: (2026)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
por: Gonzalez, Alberto Andres Valdes
Publicado: (2026)
por: Gonzalez, Alberto Andres Valdes
Publicado: (2026)
Energy-Efficient Information Representation in MNIST Classification Using Biologically Inspired Learning
por: Stricker, Patrick, et al.
Publicado: (2026)
por: Stricker, Patrick, et al.
Publicado: (2026)
Improving Large-Scale k-Nearest Neighbor Text Categorization with Label Autoencoders
por: Ribadas-Pena, Francisco J., et al.
Publicado: (2024)
por: Ribadas-Pena, Francisco J., et al.
Publicado: (2024)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
por: Wang, Yihao, et al.
Publicado: (2026)
por: Wang, Yihao, et al.
Publicado: (2026)
A Comparative Study of Feature Selection in Tsetlin Machines
por: Halenka, Vojtech, et al.
Publicado: (2025)
por: Halenka, Vojtech, et al.
Publicado: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
por: Schneider, Felix, et al.
Publicado: (2026)
por: Schneider, Felix, et al.
Publicado: (2026)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
por: Hua, Wenjie, et al.
Publicado: (2025)
por: Hua, Wenjie, et al.
Publicado: (2025)
SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats
por: Gondhalekar, Chinmay, et al.
Publicado: (2025)
por: Gondhalekar, Chinmay, et al.
Publicado: (2025)
Revisiting Non-separable Binary Classification and its Applications in Anomaly Detection
por: Lau, Matthew, et al.
Publicado: (2023)
por: Lau, Matthew, et al.
Publicado: (2023)
Graded Transformers
por: Shaska Sr, Tony
Publicado: (2025)
por: Shaska Sr, Tony
Publicado: (2025)
Is Cambodia the World's Largest Cashew Producer?
por: Chaya, Veasna, et al.
Publicado: (2024)
por: Chaya, Veasna, et al.
Publicado: (2024)
How GenAI Mentor Configurations Shape Early Collaborative Dynamics: A Classroom Comparison of Individual and Shared Agents
por: Zha, Siyu, et al.
Publicado: (2026)
por: Zha, Siyu, et al.
Publicado: (2026)
Predicting When to Trust Vision-Language Models for Spatial Reasoning
por: Imran, Muhammad, et al.
Publicado: (2026)
por: Imran, Muhammad, et al.
Publicado: (2026)
An Ensemble Embedding Approach for Improving Semantic Caching Performance in LLM-based Systems
por: Ghaffari, Shervin, et al.
Publicado: (2025)
por: Ghaffari, Shervin, et al.
Publicado: (2025)
A Lightweight Multi-Expert Generative Language Model System for Engineering Information and Knowledge Extraction
por: Bogachov, Bogdan, et al.
Publicado: (2025)
por: Bogachov, Bogdan, et al.
Publicado: (2025)
Argument Quality Annotation and Gender Bias Detection in Financial Communication through Large Language Models
por: Alhamzeh, Alaa, et al.
Publicado: (2025)
por: Alhamzeh, Alaa, et al.
Publicado: (2025)
Benchmarking Catastrophic Forgetting Mitigation Methods in Federated Time Series Forecasting
por: Hallak, Khaled, et al.
Publicado: (2025)
por: Hallak, Khaled, et al.
Publicado: (2025)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
por: Okpala, Izunna, et al.
Publicado: (2023)
por: Okpala, Izunna, et al.
Publicado: (2023)
CANAL -- Cyber Activity News Alerting Language Model: Empirical Approach vs. Expensive LLM
por: Patel, Urjitkumar, et al.
Publicado: (2024)
por: Patel, Urjitkumar, et al.
Publicado: (2024)
FinTextSim: Enhancing Financial Text Analysis with BERTopic
por: Jehnen, Simon, et al.
Publicado: (2025)
por: Jehnen, Simon, et al.
Publicado: (2025)
Learning Sign Language Representation using CNN LSTM, 3DCNN, CNN RNN LSTM and CCN TD
por: Louison, Nikita, et al.
Publicado: (2024)
por: Louison, Nikita, et al.
Publicado: (2024)
MIMIC-SR-ICD11: A Dataset for Narrative-Based Diagnosis
por: Wu, Yuexin, et al.
Publicado: (2025)
por: Wu, Yuexin, et al.
Publicado: (2025)
DROID: Dual Representation for Out-of-Scope Intent Detection
por: Rashwan, Wael, et al.
Publicado: (2025)
por: Rashwan, Wael, et al.
Publicado: (2025)
Linguistic Collapse: Neural Collapse in (Large) Language Models
por: Wu, Robert, et al.
Publicado: (2024)
por: Wu, Robert, et al.
Publicado: (2024)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
por: Moradbeiki, Pardis, et al.
Publicado: (2024)
por: Moradbeiki, Pardis, et al.
Publicado: (2024)
GATher: Graph Attention Based Predictions of Gene-Disease Links
por: Narganes-Carlon, David, et al.
Publicado: (2024)
por: Narganes-Carlon, David, et al.
Publicado: (2024)
FANAL -- Financial Activity News Alerting Language Modeling Framework
por: Patel, Urjitkumar, et al.
Publicado: (2024)
por: Patel, Urjitkumar, et al.
Publicado: (2024)
LLM-based Extraction of Contradictions from Patents
por: Trapp, Stefan, et al.
Publicado: (2024)
por: Trapp, Stefan, et al.
Publicado: (2024)
Agentic AI Systems Applied to tasks in Financial Services: Modeling and model risk management crews
por: Okpala, Izunna, et al.
Publicado: (2025)
por: Okpala, Izunna, et al.
Publicado: (2025)
Ejemplares similares
-
Advanced Stock Market Prediction Using Long Short-Term Memory Networks: A Comprehensive Deep Learning Framework
por: Chaudhary, Rajneesh
Publicado: (2025) -
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
por: Nauen, Tobias Christian, et al.
Publicado: (2024) -
Graph Connectionist Temporal Classification for Phoneme Recognition
por: Grafé, Henry, et al.
Publicado: (2025) -
GradES: Significantly Faster Training in Transformers with Gradient-Based Early Stopping
por: Wen, Qifu, et al.
Publicado: (2025) -
Semantic Retention and Extreme Compression in LLMs: Can We Have Both?
por: Laborde, Stanislas, et al.
Publicado: (2025)