Transformers perform adaptive partial pooling
Fuente:
arXiv
Salvato in:
| Autore principale: | Kapatsinski, Vsevolod |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Transformers for molecular property prediction: Domain adaptation efficiently improves performance
di: Sultan, Afnan, et al.
Pubblicazione: (2025)
di: Sultan, Afnan, et al.
Pubblicazione: (2025)
SemPool: Simple, robust, and interpretable KG pooling for enhancing language models
di: Mavromatis, Costas, et al.
Pubblicazione: (2024)
di: Mavromatis, Costas, et al.
Pubblicazione: (2024)
Transformer-Squared: Self-adaptive LLMs
di: Sun, Qi, et al.
Pubblicazione: (2025)
di: Sun, Qi, et al.
Pubblicazione: (2025)
Optimizing Transformer based on high-performance optimizer for predicting employment sentiment in American social media content
di: Wang, Feiyang, et al.
Pubblicazione: (2024)
di: Wang, Feiyang, et al.
Pubblicazione: (2024)
REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning
di: Lee, Seungmin, et al.
Pubblicazione: (2026)
di: Lee, Seungmin, et al.
Pubblicazione: (2026)
Facilitating large language model Russian adaptation with Learned Embedding Propagation
di: Tikhomirov, Mikhail, et al.
Pubblicazione: (2024)
di: Tikhomirov, Mikhail, et al.
Pubblicazione: (2024)
Ada-LEval: Evaluating long-context LLMs with length-adaptable benchmarks
di: Wang, Chonghua, et al.
Pubblicazione: (2024)
di: Wang, Chonghua, et al.
Pubblicazione: (2024)
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
di: Barboule, Camille, et al.
Pubblicazione: (2024)
di: Barboule, Camille, et al.
Pubblicazione: (2024)
LengClaro2023: A Dataset of Administrative Texts in Spanish with Plain Language adaptations
di: Agüera-Marco, Belén, et al.
Pubblicazione: (2025)
di: Agüera-Marco, Belén, et al.
Pubblicazione: (2025)
adaptNMT: an open-source, language-agnostic development environment for Neural Machine Translation
di: Lankford, Séamus, et al.
Pubblicazione: (2024)
di: Lankford, Séamus, et al.
Pubblicazione: (2024)
Can LLMs perform structured graph reasoning?
di: Agrawal, Palaash, et al.
Pubblicazione: (2024)
di: Agrawal, Palaash, et al.
Pubblicazione: (2024)
Dependency Transformer Grammars: Integrating Dependency Structures into Transformer Language Models
di: Zhao, Yida, et al.
Pubblicazione: (2024)
di: Zhao, Yida, et al.
Pubblicazione: (2024)
Transformers, Contextualism, and Polysemy
di: Grindrod, Jumbly
Pubblicazione: (2024)
di: Grindrod, Jumbly
Pubblicazione: (2024)
Can formal argumentative reasoning enhance LLMs performances?
di: Castagna, Federico, et al.
Pubblicazione: (2024)
di: Castagna, Federico, et al.
Pubblicazione: (2024)
DepressLLM: Interpretable domain-adapted language model for depression detection from real-world narratives
di: Moon, Sehwan, et al.
Pubblicazione: (2025)
di: Moon, Sehwan, et al.
Pubblicazione: (2025)
RigoChat 2: an adapted language model to Spanish using a bounded dataset and reduced hardware
di: Gómez, Gonzalo Santamaría, et al.
Pubblicazione: (2025)
di: Gómez, Gonzalo Santamaría, et al.
Pubblicazione: (2025)
adaptMLLM: Fine-Tuning Multilingual Language Models on Low-Resource Languages with Integrated LLM Playgrounds
di: Lankford, Séamus, et al.
Pubblicazione: (2024)
di: Lankford, Séamus, et al.
Pubblicazione: (2024)
SeDT: Sentence-Transformer Decision-Transformer Conditioning for Multi-Turn Conversation Reliability
di: Setti, Ramakrishna Vamsi, et al.
Pubblicazione: (2026)
di: Setti, Ramakrishna Vamsi, et al.
Pubblicazione: (2026)
HarmTransform: Transforming Explicit Harmful Queries into Stealthy via Multi-Agent Debate
di: Zhu, Shenzhe
Pubblicazione: (2025)
di: Zhu, Shenzhe
Pubblicazione: (2025)
TreeCoders: Trees of Transformers
di: D'Istria, Pierre Colonna, et al.
Pubblicazione: (2024)
di: D'Istria, Pierre Colonna, et al.
Pubblicazione: (2024)
Memorization in Attention-only Transformers
di: Dana, Léo, et al.
Pubblicazione: (2024)
di: Dana, Léo, et al.
Pubblicazione: (2024)
Are LLMs effective psychological assessors? Leveraging adaptive RAG for interpretable mental health screening through psychometric practice
di: Ravenda, Federico, et al.
Pubblicazione: (2025)
di: Ravenda, Federico, et al.
Pubblicazione: (2025)
Large language models management of medications: three performance analyses
di: Henry, Kelli, et al.
Pubblicazione: (2025)
di: Henry, Kelli, et al.
Pubblicazione: (2025)
Does quantization affect models' performance on long-context tasks?
di: Mekala, Anmol, et al.
Pubblicazione: (2025)
di: Mekala, Anmol, et al.
Pubblicazione: (2025)
Development and multi-center evaluation of domain-adapted speech recognition for human-AI teaming in real-world gastrointestinal endoscopy
di: Yang, Ruijie, et al.
Pubblicazione: (2026)
di: Yang, Ruijie, et al.
Pubblicazione: (2026)
Tandem Transformers for Inference Efficient LLMs
di: S, Aishwarya P, et al.
Pubblicazione: (2024)
di: S, Aishwarya P, et al.
Pubblicazione: (2024)
Word Meanings in Transformer Language Models
di: Grindrod, Jumbly, et al.
Pubblicazione: (2025)
di: Grindrod, Jumbly, et al.
Pubblicazione: (2025)
Question Classification with Deep Contextualized Transformer
di: Luo, Haozheng, et al.
Pubblicazione: (2019)
di: Luo, Haozheng, et al.
Pubblicazione: (2019)
Combining Transformers with Natural Language Explanations
di: Ruggeri, Federico, et al.
Pubblicazione: (2021)
di: Ruggeri, Federico, et al.
Pubblicazione: (2021)
Weighted Grouped Query Attention in Transformers
di: Chinnakonduru, Sai Sena, et al.
Pubblicazione: (2024)
di: Chinnakonduru, Sai Sena, et al.
Pubblicazione: (2024)
Superhuman performance of a large language model on the reasoning tasks of a physician
di: Brodeur, Peter G., et al.
Pubblicazione: (2024)
di: Brodeur, Peter G., et al.
Pubblicazione: (2024)
Language Models can perform Single-Utterance Self-Correction of Perturbed Reasoning
di: Silver, Sam, et al.
Pubblicazione: (2025)
di: Silver, Sam, et al.
Pubblicazione: (2025)
Reactor Mk.1 performances: MMLU, HumanEval and BBH test results
di: Dunham, TJ, et al.
Pubblicazione: (2024)
di: Dunham, TJ, et al.
Pubblicazione: (2024)
Outliers Dimensions that Disrupt Transformers Are Driven by Frequency
di: Puccetti, Giovanni, et al.
Pubblicazione: (2022)
di: Puccetti, Giovanni, et al.
Pubblicazione: (2022)
A Comparative Study on Code Generation with Transformers
di: Das, Namrata, et al.
Pubblicazione: (2024)
di: Das, Namrata, et al.
Pubblicazione: (2024)
Transformers in the Service of Description Logic-based Contexts
di: Poulis, Angelos, et al.
Pubblicazione: (2023)
di: Poulis, Angelos, et al.
Pubblicazione: (2023)
Heterogeneous Subgraph Transformer for Fake News Detection
di: Zhang, Yuchen, et al.
Pubblicazione: (2024)
di: Zhang, Yuchen, et al.
Pubblicazione: (2024)
ModRWKV: Transformer Multimodality in Linear Time
di: Kang, Jiale, et al.
Pubblicazione: (2025)
di: Kang, Jiale, et al.
Pubblicazione: (2025)
Intra-Layer Recurrence in Transformers for Language Modeling
di: Nguyen, Anthony, et al.
Pubblicazione: (2025)
di: Nguyen, Anthony, et al.
Pubblicazione: (2025)
TrInk: Ink Generation with Transformer Network
di: Jin, Zezhong, et al.
Pubblicazione: (2025)
di: Jin, Zezhong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Transformers for molecular property prediction: Domain adaptation efficiently improves performance
di: Sultan, Afnan, et al.
Pubblicazione: (2025) -
SemPool: Simple, robust, and interpretable KG pooling for enhancing language models
di: Mavromatis, Costas, et al.
Pubblicazione: (2024) -
Transformer-Squared: Self-adaptive LLMs
di: Sun, Qi, et al.
Pubblicazione: (2025) -
Optimizing Transformer based on high-performance optimizer for predicting employment sentiment in American social media content
di: Wang, Feiyang, et al.
Pubblicazione: (2024) -
REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning
di: Lee, Seungmin, et al.
Pubblicazione: (2026)