Can bidirectional encoder become the ultimate winner for downstream applications of foundation models?
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Lewen, Zhou, Xuanyu, Fan, Juao, Xie, Xinyi, Zhu, Shengxin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An overview of domain-specific foundation model: key technologies, applications and challenges
por: Chen, Haolong, et al.
Publicado: (2024)
por: Chen, Haolong, et al.
Publicado: (2024)
When AI companions become witty: Can human brain recognize AI-generated irony?
por: Rao, Xiaohui, et al.
Publicado: (2025)
por: Rao, Xiaohui, et al.
Publicado: (2025)
Language models scale reliably with over-training and on downstream tasks
por: Gadre, Samir Yitzhak, et al.
Publicado: (2024)
por: Gadre, Samir Yitzhak, et al.
Publicado: (2024)
Unleashing the potential of prompt engineering for large language models
por: Chen, Banghao, et al.
Publicado: (2023)
por: Chen, Banghao, et al.
Publicado: (2023)
Developing ChemDFM as a large language foundation model for chemistry
por: Zhao, Zihan, et al.
Publicado: (2024)
por: Zhao, Zihan, et al.
Publicado: (2024)
Open foundation models for Azerbaijani language
por: Isbarov, Jafar, et al.
Publicado: (2024)
por: Isbarov, Jafar, et al.
Publicado: (2024)
Effects of diversity incentives on sample diversity and downstream model performance in LLM-based text augmentation
por: Cegin, Jan, et al.
Publicado: (2024)
por: Cegin, Jan, et al.
Publicado: (2024)
Canonical bidirectional typechecking
por: Mihejevs, Zanzi, et al.
Publicado: (2025)
por: Mihejevs, Zanzi, et al.
Publicado: (2025)
Energy-Aware LLMs: A step towards sustainable AI for downstream applications
por: Tran, Nguyen Phuc, et al.
Publicado: (2025)
por: Tran, Nguyen Phuc, et al.
Publicado: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
por: Kumar, Aswini, et al.
Publicado: (2025)
por: Kumar, Aswini, et al.
Publicado: (2025)
As Firm As Their Foundations: Can open-sourced foundation models be used to create adversarial examples for downstream tasks?
por: Hu, Anjun, et al.
Publicado: (2024)
por: Hu, Anjun, et al.
Publicado: (2024)
Contextual morphologically-guided tokenization for Latin encoder models
por: Hudspeth, Marisa, et al.
Publicado: (2025)
por: Hudspeth, Marisa, et al.
Publicado: (2025)
The language of time: a language model perspective on time-series foundation models
por: Xie, Yi, et al.
Publicado: (2025)
por: Xie, Yi, et al.
Publicado: (2025)
The Realignment Problem: When Right becomes Wrong in LLMs
por: Sharma, Aakash Sen, et al.
Publicado: (2025)
por: Sharma, Aakash Sen, et al.
Publicado: (2025)
Word2winners at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-Checked Claim Retrieval
por: Azadi, Amirmohammad, et al.
Publicado: (2025)
por: Azadi, Amirmohammad, et al.
Publicado: (2025)
A decoder-only foundation model for time-series forecasting
por: Das, Abhimanyu, et al.
Publicado: (2023)
por: Das, Abhimanyu, et al.
Publicado: (2023)
A foundation model for human-AI collaboration in medical literature mining
por: Wang, Zifeng, et al.
Publicado: (2025)
por: Wang, Zifeng, et al.
Publicado: (2025)
A cross-species neural foundation model for end-to-end speech decoding
por: Zhang, Yizi, et al.
Publicado: (2025)
por: Zhang, Yizi, et al.
Publicado: (2025)
Scaling laws for language encoding models in fMRI
por: Antonello, Richard, et al.
Publicado: (2023)
por: Antonello, Richard, et al.
Publicado: (2023)
What's in a prompt? Language models encode literary style in prompt embeddings
por: Sarfati, Raphaël, et al.
Publicado: (2025)
por: Sarfati, Raphaël, et al.
Publicado: (2025)
Towards a clinically accessible radiology foundation model: open-access and lightweight, with automated evaluation
por: Chaves, Juan Manuel Zambrano, et al.
Publicado: (2024)
por: Chaves, Juan Manuel Zambrano, et al.
Publicado: (2024)
Modeling citation worthiness by using attention-based bidirectional long short-term memory networks and interpretable models
por: Zeng, Tong, et al.
Publicado: (2024)
por: Zeng, Tong, et al.
Publicado: (2024)
Practical token pruning for foundation models in few-shot conversational virtual assistant systems
por: Qi, Haode, et al.
Publicado: (2024)
por: Qi, Haode, et al.
Publicado: (2024)
SoftHateBench: Evaluating Moderation Models Against Reasoning-Driven, Policy-Compliant Hostility
por: Su, Xuanyu, et al.
Publicado: (2026)
por: Su, Xuanyu, et al.
Publicado: (2026)
Simpler becomes Harder: Do LLMs Exhibit a Coherent Behavior on Simplified Corpora?
por: Anschütz, Miriam, et al.
Publicado: (2024)
por: Anschütz, Miriam, et al.
Publicado: (2024)
Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control
por: Sun, Lihao, et al.
Publicado: (2026)
por: Sun, Lihao, et al.
Publicado: (2026)
Are ASR foundation models generalized enough to capture features of regional dialects for low-resource languages?
por: Dipto, Tawsif Tashwar, et al.
Publicado: (2025)
por: Dipto, Tawsif Tashwar, et al.
Publicado: (2025)
PokerBench: Training Large Language Models to become Professional Poker Players
por: Zhuang, Richard, et al.
Publicado: (2025)
por: Zhuang, Richard, et al.
Publicado: (2025)
Can Large Language Models Always Solve Easy Problems if They Can Solve Harder Ones?
por: Yang, Zhe, et al.
Publicado: (2024)
por: Yang, Zhe, et al.
Publicado: (2024)
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
por: Guan, Xinyu, et al.
Publicado: (2025)
por: Guan, Xinyu, et al.
Publicado: (2025)
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts
por: Lamsal, Rabindra, et al.
Publicado: (2023)
por: Lamsal, Rabindra, et al.
Publicado: (2023)
Fine-tuning language encoding models on slow fMRI improves prediction for fast ECoG
por: Vaidya, Aditya R., et al.
Publicado: (2026)
por: Vaidya, Aditya R., et al.
Publicado: (2026)
SurveyBench: Can LLM(-Agents) Write Academic Surveys that Align with Reader Needs?
por: Sun, Zhaojun, et al.
Publicado: (2025)
por: Sun, Zhaojun, et al.
Publicado: (2025)
Can a large language model be a gaslighter?
por: Li, Wei, et al.
Publicado: (2024)
por: Li, Wei, et al.
Publicado: (2024)
Panacea: A foundation model for clinical trial search, summarization, design, and recruitment
por: Lin, Jiacheng, et al.
Publicado: (2024)
por: Lin, Jiacheng, et al.
Publicado: (2024)
What's Mine becomes Yours: Defining, Annotating and Detecting Context-Dependent Paraphrases in News Interview Dialogs
por: Wegmann, Anna, et al.
Publicado: (2024)
por: Wegmann, Anna, et al.
Publicado: (2024)
Peacemaker at ATE-IT: Automatic term extraction from Italian text for waste management data using encoder model
por: Bakhtiyarzadeh, Mahdi, et al.
Publicado: (2026)
por: Bakhtiyarzadeh, Mahdi, et al.
Publicado: (2026)
Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text
por: Jackson, Paul, et al.
Publicado: (2026)
por: Jackson, Paul, et al.
Publicado: (2026)
Scaffolding Coordinates to Promote Vision-Language Coordination in Large Multi-Modal Models
por: Lei, Xuanyu, et al.
Publicado: (2024)
por: Lei, Xuanyu, et al.
Publicado: (2024)
Python is Not Always the Best Choice: Embracing Multilingual Program of Thoughts
por: Luo, Xianzhen, et al.
Publicado: (2024)
por: Luo, Xianzhen, et al.
Publicado: (2024)
Ejemplares similares
-
An overview of domain-specific foundation model: key technologies, applications and challenges
por: Chen, Haolong, et al.
Publicado: (2024) -
When AI companions become witty: Can human brain recognize AI-generated irony?
por: Rao, Xiaohui, et al.
Publicado: (2025) -
Language models scale reliably with over-training and on downstream tasks
por: Gadre, Samir Yitzhak, et al.
Publicado: (2024) -
Unleashing the potential of prompt engineering for large language models
por: Chen, Banghao, et al.
Publicado: (2023) -
Developing ChemDFM as a large language foundation model for chemistry
por: Zhao, Zihan, et al.
Publicado: (2024)