B-cos LM: Efficiently Transforming Pre-trained Language Models for Improved Explainability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yifan, Rao, Sukrut, Lee, Ji-Ung, Jobanputra, Mayank, Demberg, Vera |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
RSA-Control: A Pragmatics-Grounded Lightweight Controllable Text Generation Framework
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization
von: Liu, Dongqi, et al.
Veröffentlicht: (2023)
von: Liu, Dongqi, et al.
Veröffentlicht: (2023)
PhoneLM:an Efficient and Capable Small Language Model Family through Principled Pre-training
von: Yi, Rongjie, et al.
Veröffentlicht: (2024)
von: Yi, Rongjie, et al.
Veröffentlicht: (2024)
Modeling Orthographic Variation Improves NLP Performance for Nigerian Pidgin
von: Lin, Pin-Jie, et al.
Veröffentlicht: (2024)
von: Lin, Pin-Jie, et al.
Veröffentlicht: (2024)
ReasoningLM: Enabling Structural Subgraph Reasoning in Pre-trained Language Models for Question Answering over Knowledge Graph
von: Jiang, Jinhao, et al.
Veröffentlicht: (2023)
von: Jiang, Jinhao, et al.
Veröffentlicht: (2023)
SciNews: From Scholarly Complexities to Public Narratives -- A Dataset for Scientific News Report Generation
von: Liu, Dongqi, et al.
Veröffentlicht: (2024)
von: Liu, Dongqi, et al.
Veröffentlicht: (2024)
Boosting Explainability through Selective Rationalization in Pre-trained Language Models
von: Yuan, Libing, et al.
Veröffentlicht: (2025)
von: Yuan, Libing, et al.
Veröffentlicht: (2025)
ChatGPT vs Human-authored Text: Insights into Controllable Text Summarization and Sentence Style Transfer
von: Liu, Dongqi, et al.
Veröffentlicht: (2023)
von: Liu, Dongqi, et al.
Veröffentlicht: (2023)
RST-LoRA: A Discourse-Aware Low-Rank Adaptation for Long Document Abstractive Summarization
von: Liu, Dongqi, et al.
Veröffentlicht: (2024)
von: Liu, Dongqi, et al.
Veröffentlicht: (2024)
Temperature-scaling surprisal estimates improve fit to human reading times -- but does it do so for the "right reasons"?
von: Liu, Tong, et al.
Veröffentlicht: (2023)
von: Liu, Tong, et al.
Veröffentlicht: (2023)
Born a Transformer -- Always a Transformer? On the Effect of Pretraining on Architectural Abilities
von: Jobanputra, Mayank, et al.
Veröffentlicht: (2025)
von: Jobanputra, Mayank, et al.
Veröffentlicht: (2025)
Pre-trained Language Models Improve the Few-shot Prompt Ability of Decision Transformer
von: Yang, Yu, et al.
Veröffentlicht: (2024)
von: Yang, Yu, et al.
Veröffentlicht: (2024)
Planning Ahead with RSA: Efficient Signalling in Dynamic Environments by Projecting User Awareness across Future Timesteps
von: Das, Anwesha, et al.
Veröffentlicht: (2025)
von: Das, Anwesha, et al.
Veröffentlicht: (2025)
Efficient Data Learning for Open Information Extraction with Pre-trained Language Models
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2023)
Prompting Implicit Discourse Relation Annotation
von: Yung, Frances, et al.
Veröffentlicht: (2024)
von: Yung, Frances, et al.
Veröffentlicht: (2024)
Efficient Language Adaptive Pre-training: Extending State-of-the-Art Large Language Models for Polish
von: Ruciński, Szymon
Veröffentlicht: (2024)
von: Ruciński, Szymon
Veröffentlicht: (2024)
ShishuLM : Achieving Optimal and Efficient Parameterization with Low Attention Transformer Models
von: Kumar, Shivanshu, et al.
Veröffentlicht: (2025)
von: Kumar, Shivanshu, et al.
Veröffentlicht: (2025)
LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
von: Han, Chi, et al.
Veröffentlicht: (2023)
von: Han, Chi, et al.
Veröffentlicht: (2023)
UrduLM: A Resource-Efficient Monolingual Urdu Language Model
von: Ali, Syed Muhammad, et al.
Veröffentlicht: (2026)
von: Ali, Syed Muhammad, et al.
Veröffentlicht: (2026)
Probing Language Models for Pre-training Data Detection
von: Liu, Zhenhua, et al.
Veröffentlicht: (2024)
von: Liu, Zhenhua, et al.
Veröffentlicht: (2024)
SoftDedup: an Efficient Data Reweighting Method for Speeding Up Language Model Pre-training
von: He, Nan, et al.
Veröffentlicht: (2024)
von: He, Nan, et al.
Veröffentlicht: (2024)
CASE: Efficient Curricular Data Pre-training for Building Assistive Psychology Expert Models
von: Harne, Sarthak, et al.
Veröffentlicht: (2024)
von: Harne, Sarthak, et al.
Veröffentlicht: (2024)
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer
von: Lee, Seungyoon, et al.
Veröffentlicht: (2025)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2025)
MELT: Materials-aware Continued Pre-training for Language Model Adaptation to Materials Science
von: Kim, Junho, et al.
Veröffentlicht: (2024)
von: Kim, Junho, et al.
Veröffentlicht: (2024)
SQL-PaLM: Improved Large Language Model Adaptation for Text-to-SQL (extended)
von: Sun, Ruoxi, et al.
Veröffentlicht: (2023)
von: Sun, Ruoxi, et al.
Veröffentlicht: (2023)
Tug-of-war between idioms' figurative and literal interpretations in LLMs
von: Oh, Soyoung, et al.
Veröffentlicht: (2025)
von: Oh, Soyoung, et al.
Veröffentlicht: (2025)
PreCog: Exploring the Relation between Memorization and Performance in Pre-trained Language Models
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2023)
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2023)
Can Pre-trained Language Models Understand Chinese Humor?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Pre-trained Large Language Models for Financial Sentiment Analysis
von: Luo, Wei, et al.
Veröffentlicht: (2024)
von: Luo, Wei, et al.
Veröffentlicht: (2024)
Spike No More: Stabilizing the Pre-training of Large Language Models
von: Takase, Sho, et al.
Veröffentlicht: (2023)
von: Takase, Sho, et al.
Veröffentlicht: (2023)
Pragmatic Reasoning improves LLM Code Generation
von: Cao, Zhuchen, et al.
Veröffentlicht: (2025)
von: Cao, Zhuchen, et al.
Veröffentlicht: (2025)
Explanatory Summarization with Discourse-Driven Planning
von: Liu, Dongqi, et al.
Veröffentlicht: (2025)
von: Liu, Dongqi, et al.
Veröffentlicht: (2025)
TransGPT: Multi-modal Generative Pre-trained Transformer for Transportation
von: Wang, Peng, et al.
Veröffentlicht: (2024)
von: Wang, Peng, et al.
Veröffentlicht: (2024)
DocMamba: Efficient Document Pre-training with State Space Model
von: Hu, Pengfei, et al.
Veröffentlicht: (2024)
von: Hu, Pengfei, et al.
Veröffentlicht: (2024)
SPAFIT: Stratified Progressive Adaptation Fine-tuning for Pre-trained Large Language Models
von: Arora, Samir, et al.
Veröffentlicht: (2024)
von: Arora, Samir, et al.
Veröffentlicht: (2024)
AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing
von: Thillainathan, Sarubi, et al.
Veröffentlicht: (2026)
von: Thillainathan, Sarubi, et al.
Veröffentlicht: (2026)
CogLM: Tracking Cognitive Development of Large Language Models
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
PonderLM: Pretraining Language Models to Ponder in Continuous Space
von: Zeng, Boyi, et al.
Veröffentlicht: (2025)
von: Zeng, Boyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
von: Wang, Yifan, et al.
Veröffentlicht: (2025) -
RSA-Control: A Pragmatics-Grounded Lightweight Controllable Text Generation Framework
von: Wang, Yifan, et al.
Veröffentlicht: (2024) -
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization
von: Liu, Dongqi, et al.
Veröffentlicht: (2023) -
PhoneLM:an Efficient and Capable Small Language Model Family through Principled Pre-training
von: Yi, Rongjie, et al.
Veröffentlicht: (2024) -
Modeling Orthographic Variation Improves NLP Performance for Nigerian Pidgin
von: Lin, Pin-Jie, et al.
Veröffentlicht: (2024)