YaPO: Learnable Sparse Activation Steering Vectors for Domain Adaptation
Fuente:
arXiv
Saved in:
| Main Authors: | Bounhar, Abdelaziz, Elbadry, Rania Hossam Elmohamady, Abdine, Hadi, Nakov, Preslav, Vazirgiannis, Michalis, Shang, Guokan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Nile-Chat: Egyptian Language Models for Arabic and Latin Scripts
by: Shang, Guokan, et al.
Published: (2025)
by: Shang, Guokan, et al.
Published: (2025)
Shorter but not Worse: Frugal Reasoning via Easy Samples as Length Regularizers in Math RLVR
by: Bounhar, Abdelaziz, et al.
Published: (2025)
by: Bounhar, Abdelaziz, et al.
Published: (2025)
Beyond Random Sampling: Efficient Language Model Pretraining via Curriculum Learning
by: Zhang, Yang, et al.
Published: (2025)
by: Zhang, Yang, et al.
Published: (2025)
LLM as a Broken Telephone: Iterative Generation Distorts Information
by: Mohamed, Amr, et al.
Published: (2025)
by: Mohamed, Amr, et al.
Published: (2025)
Leveraging Discourse Structure for Extractive Meeting Summarization
by: Rennard, Virgile, et al.
Published: (2024)
by: Rennard, Virgile, et al.
Published: (2024)
Markovian Generation Chains in Large Language Models
by: Geng, Mingmeng, et al.
Published: (2026)
by: Geng, Mingmeng, et al.
Published: (2026)
Atlas-Chat: Adapting Large Language Models for Low-Resource Moroccan Arabic Dialect
by: Shang, Guokan, et al.
Published: (2024)
by: Shang, Guokan, et al.
Published: (2024)
Prot2Text: Multimodal Protein's Function Generation with GNNs and Transformers
by: Abdine, Hadi, et al.
Published: (2023)
by: Abdine, Hadi, et al.
Published: (2023)
The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations
by: Elbadry, Rania, et al.
Published: (2026)
by: Elbadry, Rania, et al.
Published: (2026)
Word Sense Induction with Hierarchical Clustering and Mutual Information Maximization
by: Abdine, Hadi, et al.
Published: (2022)
by: Abdine, Hadi, et al.
Published: (2022)
Graph Linearization Methods for Reasoning on Graphs with Large Language Models
by: Xypolopoulos, Christos, et al.
Published: (2024)
by: Xypolopoulos, Christos, et al.
Published: (2024)
The Curious Decline of Linguistic Diversity: Training Language Models on Synthetic Text
by: Guo, Yanzhu, et al.
Published: (2023)
by: Guo, Yanzhu, et al.
Published: (2023)
Fast-Decoding Diffusion Language Models via Progress-Aware Confidence Schedules
by: Mohamed, Amr, et al.
Published: (2025)
by: Mohamed, Amr, et al.
Published: (2025)
Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text
by: Mohamed, Amr, et al.
Published: (2025)
by: Mohamed, Amr, et al.
Published: (2025)
EngTrace: A Symbolic Benchmark for Verifiable Process Supervision of Engineering Reasoning
by: Gull, Ayesha, et al.
Published: (2025)
by: Gull, Ayesha, et al.
Published: (2025)
Instruction-Guided Poetry Generation in Arabic and Its Dialects
by: Sadallah, Abdelrahman, et al.
Published: (2026)
by: Sadallah, Abdelrahman, et al.
Published: (2026)
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
by: Almheiri, Saeed, et al.
Published: (2025)
by: Almheiri, Saeed, et al.
Published: (2025)
PTPP-Aware Adaptation Scaling Laws: Predicting Domain-Adaptation Performance at Unseen Pre-Training Budgets
by: Goffinet, Etienne, et al.
Published: (2025)
by: Goffinet, Etienne, et al.
Published: (2025)
Neural Graph Generator: Feature-Conditioned Graph Generation using Latent Diffusion Models
by: Evdaimon, Iakovos, et al.
Published: (2024)
by: Evdaimon, Iakovos, et al.
Published: (2024)
Prot2Text-V2: Protein Function Prediction with Multimodal Contrastive Alignment
by: Fei, Xiao, et al.
Published: (2025)
by: Fei, Xiao, et al.
Published: (2025)
Explaining Predictions by Characteristic Rules
by: Alkhatib, Amr, et al.
Published: (2024)
by: Alkhatib, Amr, et al.
Published: (2024)
Bias in the Mirror: Are LLMs opinions robust to their own adversarial attacks ?
by: Rennard, Virgile, et al.
Published: (2024)
by: Rennard, Virgile, et al.
Published: (2024)
From Chaos to Clarity: Claim Normalization to Empower Fact-Checking
by: Sundriyal, Megha, et al.
Published: (2023)
by: Sundriyal, Megha, et al.
Published: (2023)
DenoiseRank: Learning to Rank by Diffusion Models
by: Wang, Ying, et al.
Published: (2026)
by: Wang, Ying, et al.
Published: (2026)
Adapting Fake News Detection to the Era of Large Language Models
by: Su, Jinyan, et al.
Published: (2023)
by: Su, Jinyan, et al.
Published: (2023)
Domain Adaptation and Multi-view Attention for Learnable Landmark Tracking with Sparse Data
by: Chase Jr, Timothy, et al.
Published: (2025)
by: Chase Jr, Timothy, et al.
Published: (2025)
Obtaining Example-Based Explanations from Deep Neural Networks
by: Dong, Genghua, et al.
Published: (2025)
by: Dong, Genghua, et al.
Published: (2025)
Interpretable Graph Neural Networks for Tabular Data
by: Alkhatib, Amr, et al.
Published: (2023)
by: Alkhatib, Amr, et al.
Published: (2023)
VISPA: Pluralistic Alignment via Automatic Value Selection and Activation
by: Zheng, Shenyan, et al.
Published: (2026)
by: Zheng, Shenyan, et al.
Published: (2026)
MixtureKit: A General Framework for Composing, Training, and Visualizing Mixture-of-Experts Models
by: Chamma, Ahmad, et al.
Published: (2025)
by: Chamma, Ahmad, et al.
Published: (2025)
Multimodal Large Language Models to Support Real-World Fact-Checking
by: Geng, Jiahui, et al.
Published: (2024)
by: Geng, Jiahui, et al.
Published: (2024)
Between Underthinking and Overthinking: An Empirical Study of Reasoning Length and correctness in LLMs
by: Su, Jinyan, et al.
Published: (2025)
by: Su, Jinyan, et al.
Published: (2025)
Fast or Better? Balancing Accuracy and Cost in Retrieval-Augmented Generation with Flexible User Control
by: Su, Jinyan, et al.
Published: (2025)
by: Su, Jinyan, et al.
Published: (2025)
Persona Vectors in Games: Measuring and Steering Strategies via Activation Vectors
by: Sun, Johnathan, et al.
Published: (2026)
by: Sun, Johnathan, et al.
Published: (2026)
Improving Steering Vectors by Targeting Sparse Autoencoder Features
by: Chalnev, Sviatoslav, et al.
Published: (2024)
by: Chalnev, Sviatoslav, et al.
Published: (2024)
Steering Large Language Model Activations in Sparse Spaces
by: Bayat, Reza, et al.
Published: (2025)
by: Bayat, Reza, et al.
Published: (2025)
Unveiling Covert Semantics: Joint Source-Channel Coding Under a Covertness Constraint
by: Bounhar, Abdelaziz, et al.
Published: (2024)
by: Bounhar, Abdelaziz, et al.
Published: (2024)
Covert Multi-Access Communication with a Non-Covert User
by: Bounhar, Abdelaziz, et al.
Published: (2024)
by: Bounhar, Abdelaziz, et al.
Published: (2024)
Whispering Secrets in a Crowd: Leveraging Non-Covert Users for Covert Communications
by: Bounhar, Abdelaziz, et al.
Published: (2024)
by: Bounhar, Abdelaziz, et al.
Published: (2024)
Distributed Detection under Stringent Resource Constraints
by: Bounhar, Abdelaziz, et al.
Published: (2026)
by: Bounhar, Abdelaziz, et al.
Published: (2026)
Similar Items
-
Nile-Chat: Egyptian Language Models for Arabic and Latin Scripts
by: Shang, Guokan, et al.
Published: (2025) -
Shorter but not Worse: Frugal Reasoning via Easy Samples as Length Regularizers in Math RLVR
by: Bounhar, Abdelaziz, et al.
Published: (2025) -
Beyond Random Sampling: Efficient Language Model Pretraining via Curriculum Learning
by: Zhang, Yang, et al.
Published: (2025) -
LLM as a Broken Telephone: Iterative Generation Distorts Information
by: Mohamed, Amr, et al.
Published: (2025) -
Leveraging Discourse Structure for Extractive Meeting Summarization
by: Rennard, Virgile, et al.
Published: (2024)