Where Does Authorship Signal Emerge in Encoder-Based Language Models?
Fuente:
arXiv
Saved in:
| Main Authors: | Kulumba, Francis, Vimont, Guillaume, Romary, Laurent, Cafiero, Florian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HALvest-Contrastive: Retrieval-Like Authorship Attribution with Patch-Level Late Interaction
by: Kulumba, Francis, et al.
Published: (2024)
by: Kulumba, Francis, et al.
Published: (2024)
Triggers Hijack Language Circuits: A Mechanistic Analysis of Backdoor Behaviors in Large Language Models
by: Lasnier, Théo, et al.
Published: (2026)
by: Lasnier, Théo, et al.
Published: (2026)
Language-Switching Triggers Take a Latent Detour Through Language Models
by: Kulumba, Francis, et al.
Published: (2026)
by: Kulumba, Francis, et al.
Published: (2026)
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection
by: Antoun, Wissam, et al.
Published: (2024)
by: Antoun, Wissam, et al.
Published: (2024)
Under-resourced studies of under-resourced languages: lemmatization and POS-tagging with LLM annotators for historical Armenian, Georgian, Greek and Syriac
by: Vidal-Gorène, Chahan, et al.
Published: (2026)
by: Vidal-Gorène, Chahan, et al.
Published: (2026)
CamemBERT-bio: Leveraging Continual Pre-training for Cost-Effective Models on French Biomedical Data
by: Touchent, Rian, et al.
Published: (2023)
by: Touchent, Rian, et al.
Published: (2023)
Conversational Grounding: Annotation and Analysis of Grounding Acts and Grounding Units
by: Mohapatra, Biswesh, et al.
Published: (2024)
by: Mohapatra, Biswesh, et al.
Published: (2024)
From Veracity to Diffusion: Adressing Operational Challenges in Moving From Fake-News Detection to Information Disorders
by: Savatteri, Francesco Paolo, et al.
Published: (2025)
by: Savatteri, Francesco Paolo, et al.
Published: (2025)
Enhancing Authorship Attribution through Embedding Fusion: A Novel Approach with Masked and Encoder-Decoder Language Models
by: Kaushik, Arjun Ramesh, et al.
Published: (2024)
by: Kaushik, Arjun Ramesh, et al.
Published: (2024)
Can Large Language Models Identify Authorship?
by: Huang, Baixiang, et al.
Published: (2024)
by: Huang, Baixiang, et al.
Published: (2024)
ALMs: Authorial Language Models for Authorship Attribution
by: Huang, Weihang, et al.
Published: (2024)
by: Huang, Weihang, et al.
Published: (2024)
Using Machine Mental Imagery for Representing Common Ground in Situated Dialogue
by: Mohapatra, Biswesh, et al.
Published: (2026)
by: Mohapatra, Biswesh, et al.
Published: (2026)
Sui Generis: Large Language Models for Authorship Attribution and Verification in Latin
by: Schmidt, Gleb, et al.
Published: (2024)
by: Schmidt, Gleb, et al.
Published: (2024)
Figuratively Speaking: Authorship Attribution via Multi-Task Figurative Language Modeling
by: Katsios, Gregorios A, et al.
Published: (2024)
by: Katsios, Gregorios A, et al.
Published: (2024)
When and Where Did it Happen? An Encoder-Decoder Model to Identify Scenario Context
by: Noriega-Atala, Enrique, et al.
Published: (2024)
by: Noriega-Atala, Enrique, et al.
Published: (2024)
Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models
by: Beniwal, Himanshu, et al.
Published: (2026)
by: Beniwal, Himanshu, et al.
Published: (2026)
XAM: Interactive Explainability for Authorship Attribution Models
by: Alshomary, Milad, et al.
Published: (2025)
by: Alshomary, Milad, et al.
Published: (2025)
Authorship Without Writing: Large Language Models and the Senior Author Analogy
by: Hurshman, Clint, et al.
Published: (2025)
by: Hurshman, Clint, et al.
Published: (2025)
Authorship Impersonation via LLM Prompting does not Evade Authorship Verification Methods
by: Zeng, Baoyi, et al.
Published: (2026)
by: Zeng, Baoyi, et al.
Published: (2026)
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
by: Nguyen, Tuc, et al.
Published: (2025)
by: Nguyen, Tuc, et al.
Published: (2025)
Diachronic Document Dataset for Semantic Layout Analysis
by: Clérice, Thibault, et al.
Published: (2024)
by: Clérice, Thibault, et al.
Published: (2024)
Encoder-Decoder or Decoder-Only? Revisiting Encoder-Decoder Large Language Model
by: Zhang, Biao, et al.
Published: (2025)
by: Zhang, Biao, et al.
Published: (2025)
Frame of Reference: Addressing the Challenges of Common Ground Representation in Situational Dialogs
by: Mohapatra, Biswesh, et al.
Published: (2026)
by: Mohapatra, Biswesh, et al.
Published: (2026)
Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information
by: Park, Yein, et al.
Published: (2025)
by: Park, Yein, et al.
Published: (2025)
On Multilingual Encoder Language Model Compression for Low-Resource Languages
by: Gurgurov, Daniil, et al.
Published: (2025)
by: Gurgurov, Daniil, et al.
Published: (2025)
Cross-Genre Authorship Attribution via LLM-Based Retrieve-and-Rerank
by: Agarwal, Shantanu, et al.
Published: (2025)
by: Agarwal, Shantanu, et al.
Published: (2025)
Leveraging Multilingual Training for Authorship Representation: Enhancing Generalization across Languages and Domains
by: Kim, Junghwan, et al.
Published: (2025)
by: Kim, Junghwan, et al.
Published: (2025)
Quantifying Misattribution Unfairness in Authorship Attribution
by: Alipoormolabashi, Pegah, et al.
Published: (2025)
by: Alipoormolabashi, Pegah, et al.
Published: (2025)
Authorship Style Transfer with Policy Optimization
by: Liu, Shuai, et al.
Published: (2024)
by: Liu, Shuai, et al.
Published: (2024)
Adapting Decoder-Based Language Models for Diverse Encoder Downstream Tasks
by: Suganthan, Paul, et al.
Published: (2025)
by: Suganthan, Paul, et al.
Published: (2025)
JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
by: Fisher, Jillian, et al.
Published: (2024)
by: Fisher, Jillian, et al.
Published: (2024)
Long-Context Encoder Models for Polish Language Understanding
by: Dadas, Sławomir, et al.
Published: (2026)
by: Dadas, Sławomir, et al.
Published: (2026)
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
Does Language Model Understand Language?
by: Acharjee, Suvojit, et al.
Published: (2025)
by: Acharjee, Suvojit, et al.
Published: (2025)
CliniBench: A Clinical Outcome Prediction Benchmark for Generative and Encoder-Based Language Models
by: Grundmann, Paul, et al.
Published: (2025)
by: Grundmann, Paul, et al.
Published: (2025)
Language Models as Hierarchy Encoders
by: He, Yuan, et al.
Published: (2024)
by: He, Yuan, et al.
Published: (2024)
InstructAV: Instruction Fine-tuning Large Language Models for Authorship Verification
by: Hu, Yujia, et al.
Published: (2024)
by: Hu, Yujia, et al.
Published: (2024)
PART: Pre-trained Authorship Representation Transformer
by: Huertas-Tato, Javier, et al.
Published: (2022)
by: Huertas-Tato, Javier, et al.
Published: (2022)
Residualized Similarity for Faithfully Explainable Authorship Verification
by: Zeng, Peter, et al.
Published: (2025)
by: Zeng, Peter, et al.
Published: (2025)
Should We Still Pretrain Encoders with Masked Language Modeling?
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)
Similar Items
-
HALvest-Contrastive: Retrieval-Like Authorship Attribution with Patch-Level Late Interaction
by: Kulumba, Francis, et al.
Published: (2024) -
Triggers Hijack Language Circuits: A Mechanistic Analysis of Backdoor Behaviors in Large Language Models
by: Lasnier, Théo, et al.
Published: (2026) -
Language-Switching Triggers Take a Latent Detour Through Language Models
by: Kulumba, Francis, et al.
Published: (2026) -
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection
by: Antoun, Wissam, et al.
Published: (2024) -
Under-resourced studies of under-resourced languages: lemmatization and POS-tagging with LLM annotators for historical Armenian, Georgian, Greek and Syriac
by: Vidal-Gorène, Chahan, et al.
Published: (2026)