Isolating authorship from content with semantic embeddings and contrastive learning
Fuente:
arXiv
Saved in:
| Main Authors: | Huertas-Tato, Javier, Girón-Jiménez, Adrián, Martín, Alejandro, Camacho, David |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PART: Pre-trained Authorship Representation Transformer
by: Huertas-Tato, Javier, et al.
Published: (2022)
by: Huertas-Tato, Javier, et al.
Published: (2022)
Camouflage is all you need: Evaluating and Enhancing Language Model Robustness Against Camouflage Adversarial Attacks
by: Huertas-García, Álvaro, et al.
Published: (2024)
by: Huertas-García, Álvaro, et al.
Published: (2024)
xList-Hate: A Checklist-Based Framework for Interpretable and Generalizable Hate Speech Detection
by: Girón, Adrián, et al.
Published: (2026)
by: Girón, Adrián, et al.
Published: (2026)
On the locality bias and results in the Long Range Arena
by: Miralles-González, Pablo, et al.
Published: (2025)
by: Miralles-González, Pablo, et al.
Published: (2025)
Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection
by: Miralles-González, Pablo, et al.
Published: (2025)
by: Miralles-González, Pablo, et al.
Published: (2025)
LLM one-shot style transfer for Authorship Attribution and Verification
by: Miralles-González, Pablo, et al.
Published: (2025)
by: Miralles-González, Pablo, et al.
Published: (2025)
Pushing the boundary on Natural Language Inference
by: Miralles-González, Pablo, et al.
Published: (2025)
by: Miralles-González, Pablo, et al.
Published: (2025)
Machine learning methods fail to provide cohesive atheoretical construction of personality traits from semantic embeddings
by: Bouguettaya, Ayoub, et al.
Published: (2025)
by: Bouguettaya, Ayoub, et al.
Published: (2025)
Improving embedding with contrastive fine-tuning on small datasets with expert-augmented scores
by: Lu, Jun, et al.
Published: (2024)
by: Lu, Jun, et al.
Published: (2024)
Critical biblical studies via word frequency analysis: unveiling text authorship
by: Faigenbaum-Golovin, Shira, et al.
Published: (2024)
by: Faigenbaum-Golovin, Shira, et al.
Published: (2024)
Dissecting embedding method: learning higher-order structures from data
by: Tupikina, Liubov, et al.
Published: (2024)
by: Tupikina, Liubov, et al.
Published: (2024)
Probing the contents of semantic representations from text, behavior, and brain data using the psychNorms metabase
by: Hussain, Zak, et al.
Published: (2024)
by: Hussain, Zak, et al.
Published: (2024)
Optimal synthesis embeddings
by: Santana, Roberto, et al.
Published: (2024)
by: Santana, Roberto, et al.
Published: (2024)
Attention-aware semantic relevance predicting Chinese sentence reading
by: Sun, Kun
Published: (2024)
by: Sun, Kun
Published: (2024)
What to align in multimodal contrastive learning?
by: Dufumier, Benoit, et al.
Published: (2024)
by: Dufumier, Benoit, et al.
Published: (2024)
Semantic similarity prediction is better than other semantic similarity measures
by: Herbold, Steffen
Published: (2023)
by: Herbold, Steffen
Published: (2023)
Topological quantification of ambiguity in semantic search
by: Barillot, Thomas Roland, et al.
Published: (2024)
by: Barillot, Thomas Roland, et al.
Published: (2024)
A quantitative analysis of semantic information in deep representations of text and images
by: Acevedo, Santiago, et al.
Published: (2025)
by: Acevedo, Santiago, et al.
Published: (2025)
A comparison of latent semantic analysis and correspondence analysis of document-term matrices
by: Qi, Qianqian, et al.
Published: (2021)
by: Qi, Qianqian, et al.
Published: (2021)
Modeling the language cortex with form-independent and enriched representations of sentence meaning reveals remarkable semantic abstractness
by: Saha, Shreya, et al.
Published: (2025)
by: Saha, Shreya, et al.
Published: (2025)
LLM generation novelty through the lens of semantic similarity
by: Davydov, Philipp, et al.
Published: (2025)
by: Davydov, Philipp, et al.
Published: (2025)
Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers
by: Saxena, Vageesh Kumar
Published: (2026)
by: Saxena, Vageesh Kumar
Published: (2026)
Few-shot learning for automated content analysis: Efficient coding of arguments and claims in the debate on arms deliveries to Ukraine
by: Rieger, Jonas, et al.
Published: (2023)
by: Rieger, Jonas, et al.
Published: (2023)
Learning without training: The implicit dynamics of in-context learning
by: Dherin, Benoit, et al.
Published: (2025)
by: Dherin, Benoit, et al.
Published: (2025)
FLARE: Task-agnostic embedding model evaluation through a normalization process
by: Jiang, Jingzhou, et al.
Published: (2026)
by: Jiang, Jingzhou, et al.
Published: (2026)
Cropping outperforms dropout as an augmentation strategy for self-supervised training of text embeddings
by: González-Márquez, Rita, et al.
Published: (2025)
by: González-Márquez, Rita, et al.
Published: (2025)
Enhancing ASD detection accuracy: a combined approach of machine learning and deep learning models with natural language processing
by: Rubio-Martín, Sergio, et al.
Published: (2024)
by: Rubio-Martín, Sergio, et al.
Published: (2024)
Latent learning: episodic memory complements parametric learning by enabling flexible reuse of experiences
by: Lampinen, Andrew Kyle, et al.
Published: (2025)
by: Lampinen, Andrew Kyle, et al.
Published: (2025)
Label-semantics Aware Generative Approach for Domain-Agnostic Multilabel Classification
by: Khatuya, Subhendu, et al.
Published: (2025)
by: Khatuya, Subhendu, et al.
Published: (2025)
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
by: Park, Ji Won, et al.
Published: (2025)
by: Park, Ji Won, et al.
Published: (2025)
Classification EM-PCA for clustering and embedding
by: Tighidet, Zineddine, et al.
Published: (2025)
by: Tighidet, Zineddine, et al.
Published: (2025)
A gentle push funziona benissimo: making instructed models in Italian via contrastive activation steering
by: Scalena, Daniel, et al.
Published: (2024)
by: Scalena, Daniel, et al.
Published: (2024)
A Systematic Analysis on the Temporal Generalization of Language Models in Social Media
by: Ushio, Asahi, et al.
Published: (2024)
by: Ushio, Asahi, et al.
Published: (2024)
Parameter Importance is Not Static: Evolving Parameter Isolation for Supervised Fine-Tuning
by: Lin, Zekai, et al.
Published: (2026)
by: Lin, Zekai, et al.
Published: (2026)
Decoding Latent Spaces: Assessing the Interpretability of Time Series Foundation Models for Visual Analytics
by: Santamaria-Valenzuela, Inmaculada, et al.
Published: (2025)
by: Santamaria-Valenzuela, Inmaculada, et al.
Published: (2025)
A comparative analysis of embedding models for patent similarity
by: Ascione, Grazia Sveva, et al.
Published: (2024)
by: Ascione, Grazia Sveva, et al.
Published: (2024)
A RelEntLess Benchmark for Modelling Graded Relations between Named Entities
by: Ushio, Asahi, et al.
Published: (2023)
by: Ushio, Asahi, et al.
Published: (2023)
From communities to interpretable network and word embedding: an unified approach
by: Prouteau, Thibault, et al.
Published: (2024)
by: Prouteau, Thibault, et al.
Published: (2024)
Open-ended VQA benchmarking of Vision-Language models by exploiting Classification datasets and their semantic hierarchy
by: Ging, Simon, et al.
Published: (2024)
by: Ging, Simon, et al.
Published: (2024)
Indication Finding: a novel use case for representation learning
by: Eckhoff, Maren, et al.
Published: (2024)
by: Eckhoff, Maren, et al.
Published: (2024)
Similar Items
-
PART: Pre-trained Authorship Representation Transformer
by: Huertas-Tato, Javier, et al.
Published: (2022) -
Camouflage is all you need: Evaluating and Enhancing Language Model Robustness Against Camouflage Adversarial Attacks
by: Huertas-García, Álvaro, et al.
Published: (2024) -
xList-Hate: A Checklist-Based Framework for Interpretable and Generalizable Hate Speech Detection
by: Girón, Adrián, et al.
Published: (2026) -
On the locality bias and results in the Long Range Arena
by: Miralles-González, Pablo, et al.
Published: (2025) -
Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection
by: Miralles-González, Pablo, et al.
Published: (2025)