HALvest-Contrastive: Retrieval-Like Authorship Attribution with Patch-Level Late Interaction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kulumba, Francis, Antoun, Wissam, Vimont, Guillaume, Romary, Laurent, Cafiero, Florian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Where Does Authorship Signal Emerge in Encoder-Based Language Models?
von: Kulumba, Francis, et al.
Veröffentlicht: (2026)
von: Kulumba, Francis, et al.
Veröffentlicht: (2026)
Triggers Hijack Language Circuits: A Mechanistic Analysis of Backdoor Behaviors in Large Language Models
von: Lasnier, Théo, et al.
Veröffentlicht: (2026)
von: Lasnier, Théo, et al.
Veröffentlicht: (2026)
The \textit{Questio de aqua et terra}: A Computational Authorship Verification Study
von: Leocata, Martina, et al.
Veröffentlicht: (2025)
von: Leocata, Martina, et al.
Veröffentlicht: (2025)
Language-Switching Triggers Take a Latent Detour Through Language Models
von: Kulumba, Francis, et al.
Veröffentlicht: (2026)
von: Kulumba, Francis, et al.
Veröffentlicht: (2026)
Decoding AI and Human Authorship: Nuances Revealed Through NLP and Statistical Analysis
von: Akinwande, Mayowa, et al.
Veröffentlicht: (2024)
von: Akinwande, Mayowa, et al.
Veröffentlicht: (2024)
Preface to the Special Issue of the TAL Journal on Scholarly Document Processing
von: Boudin, Florian, et al.
Veröffentlicht: (2025)
von: Boudin, Florian, et al.
Veröffentlicht: (2025)
SciRAG: Adaptive, Citation-Aware, and Outline-Guided Retrieval and Synthesis for Scientific Literature
von: Ding, Hang, et al.
Veröffentlicht: (2025)
von: Ding, Hang, et al.
Veröffentlicht: (2025)
How to build an Open Science Monitor based on publications? A French perspective
von: Bracco, Laetitia, et al.
Veröffentlicht: (2025)
von: Bracco, Laetitia, et al.
Veröffentlicht: (2025)
The Attribution Crisis in LLM Search Results
von: Strauss, Ilan, et al.
Veröffentlicht: (2025)
von: Strauss, Ilan, et al.
Veröffentlicht: (2025)
ILiAD: An Interactive Corpus for Linguistic Annotated Data from Twitter Posts
von: Gonzalez, Simon
Veröffentlicht: (2024)
von: Gonzalez, Simon
Veröffentlicht: (2024)
Exploring the Technical Knowledge Interaction of Global Digital Humanities: Three-decade Evidence from Bibliometric-based perspectives
von: Li, Jiayi, et al.
Veröffentlicht: (2025)
von: Li, Jiayi, et al.
Veröffentlicht: (2025)
Do We Need Bigger Models for Science? Task-Aware Retrieval with Small Language Models
von: Kelber, Florian, et al.
Veröffentlicht: (2026)
von: Kelber, Florian, et al.
Veröffentlicht: (2026)
Towards Effective Authorship Attribution: Integrating Class-Incremental Learning
von: Rahgouy, Mostafa, et al.
Veröffentlicht: (2024)
von: Rahgouy, Mostafa, et al.
Veröffentlicht: (2024)
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection
von: Antoun, Wissam, et al.
Veröffentlicht: (2024)
von: Antoun, Wissam, et al.
Veröffentlicht: (2024)
An Evaluation of GPT-4V for Transcribing the Urban Renewal Hand-Written Collection
von: Lee, Myeong, et al.
Veröffentlicht: (2024)
von: Lee, Myeong, et al.
Veröffentlicht: (2024)
Improving Citation Text Generation: Overcoming Limitations in Length Control
von: Mandal, Biswadip, et al.
Veröffentlicht: (2024)
von: Mandal, Biswadip, et al.
Veröffentlicht: (2024)
Understanding Archives: Towards New Research Interfaces Relying on the Semantic Annotation of Documents
von: Gutehrlé, Nicolas, et al.
Veröffentlicht: (2024)
von: Gutehrlé, Nicolas, et al.
Veröffentlicht: (2024)
Mapping the Past: Geographically Linking an Early 20th Century Swedish Encyclopedia with Wikidata
von: Ahlin, Axel, et al.
Veröffentlicht: (2024)
von: Ahlin, Axel, et al.
Veröffentlicht: (2024)
Evaluating the Predictive Capacity of ChatGPT for Academic Peer Review Outcomes Across Multiple Platforms
von: Thelwall, Mike, et al.
Veröffentlicht: (2024)
von: Thelwall, Mike, et al.
Veröffentlicht: (2024)
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2024)
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2024)
Enriching Social Science Research via Survey Item Linking
von: Tsereteli, Tornike, et al.
Veröffentlicht: (2024)
von: Tsereteli, Tornike, et al.
Veröffentlicht: (2024)
Recent Developments in Deep Learning-based Author Name Disambiguation
von: Cappelli, Francesca, et al.
Veröffentlicht: (2024)
von: Cappelli, Francesca, et al.
Veröffentlicht: (2024)
Developing ChemDFM as a large language foundation model for chemistry
von: Zhao, Zihan, et al.
Veröffentlicht: (2024)
von: Zhao, Zihan, et al.
Veröffentlicht: (2024)
SciEvo: A 2 Million, 30-Year Cross-disciplinary Dataset for Temporal Scientometric Analysis
von: Jin, Yiqiao, et al.
Veröffentlicht: (2024)
von: Jin, Yiqiao, et al.
Veröffentlicht: (2024)
AutoLLM-CARD: Towards a Description and Landscape of Large Language Models
von: Tian, Shengwei, et al.
Veröffentlicht: (2024)
von: Tian, Shengwei, et al.
Veröffentlicht: (2024)
Citation Amnesia: On The Recency Bias of NLP and Other Academic Fields
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2024)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2024)
PST-Bench: Tracing and Benchmarking the Source of Publications
von: Zhang, Fanjin, et al.
Veröffentlicht: (2024)
von: Zhang, Fanjin, et al.
Veröffentlicht: (2024)
A Content-Based Novelty Measure for Scholarly Publications: A Proof of Concept
von: Wang, Haining
Veröffentlicht: (2024)
von: Wang, Haining
Veröffentlicht: (2024)
LyCon: Lyrics Reconstruction from the Bag-of-Words Using Large Language Models
von: Kim, Haven, et al.
Veröffentlicht: (2024)
von: Kim, Haven, et al.
Veröffentlicht: (2024)
CLOCR-C: Context Leveraging OCR Correction with Pre-trained Language Models
von: Bourne, Jonathan
Veröffentlicht: (2024)
von: Bourne, Jonathan
Veröffentlicht: (2024)
A Network Analysis Approach to Conlang Research Literature
von: Gonzalez, Simon
Veröffentlicht: (2024)
von: Gonzalez, Simon
Veröffentlicht: (2024)
Sentiment Analysis of Citations in Scientific Articles Using ChatGPT: Identifying Potential Biases and Conflicts of Interest
von: Hariri, Walid
Veröffentlicht: (2024)
von: Hariri, Walid
Veröffentlicht: (2024)
Quantifying the Impact of CU: A Systematic Literature Review
von: Compton, Thomas
Veröffentlicht: (2025)
von: Compton, Thomas
Veröffentlicht: (2025)
Institutional Books 1.0: A 242B token dataset from Harvard Library's collections, refined for accuracy and usability
von: Cargnelutti, Matteo, et al.
Veröffentlicht: (2025)
von: Cargnelutti, Matteo, et al.
Veröffentlicht: (2025)
CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era
von: Shi, Kaiwen, et al.
Veröffentlicht: (2026)
von: Shi, Kaiwen, et al.
Veröffentlicht: (2026)
"Don't Teach Minerva": Guiding LLMs Through Complex Syntax for Faithful Latin Translation with RAG
von: Aguilar, Sergio Torres
Veröffentlicht: (2025)
von: Aguilar, Sergio Torres
Veröffentlicht: (2025)
Vidya: An AI-Driven Modular Pipeline for Archival Automation and Semantic Metadata Enrichment
von: Filho, Cloter Migliorini, et al.
Veröffentlicht: (2026)
von: Filho, Cloter Migliorini, et al.
Veröffentlicht: (2026)
Unveiling Global Narratives: A Multilingual Twitter Dataset of News Media on the Russo-Ukrainian Conflict
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2023)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2023)
Hidden Entity Detection from GitHub Leveraging Large Language Models
von: Gan, Lu, et al.
Veröffentlicht: (2025)
von: Gan, Lu, et al.
Veröffentlicht: (2025)
A review of annotation classification tools in the educational domain
von: Gayoso-Cabada, Joaquín, et al.
Veröffentlicht: (2025)
von: Gayoso-Cabada, Joaquín, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Where Does Authorship Signal Emerge in Encoder-Based Language Models?
von: Kulumba, Francis, et al.
Veröffentlicht: (2026) -
Triggers Hijack Language Circuits: A Mechanistic Analysis of Backdoor Behaviors in Large Language Models
von: Lasnier, Théo, et al.
Veröffentlicht: (2026) -
The \textit{Questio de aqua et terra}: A Computational Authorship Verification Study
von: Leocata, Martina, et al.
Veröffentlicht: (2025) -
Language-Switching Triggers Take a Latent Detour Through Language Models
von: Kulumba, Francis, et al.
Veröffentlicht: (2026) -
Decoding AI and Human Authorship: Nuances Revealed Through NLP and Statistical Analysis
von: Akinwande, Mayowa, et al.
Veröffentlicht: (2024)