GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark
Fuente:
arXiv
Guardado en:
| Autores principales: | Kiefer, Lotta, Leiter, Christoph, Takeshita, Sotaro, Schmidt, Elena, Eger, Steffen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PrExMe! Large Scale Prompt Exploration of Open Source LLMs for Machine Translation and Summarization Evaluation
por: Leiter, Christoph, et al.
Publicado: (2024)
por: Leiter, Christoph, et al.
Publicado: (2024)
DeepSeek-R1 vs. o3-mini: How Well can Reasoning LLMs Evaluate MT and Summarization?
por: Larionov, Daniil, et al.
Publicado: (2025)
por: Larionov, Daniil, et al.
Publicado: (2025)
BMX: Boosting Natural Language Generation Metrics with Explainability
por: Leiter, Christoph, et al.
Publicado: (2022)
por: Leiter, Christoph, et al.
Publicado: (2022)
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
por: Leiter, Christoph, et al.
Publicado: (2025)
por: Leiter, Christoph, et al.
Publicado: (2025)
Towards Explainable Evaluation Metrics for Machine Translation
por: Leiter, Christoph, et al.
Publicado: (2023)
por: Leiter, Christoph, et al.
Publicado: (2023)
InstructAV: Instruction Fine-tuning Large Language Models for Authorship Verification
por: Hu, Yujia, et al.
Publicado: (2024)
por: Hu, Yujia, et al.
Publicado: (2024)
ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs
por: Wang, Zhipin, et al.
Publicado: (2026)
por: Wang, Zhipin, et al.
Publicado: (2026)
ACLSum: A New Dataset for Aspect-based Summarization of Scientific Publications
por: Takeshita, Sotaro, et al.
Publicado: (2024)
por: Takeshita, Sotaro, et al.
Publicado: (2024)
Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025
por: Kunilovskaya, Maria, et al.
Publicado: (2026)
por: Kunilovskaya, Maria, et al.
Publicado: (2026)
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
por: Abdalla, M. H. I., et al.
Publicado: (2025)
por: Abdalla, M. H. I., et al.
Publicado: (2025)
How Good Are LLMs for Literary Translation, Really? Literary Translation Evaluation with Humans and LLMs
por: Zhang, Ran, et al.
Publicado: (2024)
por: Zhang, Ran, et al.
Publicado: (2024)
To MRL or not to MRL: Text Embeddings are Robust to Truncation Without Matryoshka Learning, Except In Heavy Truncation Scenarios
por: Takeshita, Sotaro, et al.
Publicado: (2026)
por: Takeshita, Sotaro, et al.
Publicado: (2026)
Fine-Grained Detection of Solidarity for Women and Migrants in 155 Years of German Parliamentary Debates
por: Kostikova, Aida, et al.
Publicado: (2022)
por: Kostikova, Aida, et al.
Publicado: (2022)
ROUGE-K: Do Your Summaries Have Keywords?
por: Takeshita, Sotaro, et al.
Publicado: (2024)
por: Takeshita, Sotaro, et al.
Publicado: (2024)
NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers?
por: Leiter, Christoph, et al.
Publicado: (2024)
por: Leiter, Christoph, et al.
Publicado: (2024)
Do Emotions Really Affect Argument Convincingness? A Dynamic Approach with LLM-based Manipulation Checks
por: Chen, Yanran, et al.
Publicado: (2025)
por: Chen, Yanran, et al.
Publicado: (2025)
ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models
por: Belouadi, Jonas, et al.
Publicado: (2022)
por: Belouadi, Jonas, et al.
Publicado: (2022)
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics
por: Larionov, Daniil, et al.
Publicado: (2024)
por: Larionov, Daniil, et al.
Publicado: (2024)
USCORE: An Effective Approach to Fully Unsupervised Evaluation Metrics for Machine Translation
por: Belouadi, Jonas, et al.
Publicado: (2022)
por: Belouadi, Jonas, et al.
Publicado: (2022)
LLM-based multi-agent poetry generation in non-cooperative environments
por: Zhang, Ran, et al.
Publicado: (2024)
por: Zhang, Ran, et al.
Publicado: (2024)
BatchGEMBA: Token-Efficient Machine Translation Evaluation with Batched Prompting and Prompt Compression
por: Larionov, Daniil, et al.
Publicado: (2025)
por: Larionov, Daniil, et al.
Publicado: (2025)
Sui Generis: Large Language Models for Authorship Attribution and Verification in Latin
por: Schmidt, Gleb, et al.
Publicado: (2024)
por: Schmidt, Gleb, et al.
Publicado: (2024)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
por: Gekhman, Zorik, et al.
Publicado: (2024)
por: Gekhman, Zorik, et al.
Publicado: (2024)
Is there really a Citation Age Bias in NLP?
por: Nguyen, Hoa, et al.
Publicado: (2024)
por: Nguyen, Hoa, et al.
Publicado: (2024)
SpeakGer: A meta-data enriched speech corpus of German state and federal parliaments
por: Lange, Kai-Robin, et al.
Publicado: (2024)
por: Lange, Kai-Robin, et al.
Publicado: (2024)
Syntactic Language Change in English and German: Metrics, Parsers, and Convergences
por: Chen, Yanran, et al.
Publicado: (2024)
por: Chen, Yanran, et al.
Publicado: (2024)
Authorship Impersonation via LLM Prompting does not Evade Authorship Verification Methods
por: Zeng, Baoyi, et al.
Publicado: (2026)
por: Zeng, Baoyi, et al.
Publicado: (2026)
TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
por: Greisinger, Christian, et al.
Publicado: (2026)
por: Greisinger, Christian, et al.
Publicado: (2026)
Distinguishing Fictional Voices: a Study of Authorship Verification Models for Quotation Attribution
por: Michel, Gaspard, et al.
Publicado: (2024)
por: Michel, Gaspard, et al.
Publicado: (2024)
CAVE: Controllable Authorship Verification Explanations
por: Ramnath, Sahana, et al.
Publicado: (2024)
por: Ramnath, Sahana, et al.
Publicado: (2024)
Residualized Similarity for Faithfully Explainable Authorship Verification
por: Zeng, Peter, et al.
Publicado: (2025)
por: Zeng, Peter, et al.
Publicado: (2025)
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation
por: Zhang, Ran, et al.
Publicado: (2023)
por: Zhang, Ran, et al.
Publicado: (2023)
LLM Analysis of 150+ years of German Parliamentary Debates on Migration Reveals Shift from Post-War Solidarity to Anti-Solidarity in the Last Decade
por: Kostikova, Aida, et al.
Publicado: (2025)
por: Kostikova, Aida, et al.
Publicado: (2025)
ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation?
por: Zhang, Leixin, et al.
Publicado: (2024)
por: Zhang, Leixin, et al.
Publicado: (2024)
BARD10: A New Benchmark Reveals Significance of Bangla Stop-Words in Authorship Attribution
por: Moosa, Abdullah Muhammad, et al.
Publicado: (2025)
por: Moosa, Abdullah Muhammad, et al.
Publicado: (2025)
Towards Pedagogical LLMs with Supervised Fine Tuning for Computing Education
por: Vassar, Alexandra, et al.
Publicado: (2024)
por: Vassar, Alexandra, et al.
Publicado: (2024)
Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2024)
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2024)
A Comparative Study of LLM Prompting and Fine-Tuning for Cross-genre Authorship Attribution on Chinese Lyrics
por: Li, Yuxin, et al.
Publicado: (2025)
por: Li, Yuxin, et al.
Publicado: (2025)
Addressing Topic Leakage in Cross-Topic Evaluation for Authorship Verification
por: Sawatphol, Jitkapat, et al.
Publicado: (2024)
por: Sawatphol, Jitkapat, et al.
Publicado: (2024)
Masks and Mimicry: Strategic Obfuscation and Impersonation Attacks on Authorship Verification
por: Alperin, Kenneth, et al.
Publicado: (2025)
por: Alperin, Kenneth, et al.
Publicado: (2025)
Ejemplares similares
-
PrExMe! Large Scale Prompt Exploration of Open Source LLMs for Machine Translation and Summarization Evaluation
por: Leiter, Christoph, et al.
Publicado: (2024) -
DeepSeek-R1 vs. o3-mini: How Well can Reasoning LLMs Evaluate MT and Summarization?
por: Larionov, Daniil, et al.
Publicado: (2025) -
BMX: Boosting Natural Language Generation Metrics with Explainability
por: Leiter, Christoph, et al.
Publicado: (2022) -
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
por: Leiter, Christoph, et al.
Publicado: (2025) -
Towards Explainable Evaluation Metrics for Machine Translation
por: Leiter, Christoph, et al.
Publicado: (2023)