Stylometry recognizes human and LLM-generated texts in short samples
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Przystalski, Karol, Argasiński, Jan K., Grabska-Gradzińska, Iwona, Ochab, Jeremi K. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
StylOch at PAN: Gradient-Boosted Trees with Frequency-Based Stylometric Features
par: Ochab, Jeremi K., et autres
Publié: (2025)
par: Ochab, Jeremi K., et autres
Publié: (2025)
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters
par: Haim, Edith, et autres
Publié: (2024)
par: Haim, Edith, et autres
Publié: (2024)
CreditAudit: 2$^\text{nd}$ Dimension for LLM Evaluation and Selection
par: Song, Yiliang, et autres
Publié: (2026)
par: Song, Yiliang, et autres
Publié: (2026)
$\text{M}^{2}$LLM: Multi-view Molecular Representation Learning with Large Language Models
par: Ju, Jiaxin, et autres
Publié: (2025)
par: Ju, Jiaxin, et autres
Publié: (2025)
FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts
par: Shin, Hagyeong, et autres
Publié: (2025)
par: Shin, Hagyeong, et autres
Publié: (2025)
LLM generation novelty through the lens of semantic similarity
par: Davydov, Philipp, et autres
Publié: (2025)
par: Davydov, Philipp, et autres
Publié: (2025)
Set-LLM: A Permutation-Invariant LLM
par: Egressy, Beni, et autres
Publié: (2025)
par: Egressy, Beni, et autres
Publié: (2025)
Reveal and Release: Iterative LLM Unlearning with Self-generated Data
par: Xie, Linxi, et autres
Publié: (2025)
par: Xie, Linxi, et autres
Publié: (2025)
Small sample-based adaptive text classification through iterative and contrastive description refinement
par: Rajeev, Amrit, et autres
Publié: (2025)
par: Rajeev, Amrit, et autres
Publié: (2025)
Efficient Detection of LLM-generated Texts with a Bayesian Surrogate Model
par: Miao, Yibo, et autres
Publié: (2023)
par: Miao, Yibo, et autres
Publié: (2023)
Aligned at the Start: Conceptual Groupings in LLM Embeddings
par: Khatir, Mehrdad, et autres
Publié: (2024)
par: Khatir, Mehrdad, et autres
Publié: (2024)
Advancing Chinese biomedical text mining with community challenges
par: Zong, Hui, et autres
Publié: (2024)
par: Zong, Hui, et autres
Publié: (2024)
Clinical ModernBERT: An efficient and long context encoder for biomedical text
par: Lee, Simon A., et autres
Publié: (2025)
par: Lee, Simon A., et autres
Publié: (2025)
Language models show human-like content effects on reasoning tasks
par: Dasgupta, Ishita, et autres
Publié: (2022)
par: Dasgupta, Ishita, et autres
Publié: (2022)
LLMs can hide text in other text of the same length
par: Norelli, Antonio, et autres
Publié: (2025)
par: Norelli, Antonio, et autres
Publié: (2025)
Will we run out of data? Limits of LLM scaling based on human-generated data
par: Villalobos, Pablo, et autres
Publié: (2022)
par: Villalobos, Pablo, et autres
Publié: (2022)
Explained anomaly detection in text reviews: Can subjective scenarios be correctly evaluated?
par: Novoa-Paradela, David, et autres
Publié: (2023)
par: Novoa-Paradela, David, et autres
Publié: (2023)
Assessing Deanonymization Risks with Stylometry-Assisted LLM Agent
par: Zhang, Boyang, et autres
Publié: (2026)
par: Zhang, Boyang, et autres
Publié: (2026)
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution
par: Bisztray, Tamas, et autres
Publié: (2025)
par: Bisztray, Tamas, et autres
Publié: (2025)
Communication Compression for Tensor Parallel LLM Inference
par: Hansen-Palmus, Jan, et autres
Publié: (2024)
par: Hansen-Palmus, Jan, et autres
Publié: (2024)
Effects of term weighting approach with and without stop words removing on Arabic text classification
par: Alhenawi, Esra'a, et autres
Publié: (2024)
par: Alhenawi, Esra'a, et autres
Publié: (2024)
TrICy: Trigger-guided Data-to-text Generation with Intent aware Attention-Copy
par: Agarwal, Vibhav, et autres
Publié: (2024)
par: Agarwal, Vibhav, et autres
Publié: (2024)
ToxiGAN: Toxic Data Augmentation via LLM-Guided Directional Adversarial Generation
par: Li, Peiran, et autres
Publié: (2026)
par: Li, Peiran, et autres
Publié: (2026)
TTKV: Temporal-Tiered KV Cache for Long-Context LLM Inference
par: Dzikanyanga, Gradwell, et autres
Publié: (2026)
par: Dzikanyanga, Gradwell, et autres
Publié: (2026)
Aleph-Alpha-GermanWeb: Improving German-language LLM pre-training with model-based data curation and synthetic data generation
par: Burns, Thomas F, et autres
Publié: (2025)
par: Burns, Thomas F, et autres
Publié: (2025)
Stylometry Analysis of Human and Machine Text for Academic Integrity
par: Albaqami, Hezam, et autres
Publié: (2026)
par: Albaqami, Hezam, et autres
Publié: (2026)
Auto-Cypher: Improving LLMs on Cypher generation via LLM-supervised generation-verification framework
par: Tiwari, Aman, et autres
Publié: (2024)
par: Tiwari, Aman, et autres
Publié: (2024)
Probing the contents of semantic representations from text, behavior, and brain data using the psychNorms metabase
par: Hussain, Zak, et autres
Publié: (2024)
par: Hussain, Zak, et autres
Publié: (2024)
BP-Seg: A graphical model approach to unsupervised and non-contiguous text segmentation using belief propagation
par: Li, Fengyi, et autres
Publié: (2025)
par: Li, Fengyi, et autres
Publié: (2025)
LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
par: Shojaee, Parshin, et autres
Publié: (2025)
par: Shojaee, Parshin, et autres
Publié: (2025)
SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning
par: Chen, Jiaqi, et autres
Publié: (2025)
par: Chen, Jiaqi, et autres
Publié: (2025)
Benchmark of stylistic variation in LLM-generated texts
par: Milička, Jiří, et autres
Publié: (2025)
par: Milička, Jiří, et autres
Publié: (2025)
New Encoders for German Trained from Scratch: Comparing ModernGBERT with Converted LLM2Vec Models
par: Wunderle, Julia, et autres
Publié: (2025)
par: Wunderle, Julia, et autres
Publié: (2025)
MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models
par: Fang, Gongfan, et autres
Publié: (2024)
par: Fang, Gongfan, et autres
Publié: (2024)
H-STAR: LLM-driven Hybrid SQL-Text Adaptive Reasoning on Tables
par: Abhyankar, Nikhil, et autres
Publié: (2024)
par: Abhyankar, Nikhil, et autres
Publié: (2024)
Lightweight reranking for language model generations
par: Jain, Siddhartha, et autres
Publié: (2023)
par: Jain, Siddhartha, et autres
Publié: (2023)
An energy-based comparative analysis of common approaches to text classification in the Legal domain
par: Gultekin, Sinan, et autres
Publié: (2023)
par: Gultekin, Sinan, et autres
Publié: (2023)
Don't throw away your value model! Generating more preferable text with Value-Guided Monte-Carlo Tree Search decoding
par: Liu, Jiacheng, et autres
Publié: (2023)
par: Liu, Jiacheng, et autres
Publié: (2023)
Post-training makes large language models less human-like
par: Binz, Marcel, et autres
Publié: (2026)
par: Binz, Marcel, et autres
Publié: (2026)
Extraction of Research Objectives, Machine Learning Model Names, and Dataset Names from Academic Papers and Analysis of Their Interrelationships Using LLM and Network Analysis
par: Nishio, S., et autres
Publié: (2024)
par: Nishio, S., et autres
Publié: (2024)
Documents similaires
-
StylOch at PAN: Gradient-Boosted Trees with Frequency-Based Stylometric Features
par: Ochab, Jeremi K., et autres
Publié: (2025) -
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters
par: Haim, Edith, et autres
Publié: (2024) -
CreditAudit: 2$^\text{nd}$ Dimension for LLM Evaluation and Selection
par: Song, Yiliang, et autres
Publié: (2026) -
$\text{M}^{2}$LLM: Multi-view Molecular Representation Learning with Large Language Models
par: Ju, Jiaxin, et autres
Publié: (2025) -
FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts
par: Shin, Hagyeong, et autres
Publié: (2025)