Guardado en:
| Autor principal: | Orekhov, Boris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2407.08099 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Does Burrows' Delta really confirm that Rowling and Galbraith are the same author?
por: Orekhov, Boris
Publicado: (2024)
por: Orekhov, Boris
Publicado: (2024)
You shall know a piece by the company it keeps. Chess plays as a data for word2vec models
por: Orekhov, Boris
Publicado: (2024)
por: Orekhov, Boris
Publicado: (2024)
Is text normalization relevant for classifying medieval charters?
por: Atzenhofer-Baumgartner, Florian, et al.
Publicado: (2024)
por: Atzenhofer-Baumgartner, Florian, et al.
Publicado: (2024)
Metronome: tracing variation in poetic meters via local sequence alignment
por: Nagy, Ben, et al.
Publicado: (2024)
por: Nagy, Ben, et al.
Publicado: (2024)
Markov reads Pushkin, again: A statistical journey into the poetic world of Evgenij Onegin
por: Sabatini, Angelo Maria
Publicado: (2026)
por: Sabatini, Angelo Maria
Publicado: (2026)
Why mask diffusion does not work
por: Sun, Haocheng, et al.
Publicado: (2025)
por: Sun, Haocheng, et al.
Publicado: (2025)
How do we measure privacy in text? A survey of text anonymization metrics
por: Ren, Yaxuan, et al.
Publicado: (2025)
por: Ren, Yaxuan, et al.
Publicado: (2025)
Advancing Chinese biomedical text mining with community challenges
por: Zong, Hui, et al.
Publicado: (2024)
por: Zong, Hui, et al.
Publicado: (2024)
How and where does CLIP process negation?
por: Quantmeyer, Vincent, et al.
Publicado: (2024)
por: Quantmeyer, Vincent, et al.
Publicado: (2024)
Translating scientific Latin texts with artificial intelligence: the works of Euler and contemporaries
por: Bistafa, Sylvio R.
Publicado: (2023)
por: Bistafa, Sylvio R.
Publicado: (2023)
How does a Language-Specific Tokenizer affect LLMs?
por: Seo, Jean, et al.
Publicado: (2025)
por: Seo, Jean, et al.
Publicado: (2025)
Thematic Analysis with Large Language Models: does it work with languages other than English? A targeted test in Italian
por: De Paoli, Stefano
Publicado: (2024)
por: De Paoli, Stefano
Publicado: (2024)
Full-text Error Correction for Chinese Speech Recognition with Large Language Model
por: Tang, Zhiyuan, et al.
Publicado: (2024)
por: Tang, Zhiyuan, et al.
Publicado: (2024)
How Sampling Affects the Detectability of Machine-written texts: A Comprehensive Study
por: Dubois, Matthieu, et al.
Publicado: (2025)
por: Dubois, Matthieu, et al.
Publicado: (2025)
Synthetically generated text for supervised text analysis
por: Halterman, Andrew
Publicado: (2023)
por: Halterman, Andrew
Publicado: (2023)
How does a Multilingual LM Handle Multiple Languages?
por: Kakarla, Santhosh, et al.
Publicado: (2025)
por: Kakarla, Santhosh, et al.
Publicado: (2025)
Democratizing the medieval English legal tradition
por: Zhang, Michael, et al.
Publicado: (2026)
por: Zhang, Michael, et al.
Publicado: (2026)
How Chinese are Chinese Language Models? The Puzzling Lack of Language Policy in China's LLMs
por: Wen-Yi, Andrea W, et al.
Publicado: (2024)
por: Wen-Yi, Andrea W, et al.
Publicado: (2024)
Domain Regeneration: How well do LLMs match syntactic properties of text domains?
por: Ju, Da, et al.
Publicado: (2025)
por: Ju, Da, et al.
Publicado: (2025)
Diagnosing our datasets: How does my language model learn clinical information?
por: Jia, Furong, et al.
Publicado: (2025)
por: Jia, Furong, et al.
Publicado: (2025)
How reparametrization trick broke differentially-private text representation learning
por: Habernal, Ivan
Publicado: (2022)
por: Habernal, Ivan
Publicado: (2022)
How does Misinformation Affect Large Language Model Behaviors and Preferences?
por: Peng, Miao, et al.
Publicado: (2025)
por: Peng, Miao, et al.
Publicado: (2025)
How does fine-tuning improve sensorimotor representations in large language models?
por: Wu, Minghua, et al.
Publicado: (2026)
por: Wu, Minghua, et al.
Publicado: (2026)
A Chat About Boring Problems: Studying GPT-based text normalization
por: Zhang, Yang, et al.
Publicado: (2023)
por: Zhang, Yang, et al.
Publicado: (2023)
¡¿Qué, qué?!Transculturación and Tato Laviera's Spanglish poetics
por: Stephanie Álvarez Martínez
Publicado: (2006)
por: Stephanie Álvarez Martínez
Publicado: (2006)
How Much Do LLMs Know About Chinese Zero Pronouns?
por: Li, Yifei, et al.
Publicado: (2026)
por: Li, Yifei, et al.
Publicado: (2026)
How does Multi-Task Training Affect Transformer In-Context Capabilities? Investigations with Function Classes
por: Bhasin, Harmon, et al.
Publicado: (2024)
por: Bhasin, Harmon, et al.
Publicado: (2024)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
por: Chen, Xi, et al.
Publicado: (2025)
por: Chen, Xi, et al.
Publicado: (2025)
Can reasoning models comprehend mathematical problems in Chinese ancient texts? An empirical study based on data from Suanjing Shishu
por: Liu, Chang, et al.
Publicado: (2025)
por: Liu, Chang, et al.
Publicado: (2025)
История стиховедения и формализм
por: Orekhov, Boris
Publicado: (2024)
por: Orekhov, Boris
Publicado: (2024)
Identifying social isolation themes in NVDRS text narratives using topic modeling and text-classification methods
por: Walker, Drew, et al.
Publicado: (2025)
por: Walker, Drew, et al.
Publicado: (2025)
Identifying attributions of causality in political text
por: Garcia-Corral, Paulina
Publicado: (2025)
por: Garcia-Corral, Paulina
Publicado: (2025)
Qwen it detect machine-generated text?
por: Marchitan, Teodor-George, et al.
Publicado: (2025)
por: Marchitan, Teodor-George, et al.
Publicado: (2025)
What does it mean to understand language?
por: Casto, Colton, et al.
Publicado: (2025)
por: Casto, Colton, et al.
Publicado: (2025)
LUQ: Long-text Uncertainty Quantification for LLMs
por: Zhang, Caiqi, et al.
Publicado: (2024)
por: Zhang, Caiqi, et al.
Publicado: (2024)
Leveraging the power of transformers for guilt detection in text
por: Meque, Abdul Gafar Manuel, et al.
Publicado: (2024)
por: Meque, Abdul Gafar Manuel, et al.
Publicado: (2024)
Few-shot text-based emotion detection
por: Marchitan, Teodor-George, et al.
Publicado: (2025)
por: Marchitan, Teodor-George, et al.
Publicado: (2025)
Transferable text data distillation by trajectory matching
por: Yao, Rong, et al.
Publicado: (2025)
por: Yao, Rong, et al.
Publicado: (2025)
Where does an LLM begin computing an instruction?
por: Pola, Aditya, et al.
Publicado: (2025)
por: Pola, Aditya, et al.
Publicado: (2025)
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts
por: Gokceoglu, Gokcen, et al.
Publicado: (2024)
por: Gokceoglu, Gokcen, et al.
Publicado: (2024)
Ejemplares similares
-
Does Burrows' Delta really confirm that Rowling and Galbraith are the same author?
por: Orekhov, Boris
Publicado: (2024) -
You shall know a piece by the company it keeps. Chess plays as a data for word2vec models
por: Orekhov, Boris
Publicado: (2024) -
Is text normalization relevant for classifying medieval charters?
por: Atzenhofer-Baumgartner, Florian, et al.
Publicado: (2024) -
Metronome: tracing variation in poetic meters via local sequence alignment
por: Nagy, Ben, et al.
Publicado: (2024) -
Markov reads Pushkin, again: A statistical journey into the poetic world of Evgenij Onegin
por: Sabatini, Angelo Maria
Publicado: (2026)