Gespeichert in:
| Hauptverfasser: | Krylov, Aleksei S., Somov, Oleg D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2409.13739 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Confidence Estimation for Error Detection in Text-to-SQL Systems
von: Somov, Oleg, et al.
Veröffentlicht: (2025)
von: Somov, Oleg, et al.
Veröffentlicht: (2025)
Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers
von: Gong, Linyuan, et al.
Veröffentlicht: (2023)
von: Gong, Linyuan, et al.
Veröffentlicht: (2023)
The benefits of query-based KGQA systems for complex and temporal questions in LLM era
von: Alekseev, Artem, et al.
Veröffentlicht: (2025)
von: Alekseev, Artem, et al.
Veröffentlicht: (2025)
When Punctuation Matters: A Large-Scale Comparison of Prompt Robustness Methods for LLMs
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2025)
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2025)
Evolutionary Search for Automated Design of Uncertainty Quantification Methods
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2026)
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2026)
On The Role of Pretrained Language Models in General-Purpose Text Embeddings: A Survey
von: Zhang, Meishan, et al.
Veröffentlicht: (2025)
von: Zhang, Meishan, et al.
Veröffentlicht: (2025)
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
von: Sarti, Gabriele, et al.
Veröffentlicht: (2022)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2022)
Private Synthetic Text Generation with Diffusion Models
von: Ochs, Sebastian, et al.
Veröffentlicht: (2024)
von: Ochs, Sebastian, et al.
Veröffentlicht: (2024)
Rethinking the Role of Text Complexity in Language Model Pretraining
von: Velasco, Dan John, et al.
Veröffentlicht: (2025)
von: Velasco, Dan John, et al.
Veröffentlicht: (2025)
HAMSA: Hijacking Aligned Compact Models via Stealthy Automation
von: Krylov, Alexey, et al.
Veröffentlicht: (2025)
von: Krylov, Alexey, et al.
Veröffentlicht: (2025)
HRM-Text: Efficient Pretraining Beyond Scaling
von: Wang, Guan, et al.
Veröffentlicht: (2026)
von: Wang, Guan, et al.
Veröffentlicht: (2026)
Challenges in Explaining Pretrained Clinical Text Classifiers
von: Miok, Kristian, et al.
Veröffentlicht: (2026)
von: Miok, Kristian, et al.
Veröffentlicht: (2026)
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
MedSyn: LLM-based Synthetic Medical Text Generation Framework
von: Kumichev, Gleb, et al.
Veröffentlicht: (2024)
von: Kumichev, Gleb, et al.
Veröffentlicht: (2024)
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
von: Ginn, Michael, et al.
Veröffentlicht: (2024)
von: Ginn, Michael, et al.
Veröffentlicht: (2024)
Energy-Based Diffusion Language Models for Text Generation
von: Xu, Minkai, et al.
Veröffentlicht: (2024)
von: Xu, Minkai, et al.
Veröffentlicht: (2024)
Differences in Text Generated by Diffusion and Autoregressive Language Models
von: Zhang, Zeyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyang, et al.
Veröffentlicht: (2026)
A Reparameterized Discrete Diffusion Model for Text Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
Prune or Retrain: Optimizing the Vocabulary of Multilingual Models for Estonian
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2025)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2025)
Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)
Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation
von: Zhang, Yuanhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yuanhe, et al.
Veröffentlicht: (2025)
Harnessing the Intrinsic Knowledge of Pretrained Language Models for Challenging Text Classification Settings
von: Gao, Lingyu
Veröffentlicht: (2024)
von: Gao, Lingyu
Veröffentlicht: (2024)
Transfer Learning for Text Diffusion Models
von: Han, Kehang, et al.
Veröffentlicht: (2024)
von: Han, Kehang, et al.
Veröffentlicht: (2024)
Sõnajaht: Definition Embeddings and Semantic Search for Reverse Dictionary Creation
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
TartuNLP @ SIGTYP 2024 Shared Task: Adapting XLM-RoBERTa for Ancient and Historical Languages
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
Cross-lingual paraphrase identification
von: Fedorova, Inessa, et al.
Veröffentlicht: (2024)
von: Fedorova, Inessa, et al.
Veröffentlicht: (2024)
TartuNLP @ AXOLOTL-24: Leveraging Classifier Output for New Sense Detection in Lexical Semantics
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
Comparison of Current Approaches to Lemmatization: A Case Study in Estonian
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
TartuNLP at EvaLatin 2024: Emotion Polarity Detection
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
GliLem: Leveraging GliNER for Contextualized Lemmatization in Estonian
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2025)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2025)
Generalizable and Stable Finetuning of Pretrained Language Models on Low-Resource Texts
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
von: Somayajula, Sai Ashish, et al.
Veröffentlicht: (2024)
Comparison of End-to-end Speech Assessment Models for the NOCASA 2025 Challenge
von: Žavoronkov, Aleksei, et al.
Veröffentlicht: (2025)
von: Žavoronkov, Aleksei, et al.
Veröffentlicht: (2025)
Smoothie: Smoothing Diffusion on Token Embeddings for Text Generation
von: Shabalin, Alexander, et al.
Veröffentlicht: (2025)
von: Shabalin, Alexander, et al.
Veröffentlicht: (2025)
Empowering Diffusion Models on the Embedding Space for Text Generation
von: Gao, Zhujin, et al.
Veröffentlicht: (2022)
von: Gao, Zhujin, et al.
Veröffentlicht: (2022)
Diffusion-Pretrained Dense and Contextual Embeddings
von: Eslami, Sedigheh, et al.
Veröffentlicht: (2026)
von: Eslami, Sedigheh, et al.
Veröffentlicht: (2026)
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
von: Brannon, William, et al.
Veröffentlicht: (2023)
von: Brannon, William, et al.
Veröffentlicht: (2023)
EstLLM: Enhancing Estonian Capabilities in Multilingual LLMs via Continued Pretraining and Post-Training
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2026)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2026)
NoteContrast: Contrastive Language-Diagnostic Pretraining for Medical Text
von: Kailas, Prajwal, et al.
Veröffentlicht: (2024)
von: Kailas, Prajwal, et al.
Veröffentlicht: (2024)
TextOmics-Guided Diffusion for Hit-like Molecular Generation
von: Yuan, Hang, et al.
Veröffentlicht: (2025)
von: Yuan, Hang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Confidence Estimation for Error Detection in Text-to-SQL Systems
von: Somov, Oleg, et al.
Veröffentlicht: (2025) -
Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers
von: Gong, Linyuan, et al.
Veröffentlicht: (2023) -
The benefits of query-based KGQA systems for complex and temporal questions in LLM era
von: Alekseev, Artem, et al.
Veröffentlicht: (2025) -
When Punctuation Matters: A Large-Scale Comparison of Prompt Robustness Methods for LLMs
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2025) -
Evolutionary Search for Automated Design of Uncertainty Quantification Methods
von: Seleznyov, Mikhail, et al.
Veröffentlicht: (2026)