The Russian Legislative Corpus
Fuente:
arXiv
Saved in:
| Main Authors: | Saveliev, Denis, Kuchakov, Ruslan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SinhaLegal: A Benchmark Corpus for Information Extraction and Analysis in Sinhala Legislative Texts
by: Lasandi, Minduli, et al.
Published: (2026)
by: Lasandi, Minduli, et al.
Published: (2026)
Benchmarking terminology building capabilities of ChatGPT on an English-Russian Fashion Corpus
by: Bezobrazova, Anastasiia, et al.
Published: (2024)
by: Bezobrazova, Anastasiia, et al.
Published: (2024)
The GELATO Dataset for Legislative NER
by: Flynn, Matthew, et al.
Published: (2026)
by: Flynn, Matthew, et al.
Published: (2026)
Long Input Benchmark for Russian Analysis
by: Churin, Igor, et al.
Published: (2024)
by: Churin, Igor, et al.
Published: (2024)
Algorithm for Automatic Legislative Text Consolidation
by: Etcheverry, Matias, et al.
Published: (2025)
by: Etcheverry, Matias, et al.
Published: (2025)
MERA: A Comprehensive LLM Evaluation in Russian
by: Fenogenova, Alena, et al.
Published: (2024)
by: Fenogenova, Alena, et al.
Published: (2024)
Contextual Clarity: Generating Sentences with Transformer Models using Context-Reverso Data
by: Musaev, Ruslan
Published: (2024)
by: Musaev, Ruslan
Published: (2024)
Can Legislation Be Made Machine-Readable in PROLEG?
by: Zin, May-Myo, et al.
Published: (2026)
by: Zin, May-Myo, et al.
Published: (2026)
Creating an Aligned Corpus of Sound and Text: The Multimodal Corpus of Shakespeare and Milton
by: Agirrezabal, Manex
Published: (2024)
by: Agirrezabal, Manex
Published: (2024)
The Material Contracts Corpus
by: Adelson, Peter, et al.
Published: (2025)
by: Adelson, Peter, et al.
Published: (2025)
Is Semi-Automatic Transcription Useful in Corpus Creation? Preliminary Considerations on the KIParla Corpus
by: Simonotti, Martina, et al.
Published: (2026)
by: Simonotti, Martina, et al.
Published: (2026)
The Medical Metaphors Corpus (MCC)
by: Lippolis, Anna Sofia, et al.
Published: (2025)
by: Lippolis, Anna Sofia, et al.
Published: (2025)
Speak & Improve Corpus 2025: an L2 English Speech Corpus for Language Assessment and Feedback
by: Knill, Kate, et al.
Published: (2024)
by: Knill, Kate, et al.
Published: (2024)
RusCode: Russian Cultural Code Benchmark for Text-to-Image Generation
by: Vasilev, Viacheslav, et al.
Published: (2025)
by: Vasilev, Viacheslav, et al.
Published: (2025)
TypedCSIP: Typed Counterfactual Pretraining for Chinese Legislative Conflict Classification
by: Liu, Yao
Published: (2026)
by: Liu, Yao
Published: (2026)
Multimodal Evaluation of Russian-language Architectures
by: Chervyakov, Artem, et al.
Published: (2025)
by: Chervyakov, Artem, et al.
Published: (2025)
The SAMER Arabic Text Simplification Corpus
by: Alhafni, Bashar, et al.
Published: (2024)
by: Alhafni, Bashar, et al.
Published: (2024)
The TUB Sign Language Corpus Collection
by: Avramidis, Eleftherios, et al.
Published: (2025)
by: Avramidis, Eleftherios, et al.
Published: (2025)
The Pilot Corpus of the English Semantic Sketches
by: Petrova, Maria, et al.
Published: (2025)
by: Petrova, Maria, et al.
Published: (2025)
RLSR: Reinforcement Learning with Supervised Reward Outperforms SFT in Instruction Following
by: Wang, Zhichao, et al.
Published: (2025)
by: Wang, Zhichao, et al.
Published: (2025)
LexDrafter: Terminology Drafting for Legislative Documents using Retrieval Augmented Generation
by: Chouhan, Ashish, et al.
Published: (2024)
by: Chouhan, Ashish, et al.
Published: (2024)
The Russian-focused embedders' exploration: ruMTEB benchmark and Russian embedding model design
by: Snegirev, Artem, et al.
Published: (2024)
by: Snegirev, Artem, et al.
Published: (2024)
Computational Identification of Regulatory Statements in EU Legislation
by: Brandsma, Gijs Jan, et al.
Published: (2025)
by: Brandsma, Gijs Jan, et al.
Published: (2025)
Large Language Models in Legislative Content Analysis: A Dataset from the Polish Parliament
by: Bryłkowski, Arkadiusz, et al.
Published: (2025)
by: Bryłkowski, Arkadiusz, et al.
Published: (2025)
FFSTC: Fongbe to French Speech Translation Corpus
by: Kponou, D. Fortune, et al.
Published: (2024)
by: Kponou, D. Fortune, et al.
Published: (2024)
A French Version of the OLDI Seed Corpus
by: Marmonier, Malik, et al.
Published: (2025)
by: Marmonier, Malik, et al.
Published: (2025)
MDC-R: The Minecraft Dialogue Corpus with Reference
by: Madge, Chris, et al.
Published: (2025)
by: Madge, Chris, et al.
Published: (2025)
SiDiaC: Sinhala Diachronic Corpus
by: Jayatilleke, Nevidu, et al.
Published: (2025)
by: Jayatilleke, Nevidu, et al.
Published: (2025)
Corpus Frequencies in Morphological Inflection: Do They Matter?
by: Sourada, Tomáš, et al.
Published: (2025)
by: Sourada, Tomáš, et al.
Published: (2025)
Vacaspati: A Diverse Corpus of Bangla Literature
by: Bhattacharyya, Pramit, et al.
Published: (2023)
by: Bhattacharyya, Pramit, et al.
Published: (2023)
GigaEmbeddings: Efficient Russian Language Embedding Model
by: Kolodin, Egor, et al.
Published: (2025)
by: Kolodin, Egor, et al.
Published: (2025)
Light Coreference Resolution for Russian with Hierarchical Discourse Features
by: Chistova, Elena, et al.
Published: (2023)
by: Chistova, Elena, et al.
Published: (2023)
Detecting Spelling and Grammatical Anomalies in Russian Poetry Texts
by: Koziev, Ilya
Published: (2025)
by: Koziev, Ilya
Published: (2025)
A Family of Pretrained Transformer Language Models for Russian
by: Zmitrovich, Dmitry, et al.
Published: (2023)
by: Zmitrovich, Dmitry, et al.
Published: (2023)
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks
by: Li, Xiaoxi, et al.
Published: (2024)
by: Li, Xiaoxi, et al.
Published: (2024)
CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning
by: Lu, Zhiyuan, et al.
Published: (2026)
by: Lu, Zhiyuan, et al.
Published: (2026)
The Moral Foundations Reddit Corpus
by: Trager, Jackson, et al.
Published: (2022)
by: Trager, Jackson, et al.
Published: (2022)
Leveraging the Cross-Domain & Cross-Linguistic Corpus for Low Resource NMT: A Case Study On Bhili-Hindi-English Parallel Corpus
by: Singh, Pooja, et al.
Published: (2025)
by: Singh, Pooja, et al.
Published: (2025)
Cleaner Pretraining Corpus Curation with Neural Web Scraping
by: Xu, Zhipeng, et al.
Published: (2024)
by: Xu, Zhipeng, et al.
Published: (2024)
A Topic-aware Comparable Corpus of Chinese Variations
by: Lian, Da-Chen, et al.
Published: (2024)
by: Lian, Da-Chen, et al.
Published: (2024)
Similar Items
-
SinhaLegal: A Benchmark Corpus for Information Extraction and Analysis in Sinhala Legislative Texts
by: Lasandi, Minduli, et al.
Published: (2026) -
Benchmarking terminology building capabilities of ChatGPT on an English-Russian Fashion Corpus
by: Bezobrazova, Anastasiia, et al.
Published: (2024) -
The GELATO Dataset for Legislative NER
by: Flynn, Matthew, et al.
Published: (2026) -
Long Input Benchmark for Russian Analysis
by: Churin, Igor, et al.
Published: (2024) -
Algorithm for Automatic Legislative Text Consolidation
by: Etcheverry, Matias, et al.
Published: (2025)