A document processing pipeline for the construction of a dataset for topic modeling based on the judgments of the Italian Supreme Court
Fuente:
arXiv
Salvato in:
| Autori principali: | Marulli, Matteo, Panattoni, Glauco, Bertini, Marco |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Language models align with human judgments on key grammatical constructions
di: Hu, Jennifer, et al.
Pubblicazione: (2024)
di: Hu, Jennifer, et al.
Pubblicazione: (2024)
The simulation of judgment in LLMs
di: Loru, Edoardo, et al.
Pubblicazione: (2025)
di: Loru, Edoardo, et al.
Pubblicazione: (2025)
Heard or Halted? Gender, Interruptions, and Emotional Tone in U.S. Supreme Court Oral Arguments
di: Tong, Yifei
Pubblicazione: (2025)
di: Tong, Yifei
Pubblicazione: (2025)
Summarizing long regulatory documents with a multi-step pipeline
di: Sie, Mika, et al.
Pubblicazione: (2024)
di: Sie, Mika, et al.
Pubblicazione: (2024)
Better Aligned with Survey Respondents or Training Data? Unveiling Political Leanings of LLMs on U.S. Supreme Court Cases
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
Large-Language Memorization During the Classification of United States Supreme Court Cases
di: Ortega, John E., et al.
Pubblicazione: (2025)
di: Ortega, John E., et al.
Pubblicazione: (2025)
CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions
di: Heddaya, Mourad, et al.
Pubblicazione: (2024)
di: Heddaya, Mourad, et al.
Pubblicazione: (2024)
OneLove beyond the field -- A few-shot pipeline for topic and sentiment analysis during the FIFA World Cup in Qatar
di: Rauchegger, Christoph, et al.
Pubblicazione: (2024)
di: Rauchegger, Christoph, et al.
Pubblicazione: (2024)
DIETA: A Decoder-only transformer-based model for Italian-English machine TrAnslation
di: Kasela, Pranav, et al.
Pubblicazione: (2026)
di: Kasela, Pranav, et al.
Pubblicazione: (2026)
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
Truth-value judgment in language models: 'truth directions' are context sensitive
di: Schouten, Stefan F., et al.
Pubblicazione: (2024)
di: Schouten, Stefan F., et al.
Pubblicazione: (2024)
A Human Word Association based model for topic detection in social networks
di: Khadivi, Mehrdad Ranjbar, et al.
Pubblicazione: (2023)
di: Khadivi, Mehrdad Ranjbar, et al.
Pubblicazione: (2023)
A large-scale pipeline for automatic corpus annotation using LLMs: variation and change in the English consider construction
di: Morin, Cameron, et al.
Pubblicazione: (2025)
di: Morin, Cameron, et al.
Pubblicazione: (2025)
Quality-Aware Image-Text Alignment for Opinion-Unaware Image Quality Assessment
di: Agnolucci, Lorenzo, et al.
Pubblicazione: (2024)
di: Agnolucci, Lorenzo, et al.
Pubblicazione: (2024)
PRODIGy: a PROfile-based DIalogue Generation dataset
di: Occhipinti, Daniela, et al.
Pubblicazione: (2023)
di: Occhipinti, Daniela, et al.
Pubblicazione: (2023)
Empirical analysis of binding precedent efficiency in Brazilian Supreme Court via case classification
di: Tinarrage, Raphaël, et al.
Pubblicazione: (2024)
di: Tinarrage, Raphaël, et al.
Pubblicazione: (2024)
EventNet-ITA: Italian Frame Parsing for Events
di: Rovera, Marco
Pubblicazione: (2023)
di: Rovera, Marco
Pubblicazione: (2023)
Constraining constructions with WordNet: pros and cons for the semantic annotation of fillers in the Italian Constructicon
di: Pisciotta, Flavio, et al.
Pubblicazione: (2025)
di: Pisciotta, Flavio, et al.
Pubblicazione: (2025)
GiusBERTo: A Legal Language Model for Personal Data De-identification in Italian Court of Auditors Decisions
di: Salierno, Giulio, et al.
Pubblicazione: (2024)
di: Salierno, Giulio, et al.
Pubblicazione: (2024)
A Bayesian approach to modeling topic-metadata relationships
di: Schulze, P., et al.
Pubblicazione: (2021)
di: Schulze, P., et al.
Pubblicazione: (2021)
Capturing research literature attitude towards Sustainable Development Goals: an LLM-based topic modeling approach
di: Invernici, Francesco, et al.
Pubblicazione: (2024)
di: Invernici, Francesco, et al.
Pubblicazione: (2024)
Gender-Neutral Rewriting in Italian: Models, Approaches, and Trade-offs
di: Piergentili, Andrea, et al.
Pubblicazione: (2025)
di: Piergentili, Andrea, et al.
Pubblicazione: (2025)
pytopicgram: A library for data extraction and topic modeling from Telegram channels
di: Gómez-Romero, J., et al.
Pubblicazione: (2025)
di: Gómez-Romero, J., et al.
Pubblicazione: (2025)
NEU-ESC: A Comprehensive Vietnamese dataset for Educational Sentiment analysis and topic Classification toward multitask learning
di: Mai, Phan Quoc Hung, et al.
Pubblicazione: (2025)
di: Mai, Phan Quoc Hung, et al.
Pubblicazione: (2025)
The ProLiFIC dataset: Leveraging LLMs to Unveil the Italian Lawmaking Process
di: Contestabile, Matilde, et al.
Pubblicazione: (2025)
di: Contestabile, Matilde, et al.
Pubblicazione: (2025)
CFTM: Continuous time fractional topic model
di: Nakagawa, Kei, et al.
Pubblicazione: (2024)
di: Nakagawa, Kei, et al.
Pubblicazione: (2024)
A comparative study of transformer-based embeddings for topic coherence
di: Ding, Alex, et al.
Pubblicazione: (2026)
di: Ding, Alex, et al.
Pubblicazione: (2026)
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets
di: Münker, Simon, et al.
Pubblicazione: (2024)
di: Münker, Simon, et al.
Pubblicazione: (2024)
IMB: An Italian Medical Benchmark for Question Answering
di: Romano, Antonio, et al.
Pubblicazione: (2025)
di: Romano, Antonio, et al.
Pubblicazione: (2025)
Testimole-Conversational: A 30-Billion-Word Italian Discussion Board Corpus (1996-2024) for Language Modeling and Sociolinguistic Research
di: Rinaldi, Matteo, et al.
Pubblicazione: (2026)
di: Rinaldi, Matteo, et al.
Pubblicazione: (2026)
FAMA: The First Large-Scale Open-Science Speech Foundation Model for English and Italian
di: Papi, Sara, et al.
Pubblicazione: (2025)
di: Papi, Sara, et al.
Pubblicazione: (2025)
Combining topic modelling and citation network analysis to study case law from the European Court on Human Rights on the right to respect for private and family life
di: Mohammadi, M., et al.
Pubblicazione: (2024)
di: Mohammadi, M., et al.
Pubblicazione: (2024)
Dynamic embedded topic models and change-point detection for exploring literary-historical hypotheses
di: Sirin, Hale, et al.
Pubblicazione: (2024)
di: Sirin, Hale, et al.
Pubblicazione: (2024)
Culturally Grounded Physical Commonsense Reasoning in Italian and English: A Submission to the MRL 2025 Shared Task
di: De Santis, Marco, et al.
Pubblicazione: (2025)
di: De Santis, Marco, et al.
Pubblicazione: (2025)
ComiCap: A VLMs pipeline for dense captioning of Comic Panels
di: Vivoli, Emanuele, et al.
Pubblicazione: (2024)
di: Vivoli, Emanuele, et al.
Pubblicazione: (2024)
ChatSchema: A pipeline of extracting structured information with Large Multimodal Models based on schema
di: Wang, Fei, et al.
Pubblicazione: (2024)
di: Wang, Fei, et al.
Pubblicazione: (2024)
Harnessing LLMs for Educational Content-Driven Italian Crossword Generation
di: Zeinalipour, Kamyar, et al.
Pubblicazione: (2024)
di: Zeinalipour, Kamyar, et al.
Pubblicazione: (2024)
DART: A Structured Dataset of Regulatory Drug Documents in Italian for Clinical NLP
di: Barone, Mariano, et al.
Pubblicazione: (2025)
di: Barone, Mariano, et al.
Pubblicazione: (2025)
The study of short texts in digital politics: Document aggregation for topic modeling
di: Nakka, Nitheesha, et al.
Pubblicazione: (2025)
di: Nakka, Nitheesha, et al.
Pubblicazione: (2025)
Evalita-LLM: Benchmarking Large Language Models on Italian
di: Magnini, Bernardo, et al.
Pubblicazione: (2025)
di: Magnini, Bernardo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Language models align with human judgments on key grammatical constructions
di: Hu, Jennifer, et al.
Pubblicazione: (2024) -
The simulation of judgment in LLMs
di: Loru, Edoardo, et al.
Pubblicazione: (2025) -
Heard or Halted? Gender, Interruptions, and Emotional Tone in U.S. Supreme Court Oral Arguments
di: Tong, Yifei
Pubblicazione: (2025) -
Summarizing long regulatory documents with a multi-step pipeline
di: Sie, Mika, et al.
Pubblicazione: (2024) -
Better Aligned with Survey Respondents or Training Data? Unveiling Political Leanings of LLMs on U.S. Supreme Court Cases
di: Xu, Shanshan, et al.
Pubblicazione: (2025)