Select, Label, Evaluate: Active Testing in NLP
Fuente:
arXiv
Salvato in:
| Autori principali: | Purificato, Antonio, Bucarelli, Maria Sofia, Bacciu, Andrea, Mantrach, Amin, Silvestri, Fabrizio |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Majority Vote Paradigm Shift: When Popular Meets Optimal
di: Purificato, Antonio, et al.
Pubblicazione: (2025)
di: Purificato, Antonio, et al.
Pubblicazione: (2025)
$\nabla τ$: Gradient-based and Task-Agnostic machine Unlearning
di: Trippa, Daniel, et al.
Pubblicazione: (2024)
di: Trippa, Daniel, et al.
Pubblicazione: (2024)
Learning with Noisy Labels through Learnable Weighting and Centroid Similarity
di: Wani, Farooq Ahmad, et al.
Pubblicazione: (2023)
di: Wani, Farooq Ahmad, et al.
Pubblicazione: (2023)
Eco-Aware Graph Neural Networks for Sustainable Recommendations
di: Purificato, Antonio, et al.
Pubblicazione: (2024)
di: Purificato, Antonio, et al.
Pubblicazione: (2024)
Monte Carlo Temperature: a robust sampling strategy for LLM's uncertainty quantification methods
di: Cecere, Nicola, et al.
Pubblicazione: (2025)
di: Cecere, Nicola, et al.
Pubblicazione: (2025)
Evaluating Latent Knowledge of Public Tabular Datasets in Large Language Models
di: Silvestri, Matteo, et al.
Pubblicazione: (2025)
di: Silvestri, Matteo, et al.
Pubblicazione: (2025)
One Search Fits All: Pareto-Optimal Eco-Friendly Model Selection
di: Betello, Filippo, et al.
Pubblicazione: (2025)
di: Betello, Filippo, et al.
Pubblicazione: (2025)
MASS: MoErging through Adaptive Subspace Selection
di: Crisostomi, Donato, et al.
Pubblicazione: (2025)
di: Crisostomi, Donato, et al.
Pubblicazione: (2025)
Handling Ontology Gaps in Semantic Parsing
di: Bacciu, Andrea, et al.
Pubblicazione: (2024)
di: Bacciu, Andrea, et al.
Pubblicazione: (2024)
ATM: Improving Model Merging by Alternating Tuning and Merging
di: Zhou, Luca, et al.
Pubblicazione: (2024)
di: Zhou, Luca, et al.
Pubblicazione: (2024)
Natural Language Counterfactual Explanations for Graphs Using Large Language Models
di: Giorgi, Flavio, et al.
Pubblicazione: (2024)
di: Giorgi, Flavio, et al.
Pubblicazione: (2024)
What are you sinking? A geometric approach on attention sink
di: Ruscio, Valeria, et al.
Pubblicazione: (2025)
di: Ruscio, Valeria, et al.
Pubblicazione: (2025)
Self-generated Replay Memories for Continual Neural Machine Translation
di: Resta, Michele, et al.
Pubblicazione: (2024)
di: Resta, Michele, et al.
Pubblicazione: (2024)
Evaluation Metrics for Text Data Augmentation in NLP
di: Amadeus, Marcellus, et al.
Pubblicazione: (2024)
di: Amadeus, Marcellus, et al.
Pubblicazione: (2024)
Integrating Item Relevance in Training Loss for Sequential Recommender Systems
di: Bacciu, Andrea, et al.
Pubblicazione: (2023)
di: Bacciu, Andrea, et al.
Pubblicazione: (2023)
Optimizing Sentence Embedding with Pseudo-Labeling and Model Ensembles: A Hierarchical Framework for Enhanced NLP Tasks
di: Liu, Ziwei, et al.
Pubblicazione: (2025)
di: Liu, Ziwei, et al.
Pubblicazione: (2025)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
Attention Sinks in Diffusion Language Models
di: Rulli, Maximo Eduardo, et al.
Pubblicazione: (2025)
di: Rulli, Maximo Eduardo, et al.
Pubblicazione: (2025)
ERAS: Evaluating the Robustness of Chinese NLP Models to Morphological Garden Path Errors
di: Li, Qinchan, et al.
Pubblicazione: (2024)
di: Li, Qinchan, et al.
Pubblicazione: (2024)
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
di: Srinivasan, Sudarshan, et al.
Pubblicazione: (2024)
di: Srinivasan, Sudarshan, et al.
Pubblicazione: (2024)
BlackboxNLP-2025 MIB Shared Task: Improving Circuit Faithfulness via Better Edge Selection
di: Nikankin, Yaniv, et al.
Pubblicazione: (2025)
di: Nikankin, Yaniv, et al.
Pubblicazione: (2025)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
NLP Security and Ethics, in the Wild
di: Lent, Heather, et al.
Pubblicazione: (2025)
di: Lent, Heather, et al.
Pubblicazione: (2025)
OptiHive: Ensemble Selection for LLM-Based Optimization via Statistical Modeling
di: Bouscary, Maxime, et al.
Pubblicazione: (2025)
di: Bouscary, Maxime, et al.
Pubblicazione: (2025)
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
di: Sheth, Ivaxi, et al.
Pubblicazione: (2026)
di: Sheth, Ivaxi, et al.
Pubblicazione: (2026)
When Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLP
di: Papi, Sara, et al.
Pubblicazione: (2023)
di: Papi, Sara, et al.
Pubblicazione: (2023)
Beyond Black-Box Labels: Interpretable Criteria for Diagnosing Subjective NLP Tasks
di: Rair, Nisrine, et al.
Pubblicazione: (2026)
di: Rair, Nisrine, et al.
Pubblicazione: (2026)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2024)
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2024)
State of NLP in Kenya: A Survey
di: Amol, Cynthia Jayne, et al.
Pubblicazione: (2024)
di: Amol, Cynthia Jayne, et al.
Pubblicazione: (2024)
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
di: Chaleshtori, Fateme Hashemi, et al.
Pubblicazione: (2024)
di: Chaleshtori, Fateme Hashemi, et al.
Pubblicazione: (2024)
Maintaining Journalistic Integrity in the Digital Age: A Comprehensive NLP Framework for Evaluating Online News Content
di: Bojic, Ljubisa, et al.
Pubblicazione: (2024)
di: Bojic, Ljubisa, et al.
Pubblicazione: (2024)
Speaking of Language: Reflections on Metalanguage Research in NLP
di: Schneider, Nathan, et al.
Pubblicazione: (2026)
di: Schneider, Nathan, et al.
Pubblicazione: (2026)
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
di: Manna, Supriya, et al.
Pubblicazione: (2024)
di: Manna, Supriya, et al.
Pubblicazione: (2024)
Undesirable Biases in NLP: Addressing Challenges of Measurement
di: van der Wal, Oskar, et al.
Pubblicazione: (2022)
di: van der Wal, Oskar, et al.
Pubblicazione: (2022)
Evaluating Deduplication Techniques for Economic Research Paper Titles with a Focus on Semantic Similarity using NLP and LLMs
di: You, Doohee, et al.
Pubblicazione: (2024)
di: You, Doohee, et al.
Pubblicazione: (2024)
State-of-the-art generalisation research in NLP: A taxonomy and review
di: Hupkes, Dieuwke, et al.
Pubblicazione: (2022)
di: Hupkes, Dieuwke, et al.
Pubblicazione: (2022)
Comparative Evaluation of ChatGPT and DeepSeek Across Key NLP Tasks: Strengths, Weaknesses, and Domain-Specific Performance
di: Etaiwi, Wael, et al.
Pubblicazione: (2025)
di: Etaiwi, Wael, et al.
Pubblicazione: (2025)
Becoming Experienced Judges: Selective Test-Time Learning for Evaluators
di: Jwa, Seungyeon, et al.
Pubblicazione: (2025)
di: Jwa, Seungyeon, et al.
Pubblicazione: (2025)
Practising responsibility: Ethics in NLP as a hands-on course
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
Intertwining CP and NLP: The Generation of Unreasonably Constrained Sentences
di: Bonlarron, Alexandre, et al.
Pubblicazione: (2024)
di: Bonlarron, Alexandre, et al.
Pubblicazione: (2024)
Documenti analoghi
-
The Majority Vote Paradigm Shift: When Popular Meets Optimal
di: Purificato, Antonio, et al.
Pubblicazione: (2025) -
$\nabla τ$: Gradient-based and Task-Agnostic machine Unlearning
di: Trippa, Daniel, et al.
Pubblicazione: (2024) -
Learning with Noisy Labels through Learnable Weighting and Centroid Similarity
di: Wani, Farooq Ahmad, et al.
Pubblicazione: (2023) -
Eco-Aware Graph Neural Networks for Sustainable Recommendations
di: Purificato, Antonio, et al.
Pubblicazione: (2024) -
Monte Carlo Temperature: a robust sampling strategy for LLM's uncertainty quantification methods
di: Cecere, Nicola, et al.
Pubblicazione: (2025)