Assessing generalization capability of text ranking models in Polish
Fuente:
arXiv
Guardado en:
| Autores principales: | Dadas, Sławomir, Grębowiec, Małgorzata |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Polish linguistic and cultural competency in large language models
por: Dadas, Sławomir, et al.
Publicado: (2025)
por: Dadas, Sławomir, et al.
Publicado: (2025)
Long-Context Encoder Models for Polish Language Understanding
por: Dadas, Sławomir, et al.
Publicado: (2026)
por: Dadas, Sławomir, et al.
Publicado: (2026)
Unveiling Dual Quality in Product Reviews: An NLP-Based Approach
por: Poświata, Rafał, et al.
Publicado: (2025)
por: Poświata, Rafał, et al.
Publicado: (2025)
PL-MTEB: Polish Massive Text Embedding Benchmark
por: Poświata, Rafał, et al.
Publicado: (2024)
por: Poświata, Rafał, et al.
Publicado: (2024)
PIRB: A Comprehensive Benchmark of Polish Dense and Hybrid Text Retrieval Methods
por: Dadas, Sławomir, et al.
Publicado: (2024)
por: Dadas, Sławomir, et al.
Publicado: (2024)
SMCLM: Semantically Meaningful Causal Language Modeling for Autoregressive Paraphrase Generation
por: Perełkiewicz, Michał, et al.
Publicado: (2025)
por: Perełkiewicz, Michał, et al.
Publicado: (2025)
PRODIS -- a speech database and a phoneme-based language model for the study of predictability effects in Polish
por: Malisz, Zofia, et al.
Publicado: (2024)
por: Malisz, Zofia, et al.
Publicado: (2024)
Evaluating the capability of large language models to personalize science texts for diverse middle-school-age learners
por: Vaccaro Jr, Michael, et al.
Publicado: (2024)
por: Vaccaro Jr, Michael, et al.
Publicado: (2024)
Analysis of instruction-based LLMs' capabilities to score and judge text-input problems in an academic setting
por: Ramirez-Garcia, Valeria, et al.
Publicado: (2025)
por: Ramirez-Garcia, Valeria, et al.
Publicado: (2025)
Assessing socio-economic climate impacts from text data
por: de Brito, Mariana Madruga, et al.
Publicado: (2026)
por: de Brito, Mariana Madruga, et al.
Publicado: (2026)
Assessing SPARQL capabilities of Large Language Models
por: Meyer, Lars-Peter, et al.
Publicado: (2024)
por: Meyer, Lars-Peter, et al.
Publicado: (2024)
PLLuM: A Family of Polish Large Language Models
por: Kocoń, Jan, et al.
Publicado: (2025)
por: Kocoń, Jan, et al.
Publicado: (2025)
Evidence of interrelated cognitive-like capabilities in large language models: Indications of artificial general intelligence or achievement?
por: Ilić, David, et al.
Publicado: (2023)
por: Ilić, David, et al.
Publicado: (2023)
Bi-reachability in Petri nets with data
por: Kamiński, Łukasz, et al.
Publicado: (2024)
por: Kamiński, Łukasz, et al.
Publicado: (2024)
Reachability in symmetric VASS
por: Kamiński, Łukasz, et al.
Publicado: (2025)
por: Kamiński, Łukasz, et al.
Publicado: (2025)
Polish-ASTE: Aspect-Sentiment Triplet Extraction Datasets for Polish
por: Lango, Marta, et al.
Publicado: (2025)
por: Lango, Marta, et al.
Publicado: (2025)
MedMobile: A mobile-sized language model with clinical capabilities
por: Vishwanath, Krithik, et al.
Publicado: (2024)
por: Vishwanath, Krithik, et al.
Publicado: (2024)
Identifying social isolation themes in NVDRS text narratives using topic modeling and text-classification methods
por: Walker, Drew, et al.
Publicado: (2025)
por: Walker, Drew, et al.
Publicado: (2025)
POLygraph: Polish Fake News Dataset
por: Dzienisiewicz, Daniel, et al.
Publicado: (2024)
por: Dzienisiewicz, Daniel, et al.
Publicado: (2024)
LLMs left, right, and center: Assessing GPT's capabilities to label political bias from web domains
por: Hernandes, Raphael, et al.
Publicado: (2024)
por: Hernandes, Raphael, et al.
Publicado: (2024)
AIDBench: A benchmark for evaluating the authorship identification capability of large language models
por: Wen, Zichen, et al.
Publicado: (2024)
por: Wen, Zichen, et al.
Publicado: (2024)
Synthetically generated text for supervised text analysis
por: Halterman, Andrew
Publicado: (2023)
por: Halterman, Andrew
Publicado: (2023)
Two Approaches to Diachronic Normalization of Polish Texts
por: Dudzic, Kacper, et al.
Publicado: (2024)
por: Dudzic, Kacper, et al.
Publicado: (2024)
Punctuation Prediction for Polish Texts using Transformers
por: Pokrywka, Jakub
Publicado: (2024)
por: Pokrywka, Jakub
Publicado: (2024)
PolQA: Polish Question Answering Dataset
por: Rybak, Piotr, et al.
Publicado: (2022)
por: Rybak, Piotr, et al.
Publicado: (2022)
Auxiliary task demands mask the capabilities of smaller language models
por: Hu, Jennifer, et al.
Publicado: (2024)
por: Hu, Jennifer, et al.
Publicado: (2024)
Solvability of orbit-finite systems of linear equations
por: Ghosh, Arka, et al.
Publicado: (2022)
por: Ghosh, Arka, et al.
Publicado: (2022)
Assessing News Thumbnail Representativeness: Counterfactual text can enhance the cross-modal matching ability
por: Yoon, Yejun, et al.
Publicado: (2024)
por: Yoon, Yejun, et al.
Publicado: (2024)
Polish phonology and morphology through the lens of distributional semantics
por: Orzechowska, Paula, et al.
Publicado: (2026)
por: Orzechowska, Paula, et al.
Publicado: (2026)
ZNO-Eval: Benchmarking reasoning capabilities of large language models in Ukrainian
por: Syromiatnikov, Mykyta, et al.
Publicado: (2025)
por: Syromiatnikov, Mykyta, et al.
Publicado: (2025)
Evaluation of Few-Shot Learning for Classification Tasks in the Polish Language
por: Hadeliya, Tsimur, et al.
Publicado: (2024)
por: Hadeliya, Tsimur, et al.
Publicado: (2024)
Explanation sensitivity to the randomness of large language models: the case of journalistic text classification
por: Bogaert, Jeremie, et al.
Publicado: (2024)
por: Bogaert, Jeremie, et al.
Publicado: (2024)
Automatic detection of Gen-AI texts: A comparative framework of neural models
por: Buttaro, Cristian, et al.
Publicado: (2026)
por: Buttaro, Cristian, et al.
Publicado: (2026)
Large language models struggle with ethnographic text annotation
por: Goodall, Leonardo S., et al.
Publicado: (2026)
por: Goodall, Leonardo S., et al.
Publicado: (2026)
How do we measure privacy in text? A survey of text anonymization metrics
por: Ren, Yaxuan, et al.
Publicado: (2025)
por: Ren, Yaxuan, et al.
Publicado: (2025)
Analysis of child development facts and myths using text mining techniques and classification models
por: Tajrian, Mehedi, et al.
Publicado: (2024)
por: Tajrian, Mehedi, et al.
Publicado: (2024)
Can human clinical rationales improve the performance and explainability of clinical text classification models?
por: Metzner, Christoph, et al.
Publicado: (2025)
por: Metzner, Christoph, et al.
Publicado: (2025)
Machine-generated text detection prevents language model collapse
por: Drayson, George, et al.
Publicado: (2025)
por: Drayson, George, et al.
Publicado: (2025)
Do LLMs exhibit the same commonsense capabilities across languages?
por: Martínez-Murillo, Ivan, et al.
Publicado: (2025)
por: Martínez-Murillo, Ivan, et al.
Publicado: (2025)
Achieving Operational Universality through a Turing Complete Chemputer
por: Gahler, Daniel, et al.
Publicado: (2025)
por: Gahler, Daniel, et al.
Publicado: (2025)
Ejemplares similares
-
Evaluating Polish linguistic and cultural competency in large language models
por: Dadas, Sławomir, et al.
Publicado: (2025) -
Long-Context Encoder Models for Polish Language Understanding
por: Dadas, Sławomir, et al.
Publicado: (2026) -
Unveiling Dual Quality in Product Reviews: An NLP-Based Approach
por: Poświata, Rafał, et al.
Publicado: (2025) -
PL-MTEB: Polish Massive Text Embedding Benchmark
por: Poświata, Rafał, et al.
Publicado: (2024) -
PIRB: A Comprehensive Benchmark of Polish Dense and Hybrid Text Retrieval Methods
por: Dadas, Sławomir, et al.
Publicado: (2024)