Indexing Portuguese NLP Resources with PT-Pump-Up
Fuente:
arXiv
Saved in:
| Main Authors: | Almeida, Rúben, Campos, Ricardo, Jorge, Alípio, Nunes, Sérgio |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Question Answering Dataset for Temporal-Sensitive Retrieval-Augmented Generation
by: Chen, Ziyang, et al.
Published: (2025)
by: Chen, Ziyang, et al.
Published: (2025)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026)
by: Teixeira, Tiago, et al.
Published: (2026)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
by: Okpala, Izunna, et al.
Published: (2023)
by: Okpala, Izunna, et al.
Published: (2023)
FAIR-RAG: Faithful Adaptive Iterative Refinement for Retrieval-Augmented Generation
by: Asl, Mohammad Aghajani, et al.
Published: (2025)
by: Asl, Mohammad Aghajani, et al.
Published: (2025)
CoLe and LYS at BioASQ MESINESP8 Task: similarity based descriptor assignment in Spanish
by: Ribadas-Pena, Francisco J., et al.
Published: (2024)
by: Ribadas-Pena, Francisco J., et al.
Published: (2024)
Physio: An LLM-Based Physiotherapy Advisor
by: Almeida, Rúben, et al.
Published: (2024)
by: Almeida, Rúben, et al.
Published: (2024)
Knowledge Distillation for Low-Resource Open-source Text-to-SQL Model
by: Qiu, Tianhao, et al.
Published: (2026)
by: Qiu, Tianhao, et al.
Published: (2026)
KNIGHT: Knowledge Graph-Driven Multiple-Choice Question Generation with Adaptive Hardness Calibration
by: Amanlou, Mohammad, et al.
Published: (2026)
by: Amanlou, Mohammad, et al.
Published: (2026)
Efficient fine-tuning methodology of text embedding models for information retrieval: contrastive learning penalty (clp)
by: Yu, Jeongsu
Published: (2024)
by: Yu, Jeongsu
Published: (2024)
SegNSP: Revisiting Next Sentence Prediction for Linear Text Segmentation
by: Isidro, José, et al.
Published: (2026)
by: Isidro, José, et al.
Published: (2026)
EnterpriseRAG-Bench: A RAG Benchmark for Company Internal Knowledge
by: Sun, Yuhong, et al.
Published: (2026)
by: Sun, Yuhong, et al.
Published: (2026)
Regime-Conditional Retrieval: Theory and a Transferable Router for Two-Hop QA
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
Democratizing GraphRAG: Linear, CPU-Only Graph Retrieval for Multi-Hop QA
by: Wang, Qizhi
Published: (2025)
by: Wang, Qizhi
Published: (2025)
Agentic AI Systems Applied to tasks in Financial Services: Modeling and model risk management crews
by: Okpala, Izunna, et al.
Published: (2025)
by: Okpala, Izunna, et al.
Published: (2025)
SPARQL Generation with Entity Pre-trained GPT for KG Question Answering
by: Bustamante, Diego, et al.
Published: (2024)
by: Bustamante, Diego, et al.
Published: (2024)
DCD: Domain-Oriented Design for Controlled Retrieval-Augmented Generation
by: Kovalskiy, Valeriy, et al.
Published: (2026)
by: Kovalskiy, Valeriy, et al.
Published: (2026)
Efficient $k$-NN Search in IoT Data: Overlap Optimization in Tree-Based Indexing Structures
by: Benrazek, Ala-Eddine, et al.
Published: (2024)
by: Benrazek, Ala-Eddine, et al.
Published: (2024)
SPD-RAG: Sub-Agent Per Document Retrieval-Augmented Generation
by: Akay, Yagiz Can, et al.
Published: (2026)
by: Akay, Yagiz Can, et al.
Published: (2026)
A Reproducible, Scalable Pipeline for Synthesizing Autoregressive Model Literature
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Supercharging Federated Intelligence Retrieval
by: Stripelis, Dimitris, et al.
Published: (2026)
by: Stripelis, Dimitris, et al.
Published: (2026)
LLM-based IR-system for Bank Supervisors
by: Aarab, Ilias
Published: (2025)
by: Aarab, Ilias
Published: (2025)
A Study into Investigating Temporal Robustness of LLMs
by: Wallat, Jonas, et al.
Published: (2025)
by: Wallat, Jonas, et al.
Published: (2025)
Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models
by: Günther, Michael, et al.
Published: (2024)
by: Günther, Michael, et al.
Published: (2024)
Triplètoile: Extraction of Knowledge from Microblogging Text
by: Zavarella, Vanni, et al.
Published: (2024)
by: Zavarella, Vanni, et al.
Published: (2024)
Enabling Low-Resource Language Retrieval: Establishing Baselines for Urdu MS MARCO
by: Butt, Umer, et al.
Published: (2024)
by: Butt, Umer, et al.
Published: (2024)
PLUGH: A Benchmark for Spatial Understanding and Reasoning in Large Language Models
by: Tikhonov, Alexey
Published: (2024)
by: Tikhonov, Alexey
Published: (2024)
Perception-Aware Bias Detection for Query Suggestions
by: Haak, Fabian, et al.
Published: (2026)
by: Haak, Fabian, et al.
Published: (2026)
CAG: Chunked Augmented Generation for Google Chrome's Built-in Gemini Nano
by: Surulimuthu, Vivek Vellaiyappan, et al.
Published: (2024)
by: Surulimuthu, Vivek Vellaiyappan, et al.
Published: (2024)
NLCTables: A Dataset for Marrying Natural Language Conditions with Table Discovery
by: Cui, Lingxi, et al.
Published: (2025)
by: Cui, Lingxi, et al.
Published: (2025)
Efficient Fine-Tuning Methods for Portuguese Question Answering: A Comparative Study of PEFT on BERTimbau and Exploratory Evaluation of Generative LLMs
by: Nina, Mariela M., et al.
Published: (2026)
by: Nina, Mariela M., et al.
Published: (2026)
Rational Retrieval Acts: Leveraging Pragmatic Reasoning to Improve Sparse Retrieval
by: Satouf, Arthur, et al.
Published: (2025)
by: Satouf, Arthur, et al.
Published: (2025)
Deploying Large Language Models With Retrieval Augmented Generation
by: Prabhune, Sonal, et al.
Published: (2024)
by: Prabhune, Sonal, et al.
Published: (2024)
Graphemic Normalization of the Perso-Arabic Script
by: Doctor, Raiomond, et al.
Published: (2022)
by: Doctor, Raiomond, et al.
Published: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
by: Gutkin, Alexander, et al.
Published: (2023)
by: Gutkin, Alexander, et al.
Published: (2023)
Latent Objective Induction and Diversity-Constrained Selection: Algorithms for Multi-Locale Retrieval Pipelines
by: Alpay, Faruk, et al.
Published: (2026)
by: Alpay, Faruk, et al.
Published: (2026)
Improving Large-Scale k-Nearest Neighbor Text Categorization with Label Autoencoders
by: Ribadas-Pena, Francisco J., et al.
Published: (2024)
by: Ribadas-Pena, Francisco J., et al.
Published: (2024)
The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection
by: Repantis, Vyzantinos, et al.
Published: (2026)
by: Repantis, Vyzantinos, et al.
Published: (2026)
Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
by: Hossain, Ariyan, et al.
Published: (2025)
by: Hossain, Ariyan, et al.
Published: (2025)
Knowledge Distillation of Domain-adapted LLMs for Question-Answering in Telecom
by: Sen, Rishika, et al.
Published: (2025)
by: Sen, Rishika, et al.
Published: (2025)
Multi-Task Contrastive Learning for 8192-Token Bilingual Text Embeddings
by: Mohr, Isabelle, et al.
Published: (2024)
by: Mohr, Isabelle, et al.
Published: (2024)
Similar Items
-
A Question Answering Dataset for Temporal-Sensitive Retrieval-Augmented Generation
by: Chen, Ziyang, et al.
Published: (2025) -
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026) -
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
by: Okpala, Izunna, et al.
Published: (2023) -
FAIR-RAG: Faithful Adaptive Iterative Refinement for Retrieval-Augmented Generation
by: Asl, Mohammad Aghajani, et al.
Published: (2025) -
CoLe and LYS at BioASQ MESINESP8 Task: similarity based descriptor assignment in Spanish
by: Ribadas-Pena, Francisco J., et al.
Published: (2024)