NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Ming, Cong, Shi, Ruixin, Hu, Yifan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Named Entity Recognition Models for Russian Cultural News Texts: From BERT to LLM
por: Levchenko, Maria
Publicado: (2025)
por: Levchenko, Maria
Publicado: (2025)
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
por: Whittaker, Edward, et al.
Publicado: (2024)
por: Whittaker, Edward, et al.
Publicado: (2024)
Local Hybrid Retrieval-Augmented Document QA
por: Astrino, Paolo
Publicado: (2025)
por: Astrino, Paolo
Publicado: (2025)
A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering
por: Ros, Sereiwathna, et al.
Publicado: (2026)
por: Ros, Sereiwathna, et al.
Publicado: (2026)
NewsScope: Schema-Grounded Cross-Domain News Claim Extraction with Open Models
por: Pandya, Nidhi
Publicado: (2025)
por: Pandya, Nidhi
Publicado: (2025)
Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems
por: Cirillo, Stefano, et al.
Publicado: (2026)
por: Cirillo, Stefano, et al.
Publicado: (2026)
NCTB-QA: A Large-Scale Bangla Educational Question Answering Dataset and Benchmarking Performance
por: Eyasir, Abrar, et al.
Publicado: (2026)
por: Eyasir, Abrar, et al.
Publicado: (2026)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
por: Pradhan, Anu, et al.
Publicado: (2025)
por: Pradhan, Anu, et al.
Publicado: (2025)
IndoBERT-Relevancy: A Context-Conditioned Relevancy Classifier for Indonesian Text
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
Machine Learning for Coding Retail Product Names to Consumer-Price Categories: A Rule-plus-Bag-of-Words Pipeline with Reliability-Weighted Human-in-the-Loop Labeling
por: Beskorovainyi, Vladimir
Publicado: (2026)
por: Beskorovainyi, Vladimir
Publicado: (2026)
Prompt Perturbation in Retrieval-Augmented Generation based Large Language Models
por: Hu, Zhibo, et al.
Publicado: (2024)
por: Hu, Zhibo, et al.
Publicado: (2024)
Flippi: End To End GenAI Assistant for E-Commerce
por: Rajasekar, Anand A., et al.
Publicado: (2025)
por: Rajasekar, Anand A., et al.
Publicado: (2025)
Evaluation of Chunking Strategies for Effective Text Embedding in Low-Resource Language on Agricultural Documents
por: Chhoun, Sovandara, et al.
Publicado: (2026)
por: Chhoun, Sovandara, et al.
Publicado: (2026)
Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables
por: Shen, Chen
Publicado: (2026)
por: Shen, Chen
Publicado: (2026)
Towards Adaptive Context Management for Intelligent Conversational Question Answering
por: Perera, Manoj Madushanka, et al.
Publicado: (2025)
por: Perera, Manoj Madushanka, et al.
Publicado: (2025)
FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAG
por: Dassen, Maxime, et al.
Publicado: (2026)
por: Dassen, Maxime, et al.
Publicado: (2026)
Contextually Aware E-Commerce Product Question Answering using RAG
por: Tangarajan, Praveen, et al.
Publicado: (2025)
por: Tangarajan, Praveen, et al.
Publicado: (2025)
Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship
por: Pan, Yating, et al.
Publicado: (2026)
por: Pan, Yating, et al.
Publicado: (2026)
GISTBench: Evaluating LLM User Understanding via Evidence-Based Interest Verification
por: Fostiropoulos, Iordanis, et al.
Publicado: (2026)
por: Fostiropoulos, Iordanis, et al.
Publicado: (2026)
Query-Centric Graph Retrieval Augmented Generation
por: Wu, Yaxiong, et al.
Publicado: (2025)
por: Wu, Yaxiong, et al.
Publicado: (2025)
Augmented Relevance Datasets with Fine-Tuned Small LLMs
por: Fitte-Rey, Quentin, et al.
Publicado: (2025)
por: Fitte-Rey, Quentin, et al.
Publicado: (2025)
Agentic Retrieval-Augmented Generation for Financial Document Question Answering
por: Shu, Yang, et al.
Publicado: (2026)
por: Shu, Yang, et al.
Publicado: (2026)
RAGged Edges: The Double-Edged Sword of Retrieval-Augmented Chatbots
por: Feldman, Philip, et al.
Publicado: (2024)
por: Feldman, Philip, et al.
Publicado: (2024)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
por: Zhang, Yongyue, et al.
Publicado: (2026)
por: Zhang, Yongyue, et al.
Publicado: (2026)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
por: Verhoeff, Tom
Publicado: (2026)
por: Verhoeff, Tom
Publicado: (2026)
ArcheType: A Novel Framework for Open-Source Column Type Annotation using Large Language Models
por: Feuer, Benjamin, et al.
Publicado: (2023)
por: Feuer, Benjamin, et al.
Publicado: (2023)
Incorporating Legal Structure in Retrieval-Augmented Generation: A Case Study on Copyright Fair Use
por: Ho, Justin, et al.
Publicado: (2025)
por: Ho, Justin, et al.
Publicado: (2025)
Criteria-Based LLM Relevance Judgments
por: Farzi, Naghmeh, et al.
Publicado: (2025)
por: Farzi, Naghmeh, et al.
Publicado: (2025)
Less LLM, More Documents: Searching for Improved RAG
por: Ning, Jingjie, et al.
Publicado: (2025)
por: Ning, Jingjie, et al.
Publicado: (2025)
Walk&Retrieve: Simple Yet Effective Zero-shot Retrieval-Augmented Generation via Knowledge Graph Walks
por: Böckling, Martin, et al.
Publicado: (2025)
por: Böckling, Martin, et al.
Publicado: (2025)
Promoting Research Collaboration with Open Data Driven Team Recommendation in Response to Call for Proposals
por: Valluru, Siva Likitha, et al.
Publicado: (2023)
por: Valluru, Siva Likitha, et al.
Publicado: (2023)
Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study
por: Tummalapenta, Ravi Kumar, et al.
Publicado: (2026)
por: Tummalapenta, Ravi Kumar, et al.
Publicado: (2026)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
por: Liu, Jianan, et al.
Publicado: (2026)
por: Liu, Jianan, et al.
Publicado: (2026)
FinBERT-QA: Financial Question Answering with pre-trained BERT Language Models
por: Yuan, Bithiah
Publicado: (2025)
por: Yuan, Bithiah
Publicado: (2025)
EQUATOR: A Deterministic Framework for Evaluating LLM Reasoning with Open-Ended Questions. # v1.0.0-beta
por: Bernard, Raymond, et al.
Publicado: (2024)
por: Bernard, Raymond, et al.
Publicado: (2024)
The Case for Intent-Based Query Rewriting
por: Nicolai, Gianna Lisa, et al.
Publicado: (2025)
por: Nicolai, Gianna Lisa, et al.
Publicado: (2025)
Optimizing Retrieval-Augmented Generation (RAG) for Colloquial Cantonese: A LoRA-Based Systematic Review
por: Calonge, David Santandreu, et al.
Publicado: (2025)
por: Calonge, David Santandreu, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Language Models via Self-Refinement-Enhanced Knowledge Retrieval
por: Niu, Mengjia, et al.
Publicado: (2024)
por: Niu, Mengjia, et al.
Publicado: (2024)
Improving and Evaluating Open Deep Research Agents
por: Allabadi, Doaa, et al.
Publicado: (2025)
por: Allabadi, Doaa, et al.
Publicado: (2025)
Ejemplares similares
-
Evaluating Named Entity Recognition Models for Russian Cultural News Texts: From BERT to LLM
por: Levchenko, Maria
Publicado: (2025) -
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026) -
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
por: Whittaker, Edward, et al.
Publicado: (2024) -
Local Hybrid Retrieval-Augmented Document QA
por: Astrino, Paolo
Publicado: (2025) -
A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering
por: Ros, Sereiwathna, et al.
Publicado: (2026)