Machine Learning for Coding Retail Product Names to Consumer-Price Categories: A Rule-plus-Bag-of-Words Pipeline with Reliability-Weighted Human-in-the-Loop Labeling
Fuente:
arXiv
Salvato in:
| Autore principale: | Beskorovainyi, Vladimir |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Real-World En Call Center Transcripts Dataset with PII Redaction
di: Dao, Ha, et al.
Pubblicazione: (2025)
di: Dao, Ha, et al.
Pubblicazione: (2025)
VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio
di: Basha, Maris, et al.
Pubblicazione: (2025)
di: Basha, Maris, et al.
Pubblicazione: (2025)
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommendation
di: Zeng, Yangchen, et al.
Pubblicazione: (2026)
di: Zeng, Yangchen, et al.
Pubblicazione: (2026)
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025)
INESC-ID @ eRisk 2025: Exploring Fine-Tuned, Similarity-Based, and Prompt-Based Approaches to Depression Symptom Identification
di: Nunes, Diogo A. P., et al.
Pubblicazione: (2025)
di: Nunes, Diogo A. P., et al.
Pubblicazione: (2025)
NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data
di: Ming, Cong, et al.
Pubblicazione: (2026)
di: Ming, Cong, et al.
Pubblicazione: (2026)
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
di: Chen, Hao, et al.
Pubblicazione: (2025)
di: Chen, Hao, et al.
Pubblicazione: (2025)
Enhancing Document AI Data Generation Through Graph-Based Synthetic Layouts
di: Agarwal, Amit, et al.
Pubblicazione: (2024)
di: Agarwal, Amit, et al.
Pubblicazione: (2024)
A ripple in time: a discontinuity in American history
di: Kolpakov, Alexander, et al.
Pubblicazione: (2023)
di: Kolpakov, Alexander, et al.
Pubblicazione: (2023)
Session Context Embedding for Intent Understanding in Product Search
di: Mehrdad, Navid, et al.
Pubblicazione: (2024)
di: Mehrdad, Navid, et al.
Pubblicazione: (2024)
Contextually Aware E-Commerce Product Question Answering using RAG
di: Tangarajan, Praveen, et al.
Pubblicazione: (2025)
di: Tangarajan, Praveen, et al.
Pubblicazione: (2025)
Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship
di: Pan, Yating, et al.
Pubblicazione: (2026)
di: Pan, Yating, et al.
Pubblicazione: (2026)
Comparison between the Structures of Word Co-occurrence and Word Similarity Networks for Ill-formed and Well-formed Texts in Taiwan Mandarin
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2024)
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2024)
The Case for Intent-Based Query Rewriting
di: Nicolai, Gianna Lisa, et al.
Pubblicazione: (2025)
di: Nicolai, Gianna Lisa, et al.
Pubblicazione: (2025)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
di: Shi, Kainan, et al.
Pubblicazione: (2025)
di: Shi, Kainan, et al.
Pubblicazione: (2025)
Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature
di: Maiti, Agniva, et al.
Pubblicazione: (2025)
di: Maiti, Agniva, et al.
Pubblicazione: (2025)
Detection of Personal Data in Structured Datasets Using a Large Language Model
di: Ntwali, Albert Agisha, et al.
Pubblicazione: (2025)
di: Ntwali, Albert Agisha, et al.
Pubblicazione: (2025)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
di: Whittaker, Edward, et al.
Pubblicazione: (2024)
di: Whittaker, Edward, et al.
Pubblicazione: (2024)
A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web
di: Egami, Shusaku, et al.
Pubblicazione: (2026)
di: Egami, Shusaku, et al.
Pubblicazione: (2026)
FinQAPT: Empowering Financial Decisions with End-to-End LLM-driven Question Answering Pipeline
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
di: Zhang, Yiqing, et al.
Pubblicazione: (2026)
di: Zhang, Yiqing, et al.
Pubblicazione: (2026)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Flippi: End To End GenAI Assistant for E-Commerce
di: Rajasekar, Anand A., et al.
Pubblicazione: (2025)
di: Rajasekar, Anand A., et al.
Pubblicazione: (2025)
Fine-Grained Emotion Recognition via In-Context Learning
di: Ren, Zhaochun, et al.
Pubblicazione: (2025)
di: Ren, Zhaochun, et al.
Pubblicazione: (2025)
Adaptive ToR: Complexity-Aware Tree-Based Retrieval for Pareto-Optimal Multi-Intent NLU
di: Yoo, Hee-Kyong, et al.
Pubblicazione: (2026)
di: Yoo, Hee-Kyong, et al.
Pubblicazione: (2026)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
di: Ratul, Md Toyaha Rahman, et al.
Pubblicazione: (2026)
di: Ratul, Md Toyaha Rahman, et al.
Pubblicazione: (2026)
NewsScope: Schema-Grounded Cross-Domain News Claim Extraction with Open Models
di: Pandya, Nidhi
Pubblicazione: (2025)
di: Pandya, Nidhi
Pubblicazione: (2025)
Evaluation of Chunking Strategies for Effective Text Embedding in Low-Resource Language on Agricultural Documents
di: Chhoun, Sovandara, et al.
Pubblicazione: (2026)
di: Chhoun, Sovandara, et al.
Pubblicazione: (2026)
Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables
di: Shen, Chen
Pubblicazione: (2026)
di: Shen, Chen
Pubblicazione: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
di: Bacellar, Andre
Pubblicazione: (2026)
di: Bacellar, Andre
Pubblicazione: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
di: Iana, Andreea, et al.
Pubblicazione: (2023)
di: Iana, Andreea, et al.
Pubblicazione: (2023)
Towards Adaptive Context Management for Intelligent Conversational Question Answering
di: Perera, Manoj Madushanka, et al.
Pubblicazione: (2025)
di: Perera, Manoj Madushanka, et al.
Pubblicazione: (2025)
Scaling Multilingual Semantic Search in Uber Eats Delivery
di: Ling, Bo, et al.
Pubblicazione: (2026)
di: Ling, Bo, et al.
Pubblicazione: (2026)
When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory
di: Shao, Jiaqi, et al.
Pubblicazione: (2026)
di: Shao, Jiaqi, et al.
Pubblicazione: (2026)
Does UMBRELA Work on Other LLMs?
di: Farzi, Naghmeh, et al.
Pubblicazione: (2025)
di: Farzi, Naghmeh, et al.
Pubblicazione: (2025)
Optimizing Retrieval-Augmented Generation for Electrical Engineering: A Case Study on ABB Circuit Breakers
di: Alawadhi, Salahuddin, et al.
Pubblicazione: (2025)
di: Alawadhi, Salahuddin, et al.
Pubblicazione: (2025)
2024 Google Scholar Research Interest Ranking for Top 3260 Computer Science Authors
di: Rasane, Atharva
Pubblicazione: (2024)
di: Rasane, Atharva
Pubblicazione: (2024)
Local Hybrid Retrieval-Augmented Document QA
di: Astrino, Paolo
Pubblicazione: (2025)
di: Astrino, Paolo
Pubblicazione: (2025)
Documenti analoghi
-
Real-World En Call Center Transcripts Dataset with PII Redaction
di: Dao, Ha, et al.
Pubblicazione: (2025) -
VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio
di: Basha, Maris, et al.
Pubblicazione: (2025) -
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommendation
di: Zeng, Yangchen, et al.
Pubblicazione: (2026) -
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking
di: Seetharaman, Rahul, et al.
Pubblicazione: (2025) -
INESC-ID @ eRisk 2025: Exploring Fine-Tuned, Similarity-Based, and Prompt-Based Approaches to Depression Symptom Identification
di: Nunes, Diogo A. P., et al.
Pubblicazione: (2025)