BERTopic for Topic Modeling of Hindi Short Texts: A Comparative Study
Fuente:
arXiv
Saved in:
| Main Authors: | Mutsaddi, Atharva, Jamkhande, Anvi, Thakre, Aryan, Haribhakta, Yashodhara |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
by: Yang, Zhenye, et al.
Published: (2025)
by: Yang, Zhenye, et al.
Published: (2025)
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
by: Du, Yuchen, et al.
Published: (2026)
by: Du, Yuchen, et al.
Published: (2026)
Comparison of Unsupervised Metrics for Evaluating Judicial Decision Extraction
by: Litvak, Ivan Leonidovich, et al.
Published: (2025)
by: Litvak, Ivan Leonidovich, et al.
Published: (2025)
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
HySemRAG: A Hybrid Semantic Retrieval-Augmented Generation Framework for Automated Literature Synthesis and Methodological Gap Analysis
by: Godinez, Alejandro
Published: (2025)
by: Godinez, Alejandro
Published: (2025)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
by: Jiang, Yuning, et al.
Published: (2025)
by: Jiang, Yuning, et al.
Published: (2025)
Enhancing Plagiarism Detection in Marathi with a Weighted Ensemble of TF-IDF and BERT Embeddings for Low-Resource Language Processing
by: Mutsaddi, Atharva, et al.
Published: (2025)
by: Mutsaddi, Atharva, et al.
Published: (2025)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026)
by: Liu, Jianan, et al.
Published: (2026)
HiFi-RAG: Hierarchical Content Filtering and Two-Pass Generation for Open-Domain RAG
by: Nuengsigkapian, Cattalyya
Published: (2025)
by: Nuengsigkapian, Cattalyya
Published: (2025)
Method for Aggregating Unstructured Data Using Large Language Models
by: Lazebnyi, Vsevolod, et al.
Published: (2026)
by: Lazebnyi, Vsevolod, et al.
Published: (2026)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
by: Pawar, Tejas, et al.
Published: (2025)
by: Pawar, Tejas, et al.
Published: (2025)
Free Access to World News: Reconstructing Full-Text Articles from GDELT
by: Colladon, A. Fronzetti, et al.
Published: (2025)
by: Colladon, A. Fronzetti, et al.
Published: (2025)
Topic Is Not Agenda: A Citation-Community Audit of Text Embeddings
by: Yoo, Junseon
Published: (2026)
by: Yoo, Junseon
Published: (2026)
DySK-Attn: A Framework for Efficient, Real-Time Knowledge Updating in Large Language Models via Dynamic Sparse Knowledge Attention
by: Khan, Kabir, et al.
Published: (2025)
by: Khan, Kabir, et al.
Published: (2025)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
by: Iana, Andreea, et al.
Published: (2023)
by: Iana, Andreea, et al.
Published: (2023)
Session Context Embedding for Intent Understanding in Product Search
by: Mehrdad, Navid, et al.
Published: (2024)
by: Mehrdad, Navid, et al.
Published: (2024)
Scaling Multilingual Semantic Search in Uber Eats Delivery
by: Ling, Bo, et al.
Published: (2026)
by: Ling, Bo, et al.
Published: (2026)
Does UMBRELA Work on Other LLMs?
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines
by: Oruesagasti, Julen
Published: (2026)
by: Oruesagasti, Julen
Published: (2026)
Criteria-Based LLM Relevance Judgments
by: Farzi, Naghmeh, et al.
Published: (2025)
by: Farzi, Naghmeh, et al.
Published: (2025)
Intent-Driven Dynamic Chunking: Segmenting Documents to Reflect Predicted Information Needs
by: Koutsiaris, Christos
Published: (2026)
by: Koutsiaris, Christos
Published: (2026)
Graph-GRPO: Dependency-Aware Credit Assignment for Generative E-commerce Search Relevance
by: Che, Jiarui, et al.
Published: (2026)
by: Che, Jiarui, et al.
Published: (2026)
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025)
by: Cahoon, Joyce, et al.
Published: (2025)
Retrieval Augmented Thought Process for Private Data Handling in Healthcare
by: Pouplin, Thomas, et al.
Published: (2024)
by: Pouplin, Thomas, et al.
Published: (2024)
Retrieval and Augmentation of Domain Knowledge for Text-to-SQL Semantic Parsing
by: Patwardhan, Manasi, et al.
Published: (2025)
by: Patwardhan, Manasi, et al.
Published: (2025)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
by: Guo, David, et al.
Published: (2025)
by: Guo, David, et al.
Published: (2025)
VIRAASAT: Traversing Novel Paths for Indian Cultural Reasoning
by: Surana, Harshul Raj, et al.
Published: (2026)
by: Surana, Harshul Raj, et al.
Published: (2026)
Diversification as Risk Minimization
by: Takehi, Rikiya, et al.
Published: (2025)
by: Takehi, Rikiya, et al.
Published: (2025)
From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms
by: Kai, Zhang, et al.
Published: (2026)
by: Kai, Zhang, et al.
Published: (2026)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
by: Egami, Shusaku, et al.
Published: (2026)
by: Egami, Shusaku, et al.
Published: (2026)
Uncovering the Limitations of Query Performance Prediction: Failures, Insights, and Implications for Selective Query Processing
by: Chifu, Adrian-Gabriel, et al.
Published: (2025)
by: Chifu, Adrian-Gabriel, et al.
Published: (2025)
2024 Google Scholar Research Interest Ranking for Top 3260 Computer Science Authors
by: Rasane, Atharva
Published: (2024)
by: Rasane, Atharva
Published: (2024)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
by: Shi, Kainan, et al.
Published: (2025)
by: Shi, Kainan, et al.
Published: (2025)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
by: Budigi, Venkata Krishna Prasanth, et al.
Published: (2026)
by: Budigi, Venkata Krishna Prasanth, et al.
Published: (2026)
Cross-Subreddit Behavior as Open-Source Indicators of Coordinated Influence: A Case Study of r/Sino & r/China
by: Pilaud, Manon, et al.
Published: (2025)
by: Pilaud, Manon, et al.
Published: (2025)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
by: Yao, Zhihui, et al.
Published: (2026)
by: Yao, Zhihui, et al.
Published: (2026)
Exploring Information Retrieval Landscapes: An Investigation of a Novel Evaluation Techniques and Comparative Document Splitting Methods
by: Narimissa, Esmaeil, et al.
Published: (2024)
by: Narimissa, Esmaeil, et al.
Published: (2024)
Similar Items
-
STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
by: Yang, Zhenye, et al.
Published: (2025) -
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
by: Du, Yuchen, et al.
Published: (2026) -
Comparison of Unsupervised Metrics for Evaluating Judicial Decision Extraction
by: Litvak, Ivan Leonidovich, et al.
Published: (2025) -
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
by: Bose, Joy
Published: (2026) -
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
by: Chen, Hao, et al.
Published: (2025)