KnowsLM: A framework for evaluation of small language models for knowledge augmentation and humanised conversations
Fuente:
arXiv
Saved in:
| Main Authors: | Harbola, Chitranshu, Purwar, Anupam |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prescriptive Agents based on RAG for Automated Maintenance (PARAM)
by: Harbola, Chitranshu, et al.
Published: (2025)
by: Harbola, Chitranshu, et al.
Published: (2025)
VidyaRANG: Conversational Learning Based Platform powered by Large Language Model
by: Harbola, Chitranshu, et al.
Published: (2024)
by: Harbola, Chitranshu, et al.
Published: (2024)
InkubaLM: A small language model for low-resource African languages
by: Tonja, Atnafu Lambebo, et al.
Published: (2024)
by: Tonja, Atnafu Lambebo, et al.
Published: (2024)
NLD-LLM: A systematic framework for evaluating small language transformer models on natural language description
by: Jelodar, Hamed, et al.
Published: (2025)
by: Jelodar, Hamed, et al.
Published: (2025)
Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability
by: B, Gautam, et al.
Published: (2024)
by: B, Gautam, et al.
Published: (2024)
VLMs-in-the-Wild: Bridging the Gap Between Academic Benchmarks and Enterprise Reality
by: Bandraupalli, Srihari, et al.
Published: (2025)
by: Bandraupalli, Srihari, et al.
Published: (2025)
E-ARMOR: Edge case Assessment and Review of Multilingual Optical Character Recognition
by: Gupta, Aryan, et al.
Published: (2025)
by: Gupta, Aryan, et al.
Published: (2025)
Introducing a new hyper-parameter for RAG: Context Window Utilization
by: Juvekar, Kush, et al.
Published: (2024)
by: Juvekar, Kush, et al.
Published: (2024)
CultureVo: The Serious Game of Utilizing Gen AI for Enhancing Cultural Intelligence
by: Agarwala, Ajita, et al.
Published: (2024)
by: Agarwala, Ajita, et al.
Published: (2024)
Retrieval augmentation of large language models for lay language generation
by: Guo, Yue, et al.
Published: (2022)
by: Guo, Yue, et al.
Published: (2022)
Retrieval-augmented reasoning with lean language models
by: Chan, Ryan Sze-Yin, et al.
Published: (2025)
by: Chan, Ryan Sze-Yin, et al.
Published: (2025)
Automated test generation to evaluate tool-augmented LLMs as conversational AI agents
by: Arcadinho, Samuel, et al.
Published: (2024)
by: Arcadinho, Samuel, et al.
Published: (2024)
TopoLM: brain-like spatio-functional organization in a topographic language model
by: Rathi, Neil, et al.
Published: (2024)
by: Rathi, Neil, et al.
Published: (2024)
ADMEDTAGGER: an annotation framework for distillation of expert knowledge for the Polish medical language
by: Górski, Franciszek, et al.
Published: (2025)
by: Górski, Franciszek, et al.
Published: (2025)
Anthropocentric bias in language model evaluation
by: Millière, Raphaël, et al.
Published: (2024)
by: Millière, Raphaël, et al.
Published: (2024)
HeLM: Highlighted Evidence augmented Language Model for Enhanced Table-to-Text Generation
by: Bian, Junyi, et al.
Published: (2023)
by: Bian, Junyi, et al.
Published: (2023)
DataComp-LM: In search of the next generation of training sets for language models
by: Li, Jeffrey, et al.
Published: (2024)
by: Li, Jeffrey, et al.
Published: (2024)
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
by: Barboule, Camille, et al.
Published: (2024)
by: Barboule, Camille, et al.
Published: (2024)
Infusing clinical knowledge into tokenisers for language models
by: Hasan, Abul, et al.
Published: (2024)
by: Hasan, Abul, et al.
Published: (2024)
M-PACE: Mother Child Framework for Multimodal Compliance
by: Verma, Shreyash, et al.
Published: (2025)
by: Verma, Shreyash, et al.
Published: (2025)
WizardLM: Empowering large pre-trained language models to follow complex instructions
by: Xu, Can, et al.
Published: (2023)
by: Xu, Can, et al.
Published: (2023)
Biomedical knowledge graph-optimized prompt generation for large language models
by: Soman, Karthik, et al.
Published: (2023)
by: Soman, Karthik, et al.
Published: (2023)
Uncovering inequalities in new knowledge learning by large language models across different languages
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
MLRIP: Pre-training a military language representation model with informative factual knowledge and professional knowledge base
by: Li, Hui, et al.
Published: (2022)
by: Li, Hui, et al.
Published: (2022)
When the LM misunderstood the human chuckled: Analyzing garden path effects in humans and language models
by: Amouyal, Samuel Joseph, et al.
Published: (2025)
by: Amouyal, Samuel Joseph, et al.
Published: (2025)
Re-evaluating Theory of Mind evaluation in large language models
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
A unified foundational framework for knowledge injection and evaluation of Large Language Models in Combustion Science
by: Yang, Zonglin, et al.
Published: (2026)
by: Yang, Zonglin, et al.
Published: (2026)
Large Language Models, scientific knowledge and factuality: A framework to streamline human expert evaluation
by: Wysocka, Magdalena, et al.
Published: (2023)
by: Wysocka, Magdalena, et al.
Published: (2023)
GRASP: A novel benchmark for evaluating language GRounding And Situated Physics understanding in multimodal language models
by: Jassim, Serwan, et al.
Published: (2023)
by: Jassim, Serwan, et al.
Published: (2023)
LLMs syntactically adapt their language use to their conversational partner
by: Kandra, Florian, et al.
Published: (2025)
by: Kandra, Florian, et al.
Published: (2025)
Detecting out-of-distribution text using topological features of transformer-based language models
by: Pollano, Andres, et al.
Published: (2023)
by: Pollano, Andres, et al.
Published: (2023)
Enhancing textual textbook question answering with large language models and retrieval augmented generation
by: Alawwad, Hessa Abdulrahman, et al.
Published: (2024)
by: Alawwad, Hessa Abdulrahman, et al.
Published: (2024)
Factual consistency evaluation of summarization in the Era of large language models
by: Luo, Zheheng, et al.
Published: (2024)
by: Luo, Zheheng, et al.
Published: (2024)
A safety realignment framework via subspace-oriented model fusion for large language models
by: Yi, Xin, et al.
Published: (2024)
by: Yi, Xin, et al.
Published: (2024)
Slm-mux: Orchestrating small language models for reasoning
by: Wang, Chenyu, et al.
Published: (2025)
by: Wang, Chenyu, et al.
Published: (2025)
Linear representations in language models can change dramatically over a conversation
by: Lampinen, Andrew Kyle, et al.
Published: (2026)
by: Lampinen, Andrew Kyle, et al.
Published: (2026)
Unified attacks to large language model watermarks: spoofing and scrubbing in unauthorized knowledge distillation
by: Yi, Xin, et al.
Published: (2025)
by: Yi, Xin, et al.
Published: (2025)
A benchmark dataset for evaluating Syndrome Differentiation and Treatment in large language models
by: Li, Kunning, et al.
Published: (2025)
by: Li, Kunning, et al.
Published: (2025)
AIDBench: A benchmark for evaluating the authorship identification capability of large language models
by: Wen, Zichen, et al.
Published: (2024)
by: Wen, Zichen, et al.
Published: (2024)
Frame Sampling Strategies Matter: A Benchmark for small vision language models
by: Brkic, Marija, et al.
Published: (2025)
by: Brkic, Marija, et al.
Published: (2025)
Similar Items
-
Prescriptive Agents based on RAG for Automated Maintenance (PARAM)
by: Harbola, Chitranshu, et al.
Published: (2025) -
VidyaRANG: Conversational Learning Based Platform powered by Large Language Model
by: Harbola, Chitranshu, et al.
Published: (2024) -
InkubaLM: A small language model for low-resource African languages
by: Tonja, Atnafu Lambebo, et al.
Published: (2024) -
NLD-LLM: A systematic framework for evaluating small language transformer models on natural language description
by: Jelodar, Hamed, et al.
Published: (2025) -
Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability
by: B, Gautam, et al.
Published: (2024)