SANSKRITI: A Comprehensive Benchmark for Evaluating Language Models' Knowledge of Indian Culture
Fuente:
arXiv
Salvato in:
| Autori principali: | Maji, Arijit, Kumar, Raghvendra, Ghosh, Akash, Anushka, Saha, Sriparna |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DRISHTIKON: A Multimodal Multilingual Benchmark for Testing Language Models' Understanding on Indian Culture
di: Maji, Arijit, et al.
Pubblicazione: (2025)
di: Maji, Arijit, et al.
Pubblicazione: (2025)
VIRAASAT: Traversing Novel Paths for Indian Cultural Reasoning
di: Surana, Harshul Raj, et al.
Pubblicazione: (2026)
di: Surana, Harshul Raj, et al.
Pubblicazione: (2026)
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
di: Kumar, Raghvendra, et al.
Pubblicazione: (2026)
di: Kumar, Raghvendra, et al.
Pubblicazione: (2026)
A Survey of Multilingual Reasoning in Language Models
di: Ghosh, Akash, et al.
Pubblicazione: (2025)
di: Ghosh, Akash, et al.
Pubblicazione: (2025)
CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare
di: Ghosh, Akash, et al.
Pubblicazione: (2025)
di: Ghosh, Akash, et al.
Pubblicazione: (2025)
From Fragments to Facts: A Curriculum-Driven DPO Approach for Generating Hindi News Veracity Explanations
di: Bansal, Pulkit, et al.
Pubblicazione: (2025)
di: Bansal, Pulkit, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
COSMMIC: Comment-Sensitive Multimodal Multilingual Indian Corpus for Summarization and Headline Generation
di: Kumar, Raghvendra, et al.
Pubblicazione: (2025)
di: Kumar, Raghvendra, et al.
Pubblicazione: (2025)
Let's Play Across Cultures: A Large Multilingual, Multicultural Benchmark for Assessing Language Models' Understanding of Sports
di: Singh, Punit Kumar, et al.
Pubblicazione: (2025)
di: Singh, Punit Kumar, et al.
Pubblicazione: (2025)
Infogen: Generating Complex Statistical Infographics from Documents
di: Ghosh, Akash, et al.
Pubblicazione: (2025)
di: Ghosh, Akash, et al.
Pubblicazione: (2025)
Relic: Enhancing Reward Model Generalization for Low-Resource Indic Languages with Few-Shot Examples
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2025)
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2025)
Two eyes, Two views, and finally, One summary! Towards Multi-modal Multi-tasking Knowledge-Infused Medical Dialogue Summarization
di: Saha, Anisha, et al.
Pubblicazione: (2024)
di: Saha, Anisha, et al.
Pubblicazione: (2024)
Towards Knowledge-Infused Automated Disease Diagnosis Assistant
di: Tomar, Mohit, et al.
Pubblicazione: (2024)
di: Tomar, Mohit, et al.
Pubblicazione: (2024)
CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning
di: Onyame, Eric, et al.
Pubblicazione: (2026)
di: Onyame, Eric, et al.
Pubblicazione: (2026)
When Background Matters: Breaking Medical Vision Language Models by Transferable Attack
di: Ghosh, Akash, et al.
Pubblicazione: (2026)
di: Ghosh, Akash, et al.
Pubblicazione: (2026)
MiMIC: Multi-Modal Indian Earnings Calls Dataset to Predict Stock Prices
di: Ghosh, Sohom, et al.
Pubblicazione: (2025)
di: Ghosh, Sohom, et al.
Pubblicazione: (2025)
Tourism Question Answer System in Indian Language using Domain-Adapted Foundation Models
di: Gatla, Praveen, et al.
Pubblicazione: (2025)
di: Gatla, Praveen, et al.
Pubblicazione: (2025)
Crossing Borders: A Multimodal Challenge for Indian Poetry Translation and Image Generation
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
Experimenting with Multi-modal Information to Predict Success of Indian IPOs
di: Ghosh, Sohom, et al.
Pubblicazione: (2024)
di: Ghosh, Sohom, et al.
Pubblicazione: (2024)
HKCanto-Eval: A Benchmark for Evaluating Cantonese Language Understanding and Cultural Comprehension in LLMs
di: Cheng, Tsz Chung, et al.
Pubblicazione: (2025)
di: Cheng, Tsz Chung, et al.
Pubblicazione: (2025)
Universal Adversarial Suffixes for Language Models Using Reinforcement Learning with Calibrated Reward
di: Soor, Sampriti, et al.
Pubblicazione: (2025)
di: Soor, Sampriti, et al.
Pubblicazione: (2025)
Yes, this is what I was looking for! Towards Multi-modal Medical Consultation Concern Summary Generation
di: Tiwari, Abhisek, et al.
Pubblicazione: (2024)
di: Tiwari, Abhisek, et al.
Pubblicazione: (2024)
How Good is Zero-Shot MT Evaluation for Low Resource Indian Languages?
di: Singh, Anushka, et al.
Pubblicazione: (2024)
di: Singh, Anushka, et al.
Pubblicazione: (2024)
MILPaC: A Novel Benchmark for Evaluating Translation of Legal Text to Indian Languages
di: Mahapatra, Sayan, et al.
Pubblicazione: (2023)
di: Mahapatra, Sayan, et al.
Pubblicazione: (2023)
Hi Model, generating 'nice' instead of 'good' is not as bad as generating 'rice'! Towards Context and Semantic Infused Dialogue Generation Loss Function and Evaluation Metric
di: Tiwari, Abhisek, et al.
Pubblicazione: (2023)
di: Tiwari, Abhisek, et al.
Pubblicazione: (2023)
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models
di: Cong, Youan, et al.
Pubblicazione: (2024)
di: Cong, Youan, et al.
Pubblicazione: (2024)
FIND: Toward Multimodal Financial Reasoning and Question Answering for Indic Languages
di: Das, Sarmistha, et al.
Pubblicazione: (2026)
di: Das, Sarmistha, et al.
Pubblicazione: (2026)
INDIC DIALECT: A Multi Task Benchmark to Evaluate and Translate in Indian Language Dialects
di: Sharma, Tarun, et al.
Pubblicazione: (2026)
di: Sharma, Tarun, et al.
Pubblicazione: (2026)
Evaluating Cultural Knowledge Processing in Large Language Models: A Cognitive Benchmarking Framework Integrating Retrieval-Augmented Generation
di: Lee, Hung-Shin, et al.
Pubblicazione: (2025)
di: Lee, Hung-Shin, et al.
Pubblicazione: (2025)
Talk, Snap, Complain: Validation-Aware Multimodal Expert Framework for Fine-Grained Customer Grievances
di: Singh, Rishu Kumar, et al.
Pubblicazione: (2025)
di: Singh, Rishu Kumar, et al.
Pubblicazione: (2025)
An EcoSage Assistant: Towards Building A Multimodal Plant Care Dialogue Assistant
di: Tomar, Mohit, et al.
Pubblicazione: (2024)
di: Tomar, Mohit, et al.
Pubblicazione: (2024)
Logical Consistency of Large Language Models in Fact-checking
di: Ghosh, Bishwamittra, et al.
Pubblicazione: (2024)
di: Ghosh, Bishwamittra, et al.
Pubblicazione: (2024)
Enhancing Adverse Drug Event Detection with Multimodal Dataset: Corpus Creation and Model Development
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
MedSumm: A Multimodal Approach to Summarizing Code-Mixed Hindi-English Clinical Queries
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
mmJEE-Eval: A Bilingual Multimodal Benchmark for Evaluating Scientific Reasoning in Vision-Language Models
di: Mukherjee, Arka, et al.
Pubblicazione: (2025)
di: Mukherjee, Arka, et al.
Pubblicazione: (2025)
When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities
di: Das, Sarmistha, et al.
Pubblicazione: (2026)
di: Das, Sarmistha, et al.
Pubblicazione: (2026)
Through the Prism of Culture: Evaluating LLMs' Understanding of Indian Subcultures and Traditions
di: Chhikara, Garima, et al.
Pubblicazione: (2025)
di: Chhikara, Garima, et al.
Pubblicazione: (2025)
A Comprehensive Evaluation of Semantic Relation Knowledge of Pretrained Language Models and Humans
di: Cao, Zhihan, et al.
Pubblicazione: (2024)
di: Cao, Zhihan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
DRISHTIKON: A Multimodal Multilingual Benchmark for Testing Language Models' Understanding on Indian Culture
di: Maji, Arijit, et al.
Pubblicazione: (2025) -
VIRAASAT: Traversing Novel Paths for Indian Cultural Reasoning
di: Surana, Harshul Raj, et al.
Pubblicazione: (2026) -
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
di: Kumar, Raghvendra, et al.
Pubblicazione: (2026) -
A Survey of Multilingual Reasoning in Language Models
di: Ghosh, Akash, et al.
Pubblicazione: (2025) -
CLINIC: Evaluating Multilingual Trustworthiness in Language Models for Healthcare
di: Ghosh, Akash, et al.
Pubblicazione: (2025)