Exploring Language Model Generalization in Low-Resource Extractive QA
Fuente:
arXiv
Guardado en:
| Autores principales: | Sengupta, Saptarshi, Yin, Wenpeng, Nakov, Preslav, Ghosh, Shreya, Wang, Suhang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TOP-Training: Target-Oriented Pretraining for Medical Extractive Question Answering
por: Sengupta, Saptarshi, et al.
Publicado: (2023)
por: Sengupta, Saptarshi, et al.
Publicado: (2023)
Can Hallucinations Be Useful? Solving Multi-Hop Questions With SLMs By Chaining System-I/II Reasoning
por: Sengupta, Saptarshi, et al.
Publicado: (2026)
por: Sengupta, Saptarshi, et al.
Publicado: (2026)
Rethinking STS and NLI in Large Language Models
por: Wang, Yuxia, et al.
Publicado: (2023)
por: Wang, Yuxia, et al.
Publicado: (2023)
Large Language Models are Few-Shot Training Example Generators: A Case Study in Fallacy Recognition
por: Alhindi, Tariq, et al.
Publicado: (2023)
por: Alhindi, Tariq, et al.
Publicado: (2023)
Toxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource Languages
por: Hu, Yujia, et al.
Publicado: (2025)
por: Hu, Yujia, et al.
Publicado: (2025)
From Multiple-Choice to Extractive QA: A Case Study for English and Arabic
por: Lynn, Teresa, et al.
Publicado: (2024)
por: Lynn, Teresa, et al.
Publicado: (2024)
Milestones in Bengali Sentiment Analysis leveraging Transformer-models: Fundamentals, Challenges and Future Directions
por: Sengupta, Saptarshi, et al.
Publicado: (2024)
por: Sengupta, Saptarshi, et al.
Publicado: (2024)
DUAL-Bench: Measuring Over-Refusal and Robustness in Vision-Language Models
por: Ren, Kaixuan, et al.
Publicado: (2025)
por: Ren, Kaixuan, et al.
Publicado: (2025)
Adapting Fake News Detection to the Era of Large Language Models
por: Su, Jinyan, et al.
Publicado: (2023)
por: Su, Jinyan, et al.
Publicado: (2023)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
por: Laiyk, Nurkhan, et al.
Publicado: (2025)
por: Laiyk, Nurkhan, et al.
Publicado: (2025)
CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings
por: Orel, Daniil, et al.
Publicado: (2025)
por: Orel, Daniil, et al.
Publicado: (2025)
Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models
por: Rubashevskii, Aleksandr, et al.
Publicado: (2026)
por: Rubashevskii, Aleksandr, et al.
Publicado: (2026)
UNCERTAINTY-LINE: Length-Invariant Estimation of Uncertainty for Large Language Models
por: Vashurin, Roman, et al.
Publicado: (2025)
por: Vashurin, Roman, et al.
Publicado: (2025)
How Does Prefix Matter in Reasoning Model Tuning?
por: Tomar, Raj Vardhan, et al.
Publicado: (2026)
por: Tomar, Raj Vardhan, et al.
Publicado: (2026)
BioMol-MQA: A Multi-Modal Question Answering Dataset For LLM Reasoning Over Bio-Molecular Interactions
por: Sengupta, Saptarshi, et al.
Publicado: (2025)
por: Sengupta, Saptarshi, et al.
Publicado: (2025)
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
por: Tomar, Raj Vardhan, et al.
Publicado: (2025)
por: Tomar, Raj Vardhan, et al.
Publicado: (2025)
ConspirED: A Dataset for Cognitive Traits of Conspiracy Theories and Large Language Model Safety
por: Bates, Luke, et al.
Publicado: (2025)
por: Bates, Luke, et al.
Publicado: (2025)
Multimodal Large Language Models to Support Real-World Fact-Checking
por: Geng, Jiahui, et al.
Publicado: (2024)
por: Geng, Jiahui, et al.
Publicado: (2024)
Applicability of Large Language Models and Generative Models for Legal Case Judgement Summarization
por: Deroy, Aniket, et al.
Publicado: (2024)
por: Deroy, Aniket, et al.
Publicado: (2024)
Atlas-Chat: Adapting Large Language Models for Low-Resource Moroccan Arabic Dialect
por: Shang, Guokan, et al.
Publicado: (2024)
por: Shang, Guokan, et al.
Publicado: (2024)
Exploring the Limitations of Detecting Machine-Generated Text
por: Doughman, Jad, et al.
Publicado: (2024)
por: Doughman, Jad, et al.
Publicado: (2024)
Shaping the Safety Boundaries: Understanding and Defending Against Jailbreaks in Large Language Models
por: Gao, Lang, et al.
Publicado: (2024)
por: Gao, Lang, et al.
Publicado: (2024)
COMMUNITYNOTES: A Dataset for Exploring the Helpfulness of Fact-Checking Explanations
por: Xing, Rui, et al.
Publicado: (2025)
por: Xing, Rui, et al.
Publicado: (2025)
Exploring Concreteness Through a Figurative Lens
por: Ghosh, Saptarshi, et al.
Publicado: (2026)
por: Ghosh, Saptarshi, et al.
Publicado: (2026)
DP-Fusion: Token-Level Differentially Private Inference for Large Language Models
por: Thareja, Rushil, et al.
Publicado: (2025)
por: Thareja, Rushil, et al.
Publicado: (2025)
State Space Models for Extractive Summarization in Low Resource Scenarios
por: Khayi, Nisrine Ait
Publicado: (2025)
por: Khayi, Nisrine Ait
Publicado: (2025)
Generating Zero-shot Abstractive Explanations for Rumour Verification
por: Bilal, Iman Munire, et al.
Publicado: (2024)
por: Bilal, Iman Munire, et al.
Publicado: (2024)
Utilising Large Language Models for Generating Effective Counter Arguments to Anti-Vaccine Tweets
por: Dhanuka, Utsav, et al.
Publicado: (2025)
por: Dhanuka, Utsav, et al.
Publicado: (2025)
Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation
por: Mukherjee, Arka, et al.
Publicado: (2025)
por: Mukherjee, Arka, et al.
Publicado: (2025)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
por: Elozeiri, Kareem, et al.
Publicado: (2025)
por: Elozeiri, Kareem, et al.
Publicado: (2025)
From Chaos to Clarity: Claim Normalization to Empower Fact-Checking
por: Sundriyal, Megha, et al.
Publicado: (2023)
por: Sundriyal, Megha, et al.
Publicado: (2023)
A Survey of Confidence Estimation and Calibration in Large Language Models
por: Geng, Jiahui, et al.
Publicado: (2023)
por: Geng, Jiahui, et al.
Publicado: (2023)
The Cylindrical Representation Hypothesis for Language Model Steering
por: Gao, Lang, et al.
Publicado: (2026)
por: Gao, Lang, et al.
Publicado: (2026)
Can Machines Resonate with Humans? Evaluating the Emotional and Empathic Comprehension of LMs
por: Manzoor, Muhammad Arslan, et al.
Publicado: (2024)
por: Manzoor, Muhammad Arslan, et al.
Publicado: (2024)
Con Instruction: Universal Jailbreaking of Multimodal Large Language Models via Non-Textual Modalities
por: Geng, Jiahui, et al.
Publicado: (2025)
por: Geng, Jiahui, et al.
Publicado: (2025)
Efficient Extractive Summarization with MAMBA-Transformer Hybrids for Low-Resource Scenarios
por: Khayi, Nisrine Ait
Publicado: (2026)
por: Khayi, Nisrine Ait
Publicado: (2026)
ExaGPT: Example-Based Machine-Generated Text Detection for Human Interpretability
por: Koike, Ryuto, et al.
Publicado: (2025)
por: Koike, Ryuto, et al.
Publicado: (2025)
ToolDreamer: Instilling LLM Reasoning Into Tool Retrievers
por: Sengupta, Saptarshi, et al.
Publicado: (2025)
por: Sengupta, Saptarshi, et al.
Publicado: (2025)
Missci: Reconstructing Fallacies in Misrepresented Science
por: Glockner, Max, et al.
Publicado: (2024)
por: Glockner, Max, et al.
Publicado: (2024)
Grounding Fallacies Misrepresenting Scientific Publications in Evidence
por: Glockner, Max, et al.
Publicado: (2024)
por: Glockner, Max, et al.
Publicado: (2024)
Ejemplares similares
-
TOP-Training: Target-Oriented Pretraining for Medical Extractive Question Answering
por: Sengupta, Saptarshi, et al.
Publicado: (2023) -
Can Hallucinations Be Useful? Solving Multi-Hop Questions With SLMs By Chaining System-I/II Reasoning
por: Sengupta, Saptarshi, et al.
Publicado: (2026) -
Rethinking STS and NLI in Large Language Models
por: Wang, Yuxia, et al.
Publicado: (2023) -
Large Language Models are Few-Shot Training Example Generators: A Case Study in Fallacy Recognition
por: Alhindi, Tariq, et al.
Publicado: (2023) -
Toxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource Languages
por: Hu, Yujia, et al.
Publicado: (2025)