SMAB: MAB based word Sensitivity Estimation Framework and its Applications in Adversarial Text Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Pandey, Saurabh Kumar, Vashistha, Sachin, Das, Debrup, Aditya, Somak, Choudhury, Monojit |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tricking LLMs into Disobedience: Formalizing, Analyzing, and Detecting Jailbreaks
di: Rao, Abhinav, et al.
Pubblicazione: (2023)
di: Rao, Abhinav, et al.
Pubblicazione: (2023)
[WIP] Jailbreak Paradox: The Achilles' Heel of LLMs
di: Rao, Abhinav, et al.
Pubblicazione: (2024)
di: Rao, Abhinav, et al.
Pubblicazione: (2024)
MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning
di: Das, Debrup, et al.
Pubblicazione: (2024)
di: Das, Debrup, et al.
Pubblicazione: (2024)
To Generate or Discriminate? Methodological Considerations for Measuring Cultural Alignment in LLMs
di: Pandey, Saurabh Kumar, et al.
Pubblicazione: (2026)
di: Pandey, Saurabh Kumar, et al.
Pubblicazione: (2026)
Meta-Cultural Competence: Climbing the Right Hill of Cultural Awareness
di: Saha, Sougata, et al.
Pubblicazione: (2025)
di: Saha, Sougata, et al.
Pubblicazione: (2025)
PragWorld: A Benchmark Evaluating LLMs' Local World Model under Minimal Linguistic Alterations and Conversational Dynamics
di: Vashistha, Sachin, et al.
Pubblicazione: (2025)
di: Vashistha, Sachin, et al.
Pubblicazione: (2025)
Reading between the Lines: Can LLMs Identify Cross-Cultural Communication Gaps?
di: Saha, Sougata, et al.
Pubblicazione: (2025)
di: Saha, Sougata, et al.
Pubblicazione: (2025)
User Behavior Prediction as a Generic, Robust, Scalable, and Low-Cost Evaluation Strategy for Estimating Generalization in LLMs
di: Saha, Sougata, et al.
Pubblicazione: (2025)
di: Saha, Sougata, et al.
Pubblicazione: (2025)
All that is English may be Hindi: Enhancing language identification through automatic ranking of likeliness of word borrowing in social media
di: Patro, Jasabanta, et al.
Pubblicazione: (2017)
di: Patro, Jasabanta, et al.
Pubblicazione: (2017)
How Deep Is Representational Bias in LLMs? The Cases of Caste and Religion
di: Seth, Agrima, et al.
Pubblicazione: (2025)
di: Seth, Agrima, et al.
Pubblicazione: (2025)
Missing Melodies: AI Music Generation and its "Nearly" Complete Omission of the Global South
di: Mehta, Atharva, et al.
Pubblicazione: (2024)
di: Mehta, Atharva, et al.
Pubblicazione: (2024)
EduVidQA: Generating and Evaluating Long-form Answers to Student Questions based on Lecture Videos
di: Ray, Sourjyadip, et al.
Pubblicazione: (2025)
di: Ray, Sourjyadip, et al.
Pubblicazione: (2025)
RaDeR: Reasoning-aware Dense Retrieval Models
di: Das, Debrup, et al.
Pubblicazione: (2025)
di: Das, Debrup, et al.
Pubblicazione: (2025)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
di: Adak, Sayantan, et al.
Pubblicazione: (2024)
di: Adak, Sayantan, et al.
Pubblicazione: (2024)
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
di: Dutta, Aritra, et al.
Pubblicazione: (2026)
di: Dutta, Aritra, et al.
Pubblicazione: (2026)
Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs
di: Puerto, Haritz, et al.
Pubblicazione: (2024)
di: Puerto, Haritz, et al.
Pubblicazione: (2024)
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test
di: Khandelwal, Aditi, et al.
Pubblicazione: (2024)
di: Khandelwal, Aditi, et al.
Pubblicazione: (2024)
Ethical Reasoning and Moral Value Alignment of LLMs Depend on the Language we Prompt them in
di: Agarwal, Utkarsh, et al.
Pubblicazione: (2024)
di: Agarwal, Utkarsh, et al.
Pubblicazione: (2024)
REFINE-AF: A Task-Agnostic Framework to Align Language Models via Self-Generated Instructions using Reinforcement Learning from Automated Feedback
di: Roy, Aniruddha, et al.
Pubblicazione: (2025)
di: Roy, Aniruddha, et al.
Pubblicazione: (2025)
Think Outside the Data: Colonial Biases and Systemic Issues in Automated Moderation Pipelines for Low-Resource Languages
di: Shahid, Farhana, et al.
Pubblicazione: (2025)
di: Shahid, Farhana, et al.
Pubblicazione: (2025)
NLKI: A lightweight Natural Language Knowledge Integration Framework for Improving Small VLMs in Commonsense VQA Tasks
di: Dutta, Aritra, et al.
Pubblicazione: (2025)
di: Dutta, Aritra, et al.
Pubblicazione: (2025)
Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models
di: Mittal, Avni, et al.
Pubblicazione: (2026)
di: Mittal, Avni, et al.
Pubblicazione: (2026)
Exploring Adapter Design Tradeoffs for Low Resource Music Generation
di: Mehta, Atharva, et al.
Pubblicazione: (2025)
di: Mehta, Atharva, et al.
Pubblicazione: (2025)
Evaluating Large Language Models for Health-related Queries with Presuppositions
di: Kaur, Navreet, et al.
Pubblicazione: (2023)
di: Kaur, Navreet, et al.
Pubblicazione: (2023)
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
di: Das, Mithun, et al.
Pubblicazione: (2024)
di: Das, Mithun, et al.
Pubblicazione: (2024)
Sacred or Synthetic? Evaluating LLM Reliability and Abstention for Religious Questions
di: Atif, Farah, et al.
Pubblicazione: (2025)
di: Atif, Farah, et al.
Pubblicazione: (2025)
Women, Infamous, and Exotic Beings: A Comparative Study of Honorific Usages in Wikipedia and LLMs for Bengali and Hindi
di: Mukherjee, Sourabrata, et al.
Pubblicazione: (2025)
di: Mukherjee, Sourabrata, et al.
Pubblicazione: (2025)
Lost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models
di: Atwany, Hanin, et al.
Pubblicazione: (2025)
di: Atwany, Hanin, et al.
Pubblicazione: (2025)
Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI
di: Singh, Saurabh K., et al.
Pubblicazione: (2026)
di: Singh, Saurabh K., et al.
Pubblicazione: (2026)
ABLEIST: Intersectional Disability Bias in LLM-Generated Hiring Scenarios
di: Phutane, Mahika, et al.
Pubblicazione: (2025)
di: Phutane, Mahika, et al.
Pubblicazione: (2025)
Fluent but Foreign: Even Regional LLMs Lack Cultural Alignment
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
The Zeno's Paradox of `Low-Resource' Languages
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024)
di: Nigatu, Hellina Hailu, et al.
Pubblicazione: (2024)
"They are uncultured": Unveiling Covert Harms and Social Threats in LLM Generated Conversations
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2024)
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2024)
Are Large Language Model-based Evaluators the Solution to Scaling Up Multilingual Evaluation?
di: Hada, Rishav, et al.
Pubblicazione: (2023)
di: Hada, Rishav, et al.
Pubblicazione: (2023)
Factual and Edit-Sensitive Graph-to-Sequence Generation via Graph-Aware Adaptive Noising
di: Shahane, Aditya Hemant, et al.
Pubblicazione: (2026)
di: Shahane, Aditya Hemant, et al.
Pubblicazione: (2026)
MAG-V: A Multi-Agent Framework for Synthetic Data Generation and Verification
di: Sengupta, Saptarshi, et al.
Pubblicazione: (2024)
di: Sengupta, Saptarshi, et al.
Pubblicazione: (2024)
CAMF: Collaborative Adversarial Multi-agent Framework for Machine Generated Text Detection
di: Wang, Yue, et al.
Pubblicazione: (2025)
di: Wang, Yue, et al.
Pubblicazione: (2025)
AURA: Affordance-Understanding and Risk-aware Alignment Technique for Large Language Models
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
di: Hong, Pengfei, et al.
Pubblicazione: (2024)
di: Hong, Pengfei, et al.
Pubblicazione: (2024)
READ: Reinforcement-based Adversarial Learning for Text Classification with Limited Labeled Data
di: Sharma, Rohit, et al.
Pubblicazione: (2025)
di: Sharma, Rohit, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Tricking LLMs into Disobedience: Formalizing, Analyzing, and Detecting Jailbreaks
di: Rao, Abhinav, et al.
Pubblicazione: (2023) -
[WIP] Jailbreak Paradox: The Achilles' Heel of LLMs
di: Rao, Abhinav, et al.
Pubblicazione: (2024) -
MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning
di: Das, Debrup, et al.
Pubblicazione: (2024) -
To Generate or Discriminate? Methodological Considerations for Measuring Cultural Alignment in LLMs
di: Pandey, Saurabh Kumar, et al.
Pubblicazione: (2026) -
Meta-Cultural Competence: Climbing the Right Hill of Cultural Awareness
di: Saha, Sougata, et al.
Pubblicazione: (2025)