InfFeed: Influence Functions as a Feedback to Improve the Performance of Subjective Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Banerjee, Somnath, Sarkar, Maulindu, Saha, Punyajoy, Mathew, Binny, Mukherjee, Animesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
di: Saha, Punyajoy, et al.
Pubblicazione: (2024)
di: Saha, Punyajoy, et al.
Pubblicazione: (2024)
On Zero-Shot Counterspeech Generation by LLMs
di: Saha, Punyajoy, et al.
Pubblicazione: (2024)
di: Saha, Punyajoy, et al.
Pubblicazione: (2024)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
di: Banerjee, Somnath, et al.
Pubblicazione: (2026)
di: Banerjee, Somnath, et al.
Pubblicazione: (2026)
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
di: Das, Mithun, et al.
Pubblicazione: (2024)
di: Das, Mithun, et al.
Pubblicazione: (2024)
Evaluating the Ebb and Flow: An In-depth Analysis of Question-Answering Trends across Diverse Platforms
di: Hazra, Rima, et al.
Pubblicazione: (2023)
di: Hazra, Rima, et al.
Pubblicazione: (2023)
Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations
di: Banerjee, Somnath, et al.
Pubblicazione: (2026)
di: Banerjee, Somnath, et al.
Pubblicazione: (2026)
Soteria: Language-Specific Functional Parameter Steering for Multilingual Safety Alignment
di: Banerjee, Somnath, et al.
Pubblicazione: (2025)
di: Banerjee, Somnath, et al.
Pubblicazione: (2025)
How (un)ethical are instruction-centric responses of LLMs? Unveiling the vulnerabilities of safety guardrails to harmful queries
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
Exploring the Limits of Zero Shot Vision Language Models for Hate Meme Detection: The Vulnerabilities and their Interpretations
di: Rizwan, Naquee, et al.
Pubblicazione: (2024)
di: Rizwan, Naquee, et al.
Pubblicazione: (2024)
DistALANER: Distantly Supervised Active Learning Augmented Named Entity Recognition in the Open Source Software Ecosystem
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
Context Matters: Pushing the Boundaries of Open-Ended Answer Generation with Graph-Structured Knowledge Context
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
AURA: Affordance-Understanding and Risk-aware Alignment Technique for Large Language Models
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
ProSocialAlign: Preference Conditioned Test Time Alignment in Language Models
di: Banerjee, Somnath, et al.
Pubblicazione: (2025)
di: Banerjee, Somnath, et al.
Pubblicazione: (2025)
SafeMath: Inference-time Safety improves Math Accuracy
di: Basu, Sagnik, et al.
Pubblicazione: (2026)
di: Basu, Sagnik, et al.
Pubblicazione: (2026)
RA-MTR: A Retrieval Augmented Multi-Task Reader based Approach for Inspirational Quote Extraction from Long Documents
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
di: Nag, Arijit, et al.
Pubblicazione: (2024)
di: Nag, Arijit, et al.
Pubblicazione: (2024)
Tutoring Large Language Models to be Domain-adaptive, Precise, and Safe
di: Banerjee, Somnath
Pubblicazione: (2026)
di: Banerjee, Somnath
Pubblicazione: (2026)
MemeSense: An Adaptive In-Context Framework for Social Commonsense Driven Meme Moderation
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
di: Adak, Sayantan, et al.
Pubblicazione: (2025)
Attributional Safety Failures in Large Language Models under Code-Mixed Perturbations
di: Banerjee, Somnath, et al.
Pubblicazione: (2025)
di: Banerjee, Somnath, et al.
Pubblicazione: (2025)
Can Safety Emerge from Weak Supervision? A Systematic Analysis of Small Language Models
di: Saha, Punyajoy, et al.
Pubblicazione: (2026)
di: Saha, Punyajoy, et al.
Pubblicazione: (2026)
Demarked: A Strategy for Enhanced Abusive Speech Moderation through Counterspeech, Detoxification, and Message Management
di: Yimam, Seid Muhie, et al.
Pubblicazione: (2024)
di: Yimam, Seid Muhie, et al.
Pubblicazione: (2024)
HatePRISM: Policies, Platforms, and Research Integration. Advancing NLP for Hate Speech Proactive Mitigation
di: Rizwan, Naquee, et al.
Pubblicazione: (2025)
di: Rizwan, Naquee, et al.
Pubblicazione: (2025)
Navigating the Cultural Kaleidoscope: A Hitchhiker's Guide to Sensitivity in Large Language Models
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
di: Banerjee, Somnath, et al.
Pubblicazione: (2024)
STEMTOX: From Social Tags to Fine-Grained Toxic Meme Detection via Entropy-Guided Multi-Task Learning
di: Swain, Subhankar, et al.
Pubblicazione: (2025)
di: Swain, Subhankar, et al.
Pubblicazione: (2025)
Sowing the Wind, Reaping the Whirlwind: The Impact of Editing Language Models
di: Hazra, Rima, et al.
Pubblicazione: (2024)
di: Hazra, Rima, et al.
Pubblicazione: (2024)
Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations
di: Hazra, Rima, et al.
Pubblicazione: (2024)
di: Hazra, Rima, et al.
Pubblicazione: (2024)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
di: Adak, Sayantan, et al.
Pubblicazione: (2024)
di: Adak, Sayantan, et al.
Pubblicazione: (2024)
Retriv at BLP-2025 Task 2: Test-Driven Feedback-Guided Framework for Bangla-to-Python Code Generation
di: Asib, K M Nafi, et al.
Pubblicazione: (2025)
di: Asib, K M Nafi, et al.
Pubblicazione: (2025)
A Community-Centric Perspective for Characterizing and Detecting Anti-Asian Violence-Provoking Speech
di: Verma, Gaurav, et al.
Pubblicazione: (2024)
di: Verma, Gaurav, et al.
Pubblicazione: (2024)
Lost without translation -- Can transformer (language models) understand mood states?
di: Shivaprakash, Prakrithi, et al.
Pubblicazione: (2025)
di: Shivaprakash, Prakrithi, et al.
Pubblicazione: (2025)
Analyzing Sentiment Polarity Reduction in News Presentation through Contextual Perturbation and Large Language Models
di: Kuila, Alapan, et al.
Pubblicazione: (2024)
di: Kuila, Alapan, et al.
Pubblicazione: (2024)
FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection
di: Bhaskar, Paramananda, et al.
Pubblicazione: (2026)
di: Bhaskar, Paramananda, et al.
Pubblicazione: (2026)
Exploring the Performance of Large Language Models on Subjective Span Identification Tasks
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2026)
di: Dmonte, Alphaeus, et al.
Pubblicazione: (2026)
Duplicate Question Retrieval and Confirmation Time Prediction in Software Communities
di: Hazra, Rima, et al.
Pubblicazione: (2023)
di: Hazra, Rima, et al.
Pubblicazione: (2023)
Improving Multi-turn Task Completion in Task-Oriented Dialog Systems via Prompt Chaining and Fine-Grained Feedback
di: Fereidouni, Moghis, et al.
Pubblicazione: (2025)
di: Fereidouni, Moghis, et al.
Pubblicazione: (2025)
FeedEval: Pedagogically Aligned Evaluation of LLM-Generated Essay Feedback
di: Chu, Seongyeub, et al.
Pubblicazione: (2026)
di: Chu, Seongyeub, et al.
Pubblicazione: (2026)
Efficient Continual Pre-training of LLMs for Low-resource Languages
di: Nag, Arijit, et al.
Pubblicazione: (2024)
di: Nag, Arijit, et al.
Pubblicazione: (2024)
InfMem: Learning System-2 Memory Control for Long-Context Agent
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
di: Saha, Punyajoy, et al.
Pubblicazione: (2024) -
On Zero-Shot Counterspeech Generation by LLMs
di: Saha, Punyajoy, et al.
Pubblicazione: (2024) -
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
di: Banerjee, Somnath, et al.
Pubblicazione: (2026) -
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
di: Das, Mithun, et al.
Pubblicazione: (2024) -
Evaluating the Ebb and Flow: An In-depth Analysis of Question-Answering Trends across Diverse Platforms
di: Hazra, Rima, et al.
Pubblicazione: (2023)