Gespeichert in:
| Hauptverfasser: | Panda, Srikant, Yadav, Sourabh Singh, Malviya, Palkesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.17348 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Who Gets Left Behind? Auditing Disability Inclusivity in Large Language Models
von: Dash, Deepika, et al.
Veröffentlicht: (2025)
von: Dash, Deepika, et al.
Veröffentlicht: (2025)
AccessEval: Benchmarking Disability Bias in Large Language Models
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
Say It Differently: Linguistic Styles as Jailbreak Vectors
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
DAIQ: Auditing Demographic Attribute Inference from Question in LLMs
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)
Privacy Auditing of Large Language Models
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025)
von: Panda, Ashwinee, et al.
Veröffentlicht: (2025)
Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2025)
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2025)
FS-DAG: Few Shot Domain Adapting Graph Networks for Visually Rich Document Understanding
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
How Does Thinking Mode Change LLM Moral Judgments? A Controlled Instant-vs-Thinking Comparison Across Five Frontier Models
von: Madur, Sai Sourabh
Veröffentlicht: (2026)
von: Madur, Sai Sourabh
Veröffentlicht: (2026)
Judging by Appearances? Auditing and Intervening Vision-Language Models for Bail Prediction
von: Basu, Sagnik, et al.
Veröffentlicht: (2025)
von: Basu, Sagnik, et al.
Veröffentlicht: (2025)
FinVerBench: Benchmark Validity and Calibration in Large Language Model Financial Statement Verification
von: Panda, Silu
Veröffentlicht: (2026)
von: Panda, Silu
Veröffentlicht: (2026)
Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores
von: Panda, Shevya, et al.
Veröffentlicht: (2026)
von: Panda, Shevya, et al.
Veröffentlicht: (2026)
Clinical QA 2.0: Multi-Task Learning for Answer Extraction and Categorization
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2025)
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2025)
Survey of Large Multimodal Model Datasets, Application Categories and Taxonomy
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2024)
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2024)
AuditLLM: A Tool for Auditing Large Language Models Using Multiprobe Approach
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
Auditing Frontier Vision-Language Models for Trustworthy Medical VQA: Grounding Failures, Format Collapse, and Domain Adaptation
von: Chen, Xupeng, et al.
Veröffentlicht: (2026)
von: Chen, Xupeng, et al.
Veröffentlicht: (2026)
AuditWen:An Open-Source Large Language Model for Audit
von: Huang, Jiajia, et al.
Veröffentlicht: (2024)
von: Huang, Jiajia, et al.
Veröffentlicht: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
von: Anjarlekar, Ameya, et al.
Veröffentlicht: (2025)
von: Anjarlekar, Ameya, et al.
Veröffentlicht: (2025)
Mechanistically Interpreting Compression in Vision-Language Models
von: Elluru, Veeraraju, et al.
Veröffentlicht: (2026)
von: Elluru, Veeraraju, et al.
Veröffentlicht: (2026)
The Geometry of Representational Failures in Vision Language Models
von: Savietto, Daniele, et al.
Veröffentlicht: (2026)
von: Savietto, Daniele, et al.
Veröffentlicht: (2026)
Alignment among Language, Vision and Action Representations
von: Milano, Nicola, et al.
Veröffentlicht: (2026)
von: Milano, Nicola, et al.
Veröffentlicht: (2026)
Lookbehind-SAM: k steps back, 1 step forward
von: Mordido, Gonçalo, et al.
Veröffentlicht: (2023)
von: Mordido, Gonçalo, et al.
Veröffentlicht: (2023)
Representation-Based Data Quality Audits for Audio
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
Auditing Language Model Unlearning via Information Decomposition
von: Goel, Anmol, et al.
Veröffentlicht: (2026)
von: Goel, Anmol, et al.
Veröffentlicht: (2026)
CALM: Curiosity-Driven Auditing for Large Language Models
von: Zheng, Xiang, et al.
Veröffentlicht: (2025)
von: Zheng, Xiang, et al.
Veröffentlicht: (2025)
When Prohibitions Become Permissions: Auditing Negation Sensitivity in Language Models
von: Elkins, Katherine, et al.
Veröffentlicht: (2026)
von: Elkins, Katherine, et al.
Veröffentlicht: (2026)
TRUST: A Decentralized Framework for Auditing Large Language Model Reasoning
von: Huang, Morris Yu-Chao, et al.
Veröffentlicht: (2025)
von: Huang, Morris Yu-Chao, et al.
Veröffentlicht: (2025)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
von: Murthy, Yashaswini, et al.
Veröffentlicht: (2023)
von: Murthy, Yashaswini, et al.
Veröffentlicht: (2023)
Question the Questions: Auditing Representation in Online Deliberative Processes
von: De, Soham, et al.
Veröffentlicht: (2025)
von: De, Soham, et al.
Veröffentlicht: (2025)
Towards Security-Auditable LLM Agents: A Unified Graph Representation
von: Li, Chaofan, et al.
Veröffentlicht: (2026)
von: Li, Chaofan, et al.
Veröffentlicht: (2026)
Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025)
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025)
VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing
von: Shi, Haoyuan, et al.
Veröffentlicht: (2026)
von: Shi, Haoyuan, et al.
Veröffentlicht: (2026)
LLMAuditor: A Framework for Auditing Large Language Models Using Human-in-the-Loop
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
IndicFairFace: Balanced Indian Face Dataset for Auditing and Mitigating Geographical Bias in Vision-Language Models
von: Mohsin, Aarish Shah, et al.
Veröffentlicht: (2026)
von: Mohsin, Aarish Shah, et al.
Veröffentlicht: (2026)
Bypassing the Rationale: Causal Auditing of Implicit Reasoning in Language Models
von: Sathyanarayanan, Anish, et al.
Veröffentlicht: (2026)
von: Sathyanarayanan, Anish, et al.
Veröffentlicht: (2026)
Disability Across Cultures: A Human-Centered Audit of Ableism in Western and Indic LLMs
von: Phutane, Mahika, et al.
Veröffentlicht: (2025)
von: Phutane, Mahika, et al.
Veröffentlicht: (2025)
Output Scouting: Auditing Large Language Models for Catastrophic Responses
von: Bell, Andrew, et al.
Veröffentlicht: (2024)
von: Bell, Andrew, et al.
Veröffentlicht: (2024)
Robust Finetuning of Vision-Language-Action Robot Policies via Parameter Merging
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
GTMA: Dynamic Representation Optimization for OOD Vision-Language Models
von: Zhang, Jensen, et al.
Veröffentlicht: (2025)
von: Zhang, Jensen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Who Gets Left Behind? Auditing Disability Inclusivity in Large Language Models
von: Dash, Deepika, et al.
Veröffentlicht: (2025) -
AccessEval: Benchmarking Disability Bias in Large Language Models
von: Panda, Srikant, et al.
Veröffentlicht: (2025) -
Say It Differently: Linguistic Styles as Jailbreak Vectors
von: Panda, Srikant, et al.
Veröffentlicht: (2025) -
DAIQ: Auditing Demographic Attribute Inference from Question in LLMs
von: Panda, Srikant, et al.
Veröffentlicht: (2025) -
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)