OrgAccess: A Benchmark for Role Based Access Control in Organization Scale LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Sanyal, Debdeep, Maharana, Umakanta, Sinha, Yash, Tan, Hong Ming, Karande, Shirish, Kankanhalli, Mohan, Mandal, Murari |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
UnStar: Unlearning with Self-Taught Anti-Sample Reasoning for LLMs
di: Sinha, Yash, et al.
Pubblicazione: (2024)
di: Sinha, Yash, et al.
Pubblicazione: (2024)
Nine Ways to Break Copyright Law and Why Our LLM Won't: A Fair Use Aligned Generation Framework
di: Sharma, Aakash Sen, et al.
Pubblicazione: (2025)
di: Sharma, Aakash Sen, et al.
Pubblicazione: (2025)
The Realignment Problem: When Right becomes Wrong in LLMs
di: Sharma, Aakash Sen, et al.
Pubblicazione: (2025)
di: Sharma, Aakash Sen, et al.
Pubblicazione: (2025)
Distill to Delete: Unlearning in Graph Networks with Knowledge Distillation
di: Sinha, Yash, et al.
Pubblicazione: (2023)
di: Sinha, Yash, et al.
Pubblicazione: (2023)
Multi-Modal Recommendation Unlearning for Legal, Licensing, and Modality Constraints
di: Sinha, Yash, et al.
Pubblicazione: (2024)
di: Sinha, Yash, et al.
Pubblicazione: (2024)
Agents Are All You Need for LLM Unlearning
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
AntiDote: Bi-level Adversarial Training for Tamper-Resistant LLMs
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
Investigating Pedagogical Teacher and Student LLM Agents: Genetic Adaptation Meets Retrieval Augmented Generation Across Learning Style
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
Guardians of Generation: Dynamic Inference-Time Copyright Shielding with Adaptive Guidance for AI Image Generation
di: Roy, Soham, et al.
Pubblicazione: (2025)
di: Roy, Soham, et al.
Pubblicazione: (2025)
Forgetting is Competition: Rethinking Unlearning as Representation Interference in Diffusion Models
di: Ranjan, Ashutosh, et al.
Pubblicazione: (2026)
di: Ranjan, Ashutosh, et al.
Pubblicazione: (2026)
Step-by-Step Reasoning Attack: Revealing 'Erased' Knowledge in Large Language Models
di: Sinha, Yash, et al.
Pubblicazione: (2025)
di: Sinha, Yash, et al.
Pubblicazione: (2025)
Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensions
di: Patil, Parth, et al.
Pubblicazione: (2026)
di: Patil, Parth, et al.
Pubblicazione: (2026)
Confidence is Not Competence
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
time2time: Causal Intervention in Hidden States to Simulate Rare Events in Time Series Foundation Models
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
EcoVal: An Efficient Data Valuation Framework for Machine Learning
di: Tarun, Ayush K, et al.
Pubblicazione: (2024)
di: Tarun, Ayush K, et al.
Pubblicazione: (2024)
Persuasion Games using Large Language Models
di: Ramani, Ganesh Prasath, et al.
Pubblicazione: (2024)
di: Ramani, Ganesh Prasath, et al.
Pubblicazione: (2024)
Policy Optimization Prefers The Path of Least Resistance
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
di: Sanyal, Debdeep, et al.
Pubblicazione: (2025)
Right Prediction, Wrong Reasoning: Uncovering LLM Misalignment in RA Disease Diagnosis
di: Maharana, Umakanta, et al.
Pubblicazione: (2025)
di: Maharana, Umakanta, et al.
Pubblicazione: (2025)
Polaris: A Gödel Agent Framework for Small Language Models through Experience-Abstracted Policy Repair
di: Kakade, Aditya, et al.
Pubblicazione: (2026)
di: Kakade, Aditya, et al.
Pubblicazione: (2026)
Towards Improving NAM-to-Speech Synthesis Intelligibility using Self-Supervised Speech Models
di: Shah, Neil, et al.
Pubblicazione: (2024)
di: Shah, Neil, et al.
Pubblicazione: (2024)
Advancing NAM-to-Speech Conversion with Novel Methods and the MultiNAM Dataset
di: Shah, Neil, et al.
Pubblicazione: (2024)
di: Shah, Neil, et al.
Pubblicazione: (2024)
I Know Therefore I Score: Label-Free Crafting of Scoring Functions using Constraints Based on Domain Expertise
di: Palakkadavath, Ragja, et al.
Pubblicazione: (2022)
di: Palakkadavath, Ragja, et al.
Pubblicazione: (2022)
VidHal: Benchmarking Temporal Hallucinations in Vision LLMs
di: Choong, Wey Yeh, et al.
Pubblicazione: (2024)
di: Choong, Wey Yeh, et al.
Pubblicazione: (2024)
MRI2Speech: Speech Synthesis from Articulatory Movements Recorded by Real-time MRI
di: Shah, Neil, et al.
Pubblicazione: (2024)
di: Shah, Neil, et al.
Pubblicazione: (2024)
The Silent Brush: Evaluating Artistic Style Leakage in AI Art Generation
di: Joshi, Ninad, et al.
Pubblicazione: (2026)
di: Joshi, Ninad, et al.
Pubblicazione: (2026)
Measuring Representation Robustness in Large Language Models for Geometry
di: Jawandhia, Vedant, et al.
Pubblicazione: (2026)
di: Jawandhia, Vedant, et al.
Pubblicazione: (2026)
CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics
di: Agarwal, Parth, et al.
Pubblicazione: (2025)
di: Agarwal, Parth, et al.
Pubblicazione: (2025)
Reasoning LLMs are Wandering Solution Explorers
di: Lu, Jiahao, et al.
Pubblicazione: (2025)
di: Lu, Jiahao, et al.
Pubblicazione: (2025)
How to Trick Your AI TA: A Systematic Study of Academic Jailbreaking in LLM Code Evaluation
di: Sahoo, Devanshu, et al.
Pubblicazione: (2025)
di: Sahoo, Devanshu, et al.
Pubblicazione: (2025)
"I Strongly Suspect This Website Is a Scam": Benchmarking PII Leakage and Detection without Defense in Autonomous Web Agents
di: Roy, Soham, et al.
Pubblicazione: (2026)
di: Roy, Soham, et al.
Pubblicazione: (2026)
Role-Aware Language Models for Secure and Contextualized Access Control in Organizations
di: Almheiri, Saeed, et al.
Pubblicazione: (2025)
di: Almheiri, Saeed, et al.
Pubblicazione: (2025)
When Reject Turns into Accept: Quantifying the Vulnerability of LLM-Based Scientific Reviewers to Indirect Prompt Injection
di: Sahoo, Devanshu, et al.
Pubblicazione: (2025)
di: Sahoo, Devanshu, et al.
Pubblicazione: (2025)
Strong Preferences Affect the Robustness of Preference Models and Value Alignment
di: Xu, Ziwei, et al.
Pubblicazione: (2024)
di: Xu, Ziwei, et al.
Pubblicazione: (2024)
DPTraj-PM: Differentially Private Trajectory Synthesis Using Prefix Tree and Markov Process
di: Wang, Nana, et al.
Pubblicazione: (2024)
di: Wang, Nana, et al.
Pubblicazione: (2024)
SCAN: Bootstrapping Contrastive Pre-training for Data Efficiency
di: Guo, Yangyang, et al.
Pubblicazione: (2024)
di: Guo, Yangyang, et al.
Pubblicazione: (2024)
Aggregating Diverse Cue Experts for AI-Generated Image Detection
di: Tan, Lei, et al.
Pubblicazione: (2026)
di: Tan, Lei, et al.
Pubblicazione: (2026)
The Compliance Paradox: Semantic-Instruction Decoupling in Automated Academic Code Evaluation
di: Sahoo, Devanshu, et al.
Pubblicazione: (2026)
di: Sahoo, Devanshu, et al.
Pubblicazione: (2026)
ARBITER: AI-Driven Filtering for Role-Based Access Control
di: Lorenzo, Michele, et al.
Pubblicazione: (2025)
di: Lorenzo, Michele, et al.
Pubblicazione: (2025)
LLMs Can Unlearn Refusal with Only 1,000 Benign Samples
di: Guo, Yangyang, et al.
Pubblicazione: (2026)
di: Guo, Yangyang, et al.
Pubblicazione: (2026)
Improving Access to Physical Activity Resources Through Multimedia Health Promotion
di: Tanvee Sinha
Pubblicazione: (2025)
di: Tanvee Sinha
Pubblicazione: (2025)
Documenti analoghi
-
UnStar: Unlearning with Self-Taught Anti-Sample Reasoning for LLMs
di: Sinha, Yash, et al.
Pubblicazione: (2024) -
Nine Ways to Break Copyright Law and Why Our LLM Won't: A Fair Use Aligned Generation Framework
di: Sharma, Aakash Sen, et al.
Pubblicazione: (2025) -
The Realignment Problem: When Right becomes Wrong in LLMs
di: Sharma, Aakash Sen, et al.
Pubblicazione: (2025) -
Distill to Delete: Unlearning in Graph Networks with Knowledge Distillation
di: Sinha, Yash, et al.
Pubblicazione: (2023) -
Multi-Modal Recommendation Unlearning for Legal, Licensing, and Modality Constraints
di: Sinha, Yash, et al.
Pubblicazione: (2024)