Position: Safety and Fairness in Agentic AI Depend on Interaction Topology, Not on Model Scale or Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bajaj, Tanav Singh, Singh, Nikhil, Anand, Karan, Singh, Eishkaran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
REAMS: Reasoning Enhanced Algorithm for Maths Solving
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025)
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025)
First Train to Generate, then Generate to Train: UnitedSynT5 for Few-Shot NLI
von: Banerjee, Sourav, et al.
Veröffentlicht: (2024)
von: Banerjee, Sourav, et al.
Veröffentlicht: (2024)
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2025)
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2025)
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025)
Matching Ranks Over Probability Yields Truly Deep Safety Alignment
von: Vega, Jason, et al.
Veröffentlicht: (2025)
von: Vega, Jason, et al.
Veröffentlicht: (2025)
Impact of Multimodal and Conversational AI on Learning Outcomes and Experience
von: Taneja, Karan, et al.
Veröffentlicht: (2026)
von: Taneja, Karan, et al.
Veröffentlicht: (2026)
Bias-Aware Agent: Enhancing Fairness in AI-Driven Knowledge Retrieval
von: Singh, Karanbir, et al.
Veröffentlicht: (2025)
von: Singh, Karanbir, et al.
Veröffentlicht: (2025)
Temporal Dependencies in In-Context Learning: The Role of Induction Heads
von: Bajaj, Anooshka, et al.
Veröffentlicht: (2026)
von: Bajaj, Anooshka, et al.
Veröffentlicht: (2026)
Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique
von: Roy, Joyjit, et al.
Veröffentlicht: (2026)
von: Roy, Joyjit, et al.
Veröffentlicht: (2026)
Securing Agentic AI: Threat Modeling and Risk Analysis for Network Monitoring Agentic AI System
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
TRAP: Targeted Redirecting of Agentic Preferences
von: Kang, Hangoo, et al.
Veröffentlicht: (2025)
von: Kang, Hangoo, et al.
Veröffentlicht: (2025)
Operationalizing Fairness: Post-Hoc Threshold Optimization Under Hard Resource Limits
von: Singh, Moirangthem Tiken, et al.
Veröffentlicht: (2026)
von: Singh, Moirangthem Tiken, et al.
Veröffentlicht: (2026)
LLM Agents Are Hypersensitive to Nudges
von: Cherep, Manuel, et al.
Veröffentlicht: (2025)
von: Cherep, Manuel, et al.
Veröffentlicht: (2025)
Towards a Multimodal Document-grounded Conversational AI System for Education
von: Taneja, Karan, et al.
Veröffentlicht: (2025)
von: Taneja, Karan, et al.
Veröffentlicht: (2025)
AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation
von: Roy, Joyjit, et al.
Veröffentlicht: (2026)
von: Roy, Joyjit, et al.
Veröffentlicht: (2026)
A Comprehensive Survey of AI-Driven Advancements and Techniques in Automated Program Repair and Code Generation
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Memoria: A Scalable Agentic Memory Framework for Personalized Conversational AI
von: Sarin, Samarth, et al.
Veröffentlicht: (2025)
von: Sarin, Samarth, et al.
Veröffentlicht: (2025)
Agentic AI in Healthcare & Medicine: A Seven-Dimensional Taxonomy for Empirical Evaluation of LLM-based Agents
von: Vatsal, Shubham, et al.
Veröffentlicht: (2026)
von: Vatsal, Shubham, et al.
Veröffentlicht: (2026)
Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
Enhancing the Detection of Coronary Artery Disease Using Machine Learning
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
Program Synthesis Dialog Agents for Interactive Decision-Making
von: Toles, Matthew, et al.
Veröffentlicht: (2025)
von: Toles, Matthew, et al.
Veröffentlicht: (2025)
HySafe-AI: Hybrid Safety Architectural Analysis Framework for AI Systems: A Case Study
von: Pitale, Mandar, et al.
Veröffentlicht: (2025)
von: Pitale, Mandar, et al.
Veröffentlicht: (2025)
Improving Group Fairness in Knowledge Distillation via Laplace Approximation of Early Exits
von: Fasth, Edvin, et al.
Veröffentlicht: (2025)
von: Fasth, Edvin, et al.
Veröffentlicht: (2025)
A Hybrid Machine Learning Model for Cerebral Palsy Detection
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
When Big Models Train Small Ones: Label-Free Model Parity Alignment for Efficient Visual Question Answering using Small VLMs
von: Penamakuri, Abhirama Subramanyam, et al.
Veröffentlicht: (2025)
von: Penamakuri, Abhirama Subramanyam, et al.
Veröffentlicht: (2025)
Quantitative Certification of Agentic Tool Selection
von: Yeon, Jehyeok, et al.
Veröffentlicht: (2025)
von: Yeon, Jehyeok, et al.
Veröffentlicht: (2025)
Arithmetic-Intensity-Aware Quantization
von: Singh, Taig, et al.
Veröffentlicht: (2025)
von: Singh, Taig, et al.
Veröffentlicht: (2025)
TechGraphRAG: An Agentic Graph-Augmented RAG Framework for Technical Literature Reasoning
von: Singh, Kanwar Bharat
Veröffentlicht: (2026)
von: Singh, Kanwar Bharat
Veröffentlicht: (2026)
Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation
von: Kim, Donghwan, et al.
Veröffentlicht: (2026)
von: Kim, Donghwan, et al.
Veröffentlicht: (2026)
Leveraging Natural Language Processing and Machine Learning for Evidence-Based Food Security Policy Decision-Making in Data-Scarce Making
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
von: Singh, Nikhil Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Nikhil Kumar, et al.
Veröffentlicht: (2024)
Stochastic Monkeys at Play: Random Augmentations Cheaply Break LLM Safety Alignment
von: Vega, Jason, et al.
Veröffentlicht: (2024)
von: Vega, Jason, et al.
Veröffentlicht: (2024)
Strongly Topology-preserving GNNs for Brain Graph Super-resolution
von: Singh, Pragya, et al.
Veröffentlicht: (2024)
von: Singh, Pragya, et al.
Veröffentlicht: (2024)
Beyond Semantics: How Temporal Biases Shape Retrieval in Transformer and State-Space Models
von: Bajaj, Anooshka, et al.
Veröffentlicht: (2025)
von: Bajaj, Anooshka, et al.
Veröffentlicht: (2025)
NetMoniAI: An Agentic AI Framework for Network Security & Monitoring
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
A Meta-learning based Stacked Regression Approach for Customer Lifetime Value Prediction
von: Gadgil, Karan, et al.
Veröffentlicht: (2023)
von: Gadgil, Karan, et al.
Veröffentlicht: (2023)
Stock Market Price Prediction: A Hybrid LSTM and Sequential Self-Attention based Approach
von: Pardeshi, Karan, et al.
Veröffentlicht: (2023)
von: Pardeshi, Karan, et al.
Veröffentlicht: (2023)
Agentic Reasoning and Tool Integration for LLMs via Reinforcement Learning
von: Singh, Joykirat, et al.
Veröffentlicht: (2025)
von: Singh, Joykirat, et al.
Veröffentlicht: (2025)
Safety Recovery in Reasoning Models Is Only a Few Early Steering Steps Away
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2026)
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2026)
MetaCues: Enabling Critical Engagement with Generative AI for Information Seeking and Sensemaking
von: Singh, Anjali, et al.
Veröffentlicht: (2026)
von: Singh, Anjali, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
REAMS: Reasoning Enhanced Algorithm for Maths Solving
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025) -
First Train to Generate, then Generate to Train: UnitedSynT5 for Few-Shot NLI
von: Banerjee, Sourav, et al.
Veröffentlicht: (2024) -
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2025) -
LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases
von: Bouchard, Dylan, et al.
Veröffentlicht: (2025) -
Matching Ranks Over Probability Yields Truly Deep Safety Alignment
von: Vega, Jason, et al.
Veröffentlicht: (2025)