Trustworthiness in Stochastic Systems: Towards Opening the Black Box
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chien, Jennifer, Danks, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Behaviorist Representational Harms: A Plan for Measurement and Mitigation
von: Chien, Jennifer, et al.
Veröffentlicht: (2024)
von: Chien, Jennifer, et al.
Veröffentlicht: (2024)
Commercial AI, Conflict, and Moral Responsibility: A theoretical analysis and practical approach to the moral responsibilities associated with dual-use AI technology
von: Trusilo, Daniel, et al.
Veröffentlicht: (2024)
von: Trusilo, Daniel, et al.
Veröffentlicht: (2024)
Application of the NIST AI Risk Management Framework to Surveillance Technology
von: Swaminathan, Nandhini, et al.
Veröffentlicht: (2024)
von: Swaminathan, Nandhini, et al.
Veröffentlicht: (2024)
Addressing the Unforeseen Harms of Technology CCC Whitepaper
von: Bliss, Nadya, et al.
Veröffentlicht: (2024)
von: Bliss, Nadya, et al.
Veröffentlicht: (2024)
AI, Pluralism, and (Social) Compensation
von: Swaminathan, Nandhini, et al.
Veröffentlicht: (2024)
von: Swaminathan, Nandhini, et al.
Veröffentlicht: (2024)
Future of Pandemic Prevention and Response CCC Workshop Report
von: Danks, David, et al.
Veröffentlicht: (2024)
von: Danks, David, et al.
Veröffentlicht: (2024)
Mysterious and Manipulative Black Boxes: A Qualitative Analysis of Perceptions on Recommender Systems
von: Ruohonen, Jukka
Veröffentlicht: (2023)
von: Ruohonen, Jukka
Veröffentlicht: (2023)
RE-centric Recommendations for the Development of Trustworthy(er) Autonomous Systems
von: Ronanki, Krishna, et al.
Veröffentlicht: (2023)
von: Ronanki, Krishna, et al.
Veröffentlicht: (2023)
Reinforcing Trustworthiness in Multimodal Emotional Support Systems
von: Le, Huy M., et al.
Veröffentlicht: (2025)
von: Le, Huy M., et al.
Veröffentlicht: (2025)
Turning Language Model Training from Black Box into a Sandbox
von: Pope, Nicolas, et al.
Veröffentlicht: (2026)
von: Pope, Nicolas, et al.
Veröffentlicht: (2026)
Trustworthy human-centric based Automated Decision-Making Systems
von: Cabrera, Marcelino, et al.
Veröffentlicht: (2023)
von: Cabrera, Marcelino, et al.
Veröffentlicht: (2023)
Navigating the sociotechnical labyrinth: Dynamic certification for responsible embodied AI
von: Bakirtzis, Georgios, et al.
Veröffentlicht: (2024)
von: Bakirtzis, Georgios, et al.
Veröffentlicht: (2024)
Enabling the AI Revolution in Healthcare
von: Singh, Mona, et al.
Veröffentlicht: (2025)
von: Singh, Mona, et al.
Veröffentlicht: (2025)
Black-Box Access is Insufficient for Rigorous AI Audits
von: Casper, Stephen, et al.
Veröffentlicht: (2024)
von: Casper, Stephen, et al.
Veröffentlicht: (2024)
Decoding the Black Box: Discerning AI Rhetorics About and Through Poetic Prompting
von: Edgar, P. D., et al.
Veröffentlicht: (2025)
von: Edgar, P. D., et al.
Veröffentlicht: (2025)
Brokerage in the Black Box: Swing States, Strategic Ambiguity, and the Global Politics of AI Governance
von: Tran, Ha-Chi
Veröffentlicht: (2026)
von: Tran, Ha-Chi
Veröffentlicht: (2026)
Towards Trustworthy AI: Characterizing User-Reported Risks across LLMs "In the Wild"
von: Li, Lingyao, et al.
Veröffentlicht: (2025)
von: Li, Lingyao, et al.
Veröffentlicht: (2025)
Audit Me If You Can: Query-Efficient Active Fairness Auditing of Black-Box LLMs
von: Hartmann, David, et al.
Veröffentlicht: (2026)
von: Hartmann, David, et al.
Veröffentlicht: (2026)
Opening Knowledge Gaps Drives Scientific Progress
von: Kedrick, Kara, et al.
Veröffentlicht: (2025)
von: Kedrick, Kara, et al.
Veröffentlicht: (2025)
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2026)
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2026)
Towards Trustworthy AI: A Review of Ethical and Robust Large Language Models
von: Ferdaus, Md Meftahul, et al.
Veröffentlicht: (2024)
von: Ferdaus, Md Meftahul, et al.
Veröffentlicht: (2024)
How to Assess Trustworthy AI in Practice
von: Zicari, Roberto V., et al.
Veröffentlicht: (2022)
von: Zicari, Roberto V., et al.
Veröffentlicht: (2022)
Trust or Bust: Ensuring Trustworthiness in Autonomous Weapon Systems
von: Cools, Kasper, et al.
Veröffentlicht: (2024)
von: Cools, Kasper, et al.
Veröffentlicht: (2024)
Inside the Black Box: Detecting and Mitigating Algorithmic Bias across Racialized Groups in College Student-Success Prediction
von: Gándara, Denisa, et al.
Veröffentlicht: (2023)
von: Gándara, Denisa, et al.
Veröffentlicht: (2023)
All in How You Ask for It: Simple Black-Box Method for Jailbreak Attacks
von: Takemoto, Kazuhiro
Veröffentlicht: (2024)
von: Takemoto, Kazuhiro
Veröffentlicht: (2024)
Assessing the Sustainability and Trustworthiness of Federated Learning Models
von: Feng, Chao, et al.
Veröffentlicht: (2023)
von: Feng, Chao, et al.
Veröffentlicht: (2023)
Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI
von: Yang, Chao, et al.
Veröffentlicht: (2024)
von: Yang, Chao, et al.
Veröffentlicht: (2024)
Leveraging Imperfect Sources to Detect Fairwashing in Black-Box Auditing
von: Bourrée, Jade Garcia, et al.
Veröffentlicht: (2023)
von: Bourrée, Jade Garcia, et al.
Veröffentlicht: (2023)
Unlocking the Black Box: Analysing the EU Artificial Intelligence Act's Framework for Explainability in AI
von: Pavlidis, Georgios
Veröffentlicht: (2025)
von: Pavlidis, Georgios
Veröffentlicht: (2025)
Toward Safe and Responsible AI Agents: A Three-Pillar Model for Transparency, Accountability, and Trustworthiness
von: Cheng, Edward C., et al.
Veröffentlicht: (2026)
von: Cheng, Edward C., et al.
Veröffentlicht: (2026)
First Analysis of the EU Artifical Intelligence Act: Towards a Global Standard for Trustworthy AI?
von: Ho-Dac, Marion
Veröffentlicht: (2024)
von: Ho-Dac, Marion
Veröffentlicht: (2024)
Trustworthy AI Posture (TAIP): A Framework for Continuous AI Assurance of Agentic Systems at Horizontal and Vertical scale
von: Lupo, Guy, et al.
Veröffentlicht: (2026)
von: Lupo, Guy, et al.
Veröffentlicht: (2026)
A Trustworthiness-based Metaphysics of Artificial Intelligence Systems
von: Ferrario, Andrea
Veröffentlicht: (2025)
von: Ferrario, Andrea
Veröffentlicht: (2025)
On the Trustworthiness of Generative Foundation Models: Guideline, Assessment, and Perspective
von: Huang, Yue, et al.
Veröffentlicht: (2025)
von: Huang, Yue, et al.
Veröffentlicht: (2025)
The Trust Calibration Maturity Model for Characterizing and Communicating Trustworthiness of AI Systems
von: Steinmetz, Scott T, et al.
Veröffentlicht: (2025)
von: Steinmetz, Scott T, et al.
Veröffentlicht: (2025)
Conceptualizing Trustworthiness and Trust in Communications
von: Fettweis, Gerhard P., et al.
Veröffentlicht: (2024)
von: Fettweis, Gerhard P., et al.
Veröffentlicht: (2024)
Visible, Trackable, Forkable: Opening the Process of Science
von: Samsonau, Sergey V.
Veröffentlicht: (2026)
von: Samsonau, Sergey V.
Veröffentlicht: (2026)
Bridging Ethical Principles and Algorithmic Methods: An Alternative Approach for Assessing Trustworthiness in AI Systems
von: Papademas, Michael, et al.
Veröffentlicht: (2025)
von: Papademas, Michael, et al.
Veröffentlicht: (2025)
White-Box Sensitivity Auditing with Steering Vectors
von: Cyberey, Hannah, et al.
Veröffentlicht: (2026)
von: Cyberey, Hannah, et al.
Veröffentlicht: (2026)
Group Fairness Meets the Black Box: Enabling Fair Algorithms on Closed LLMs via Post-Processing
von: Xian, Ruicheng, et al.
Veröffentlicht: (2025)
von: Xian, Ruicheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Beyond Behaviorist Representational Harms: A Plan for Measurement and Mitigation
von: Chien, Jennifer, et al.
Veröffentlicht: (2024) -
Commercial AI, Conflict, and Moral Responsibility: A theoretical analysis and practical approach to the moral responsibilities associated with dual-use AI technology
von: Trusilo, Daniel, et al.
Veröffentlicht: (2024) -
Application of the NIST AI Risk Management Framework to Surveillance Technology
von: Swaminathan, Nandhini, et al.
Veröffentlicht: (2024) -
Addressing the Unforeseen Harms of Technology CCC Whitepaper
von: Bliss, Nadya, et al.
Veröffentlicht: (2024) -
AI, Pluralism, and (Social) Compensation
von: Swaminathan, Nandhini, et al.
Veröffentlicht: (2024)