An Evaluation of Chat Safety Moderations in Roblox
Fuente:
arXiv
Saved in:
| Main Authors: | Kaushik, Priya, Brown, Sonja, Hasan, Rakibul, Rahaman, Sazzadur |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What's Privacy Good for? Measuring Privacy as a Shield from Harms due to Personal Data Use
by: Gajavalli, Sri Harsha, et al.
Published: (2025)
by: Gajavalli, Sri Harsha, et al.
Published: (2025)
Private Links, Public Leaks: Consequences of Frictionless User Experience on the Security and Privacy Posture of SMS-Delivered URLs
by: Danish, Muhammad, et al.
Published: (2026)
by: Danish, Muhammad, et al.
Published: (2026)
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
A Proposal for Evaluating the Operational Risk for ChatBots based on Large Language Models
by: Pinacho-Davidson, Pedro, et al.
Published: (2025)
by: Pinacho-Davidson, Pedro, et al.
Published: (2025)
Safety case template for frontier AI: A cyber inability argument
by: Goemans, Arthur, et al.
Published: (2024)
by: Goemans, Arthur, et al.
Published: (2024)
Safety and Security Analysis of Large Language Models: Benchmarking Risk Profile and Harm Potential
by: Akiri, Charankumar, et al.
Published: (2025)
by: Akiri, Charankumar, et al.
Published: (2025)
From Chat Control to Robot Control: Implications of the Chat Control Proposal for Human-Robot Interaction
by: Akalin, Neziha, et al.
Published: (2026)
by: Akalin, Neziha, et al.
Published: (2026)
TRIDENT -- A Three-Tier Privacy-Preserving Propaganda Detection Model in Mobile Networks using Transformers, Adversarial Learning, and Differential Privacy
by: Emran, Al Nahian Bin, et al.
Published: (2025)
by: Emran, Al Nahian Bin, et al.
Published: (2025)
Case Studies: Effective Approaches for Navigating Cross-Border Cloud Data Transfers Amid U.S. Government Privacy and Safety Concerns
by: Adebayo, Motunrayo
Published: (2025)
by: Adebayo, Motunrayo
Published: (2025)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
by: Tong, Haibo, et al.
Published: (2026)
by: Tong, Haibo, et al.
Published: (2026)
Chatting with Confidants or Corporations? Privacy Management with AI Companions
by: Chiu, Hsuen-Chi, et al.
Published: (2026)
by: Chiu, Hsuen-Chi, et al.
Published: (2026)
Evaluating the Impacts of Swapping on the US Decennial Census
by: Ballesteros, Maria, et al.
Published: (2025)
by: Ballesteros, Maria, et al.
Published: (2025)
Risks & Benefits of LLMs & GenAI for Platform Integrity, Healthcare Diagnostics, Financial Trust and Compliance, Cybersecurity, Privacy & AI Safety: A Comprehensive Survey, Roadmap & Implementation Blueprint
by: Ahi, Kiarash
Published: (2025)
by: Ahi, Kiarash
Published: (2025)
Assessing the influence of cybersecurity threats and risks on the adoption and growth of digital banking: a systematic literature review
by: Waliullah, Md., et al.
Published: (2025)
by: Waliullah, Md., et al.
Published: (2025)
Evaluating Privacy Measures in Healthcare Apps Predominantly Used by Older Adults
by: Saka, Suleiman, et al.
Published: (2024)
by: Saka, Suleiman, et al.
Published: (2024)
Expected Harm: Rethinking Safety Evaluation of (Mis)Aligned LLMs
by: Chen, Yen-Shan, et al.
Published: (2026)
by: Chen, Yen-Shan, et al.
Published: (2026)
Evaluating Organization Security: User Stories of European Union NIS2 Directive
by: Seeba, Mari, et al.
Published: (2025)
by: Seeba, Mari, et al.
Published: (2025)
ConVerse: Benchmarking Contextual Safety in Agent-to-Agent Conversations
by: Gomaa, Amr, et al.
Published: (2025)
by: Gomaa, Amr, et al.
Published: (2025)
Trust, Because You Can't Verify:Privacy and Security Hurdles in Education Technology Acquisition Practices
by: Kelso, Easton, et al.
Published: (2024)
by: Kelso, Easton, et al.
Published: (2024)
Abusing the Internet of Medical Things: Evaluating Threat Models and Forensic Readiness for Multi-Vector Attacks on Connected Healthcare Devices
by: Straw, Isabel, et al.
Published: (2026)
by: Straw, Isabel, et al.
Published: (2026)
AuditGPT: Auditing Smart Contracts with ChatGPT
by: Xia, Shihao, et al.
Published: (2024)
by: Xia, Shihao, et al.
Published: (2024)
Evaluation of compliance with democratic and technical standards of i-voting in elections to academic senates in Czech higher education
by: Martinek, Tomas, et al.
Published: (2025)
by: Martinek, Tomas, et al.
Published: (2025)
A Review on Internet of Things for Defense and Public Safety
by: Fraga-Lamas, Paula, et al.
Published: (2024)
by: Fraga-Lamas, Paula, et al.
Published: (2024)
A Technical Policy Blueprint for Trustworthy Decentralized AI
by: Kassem, Hasan, et al.
Published: (2025)
by: Kassem, Hasan, et al.
Published: (2025)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
by: Lin, Zhiqiang, et al.
Published: (2025)
by: Lin, Zhiqiang, et al.
Published: (2025)
Governable AI: Provable Safety Under Extreme Threat Models
by: Wang, Donglin, et al.
Published: (2025)
by: Wang, Donglin, et al.
Published: (2025)
Benchmarking and Understanding Safety Risks in AI Character Platforms
by: Wei, Yiluo, et al.
Published: (2025)
by: Wei, Yiluo, et al.
Published: (2025)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
by: Feakins, Shaun, et al.
Published: (2026)
by: Feakins, Shaun, et al.
Published: (2026)
Gaming the Metric, Not the Harm: Certifying Safety Audits against Strategic Platform Manipulation
by: Burnat, Florian A. D., et al.
Published: (2026)
by: Burnat, Florian A. D., et al.
Published: (2026)
Can LLMs Infer Conversational Agent Users' Personality Traits from Chat History?
by: Cögendez, Derya, et al.
Published: (2026)
by: Cögendez, Derya, et al.
Published: (2026)
How Generative AI Empowers Attackers and Defenders Across the Trust & Safety Landscape
by: Kelley, Patrick Gage, et al.
Published: (2025)
by: Kelley, Patrick Gage, et al.
Published: (2025)
Advancing Highway Work Zone Safety: A Comprehensive Review of Sensor Technologies for Intrusion and Proximity Hazards
by: Demeke, Ayenew Yihune, et al.
Published: (2025)
by: Demeke, Ayenew Yihune, et al.
Published: (2025)
LLM Platform Security: Applying a Systematic Evaluation Framework to OpenAI's ChatGPT Plugins
by: Iqbal, Umar, et al.
Published: (2023)
by: Iqbal, Umar, et al.
Published: (2023)
Phare: A Safety Probe for Large Language Models
by: Jeune, Pierre Le, et al.
Published: (2025)
by: Jeune, Pierre Le, et al.
Published: (2025)
How Safe Is Your Data in Connected and Autonomous Cars: A Consumer Advantage or a Privacy Nightmare ?
by: Chougule, Amit, et al.
Published: (2026)
by: Chougule, Amit, et al.
Published: (2026)
MemeChain: A Multimodal Cross-Chain Dataset for Meme Coin Forensics and Risk Analysis
by: Mongardini, Alberto Maria, et al.
Published: (2026)
by: Mongardini, Alberto Maria, et al.
Published: (2026)
Taking a Bite Out of the Forbidden Fruit: Characterizing Third-Party Iranian iOS App Stores
by: Khanlari, Amirhossein, et al.
Published: (2026)
by: Khanlari, Amirhossein, et al.
Published: (2026)
A Relay a Day Keeps the AirTag Away: Practical Relay Attacks on Apple's AirTags
by: Gegenhuber, Gabriel K., et al.
Published: (2026)
by: Gegenhuber, Gabriel K., et al.
Published: (2026)
Long-Term Risks of IoT Devices: The Case of the Smart Fridge
by: Buchmann, Erik
Published: (2026)
by: Buchmann, Erik
Published: (2026)
What's on Your Mind? Exploring Privacy of Mental Health Apps
by: Georgiou, Chloe, et al.
Published: (2026)
by: Georgiou, Chloe, et al.
Published: (2026)
Similar Items
-
What's Privacy Good for? Measuring Privacy as a Shield from Harms due to Personal Data Use
by: Gajavalli, Sri Harsha, et al.
Published: (2025) -
Private Links, Public Leaks: Consequences of Frictionless User Experience on the Security and Privacy Posture of SMS-Delivered URLs
by: Danish, Muhammad, et al.
Published: (2026) -
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025) -
A Proposal for Evaluating the Operational Risk for ChatBots based on Large Language Models
by: Pinacho-Davidson, Pedro, et al.
Published: (2025) -
Safety case template for frontier AI: A cyber inability argument
by: Goemans, Arthur, et al.
Published: (2024)