AI Propaganda factories with language models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Olejnik, Lukasz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Naming is framing: How cybersecurity's language problems are repeating in AI governance
von: Potter, Lianne
Veröffentlicht: (2025)
von: Potter, Lianne
Veröffentlicht: (2025)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025)
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
von: Barrett, Anthony M., et al.
Veröffentlicht: (2025)
von: Barrett, Anthony M., et al.
Veröffentlicht: (2025)
Private, Verifiable, and Auditable AI Systems
von: South, Tobin
Veröffentlicht: (2025)
von: South, Tobin
Veröffentlicht: (2025)
Frontier AI's Impact on the Cybersecurity Landscape
von: Potter, Yujin, et al.
Veröffentlicht: (2025)
von: Potter, Yujin, et al.
Veröffentlicht: (2025)
Red Teaming AI Red Teaming
von: Majumdar, Subhabrata, et al.
Veröffentlicht: (2025)
von: Majumdar, Subhabrata, et al.
Veröffentlicht: (2025)
Accelerating AI Development with Cyber Arenas
von: Cashman, William, et al.
Veröffentlicht: (2025)
von: Cashman, William, et al.
Veröffentlicht: (2025)
AI-Driven Cyber Threat Intelligence Automation
von: Shah, Shrit, et al.
Veröffentlicht: (2024)
von: Shah, Shrit, et al.
Veröffentlicht: (2024)
Securing the Future of GenAI: Policy and Technology
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2024)
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2024)
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
von: Hall, Peter, et al.
Veröffentlicht: (2025)
von: Hall, Peter, et al.
Veröffentlicht: (2025)
Black-Box Access is Insufficient for Rigorous AI Audits
von: Casper, Stephen, et al.
Veröffentlicht: (2024)
von: Casper, Stephen, et al.
Veröffentlicht: (2024)
Coordinated Flaw Disclosure for AI: Beyond Security Vulnerabilities
von: Cattell, Sven, et al.
Veröffentlicht: (2024)
von: Cattell, Sven, et al.
Veröffentlicht: (2024)
Interplay of ISMS and AIMS in context of the EU AI Act
von: Pötsch, Jordan
Veröffentlicht: (2024)
von: Pötsch, Jordan
Veröffentlicht: (2024)
The End Of Universal Lifelong Identifiers: Identity Systems For The AI Era
von: Palakodety, Shriphani
Veröffentlicht: (2025)
von: Palakodety, Shriphani
Veröffentlicht: (2025)
Enabling External Scrutiny of AI Systems with Privacy-Enhancing Technologies
von: Beers, Kendrea, et al.
Veröffentlicht: (2025)
von: Beers, Kendrea, et al.
Veröffentlicht: (2025)
RedTeamLLM: an Agentic AI framework for offensive security
von: Challita, Brian, et al.
Veröffentlicht: (2025)
von: Challita, Brian, et al.
Veröffentlicht: (2025)
Governable AI: Provable Safety Under Extreme Threat Models
von: Wang, Donglin, et al.
Veröffentlicht: (2025)
von: Wang, Donglin, et al.
Veröffentlicht: (2025)
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
von: Lin, Justin W., et al.
Veröffentlicht: (2025)
von: Lin, Justin W., et al.
Veröffentlicht: (2025)
The Impact of AI on the Cyber Offense-Defense Balance and the Character of Cyber Conflict
von: Lohn, Andrew J.
Veröffentlicht: (2025)
von: Lohn, Andrew J.
Veröffentlicht: (2025)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
von: Qayyum, Adnan, et al.
Veröffentlicht: (2022)
von: Qayyum, Adnan, et al.
Veröffentlicht: (2022)
Is Your AI Truly Yours? Leveraging Blockchain for Copyrights, Provenance, and Lineage
von: Wang, Qin, et al.
Veröffentlicht: (2024)
von: Wang, Qin, et al.
Veröffentlicht: (2024)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
von: Feakins, Shaun, et al.
Veröffentlicht: (2026)
von: Feakins, Shaun, et al.
Veröffentlicht: (2026)
Love, Lies, and Language Models: Investigating AI's Role in Romance-Baiting Scams
von: Gressel, Gilad, et al.
Veröffentlicht: (2025)
von: Gressel, Gilad, et al.
Veröffentlicht: (2025)
Can AI Models be Jailbroken to Phish Elderly Victims? An End-to-End Evaluation
von: Heiding, Fred, et al.
Veröffentlicht: (2025)
von: Heiding, Fred, et al.
Veröffentlicht: (2025)
Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
von: Glukhov, David, et al.
Veröffentlicht: (2024)
von: Glukhov, David, et al.
Veröffentlicht: (2024)
To Defend Against Cyber Attacks, We Must Teach AI Agents to Hack
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2026)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2026)
Large language model-powered AI systems achieve self-replication with no human intervention
von: Pan, Xudong, et al.
Veröffentlicht: (2025)
von: Pan, Xudong, et al.
Veröffentlicht: (2025)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
von: Bucknall, Ben, et al.
Veröffentlicht: (2025)
von: Bucknall, Ben, et al.
Veröffentlicht: (2025)
Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
von: Mazzocchetti, Adam Massimo
Veröffentlicht: (2026)
von: Mazzocchetti, Adam Massimo
Veröffentlicht: (2026)
Trust and Dependability in Blockchain & AI Based MedIoT Applications: Research Challenges and Future Directions
von: Solaiman, Ellis, et al.
Veröffentlicht: (2025)
von: Solaiman, Ellis, et al.
Veröffentlicht: (2025)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
von: Tong, Haibo, et al.
Veröffentlicht: (2026)
von: Tong, Haibo, et al.
Veröffentlicht: (2026)
Security, privacy, and agentic AI in a regulatory view: From definitions and distinctions to provisions and reflections
von: Zhang, Shiliang, et al.
Veröffentlicht: (2026)
von: Zhang, Shiliang, et al.
Veröffentlicht: (2026)
Can a large language model be a gaslighter?
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
A Whole New World: Creating a Parallel-Poisoned Web Only AI-Agents Can See
von: Zychlinski, Shaked
Veröffentlicht: (2025)
von: Zychlinski, Shaked
Veröffentlicht: (2025)
Authenticity Debt and the Synthetic Content Threat Landscape: A Layered Framework for Trust, Provenance, and IP Governance in the Generative AI Era
von: Sengupta, Shubhashis, et al.
Veröffentlicht: (2026)
von: Sengupta, Shubhashis, et al.
Veröffentlicht: (2026)
Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control
von: Wei, Peng, et al.
Veröffentlicht: (2026)
von: Wei, Peng, et al.
Veröffentlicht: (2026)
How Do Data Owners Say No? A Case Study of Data Consent Mechanisms in Web-Scraped Vision-Language AI Training Datasets
von: Lee, Chung Peng, et al.
Veröffentlicht: (2025)
von: Lee, Chung Peng, et al.
Veröffentlicht: (2025)
Complete Evasion, Zero Modification: PDF Attacks on AI Text Detection
von: Creo, Aldan
Veröffentlicht: (2025)
von: Creo, Aldan
Veröffentlicht: (2025)
Blueprints of Trust: AI System Cards for End to End Transparency and Governance
von: Sidhpurwala, Huzaifa, et al.
Veröffentlicht: (2025)
von: Sidhpurwala, Huzaifa, et al.
Veröffentlicht: (2025)
Securing the AI Frontier: Urgent Ethical and Regulatory Imperatives for AI-Driven Cybersecurity
von: Kulothungan, Vikram
Veröffentlicht: (2025)
von: Kulothungan, Vikram
Veröffentlicht: (2025)
Ähnliche Einträge
-
Naming is framing: How cybersecurity's language problems are repeating in AI governance
von: Potter, Lianne
Veröffentlicht: (2025) -
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025) -
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
von: Barrett, Anthony M., et al.
Veröffentlicht: (2025) -
Private, Verifiable, and Auditable AI Systems
von: South, Tobin
Veröffentlicht: (2025) -
Frontier AI's Impact on the Cybersecurity Landscape
von: Potter, Yujin, et al.
Veröffentlicht: (2025)