Policy-as-Prompt: Rethinking Content Moderation in the Age of Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Palla, Konstantina, García, José Luis Redondo, Hauff, Claudia, Fabbri, Francesco, Lindström, Henrik, Taber, Daniel R., Damianou, Andreas, Lalmas, Mounia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
by: De Nadai, Marco, et al.
Published: (2025)
by: De Nadai, Marco, et al.
Published: (2025)
Mapping Faithful Reasoning in Language Models
by: Li, Jiazheng, et al.
Published: (2025)
by: Li, Jiazheng, et al.
Published: (2025)
Towards Graph Foundation Models for Personalization
by: Damianou, Andreas, et al.
Published: (2024)
by: Damianou, Andreas, et al.
Published: (2024)
Text2Tracks: Prompt-based Music Recommendation via Generative Retrieval
by: Palumbo, Enrico, et al.
Published: (2025)
by: Palumbo, Enrico, et al.
Published: (2025)
Prompt-to-Slate: Diffusion Models for Prompt-Conditioned Slate Generation
by: Tomasi, Federico, et al.
Published: (2024)
by: Tomasi, Federico, et al.
Published: (2024)
Enhancing Content Moderation with Culturally-Aware Models
by: Chan, Alex J., et al.
Published: (2023)
by: Chan, Alex J., et al.
Published: (2023)
Do LLM-judges Align with Human Relevance in Cranfield-style Recommender Evaluation?
by: Penha, Gustavo, et al.
Published: (2025)
by: Penha, Gustavo, et al.
Published: (2025)
Evaluating Podcast Recommendations with Profile-Aware LLM-as-a-Judge
by: Fabbri, Francesco, et al.
Published: (2025)
by: Fabbri, Francesco, et al.
Published: (2025)
Embedding-to-Prefix: Parameter-Efficient Personalization for Pre-Trained Large Language Models
by: Huber, Bernd, et al.
Published: (2025)
by: Huber, Bernd, et al.
Published: (2025)
Long-term Off-Policy Evaluation and Learning
by: Saito, Yuta, et al.
Published: (2024)
by: Saito, Yuta, et al.
Published: (2024)
Content Moderation Futures
by: Blackwell, Lindsay
Published: (2025)
by: Blackwell, Lindsay
Published: (2025)
Greedy routing optimisation in hyperbolic networks
by: Sulyok, Bendegúz, et al.
Published: (2023)
by: Sulyok, Bendegúz, et al.
Published: (2023)
Improving Regulatory Oversight in Online Content Moderation
by: Tessa, Benedetta, et al.
Published: (2025)
by: Tessa, Benedetta, et al.
Published: (2025)
Subdiffusive semantic evolution in Indo-European languages
by: Asztalos, Bogdán, et al.
Published: (2022)
by: Asztalos, Bogdán, et al.
Published: (2022)
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
by: Nakanishi, Akito, et al.
Published: (2025)
by: Nakanishi, Akito, et al.
Published: (2025)
Exploring the Boundaries of Content Moderation in Text-to-Image Generation
by: Riccio, Piera, et al.
Published: (2024)
by: Riccio, Piera, et al.
Published: (2024)
Personalized Content Moderation and Emergent Outcomes
by: Gurkan, Necdet, et al.
Published: (2024)
by: Gurkan, Necdet, et al.
Published: (2024)
Algorithmic Arbitrariness in Content Moderation
by: Gomez, Juan Felipe, et al.
Published: (2024)
by: Gomez, Juan Felipe, et al.
Published: (2024)
Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework
by: Gajewska, Ewelina, et al.
Published: (2026)
by: Gajewska, Ewelina, et al.
Published: (2026)
Primal-Dual Guided Decoding for Constrained Discrete Diffusion
by: Tomasi, Federico, et al.
Published: (2026)
by: Tomasi, Federico, et al.
Published: (2026)
Community detection in hypergraphs through hyperedge percolation
by: Kovács, Bianka, et al.
Published: (2025)
by: Kovács, Bianka, et al.
Published: (2025)
Opinion polarisation in social networks driven by cognitive dissonance avoidance
by: Kovács, Zoltan, et al.
Published: (2024)
by: Kovács, Zoltan, et al.
Published: (2024)
Personalized Audiobook Recommendations at Spotify Through Graph Neural Networks
by: De Nadai, Marco, et al.
Published: (2024)
by: De Nadai, Marco, et al.
Published: (2024)
Impact of Stricter Content Moderation on Parler's Users' Discourse
by: Kumarswamy, Nihal, et al.
Published: (2023)
by: Kumarswamy, Nihal, et al.
Published: (2023)
Network geometry of the Drosophila brain
by: Sulyok, Bendegúz, et al.
Published: (2026)
by: Sulyok, Bendegúz, et al.
Published: (2026)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
by: Zhou, Yuhang, et al.
Published: (2025)
by: Zhou, Yuhang, et al.
Published: (2025)
Moderating Illicit Online Image Promotion for Unsafe User-Generated Content Games Using Large Vision-Language Models
by: Guo, Keyan, et al.
Published: (2024)
by: Guo, Keyan, et al.
Published: (2024)
Intra-community link formation and modularity in ultracold growing hyperbolic networks
by: Balogh, Sámuel G., et al.
Published: (2024)
by: Balogh, Sámuel G., et al.
Published: (2024)
Beyond Hate: Differentiating Uncivil and Intolerant Speech in Multimodal Content Moderation
by: Herrmann, Nils A., et al.
Published: (2026)
by: Herrmann, Nils A., et al.
Published: (2026)
Partis politiques et protestations au Maroc (1934-2020)
by: Bennani-Chraïbi, Mounia
Published: (2024)
by: Bennani-Chraïbi, Mounia
Published: (2024)
Social Policy of Large Language Models: How GPT, Claude, DeepSeek and Grok Allocate Social Budgets in Spain and Germany
by: Cantos, Claudia Benavides, et al.
Published: (2026)
by: Cantos, Claudia Benavides, et al.
Published: (2026)
Stream: Scaling up Mechanistic Interpretability to Long Context in LLMs via Sparse Attention
by: Rosser, J, et al.
Published: (2025)
by: Rosser, J, et al.
Published: (2025)
Hierarchy and ranking in pairwise sports contests
by: Asztalos, Bogdán, et al.
Published: (2025)
by: Asztalos, Bogdán, et al.
Published: (2025)
Iterative embedding and reweighting of complex networks reveals community structure
by: Kovács, Bianka, et al.
Published: (2024)
by: Kovács, Bianka, et al.
Published: (2024)
Longitudinal Monitoring of LLM Content Moderation of Social Issues
by: Dai, Yunlang, et al.
Published: (2025)
by: Dai, Yunlang, et al.
Published: (2025)
Towards Safer Social Media Platforms: Scalable and Performant Few-Shot Harmful Content Moderation Using Large Language Models
by: Bonagiri, Akash, et al.
Published: (2025)
by: Bonagiri, Akash, et al.
Published: (2025)
Quantifying Feature Importance for Online Content Moderation
by: Tessa, Benedetta, et al.
Published: (2025)
by: Tessa, Benedetta, et al.
Published: (2025)
A Multi-Level Strategy for Deepfake Content Moderation under EU Regulation
by: Förster, Max-Paul, et al.
Published: (2025)
by: Förster, Max-Paul, et al.
Published: (2025)
Autonomous Prompt Engineering in Large Language Models
by: Kepel, Daan, et al.
Published: (2024)
by: Kepel, Daan, et al.
Published: (2024)
AI Content Moderation in Therapy Conversations
by: Kim, Jiwon, et al.
Published: (2026)
by: Kim, Jiwon, et al.
Published: (2026)
Similar Items
-
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
by: De Nadai, Marco, et al.
Published: (2025) -
Mapping Faithful Reasoning in Language Models
by: Li, Jiazheng, et al.
Published: (2025) -
Towards Graph Foundation Models for Personalization
by: Damianou, Andreas, et al.
Published: (2024) -
Text2Tracks: Prompt-based Music Recommendation via Generative Retrieval
by: Palumbo, Enrico, et al.
Published: (2025) -
Prompt-to-Slate: Diffusion Models for Prompt-Conditioned Slate Generation
by: Tomasi, Federico, et al.
Published: (2024)