Scaling Reinforcement Learning for Content Moderation with Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Firooz, Hamed, Liu, Rui, Lu, Yuchen, Hou, Zhenyu, Xiong, Fangzhou, Zhang, Xiaoyang, Jian, Changshu, Zhu, Zhicheng, Ma, Jiayuan, Tao, Jacob, Gupta, Chaitali, Peng, Xiaochang, Mei, Shike, Cui, Hang, Qin, Yang, Tang, Shuo, Gaedtke, Jason, Mittal, Arpit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Afeto e Cuidado nas Relações Entre Humanos e seus Animais de Estimação
por: Kênia Mara Gaedtke
Publicado: (2019)
por: Kênia Mara Gaedtke
Publicado: (2019)
Evaluating LLMs Code Reasoning Under Real-World Context
por: Liu, Changshu
Publicado: (2026)
por: Liu, Changshu
Publicado: (2026)
CoT-ICL Lab: A Synthetic Framework for Studying Chain-of-Thought Learning from In-Context Demonstrations
por: Kothapalli, Vignesh, et al.
Publicado: (2025)
por: Kothapalli, Vignesh, et al.
Publicado: (2025)
A Tool for In-depth Analysis of Code Execution Reasoning of Large Language Models
por: Liu, Changshu, et al.
Publicado: (2025)
por: Liu, Changshu, et al.
Publicado: (2025)
CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation
por: Wu, Bingzhe, et al.
Publicado: (2026)
por: Wu, Bingzhe, et al.
Publicado: (2026)
To Think or Not to Think: The Hidden Cost of Meta-Training with Excessive CoT Examples
por: Kothapalli, Vignesh, et al.
Publicado: (2025)
por: Kothapalli, Vignesh, et al.
Publicado: (2025)
Lost-in-Distance: Impact of Contextual Proximity on LLM Performance in Graph Tasks
por: Firooz, Hamed, et al.
Publicado: (2024)
por: Firooz, Hamed, et al.
Publicado: (2024)
Partial independent transversals in multipartite graphs
por: Haxell, Penny, et al.
Publicado: (2025)
por: Haxell, Penny, et al.
Publicado: (2025)
The Relationship between the Use of Language Learning Strategies and Teaching Methods: A Case of Iranian EFL Learners
por: Firooz Sadighi
Publicado: (2006)
por: Firooz Sadighi
Publicado: (2006)
CodeMind: Evaluating Large Language Models for Code Reasoning
por: Liu, Changshu, et al.
Publicado: (2024)
por: Liu, Changshu, et al.
Publicado: (2024)
Assessing Coherency and Consistency of Code Execution Reasoning by Large Language Models
por: Liu, Changshu, et al.
Publicado: (2025)
por: Liu, Changshu, et al.
Publicado: (2025)
A Trust-Aware and Cost-Optimized Blockchain Oracle Selection Model with Deep Reinforcement Learning
por: Zhang, Hengyang, et al.
Publicado: (2025)
por: Zhang, Hengyang, et al.
Publicado: (2025)
Enhancing Stability for Large Language Models Training in Constrained Bandwidth Networks
por: Dai, Yun, et al.
Publicado: (2024)
por: Dai, Yun, et al.
Publicado: (2024)
Scaling Up LLM Reviews for Google Ads Content Moderation
por: Qiao, Wei, et al.
Publicado: (2024)
por: Qiao, Wei, et al.
Publicado: (2024)
Generate, Not Recommend: Personalized Multimodal Content Generation
por: Liu, Jiongnan, et al.
Publicado: (2025)
por: Liu, Jiongnan, et al.
Publicado: (2025)
Cold-RL: Learning Cache Eviction with Offline Reinforcement Learning for NGINX
por: Gupta, Aayush, et al.
Publicado: (2025)
por: Gupta, Aayush, et al.
Publicado: (2025)
Analyzing the Presentation, Content, and Utilization of References in LLM-powered Conversational AI Systems
por: Ouyang, Jianheng, et al.
Publicado: (2026)
por: Ouyang, Jianheng, et al.
Publicado: (2026)
Content Moderation Futures
por: Blackwell, Lindsay
Publicado: (2025)
por: Blackwell, Lindsay
Publicado: (2025)
Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings
por: Liu, Changshu, et al.
Publicado: (2025)
por: Liu, Changshu, et al.
Publicado: (2025)
A reconfigurable non-linear active metasurface for coherent wave down-conversion
por: Sanjari, Pouria, et al.
Publicado: (2024)
por: Sanjari, Pouria, et al.
Publicado: (2024)
IPS: In-Prompt Process Supervision for Short Video Content Moderation
por: Liu, Mingchao, et al.
Publicado: (2024)
por: Liu, Mingchao, et al.
Publicado: (2024)
Critical Challenges in Content Moderation for People Who Use Drugs (PWUD): Insights into Online Harm Reduction Practices from Moderators
por: Wang, Kaixuan, et al.
Publicado: (2025)
por: Wang, Kaixuan, et al.
Publicado: (2025)
Generative Reasoning Re-ranker
por: Liang, Mingfu, et al.
Publicado: (2026)
por: Liang, Mingfu, et al.
Publicado: (2026)
Toxic Ink on Immutable Paper: Content Moderation for Ethereum Input Data Messages (IDMs)
por: Xiong, Xihan, et al.
Publicado: (2025)
por: Xiong, Xihan, et al.
Publicado: (2025)
Collaborative Content Moderation in the Fediverse
por: Zia, Haris Bin, et al.
Publicado: (2025)
por: Zia, Haris Bin, et al.
Publicado: (2025)
Algorithmic Arbitrariness in Content Moderation
por: Gomez, Juan Felipe, et al.
Publicado: (2024)
por: Gomez, Juan Felipe, et al.
Publicado: (2024)
BingoGuard: LLM Content Moderation Tools with Risk Levels
por: Yin, Fan, et al.
Publicado: (2025)
por: Yin, Fan, et al.
Publicado: (2025)
Asynchronous Decentralized Optimization with Constraints: Achievable Speeds of Convergence for Directed Graphs
por: Shahriari-Mehr, Firooz, et al.
Publicado: (2024)
por: Shahriari-Mehr, Firooz, et al.
Publicado: (2024)
A compact scalable phase modulator with zero static power consumption for visible integrated photonics
por: MacFarlane, Neil, et al.
Publicado: (2024)
por: MacFarlane, Neil, et al.
Publicado: (2024)
A Polynomial Kernel for Vertex Deletion to the Scattered Class of Proper Interval Graph and Trees
por: Jacob, Ashwin, et al.
Publicado: (2026)
por: Jacob, Ashwin, et al.
Publicado: (2026)
Investigating the Correlation between Dark Matter Content, Ages and Mass-to-Light Ratios in Spiral Galaxies
por: Kottur, Arpit, et al.
Publicado: (2025)
por: Kottur, Arpit, et al.
Publicado: (2025)
Motherhood And The Right To Education: Harmonising Statutory Entitlements And Institutional Practice
por: Chaitali, Wadhwa, et al.
Publicado: (2026)
por: Chaitali, Wadhwa, et al.
Publicado: (2026)
Dynamic Content Moderation in Livestreams: Combining Supervised Classification with MLLM-Boosted Similarity Matching
por: Yew, Wei Chee, et al.
Publicado: (2025)
por: Yew, Wei Chee, et al.
Publicado: (2025)
Ideology-Based LLMs for Content Moderation
por: Civelli, Stefano, et al.
Publicado: (2025)
por: Civelli, Stefano, et al.
Publicado: (2025)
Selling Certification, Content Moderation, and Attention
por: Bar-Isaac, Heski, et al.
Publicado: (2025)
por: Bar-Isaac, Heski, et al.
Publicado: (2025)
AI Content Moderation in Therapy Conversations
por: Kim, Jiwon, et al.
Publicado: (2026)
por: Kim, Jiwon, et al.
Publicado: (2026)
Experimentation in Content Moderation using RWKV
por: Yildirim, Umut, et al.
Publicado: (2024)
por: Yildirim, Umut, et al.
Publicado: (2024)
Personalized Content Moderation and Emergent Outcomes
por: Gurkan, Necdet, et al.
Publicado: (2024)
por: Gurkan, Necdet, et al.
Publicado: (2024)
Improving Critical Node Detection Using Neural Network-based Initialization in a Genetic Algorithm
por: Liu, Chanjuan, et al.
Publicado: (2024)
por: Liu, Chanjuan, et al.
Publicado: (2024)
Efficacy and Safety of Oxymetazoline 1% Cream for the Treatment of Mild to Moderate Facial Rosacea
por: Fatemeh Sajdeh, et al.
Publicado: (2025)
por: Fatemeh Sajdeh, et al.
Publicado: (2025)
Ejemplares similares
-
Afeto e Cuidado nas Relações Entre Humanos e seus Animais de Estimação
por: Kênia Mara Gaedtke
Publicado: (2019) -
Evaluating LLMs Code Reasoning Under Real-World Context
por: Liu, Changshu
Publicado: (2026) -
CoT-ICL Lab: A Synthetic Framework for Studying Chain-of-Thought Learning from In-Context Demonstrations
por: Kothapalli, Vignesh, et al.
Publicado: (2025) -
A Tool for In-depth Analysis of Code Execution Reasoning of Large Language Models
por: Liu, Changshu, et al.
Publicado: (2025) -
CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation
por: Wu, Bingzhe, et al.
Publicado: (2026)