Qwerty AI: Explainable Automated Age Rating and Content Safety Assessment for Russian-Language Screenplays
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Zmanovskii, Nikita |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
par: Hossain, Ariyan, et autres
Publié: (2025)
par: Hossain, Ariyan, et autres
Publié: (2025)
Replicating TEMPEST at Scale: Multi-Turn Adversarial Attacks Against Trillion-Parameter Frontier Models
par: Young, Richard
Publié: (2025)
par: Young, Richard
Publié: (2025)
Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation
par: Hartmann, David, et autres
Publié: (2026)
par: Hartmann, David, et autres
Publié: (2026)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
par: Stewart, Ian, et autres
Publié: (2024)
par: Stewart, Ian, et autres
Publié: (2024)
Growing a Tail: Increasing Output Diversity in Large Language Models
par: Shur-Ofry, Michal, et autres
Publié: (2024)
par: Shur-Ofry, Michal, et autres
Publié: (2024)
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk
par: Wu, Shuai, et autres
Publié: (2026)
par: Wu, Shuai, et autres
Publié: (2026)
Leveraging Multi-Source Textural UGC for Neighbourhood Housing Quality Assessment: A GPT-Enhanced Framework
par: Hong, Qiyuan, et autres
Publié: (2025)
par: Hong, Qiyuan, et autres
Publié: (2025)
When Large Language Models are More PersuasiveThan Incentivized Humans, and Why
par: Schoenegger, Philipp, et autres
Publié: (2025)
par: Schoenegger, Philipp, et autres
Publié: (2025)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
par: Beltoft, Stine, et autres
Publié: (2025)
par: Beltoft, Stine, et autres
Publié: (2025)
The Company You Keep: How LLMs Respond to Dark Triad Traits
par: Lu, Zeyi, et autres
Publié: (2026)
par: Lu, Zeyi, et autres
Publié: (2026)
WSC+: Enhancing The Winograd Schema Challenge Using Tree-of-Experts
par: Zahraei, Pardis Sadat, et autres
Publié: (2024)
par: Zahraei, Pardis Sadat, et autres
Publié: (2024)
APPSI-139: A Parallel Corpus of English Application Privacy Policy Summarization and Interpretation
par: Zhu, Pengyun, et autres
Publié: (2026)
par: Zhu, Pengyun, et autres
Publié: (2026)
The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete
par: Barmettler, Joel
Publié: (2026)
par: Barmettler, Joel
Publié: (2026)
Human Values in a Single Sentence: Moral Presence, Hierarchies, and Transformer Ensembles on the Schwartz Continuum
par: Yeste, Víctor, et autres
Publié: (2026)
par: Yeste, Víctor, et autres
Publié: (2026)
REMIND: Input Loss Landscapes Reveal Residual Memorization in Post-Unlearning LLMs
par: Cohen, Liran, et autres
Publié: (2025)
par: Cohen, Liran, et autres
Publié: (2025)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
par: Wang, Xinyue, et autres
Publié: (2026)
par: Wang, Xinyue, et autres
Publié: (2026)
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
par: Arora, Sunil, et autres
Publié: (2025)
par: Arora, Sunil, et autres
Publié: (2025)
Big AI is accelerating the metacrisis: What can we do?
par: Bird, Steven
Publié: (2025)
par: Bird, Steven
Publié: (2025)
From Native Memes to Global Moderation: Cross-Cultural Evaluation of Vision-Language Models for Hateful Meme Detection
par: Wang, Mo, et autres
Publié: (2026)
par: Wang, Mo, et autres
Publié: (2026)
The Cultural Gene of Large Language Models: A Study on the Impact of Cross-Corpus Training on Model Values and Biases
par: Fenech-Borg, Emanuel Z., et autres
Publié: (2025)
par: Fenech-Borg, Emanuel Z., et autres
Publié: (2025)
Revealing Hidden Bias in AI: Lessons from Large Language Models
par: Beatty, Django, et autres
Publié: (2024)
par: Beatty, Django, et autres
Publié: (2024)
GeoGalactica: A Scientific Large Language Model in Geoscience
par: Lin, Zhouhan, et autres
Publié: (2023)
par: Lin, Zhouhan, et autres
Publié: (2023)
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
par: Hildebrand, Samuel, et autres
Publié: (2025)
par: Hildebrand, Samuel, et autres
Publié: (2025)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
par: Chen, Renmiao, et autres
Publié: (2025)
par: Chen, Renmiao, et autres
Publié: (2025)
Deep Learning based Key Information Extraction from Business Documents: Systematic Literature Review
par: Rombach, Alexander Michael, et autres
Publié: (2024)
par: Rombach, Alexander Michael, et autres
Publié: (2024)
Document Understanding for Healthcare Referrals
par: Mistry, Jimit, et autres
Publié: (2023)
par: Mistry, Jimit, et autres
Publié: (2023)
Powerful Training-Free Membership Inference Against Autoregressive Language Models
par: Ilić, David, et autres
Publié: (2026)
par: Ilić, David, et autres
Publié: (2026)
Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations
par: Wong, Qi Han
Publié: (2026)
par: Wong, Qi Han
Publié: (2026)
AI and My Values: User Perceptions of LLMs' Ability to Extract, Embody, and Explain Human Values from Casual Conversations
par: Yun, Bhada, et autres
Publié: (2026)
par: Yun, Bhada, et autres
Publié: (2026)
KidsNanny: A Two-Stage Multimodal Content Moderation Pipeline Integrating Visual Classification, Object Detection, OCR, and Contextual Reasoning for Child Safety
par: Panchal, Viraj, et autres
Publié: (2026)
par: Panchal, Viraj, et autres
Publié: (2026)
Automated Theorem Provers Help Improve Large Language Model Reasoning
par: McGinness, Lachlan, et autres
Publié: (2024)
par: McGinness, Lachlan, et autres
Publié: (2024)
The Fragility Of Moral Judgment In Large Language Models
par: van Nuenen, Tom, et autres
Publié: (2026)
par: van Nuenen, Tom, et autres
Publié: (2026)
Pipeline and Dataset Generation for Automated Fact-checking in Almost Any Language
par: Drchal, Jan, et autres
Publié: (2023)
par: Drchal, Jan, et autres
Publié: (2023)
Cultural Encoding in Large Language Models: The Existence Gap in AI-Mediated Brand Discovery
par: Junyao, Huang, et autres
Publié: (2025)
par: Junyao, Huang, et autres
Publié: (2025)
Terminal Wrench: A Dataset of 331 Reward-Hackable Environments and 3,632 Exploit Trajectories
par: Bercovich, Ivan, et autres
Publié: (2026)
par: Bercovich, Ivan, et autres
Publié: (2026)
More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts
par: Yeste, Víctor, et autres
Publié: (2026)
par: Yeste, Víctor, et autres
Publié: (2026)
Do Schwartz Higher-Order Values Help Sentence-Level Human Value Detection? A Study of Hierarchical Gating and Calibration
par: Yeste, Víctor, et autres
Publié: (2026)
par: Yeste, Víctor, et autres
Publié: (2026)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
par: Whittaker, Edward, et autres
Publié: (2024)
par: Whittaker, Edward, et autres
Publié: (2024)
HiPS: Hierarchical PDF Segmentation of Textbooks
par: Wehnert, Sabine, et autres
Publié: (2025)
par: Wehnert, Sabine, et autres
Publié: (2025)
LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability
par: Lucas, Tom, et autres
Publié: (2026)
par: Lucas, Tom, et autres
Publié: (2026)
Documents similaires
-
Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
par: Hossain, Ariyan, et autres
Publié: (2025) -
Replicating TEMPEST at Scale: Multi-Turn Adversarial Attacks Against Trillion-Parameter Frontier Models
par: Young, Richard
Publié: (2025) -
Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation
par: Hartmann, David, et autres
Publié: (2026) -
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
par: Stewart, Ian, et autres
Publié: (2024) -
Growing a Tail: Increasing Output Diversity in Large Language Models
par: Shur-Ofry, Michal, et autres
Publié: (2024)