JiraiBench: A Bilingual Benchmark for Evaluating Large Language Models' Detection of Human Self-Destructive Behavior Content in Jirai Community
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Yunze, He, Tingyu, Wang, Lionel Z., Ma, Yiming, Song, Xingyu, Xu, Xiaohang, Diab, Mona, Li, Irene, Ng, Ka Chung |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei
por: Wang, Peng, et al.
Publicado: (2026)
por: Wang, Peng, et al.
Publicado: (2026)
Humanizing Machines: Rethinking LLM Anthropomorphism Through a Multi-Level Framework of Design
por: Xiao, Yunze, et al.
Publicado: (2025)
por: Xiao, Yunze, et al.
Publicado: (2025)
Hire Your Anthropologist! Rethinking Culture Benchmarks Through an Anthropological Lens
por: AlKhamissi, Mai, et al.
Publicado: (2025)
por: AlKhamissi, Mai, et al.
Publicado: (2025)
MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models
por: Wang, Lionel Z., et al.
Publicado: (2024)
por: Wang, Lionel Z., et al.
Publicado: (2024)
Towards Valid Student Simulation with Large Language Models
por: Yuan, Zhihao, et al.
Publicado: (2026)
por: Yuan, Zhihao, et al.
Publicado: (2026)
SimBA: Simplifying Benchmark Analysis Using Performance Matrices Alone
por: Subramani, Nishant, et al.
Publicado: (2025)
por: Subramani, Nishant, et al.
Publicado: (2025)
A Note on Bias to Complete
por: Xu, Jia, et al.
Publicado: (2024)
por: Xu, Jia, et al.
Publicado: (2024)
Sentipolis: Emotion-Aware Agents for Social Simulations
por: Fu, Chiyuan, et al.
Publicado: (2026)
por: Fu, Chiyuan, et al.
Publicado: (2026)
LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
por: Bai, Yushi, et al.
Publicado: (2023)
por: Bai, Yushi, et al.
Publicado: (2023)
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
por: Salama, Rana, et al.
Publicado: (2025)
por: Salama, Rana, et al.
Publicado: (2025)
Evaluating Large Language Model Biases in Persona-Steered Generation
por: Liu, Andy, et al.
Publicado: (2024)
por: Liu, Andy, et al.
Publicado: (2024)
DentalBench: Benchmarking and Advancing LLMs Capability for Bilingual Dentistry Understanding
por: Zhu, Hengchuan, et al.
Publicado: (2025)
por: Zhu, Hengchuan, et al.
Publicado: (2025)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
por: Hu, Tiancheng, et al.
Publicado: (2025)
por: Hu, Tiancheng, et al.
Publicado: (2025)
RefusalBench: Generative Evaluation of Selective Refusal in Grounded Language Models
por: Muhamed, Aashiq, et al.
Publicado: (2025)
por: Muhamed, Aashiq, et al.
Publicado: (2025)
ScholarBench: A Bilingual Benchmark for Abstraction, Comprehension, and Reasoning Evaluation in Academic Contexts
por: Noh, Dongwon, et al.
Publicado: (2025)
por: Noh, Dongwon, et al.
Publicado: (2025)
Synthetic Socratic Debates: Examining Persona Effects on Moral Decision and Persuasion Dynamics
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
por: Alqahtani, Amal, et al.
Publicado: (2025)
por: Alqahtani, Amal, et al.
Publicado: (2025)
DWTSumm: Discrete Wavelet Transform for Document Summarization
por: Salama, Rana, et al.
Publicado: (2026)
por: Salama, Rana, et al.
Publicado: (2026)
Semantic Compression for Word and Sentence Embeddings using Discrete Wavelet Transform
por: Salama, Rana Aref, et al.
Publicado: (2025)
por: Salama, Rana Aref, et al.
Publicado: (2025)
Taming Object Hallucinations with Verified Atomic Confidence Estimation
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
EigenBench: A Comparative Behavioral Measure of Value Alignment
por: Chang, Jonathn, et al.
Publicado: (2025)
por: Chang, Jonathn, et al.
Publicado: (2025)
BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data
por: Li, Wenkai, et al.
Publicado: (2024)
por: Li, Wenkai, et al.
Publicado: (2024)
Decoding Dark Matter: Specialized Sparse Autoencoders for Interpreting Rare Concepts in Foundation Models
por: Muhamed, Aashiq, et al.
Publicado: (2024)
por: Muhamed, Aashiq, et al.
Publicado: (2024)
Emotion Classification in Low and Moderate Resource Languages
por: Tafreshi, Shabnam, et al.
Publicado: (2024)
por: Tafreshi, Shabnam, et al.
Publicado: (2024)
OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
por: He, Chaoqun, et al.
Publicado: (2024)
por: He, Chaoqun, et al.
Publicado: (2024)
Biases Propagate in Encoder-based Vision-Language Models: A Systematic Analysis From Intrinsic Measures to Zero-shot Retrieval Outcomes
por: Ghate, Kshitish, et al.
Publicado: (2025)
por: Ghate, Kshitish, et al.
Publicado: (2025)
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
por: Liu, Jiarui, et al.
Publicado: (2024)
por: Liu, Jiarui, et al.
Publicado: (2024)
LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Utilization
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
LongBench Pro: A More Realistic and Comprehensive Bilingual Long-Context Evaluation Benchmark
por: Chen, Ziyang, et al.
Publicado: (2026)
por: Chen, Ziyang, et al.
Publicado: (2026)
CoRAG: Collaborative Retrieval-Augmented Generation
por: Muhamed, Aashiq, et al.
Publicado: (2025)
por: Muhamed, Aashiq, et al.
Publicado: (2025)
Personal Information Parroting in Language Models
por: Subramani, Nishant, et al.
Publicado: (2026)
por: Subramani, Nishant, et al.
Publicado: (2026)
LongGenBench: Benchmarking Long-Form Generation in Long Context LLMs
por: Wu, Yuhao, et al.
Publicado: (2024)
por: Wu, Yuhao, et al.
Publicado: (2024)
ChatBench: From Static Benchmarks to Human-AI Evaluation
por: Chang, Serina, et al.
Publicado: (2025)
por: Chang, Serina, et al.
Publicado: (2025)
ARCH2S: Dataset, Benchmark and Challenges for Learning Exterior Architectural Structures from Point Clouds
por: Cheung, Ka Lung, et al.
Publicado: (2024)
por: Cheung, Ka Lung, et al.
Publicado: (2024)
Res-Bench: Benchmarking the Robustness of Multimodal Large Language Models to Dynamic Resolution Input
por: Li, Chenxu, et al.
Publicado: (2025)
por: Li, Chenxu, et al.
Publicado: (2025)
Investigating Cultural Alignment of Large Language Models
por: AlKhamissi, Badr, et al.
Publicado: (2024)
por: AlKhamissi, Badr, et al.
Publicado: (2024)
QuantBench: Benchmarking AI Methods for Quantitative Investment
por: Wang, Saizhuo, et al.
Publicado: (2025)
por: Wang, Saizhuo, et al.
Publicado: (2025)
Depth-Wise Attention (DWAtt): A Layer Fusion Method for Data-Efficient Classification
por: ElNokrashy, Muhammad, et al.
Publicado: (2022)
por: ElNokrashy, Muhammad, et al.
Publicado: (2022)
Can Large Language Models Replace Human Coders? Introducing ContentBench
por: Haman, Michael
Publicado: (2026)
por: Haman, Michael
Publicado: (2026)
TableVQA-Bench: A Visual Question Answering Benchmark on Multiple Table Domains
por: Kim, Yoonsik, et al.
Publicado: (2024)
por: Kim, Yoonsik, et al.
Publicado: (2024)
Ejemplares similares
-
Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei
por: Wang, Peng, et al.
Publicado: (2026) -
Humanizing Machines: Rethinking LLM Anthropomorphism Through a Multi-Level Framework of Design
por: Xiao, Yunze, et al.
Publicado: (2025) -
Hire Your Anthropologist! Rethinking Culture Benchmarks Through an Anthropological Lens
por: AlKhamissi, Mai, et al.
Publicado: (2025) -
MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models
por: Wang, Lionel Z., et al.
Publicado: (2024) -
Towards Valid Student Simulation with Large Language Models
por: Yuan, Zhihao, et al.
Publicado: (2026)