AI Awareness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Xiaojian, Shi, Haoyuan, Xu, Rongwu, Xu, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Nuclear Deployed: Analyzing Catastrophic Risks in Decision-making of Autonomous LLM Agents
von: Xu, Rongwu, et al.
Veröffentlicht: (2025)
von: Xu, Rongwu, et al.
Veröffentlicht: (2025)
Walking in Others' Shoes: How Perspective-Taking Guides Large Language Models in Reducing Toxicity and Bias
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
Exploring Chinese Humor Generation: A Study on Two-Part Allegorical Sayings
von: Xu, Rongwu
Veröffentlicht: (2024)
von: Xu, Rongwu
Veröffentlicht: (2024)
The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
von: Xu, Rongwu, et al.
Veröffentlicht: (2023)
von: Xu, Rongwu, et al.
Veröffentlicht: (2023)
How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)
Preemptive Answer "Attacks" on Chain-of-Thought Reasoning
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
MCTSr-Zero: Self-Reflective Psychological Counseling Dialogues Generation via Principles and Adaptive Exploration
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
How to Protect Yourself from 5G Radiation? Investigating LLM Responses to Implicit Misinformation
von: Guo, Ruohao, et al.
Veröffentlicht: (2025)
von: Guo, Ruohao, et al.
Veröffentlicht: (2025)
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
von: Oh, Gyutaek, et al.
Veröffentlicht: (2025)
von: Oh, Gyutaek, et al.
Veröffentlicht: (2025)
Toxicity-Aware Few-Shot Prompting for Low-Resource Singlish Translation
von: Ge, Ziyu, et al.
Veröffentlicht: (2025)
von: Ge, Ziyu, et al.
Veröffentlicht: (2025)
When AI Navigates the Fog of War
von: Li, Ming, et al.
Veröffentlicht: (2026)
von: Li, Ming, et al.
Veröffentlicht: (2026)
UrbanGPT: Spatio-Temporal Large Language Models
von: Li, Zhonghang, et al.
Veröffentlicht: (2024)
von: Li, Zhonghang, et al.
Veröffentlicht: (2024)
Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
Intelligent Tutor: Leveraging ChatGPT and Microsoft Copilot Studio to Deliver a Generative AI Student Support and Feedback System within Teams
von: Chen, Wei-Yu
Veröffentlicht: (2024)
von: Chen, Wei-Yu
Veröffentlicht: (2024)
Does Socialization Emerge in AI Agent Society? A Case Study of Moltbook
von: Li, Ming, et al.
Veröffentlicht: (2026)
von: Li, Ming, et al.
Veröffentlicht: (2026)
Writing in Symbiosis: Mapping Human Creative Agency in the AI Era
von: Doshi, Vivan, et al.
Veröffentlicht: (2025)
von: Doshi, Vivan, et al.
Veröffentlicht: (2025)
From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
Reinforcing Stereotypes of Anger: Emotion AI on African American Vernacular English
von: Dorn, Rebecca, et al.
Veröffentlicht: (2025)
von: Dorn, Rebecca, et al.
Veröffentlicht: (2025)
AI-Assisted Systematization for Evaluating GenAI Systems
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2026)
von: Agarwal, Dhruv, et al.
Veröffentlicht: (2026)
Polarized Patterns of Language Toxicity and Sentiment of Debunking Posts on Social Media
von: Xu, Wentao, et al.
Veröffentlicht: (2025)
von: Xu, Wentao, et al.
Veröffentlicht: (2025)
Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap
von: Qi, Xuan, et al.
Veröffentlicht: (2025)
von: Qi, Xuan, et al.
Veröffentlicht: (2025)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
von: Xu, Shanshan, et al.
Veröffentlicht: (2025)
von: Xu, Shanshan, et al.
Veröffentlicht: (2025)
Verifying Rumors via Stance-Aware Structural Modeling
von: Nkhata, Gibson, et al.
Veröffentlicht: (2025)
von: Nkhata, Gibson, et al.
Veröffentlicht: (2025)
Evaluating Cultural Awareness of LLMs for Yoruba, Malayalam, and English
von: Dawson, Fiifi, et al.
Veröffentlicht: (2024)
von: Dawson, Fiifi, et al.
Veröffentlicht: (2024)
A Tale of Two Identities: An Ethical Audit of Human and AI-Crafted Personas
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2025)
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations
von: Li, Jiatong, et al.
Veröffentlicht: (2024)
von: Li, Jiatong, et al.
Veröffentlicht: (2024)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
von: Li, Jing-Jing, et al.
Veröffentlicht: (2026)
von: Li, Jing-Jing, et al.
Veröffentlicht: (2026)
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2026)
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2026)
Position: The Current AI Conference Model is Unsustainable! Diagnosing the Crisis of Centralized AI Conference
von: Chen, Nuo, et al.
Veröffentlicht: (2025)
von: Chen, Nuo, et al.
Veröffentlicht: (2025)
AI Literacy in Low-Resource Languages:Insights from creating AI in Yoruba videos
von: Oyewusi, Wuraola
Veröffentlicht: (2024)
von: Oyewusi, Wuraola
Veröffentlicht: (2024)
Knowledge Acquisition on Mass-shooting Events via LLMs for AI-Driven Justice
von: Ihugba, Benign John, et al.
Veröffentlicht: (2025)
von: Ihugba, Benign John, et al.
Veröffentlicht: (2025)
Meta-Cultural Competence: Climbing the Right Hill of Cultural Awareness
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning
von: Gao, Zihan, et al.
Veröffentlicht: (2025)
von: Gao, Zihan, et al.
Veröffentlicht: (2025)
Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
von: Wu, Addison J., et al.
Veröffentlicht: (2026)
von: Wu, Addison J., et al.
Veröffentlicht: (2026)
Human-AI Collaboration or Academic Misconduct? Measuring AI Use in Student Writing Through Stylometric Evidence
von: Oliveira, Eduardo Araujo, et al.
Veröffentlicht: (2025)
von: Oliveira, Eduardo Araujo, et al.
Veröffentlicht: (2025)
Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity
von: Najjar, Ayat A., et al.
Veröffentlicht: (2025)
von: Najjar, Ayat A., et al.
Veröffentlicht: (2025)
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
von: Li, Minzhi, et al.
Veröffentlicht: (2024)
von: Li, Minzhi, et al.
Veröffentlicht: (2024)
Berta: an open-source, modular tool for AI-enabled clinical documentation
von: Vaid, Samridhi, et al.
Veröffentlicht: (2026)
von: Vaid, Samridhi, et al.
Veröffentlicht: (2026)
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race
von: Sun, Lihao, et al.
Veröffentlicht: (2025)
von: Sun, Lihao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Nuclear Deployed: Analyzing Catastrophic Risks in Decision-making of Autonomous LLM Agents
von: Xu, Rongwu, et al.
Veröffentlicht: (2025) -
Walking in Others' Shoes: How Perspective-Taking Guides Large Language Models in Reducing Toxicity and Bias
von: Xu, Rongwu, et al.
Veröffentlicht: (2024) -
Exploring Chinese Humor Generation: A Study on Two-Part Allegorical Sayings
von: Xu, Rongwu
Veröffentlicht: (2024) -
The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
von: Xu, Rongwu, et al.
Veröffentlicht: (2023) -
How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)