Beyond single-channel agentic benchmarking
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Radpour, Nelu D. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Beyond Technocratic XAI: The Who, What & How in Explanation Design
par: Dhar, Ruchira, et autres
Publié: (2025)
par: Dhar, Ruchira, et autres
Publié: (2025)
Artificial Intelligence: Beyound Ocularcentrism, the New Age of Humans Beyond the Spectacle
par: Moussaoui, Mustapha El
Publié: (2026)
par: Moussaoui, Mustapha El
Publié: (2026)
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
par: Baum, Kevin
Publié: (2025)
par: Baum, Kevin
Publié: (2025)
Empathic and agentic artificial intelligence in nursing: perspectives on a human-centered framework for cancer care navigation in the United States
par: Girdwood, Tyra, et autres
Publié: (2026)
par: Girdwood, Tyra, et autres
Publié: (2026)
Beyond Detection: Governing GenAI in Academic Peer Review as a Sociotechnical Challenge
par: Chakravorti, Tatiana, et autres
Publié: (2026)
par: Chakravorti, Tatiana, et autres
Publié: (2026)
Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models
par: Singh, Divyanshu Kumar, et autres
Publié: (2026)
par: Singh, Divyanshu Kumar, et autres
Publié: (2026)
Designing Beyond Language: Sociotechnical Barriers in AI Health Technologies for Limited English Proficiency
par: Huang, Michelle, et autres
Publié: (2025)
par: Huang, Michelle, et autres
Publié: (2025)
Beyond Procedural Compliance: Human Oversight as a Dimension of Well-being Efficacy in AI Governance
par: Xie, Yao, et autres
Publié: (2025)
par: Xie, Yao, et autres
Publié: (2025)
The AI Cognitive Trojan Horse: How Large Language Models May Bypass Human Epistemic Vigilance
par: Maynard, Andrew D.
Publié: (2026)
par: Maynard, Andrew D.
Publié: (2026)
An Innovative Solution: AI-Based Digital Screen-Integrated Tables for Educational Settings
par: Tamang, S., et autres
Publié: (2024)
par: Tamang, S., et autres
Publié: (2024)
Tinker Tales: Interactive Storytelling Framework for Early Childhood Narrative Development and AI Literacy
par: Choi, Nayoung, et autres
Publié: (2025)
par: Choi, Nayoung, et autres
Publié: (2025)
Which Contributions Deserve Credit? Perceptions of Attribution in Human-AI Co-Creation
par: He, Jessica, et autres
Publié: (2025)
par: He, Jessica, et autres
Publié: (2025)
Beyond Models: A Framework for Contextual and Cultural Intelligence in African AI Deployment
par: Ndlovu, Qness
Publié: (2025)
par: Ndlovu, Qness
Publié: (2025)
Aligning Tutor Discourse Supporting Rigorous Thinking with Tutee Content Mastery for Predicting Math Achievement
par: Abdelshiheed, Mark, et autres
Publié: (2024)
par: Abdelshiheed, Mark, et autres
Publié: (2024)
Pensieve Discuss: Scalable Small-Group CS Tutoring System with AI
par: Yang, Yoonseok, et autres
Publié: (2024)
par: Yang, Yoonseok, et autres
Publié: (2024)
Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions
par: Jiang, Yuyang, et autres
Publié: (2025)
par: Jiang, Yuyang, et autres
Publié: (2025)
Artificial Intelligence as a Training Tool in Clinical Psychology: A Comparison of Text-Based and Avatar Simulations
par: Sawah, V. El, et autres
Publié: (2025)
par: Sawah, V. El, et autres
Publié: (2025)
From job titles to jawlines: Using context voids to study generative AI systems
par: Memon, Shahan Ali, et autres
Publié: (2025)
par: Memon, Shahan Ali, et autres
Publié: (2025)
Environment Scan of Generative AI Infrastructure for Clinical and Translational Science
par: Idnay, Betina, et autres
Publié: (2024)
par: Idnay, Betina, et autres
Publié: (2024)
Orchestrating LLM Agents for Scientific Research: A Pilot Study of Multiple Choice Question (MCQ) Generation and Evaluation
par: An, Yuan
Publié: (2026)
par: An, Yuan
Publié: (2026)
Privacy in Human-AI Romantic Relationships: Concerns, Boundaries, and Agency
par: Ma, Rongjun, et autres
Publié: (2026)
par: Ma, Rongjun, et autres
Publié: (2026)
Accuracy Standards for AI at Work vs. Personal Life: Evidence from an Online Survey
par: Besanson, Gaston, et autres
Publié: (2026)
par: Besanson, Gaston, et autres
Publié: (2026)
Taming Toxic Talk: Using chatbots to intervene with users posting toxic comments
par: Foote, Jeremy, et autres
Publié: (2026)
par: Foote, Jeremy, et autres
Publié: (2026)
How AI Impacts Skill Formation
par: Shen, Judy Hanwen, et autres
Publié: (2026)
par: Shen, Judy Hanwen, et autres
Publié: (2026)
Towards Apples to Apples for AI Evaluations: From Real-World Use Cases to Evaluation Scenarios
par: Choong, Yee-Yin, et autres
Publié: (2026)
par: Choong, Yee-Yin, et autres
Publié: (2026)
Unilateral Relationship Revision Power in Human-AI Companion Interaction
par: Lange, Benjamin
Publié: (2026)
par: Lange, Benjamin
Publié: (2026)
Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South
par: Rastogi, Charvi, et autres
Publié: (2026)
par: Rastogi, Charvi, et autres
Publié: (2026)
Struggle Premium : How Human Effort and Imperfection Drive Perceived Value in the Age of AI
par: Sultana, Nazneen, et autres
Publié: (2026)
par: Sultana, Nazneen, et autres
Publié: (2026)
Bureaucratic Silences: What the Canadian AI Register Reveals, Omits, and Obscures
par: Das, Dipto, et autres
Publié: (2026)
par: Das, Dipto, et autres
Publié: (2026)
Artificial intelligence can persuade people to take political actions
par: Hackenburg, Kobi, et autres
Publié: (2026)
par: Hackenburg, Kobi, et autres
Publié: (2026)
The Imbalanced User-AI Relationships as an Ethical Failure of Front-End Design in Healthcare AI
par: Mwadime, Maureen Mghambi
Publié: (2026)
par: Mwadime, Maureen Mghambi
Publié: (2026)
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers
par: Chen, Benjamin Minhao, et autres
Publié: (2026)
par: Chen, Benjamin Minhao, et autres
Publié: (2026)
Incentives shape how humans co-create with generative AI
par: Jo, Nathanael, et autres
Publié: (2026)
par: Jo, Nathanael, et autres
Publié: (2026)
Enhancing behavioral nudges with large language model-based iterative personalization: A field experiment on electricity and hot-water conservation
par: Li, Zonghan, et autres
Publié: (2026)
par: Li, Zonghan, et autres
Publié: (2026)
AI Meets Mathematics Education: A Case Study on Supporting an Instructor in a Large Mathematics Class with Context-Aware AI
par: Barghorn, Jérémy, et autres
Publié: (2026)
par: Barghorn, Jérémy, et autres
Publié: (2026)
A federated architecture for sector-led AI governance: lessons from India
par: Agarwal, Avinash, et autres
Publié: (2026)
par: Agarwal, Avinash, et autres
Publié: (2026)
Chat-Based Support Alone May Not Be Enough: Comparing Conversational and Embedded LLM Feedback for Mathematical Proof Learning
par: Chen, Eason, et autres
Publié: (2026)
par: Chen, Eason, et autres
Publié: (2026)
Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decision Support in Industrial Tasks
par: Figliè, Roberto, et autres
Publié: (2026)
par: Figliè, Roberto, et autres
Publié: (2026)
Faster Completion, Less Learning: Generative AI Reduced Study Time on Math Problems and the Knowledge They Build
par: Rismanchian, Sina, et autres
Publié: (2026)
par: Rismanchian, Sina, et autres
Publié: (2026)
Addressing the Synergy Gap: The Six Elements of the Design Space
par: Turchi, Tommaso, et autres
Publié: (2026)
par: Turchi, Tommaso, et autres
Publié: (2026)
Documents similaires
-
Beyond Technocratic XAI: The Who, What & How in Explanation Design
par: Dhar, Ruchira, et autres
Publié: (2025) -
Artificial Intelligence: Beyound Ocularcentrism, the New Age of Humans Beyond the Spectacle
par: Moussaoui, Mustapha El
Publié: (2026) -
Disentangling AI Alignment: A Structured Taxonomy Beyond Safety and Ethics
par: Baum, Kevin
Publié: (2025) -
Empathic and agentic artificial intelligence in nursing: perspectives on a human-centered framework for cancer care navigation in the United States
par: Girdwood, Tyra, et autres
Publié: (2026) -
Beyond Detection: Governing GenAI in Academic Peer Review as a Sociotechnical Challenge
par: Chakravorti, Tatiana, et autres
Publié: (2026)