Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion
Fuente:
arXiv
Saved in:
| Main Author: | Leitch, Terry |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking AI Evaluation in Education: The TEACH-AI Framework and Benchmark for Generative AI Assistants
by: Ding, Shi, et al.
Published: (2025)
by: Ding, Shi, et al.
Published: (2025)
State Your Intention to Steer Your Attention: An AI Assistant for Intentional Digital Living
by: Choi, Juheon, et al.
Published: (2025)
by: Choi, Juheon, et al.
Published: (2025)
Panza: Design and Analysis of a Fully-Local Personalized Text Writing Assistant
by: Nicolicioiu, Armand, et al.
Published: (2024)
by: Nicolicioiu, Armand, et al.
Published: (2024)
Evaluating Machine Expertise: How Graduate Students Develop Frameworks for Assessing GenAI Content
by: Chen, Celia, et al.
Published: (2025)
by: Chen, Celia, et al.
Published: (2025)
Benchmark It Yourself (BIY): Preparing a Dataset and Benchmarking AI Models for Scatterplot-Related Tasks
by: Palmeiro, João, et al.
Published: (2025)
by: Palmeiro, João, et al.
Published: (2025)
Detection of adversarial intent in Human-AI teams using LLMs
by: Musaffar, Abed K., et al.
Published: (2026)
by: Musaffar, Abed K., et al.
Published: (2026)
From Accuracy to Readiness: Metrics and Benchmarks for Human-AI Decision-Making
by: Lee, Min Hun
Published: (2026)
by: Lee, Min Hun
Published: (2026)
Combating Spatial Disorientation in a Dynamic Self-Stabilization Task Using AI Assistants
by: Mannan, Sheikh, et al.
Published: (2024)
by: Mannan, Sheikh, et al.
Published: (2024)
LLMs as Writing Assistants: Exploring Perspectives on Sense of Ownership and Reasoning
by: Wasi, Azmine Toushik, et al.
Published: (2024)
by: Wasi, Azmine Toushik, et al.
Published: (2024)
Farsight: Fostering Responsible AI Awareness During AI Application Prototyping
by: Wang, Zijie J., et al.
Published: (2024)
by: Wang, Zijie J., et al.
Published: (2024)
Why and When LLM-Based Assistants Can Go Wrong: Investigating the Effectiveness of Prompt-Based Interactions for Software Help-Seeking
by: Khurana, Anjali, et al.
Published: (2024)
by: Khurana, Anjali, et al.
Published: (2024)
AI, Meet Human: Learning Paradigms for Hybrid Decision Making Systems
by: Punzi, Clara, et al.
Published: (2024)
by: Punzi, Clara, et al.
Published: (2024)
Cloud Infrastructure Management in the Age of AI Agents
by: Yang, Zhenning, et al.
Published: (2025)
by: Yang, Zhenning, et al.
Published: (2025)
Agent Laboratory: Using LLM Agents as Research Assistants
by: Schmidgall, Samuel, et al.
Published: (2025)
by: Schmidgall, Samuel, et al.
Published: (2025)
Optimizing Feature Extraction for On-device Model Inference with User Behavior Sequences
by: Gong, Chen, et al.
Published: (2026)
by: Gong, Chen, et al.
Published: (2026)
"Excuse me, may I say something..." CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations
by: Wu, Yang, et al.
Published: (2026)
by: Wu, Yang, et al.
Published: (2026)
Interaction Dynamics as a Reward Signal for LLMs
by: Gooding, Sian, et al.
Published: (2025)
by: Gooding, Sian, et al.
Published: (2025)
Measuring What Matters: Connecting AI Ethics Evaluations to System Attributes, Hazards, and Harms
by: Rismani, Shalaleh, et al.
Published: (2025)
by: Rismani, Shalaleh, et al.
Published: (2025)
A Resilient Solution for Sewer Overflow Monitoring across Cloud and Edge
by: Singh, Vipin, et al.
Published: (2026)
by: Singh, Vipin, et al.
Published: (2026)
HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants
by: Sturgeon, Benjamin, et al.
Published: (2025)
by: Sturgeon, Benjamin, et al.
Published: (2025)
Efficient Human-in-the-Loop Active Learning: A Novel Framework for Data Labeling in AI Systems
by: Huang, Yiran, et al.
Published: (2024)
by: Huang, Yiran, et al.
Published: (2024)
MetaExplainer: A Framework to Generate Multi-Type User-Centered Explanations for AI Systems
by: Chari, Shruthi, et al.
Published: (2025)
by: Chari, Shruthi, et al.
Published: (2025)
IronEngine: Towards General AI Assistant
by: Mo, Xi
Published: (2026)
by: Mo, Xi
Published: (2026)
Model-in-the-Loop (MILO): Accelerating Multimodal AI Data Annotation with LLMs
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
MHDash: An Online Platform for Benchmarking Mental Health-Aware AI Assistants
by: Zhang, Yihe, et al.
Published: (2026)
by: Zhang, Yihe, et al.
Published: (2026)
LabelAId: Just-in-time AI Interventions for Improving Human Labeling Quality and Domain Knowledge in Crowdsourcing Systems
by: Li, Chu, et al.
Published: (2024)
by: Li, Chu, et al.
Published: (2024)
Enabling On-Device LLMs Personalization with Smartphone Sensing
by: Zhang, Shiquan, et al.
Published: (2024)
by: Zhang, Shiquan, et al.
Published: (2024)
LingVarBench: Benchmarking LLMs on Entity Recognitions and Linguistic Verbalization Patterns in Phone-Call Transcripts
by: Mohammadi, Seyedali, et al.
Published: (2025)
by: Mohammadi, Seyedali, et al.
Published: (2025)
Domain-Grounded Evaluation of LLMs in International Student Knowledge
by: Daitx, Claudinei, et al.
Published: (2025)
by: Daitx, Claudinei, et al.
Published: (2025)
Designing Interpretable ML System to Enhance Trust in Healthcare: A Systematic Review to Proposed Responsible Clinician-AI-Collaboration Framework
by: Nasarian, Elham, et al.
Published: (2023)
by: Nasarian, Elham, et al.
Published: (2023)
The case for delegated AI autonomy for Human AI teaming in healthcare
by: Jia, Yan, et al.
Published: (2025)
by: Jia, Yan, et al.
Published: (2025)
Explaining AI Without Code: A User Study on Explainable AI
by: Abarca, Natalia, et al.
Published: (2025)
by: Abarca, Natalia, et al.
Published: (2025)
Benchmarking Mobile Device Control Agents across Diverse Configurations
by: Lee, Juyong, et al.
Published: (2024)
by: Lee, Juyong, et al.
Published: (2024)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
by: Yu, Guanghui, et al.
Published: (2024)
by: Yu, Guanghui, et al.
Published: (2024)
Improving Health Professionals' Onboarding with AI and XAI for Trustworthy Human-AI Collaborative Decision Making
by: Lee, Min Hun, et al.
Published: (2024)
by: Lee, Min Hun, et al.
Published: (2024)
Human-AI Collaborative Uncertainty Quantification
by: Noorani, Sima, et al.
Published: (2025)
by: Noorani, Sima, et al.
Published: (2025)
Everyday AR through AI-in-the-Loop
by: Suzuki, Ryo, et al.
Published: (2024)
by: Suzuki, Ryo, et al.
Published: (2024)
TutoAI: A Cross-domain Framework for AI-assisted Mixed-media Tutorial Creation on Physical Tasks
by: Chen, Yuexi, et al.
Published: (2024)
by: Chen, Yuexi, et al.
Published: (2024)
Interactive Example-based Explanations to Improve Health Professionals' Onboarding with AI for Human-AI Collaborative Decision Making
by: Lee, Min Hun, et al.
Published: (2024)
by: Lee, Min Hun, et al.
Published: (2024)
HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Similar Items
-
Rethinking AI Evaluation in Education: The TEACH-AI Framework and Benchmark for Generative AI Assistants
by: Ding, Shi, et al.
Published: (2025) -
State Your Intention to Steer Your Attention: An AI Assistant for Intentional Digital Living
by: Choi, Juheon, et al.
Published: (2025) -
Panza: Design and Analysis of a Fully-Local Personalized Text Writing Assistant
by: Nicolicioiu, Armand, et al.
Published: (2024) -
Evaluating Machine Expertise: How Graduate Students Develop Frameworks for Assessing GenAI Content
by: Chen, Celia, et al.
Published: (2025) -
Benchmark It Yourself (BIY): Preparing a Dataset and Benchmarking AI Models for Scatterplot-Related Tasks
by: Palmeiro, João, et al.
Published: (2025)