TestAgent: An Adaptive and Intelligent Expert for Human Assessment
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Junhao, Zhuang, Yan, Sun, YuXuan, Gao, Weibo, Liu, Qi, Cheng, Mingyue, Huang, Zhenya, Chen, Enhong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Education Systems
by: Gao, Weibo, et al.
Published: (2025)
by: Gao, Weibo, et al.
Published: (2025)
Survey of Computerized Adaptive Testing: A Machine Learning Perspective
by: Zhuang, Yan, et al.
Published: (2024)
by: Zhuang, Yan, et al.
Published: (2024)
ARF-RLHF: Adaptive Reward-Following for RLHF through Emotion-Driven Self-Supervision and Trace-Biased Dynamic Optimization
by: Zhang, YuXuan
Published: (2025)
by: Zhang, YuXuan
Published: (2025)
A Survey of Knowledge Tracing: Models, Variants, and Applications
by: Shen, Shuanghong, et al.
Published: (2021)
by: Shen, Shuanghong, et al.
Published: (2021)
Unified Uncertainty Estimation for Cognitive Diagnosis Models
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
TestAgent: Automatic Benchmarking and Exploratory Interaction for Evaluating LLMs in Vertical Domains
by: Wang, Wanying, et al.
Published: (2024)
by: Wang, Wanying, et al.
Published: (2024)
Position: AI Evaluation Should Learn from How We Test Humans
by: Zhuang, Yan, et al.
Published: (2023)
by: Zhuang, Yan, et al.
Published: (2023)
Expert Assessment: The Systemic Environmental Risks of Artficial Intelligence
by: Schön, Julian, et al.
Published: (2025)
by: Schön, Julian, et al.
Published: (2025)
Conversational Learning Diagnosis via Reasoning Multi-Turn Interactive Learning
by: Yao, Fangzhou, et al.
Published: (2026)
by: Yao, Fangzhou, et al.
Published: (2026)
Emergent Social Intelligence Risks in Generative Multi-Agent Systems
by: Huang, Yue, et al.
Published: (2026)
by: Huang, Yue, et al.
Published: (2026)
When Agents See Humans as the Outgroup: Belief-Dependent Bias in LLM-Powered Agents
by: Wang, Zongwei, et al.
Published: (2026)
by: Wang, Zongwei, et al.
Published: (2026)
Mind2Report: A Cognitive Deep Research Agent for Expert-Level Commercial Report Synthesis
by: Cheng, Mingyue, et al.
Published: (2026)
by: Cheng, Mingyue, et al.
Published: (2026)
Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents
by: Li, Jiaxing, et al.
Published: (2026)
by: Li, Jiaxing, et al.
Published: (2026)
An Intelligent Mobile Application to Monitor and Correct Sitting Posture Using Raspberry Pi and MediaPipe Pose Detection
by: Yung-Chen, et al.
Published: (2025)
by: Yung-Chen, et al.
Published: (2025)
MIRACLE_Multi-Agent Intelligent Regulation to Advance Collaborative Learning Environment
by: Li, Shuang, et al.
Published: (2026)
by: Li, Shuang, et al.
Published: (2026)
The Lovelace Test of Intelligence: Can Humans Recognise and Esteem AI-Generated Art?
by: Gajewska, Ewelina
Published: (2025)
by: Gajewska, Ewelina
Published: (2025)
Making a Case for Research Collaboration Between Artificial Intelligence and Operations Research Experts
by: Kulkarni, Radhika, et al.
Published: (2025)
by: Kulkarni, Radhika, et al.
Published: (2025)
Cybroc: Cyborgizing Broccoli for Longevity
by: Huang, Ke, et al.
Published: (2025)
by: Huang, Ke, et al.
Published: (2025)
Explainable Ethical Assessment on Human Behaviors by Generating Conflicting Social Norms
by: Sun, Yuxi, et al.
Published: (2025)
by: Sun, Yuxi, et al.
Published: (2025)
DASKT: A Dynamic Affect Simulation Method for Knowledge Tracing
by: Sun, Xinjie, et al.
Published: (2025)
by: Sun, Xinjie, et al.
Published: (2025)
Foundation of Intelligence: Review of Math Word Problems from Human Cognition Perspective
by: Huang, Zhenya, et al.
Published: (2025)
by: Huang, Zhenya, et al.
Published: (2025)
Information Retrieval Induced Safety Degradation in AI Agents
by: Yu, Cheng, et al.
Published: (2025)
by: Yu, Cheng, et al.
Published: (2025)
Validated Hypotheses as a Lens for Human-Likeness Evaluation in AI Agents
by: Liu, Xuan, et al.
Published: (2026)
by: Liu, Xuan, et al.
Published: (2026)
S$^3$IT: A Benchmark for Spatially Situated Social Intelligence Test
by: Sun, Zhe, et al.
Published: (2025)
by: Sun, Zhe, et al.
Published: (2025)
The Use of Artificial Intelligence Tools in Assessing Content Validity: A Comparative Study with Human Experts
by: Gurdil, Hatice, et al.
Published: (2025)
by: Gurdil, Hatice, et al.
Published: (2025)
Adaptive Punishment in Social Dilemmas
by: Ke, Xingfu, et al.
Published: (2025)
by: Ke, Xingfu, et al.
Published: (2025)
TAI3: Testing Agent Integrity in Interpreting User Intent
by: Feng, Shiwei, et al.
Published: (2025)
by: Feng, Shiwei, et al.
Published: (2025)
AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios
by: Mou, Xinyi, et al.
Published: (2024)
by: Mou, Xinyi, et al.
Published: (2024)
Empowering Sequential Recommendation from Collaborative Signals and Semantic Relatedness
by: Cheng, Mingyue, et al.
Published: (2024)
by: Cheng, Mingyue, et al.
Published: (2024)
PEOAT: Personalization-Guided Evolutionary Question Assembly for One-Shot Adaptive Testing
by: Yu, Xiaoshan, et al.
Published: (2025)
by: Yu, Xiaoshan, et al.
Published: (2025)
The Adaptive Communication Framework (ACF) for Extraterrestrial Intelligence Discovery
by: Eldadi, Omer, et al.
Published: (2025)
by: Eldadi, Omer, et al.
Published: (2025)
Optimizing Student Ability Assessment: A Hierarchy Constraint-Aware Cognitive Diagnosis Framework for Educational Contexts
by: Sun, Xinjie, et al.
Published: (2024)
by: Sun, Xinjie, et al.
Published: (2024)
APS: Bias-Controlled Adaptive Prototype Simulation for Population-Scale LLM Agents
by: Zheng, Quan, et al.
Published: (2026)
by: Zheng, Quan, et al.
Published: (2026)
Exploring Heterogeneity and Uncertainty for Graph-based Cognitive Diagnosis Models in Intelligent Education
by: Shao, Pengyang, et al.
Published: (2024)
by: Shao, Pengyang, et al.
Published: (2024)
Can LLMs Help Decentralized Dispute Arbitration? A Case Study of UMA-Resolved Markets on Polymarket
by: Wen, Junhao, et al.
Published: (2026)
by: Wen, Junhao, et al.
Published: (2026)
Facilitating Cooperation in Human-Agent Hybrid Populations through Autonomous Agents
by: Guo, Hao, et al.
Published: (2023)
by: Guo, Hao, et al.
Published: (2023)
Epidemic spreading under game-based self-quarantine behaviors: The different effects of local and global information
by: Huang, Zegang, et al.
Published: (2023)
by: Huang, Zegang, et al.
Published: (2023)
A Comprehensive Survey of Artificial Intelligence Techniques for Talent Analytics
by: Qin, Chuan, et al.
Published: (2023)
by: Qin, Chuan, et al.
Published: (2023)
Infrastructure for Valuable, Tradable, and Verifiable Agent Memory
by: Li, Mengyuan, et al.
Published: (2026)
by: Li, Mengyuan, et al.
Published: (2026)
Exploring AI-Enabled Test Practice, Affect, and Test Outcomes in Language Assessment
by: Burstein, Jill, et al.
Published: (2025)
by: Burstein, Jill, et al.
Published: (2025)
Similar Items
-
Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Education Systems
by: Gao, Weibo, et al.
Published: (2025) -
Survey of Computerized Adaptive Testing: A Machine Learning Perspective
by: Zhuang, Yan, et al.
Published: (2024) -
ARF-RLHF: Adaptive Reward-Following for RLHF through Emotion-Driven Self-Supervision and Trace-Biased Dynamic Optimization
by: Zhang, YuXuan
Published: (2025) -
A Survey of Knowledge Tracing: Models, Variants, and Applications
by: Shen, Shuanghong, et al.
Published: (2021) -
Unified Uncertainty Estimation for Cognitive Diagnosis Models
by: Wang, Fei, et al.
Published: (2024)