From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Yi-Fan, Li, Hang, Song, Dingjie, Sun, Lichao, Xu, Tianlong, Wen, Qingsong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
por: Song, Dingjie, et al.
Publicado: (2026)
por: Song, Dingjie, et al.
Publicado: (2026)
AI-Driven Virtual Teacher for Enhanced Educational Efficiency: Leveraging Large Pretrain Models for Autonomous Error Analysis and Correction
por: Xu, Tianlong, et al.
Publicado: (2024)
por: Xu, Tianlong, et al.
Publicado: (2024)
Aligning Multimodal LLM with Human Preference: A Survey
por: Yu, Tao, et al.
Publicado: (2025)
por: Yu, Tao, et al.
Publicado: (2025)
Both Text and Images Leaked! A Systematic Analysis of Data Contamination in Multimodal LLM
por: Song, Dingjie, et al.
Publicado: (2024)
por: Song, Dingjie, et al.
Publicado: (2024)
ZeroMamba: Exploring Visual State Space Model for Zero-Shot Learning
por: Hou, Wenjin, et al.
Publicado: (2024)
por: Hou, Wenjin, et al.
Publicado: (2024)
SAMed-2: Selective Memory Enhanced Medical Segment Anything Model
por: Yan, Zhiling, et al.
Publicado: (2025)
por: Yan, Zhiling, et al.
Publicado: (2025)
LongLLaVA: Scaling Multi-modal LLMs to 1000 Images Efficiently via a Hybrid Architecture
por: Wang, Xidong, et al.
Publicado: (2024)
por: Wang, Xidong, et al.
Publicado: (2024)
Stable Unlearnable Example: Enhancing the Robustness of Unlearnable Examples via Stable Error-Minimizing Noise
por: Liu, Yixin, et al.
Publicado: (2023)
por: Liu, Yixin, et al.
Publicado: (2023)
LaneCorrect: Self-supervised Lane Detection
por: Nie, Ming, et al.
Publicado: (2024)
por: Nie, Ming, et al.
Publicado: (2024)
Tailored Visions: Enhancing Text-to-Image Generation with Personalized Prompt Rewriting
por: Chen, Zijie, et al.
Publicado: (2023)
por: Chen, Zijie, et al.
Publicado: (2023)
Beyond LLaVA-HD: Diving into High-Resolution Large Multimodal Models
por: Zhang, Yi-Fan, et al.
Publicado: (2024)
por: Zhang, Yi-Fan, et al.
Publicado: (2024)
LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
por: Xu, Guowei, et al.
Publicado: (2024)
por: Xu, Guowei, et al.
Publicado: (2024)
CADReview: Automatically Reviewing CAD Programs with Error Detection and Correction
por: Chen, Jiali, et al.
Publicado: (2025)
por: Chen, Jiali, et al.
Publicado: (2025)
From Preoperative CT to Postmastoidectomy Mesh Construction: Mastoidectomy Shape Prediction for Cochlear Implant Surgery
por: Zhang, Yike, et al.
Publicado: (2026)
por: Zhang, Yike, et al.
Publicado: (2026)
PersonaVlog: Personalized Multimodal Vlog Generation with Multi-Agent Collaboration and Iterative Self-Correction
por: Hou, Xiaolu, et al.
Publicado: (2025)
por: Hou, Xiaolu, et al.
Publicado: (2025)
Rethinking and Red-Teaming Protective Perturbation in Personalized Diffusion Models
por: Liu, Yixin, et al.
Publicado: (2024)
por: Liu, Yixin, et al.
Publicado: (2024)
Fairness in Multi-modal Medical Diagnosis with Demonstration Selection
por: Li, Dawei, et al.
Publicado: (2025)
por: Li, Dawei, et al.
Publicado: (2025)
TransMed: Large Language Models Enhance Vision Transformer for Biomedical Image Classification
por: Zheng, Kaipeng, et al.
Publicado: (2023)
por: Zheng, Kaipeng, et al.
Publicado: (2023)
A Comprehensive Overview of Fish-Eye Camera Distortion Correction Methods
por: Xu, Jian, et al.
Publicado: (2023)
por: Xu, Jian, et al.
Publicado: (2023)
PersONAL: Towards a Comprehensive Benchmark for Personalized Embodied Agents
por: Ziliotto, Filippo, et al.
Publicado: (2025)
por: Ziliotto, Filippo, et al.
Publicado: (2025)
Personalized Face Super-Resolution with Identity Decoupling and Fitting
por: Yang, Jiarui, et al.
Publicado: (2025)
por: Yang, Jiarui, et al.
Publicado: (2025)
CAMeL: Cross-modality Adaptive Meta-Learning for Text-based Person Retrieval
por: Yu, Hang, et al.
Publicado: (2025)
por: Yu, Hang, et al.
Publicado: (2025)
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
por: Xu, Zhiyu, et al.
Publicado: (2025)
por: Xu, Zhiyu, et al.
Publicado: (2025)
Can Synthetic Images Serve as Effective and Efficient Class Prototypes?
por: Shi, Dianxing, et al.
Publicado: (2025)
por: Shi, Dianxing, et al.
Publicado: (2025)
FGP: Feature-Gradient-Prune for Efficient Convolutional Layer Pruning
por: Lv, Qingsong, et al.
Publicado: (2024)
por: Lv, Qingsong, et al.
Publicado: (2024)
From Calibration to Refinement: Seeking Certainty via Probabilistic Evidence Propagation for Noisy-Label Person Re-Identification
por: Yuan, Xin, et al.
Publicado: (2026)
por: Yuan, Xin, et al.
Publicado: (2026)
From Prediction to Diagnosis: Reasoning-Aware AI for Photovoltaic Defect Inspection
por: Mistry, Dev, et al.
Publicado: (2026)
por: Mistry, Dev, et al.
Publicado: (2026)
Fast Person Detection Using YOLOX With AI Accelerator For Train Station Safety
por: Achmadiah, Mas Nurul, et al.
Publicado: (2026)
por: Achmadiah, Mas Nurul, et al.
Publicado: (2026)
Hypothesis Graph Refinement: Hypothesis-Driven Exploration with Cascade Error Correction for Embodied Navigation
por: Chen, Peixin, et al.
Publicado: (2026)
por: Chen, Peixin, et al.
Publicado: (2026)
Debiasing Multimodal Large Language Models via Penalization of Language Priors
por: Zhang, YiFan, et al.
Publicado: (2024)
por: Zhang, YiFan, et al.
Publicado: (2024)
Ear-Keeper: A Cross-Platform AI System for Rapid and Accurate Ear Disease Diagnosis
por: Lu, Feiyan, et al.
Publicado: (2023)
por: Lu, Feiyan, et al.
Publicado: (2023)
Multi-Agent System for Comprehensive Soccer Understanding
por: Rao, Jiayuan, et al.
Publicado: (2025)
por: Rao, Jiayuan, et al.
Publicado: (2025)
Physical Backdoor: Towards Temperature-based Backdoor Attacks in the Physical World
por: Yin, Wen, et al.
Publicado: (2024)
por: Yin, Wen, et al.
Publicado: (2024)
Adversarial Error Correction for Visual Autoregressive Generation
por: Bi, Ligong, et al.
Publicado: (2026)
por: Bi, Ligong, et al.
Publicado: (2026)
Self-supervised Mamba-based Mastoidectomy Shape Prediction for Cochlear Implant Surgery
por: Zhang, Yike, et al.
Publicado: (2024)
por: Zhang, Yike, et al.
Publicado: (2024)
Monocular Microscope to CT Registration using Pose Estimation of the Incus for Augmented Reality Cochlear Implant Surgery
por: Zhang, Yike, et al.
Publicado: (2024)
por: Zhang, Yike, et al.
Publicado: (2024)
CogAgent: A Visual Language Model for GUI Agents
por: Hong, Wenyi, et al.
Publicado: (2023)
por: Hong, Wenyi, et al.
Publicado: (2023)
A Comprehensive Survey on World Models for Embodied AI
por: Li, Xinqing, et al.
Publicado: (2025)
por: Li, Xinqing, et al.
Publicado: (2025)
LLM4Brain: Training a Large Language Model for Brain Video Understanding
por: Zheng, Ruizhe, et al.
Publicado: (2024)
por: Zheng, Ruizhe, et al.
Publicado: (2024)
LungNoduleAgent: A Collaborative Multi-Agent System for Precision Diagnosis of Lung Nodules
por: Yang, Cheng, et al.
Publicado: (2025)
por: Yang, Cheng, et al.
Publicado: (2025)
Ejemplares similares
-
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
por: Song, Dingjie, et al.
Publicado: (2026) -
AI-Driven Virtual Teacher for Enhanced Educational Efficiency: Leveraging Large Pretrain Models for Autonomous Error Analysis and Correction
por: Xu, Tianlong, et al.
Publicado: (2024) -
Aligning Multimodal LLM with Human Preference: A Survey
por: Yu, Tao, et al.
Publicado: (2025) -
Both Text and Images Leaked! A Systematic Analysis of Data Contamination in Multimodal LLM
por: Song, Dingjie, et al.
Publicado: (2024) -
ZeroMamba: Exploring Visual State Space Model for Zero-Shot Learning
por: Hou, Wenjin, et al.
Publicado: (2024)