AI-Driven Virtual Teacher for Enhanced Educational Efficiency: Leveraging Large Pretrain Models for Autonomous Error Analysis and Correction
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Tianlong, Zhang, Yi-Fan, Chu, Zhendong, Wang, Shen, Wen, Qingsong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multimodal AI Teacher: Integrating Edge Computing and Reasoning Models for Enhanced Student Error Analysis
di: Tianlong Xu, et al.
Pubblicazione: (2025)
di: Tianlong Xu, et al.
Pubblicazione: (2025)
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education
di: Zhang, Yi-Fan, et al.
Pubblicazione: (2025)
di: Zhang, Yi-Fan, et al.
Pubblicazione: (2025)
CustomContrast: A Multilevel Contrastive Perspective For Subject-Driven Text-to-Image Customization
di: Chen, Nan, et al.
Pubblicazione: (2024)
di: Chen, Nan, et al.
Pubblicazione: (2024)
TraveLLaMA: A Multimodal Travel Assistant with Large-Scale Dataset and Structured Reasoning
di: Chu, Meng, et al.
Pubblicazione: (2025)
di: Chu, Meng, et al.
Pubblicazione: (2025)
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
di: Zhou, Tian-Yi, et al.
Pubblicazione: (2026)
di: Zhou, Tian-Yi, et al.
Pubblicazione: (2026)
Reasoning Like Experts: Leveraging Multimodal Large Language Models for Drawing-based Psychoanalysis
di: Ma, Xueqi, et al.
Pubblicazione: (2025)
di: Ma, Xueqi, et al.
Pubblicazione: (2025)
FoodMLLM-JP: Leveraging Multimodal Large Language Models for Japanese Recipe Generation
di: Imajuku, Yuki, et al.
Pubblicazione: (2024)
di: Imajuku, Yuki, et al.
Pubblicazione: (2024)
Provably Secure Robust Image Steganography via Cross-Modal Error Correction
di: Qi, Yuang, et al.
Pubblicazione: (2024)
di: Qi, Yuang, et al.
Pubblicazione: (2024)
XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments
di: Qian, Kangan, et al.
Pubblicazione: (2026)
di: Qian, Kangan, et al.
Pubblicazione: (2026)
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
di: Wang, Hang, et al.
Pubblicazione: (2026)
di: Wang, Hang, et al.
Pubblicazione: (2026)
Querying Autonomous Vehicle Point Clouds: Enhanced by 3D Object Counting with CounterNet
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2025)
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
di: Li, Deng, et al.
Pubblicazione: (2024)
di: Li, Deng, et al.
Pubblicazione: (2024)
Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis
di: Hao, Jing, et al.
Pubblicazione: (2025)
di: Hao, Jing, et al.
Pubblicazione: (2025)
Waymo-3DSkelMo: A Multi-Agent 3D Skeletal Motion Dataset for Pedestrian Interaction Modeling in Autonomous Driving
di: Zhu, Guangxun, et al.
Pubblicazione: (2025)
di: Zhu, Guangxun, et al.
Pubblicazione: (2025)
Enhancing Fake News Video Detection via LLM-Driven Creative Process Simulation
di: Bu, Yuyan, et al.
Pubblicazione: (2025)
di: Bu, Yuyan, et al.
Pubblicazione: (2025)
RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
di: Huang, Zhijian, et al.
Pubblicazione: (2024)
di: Huang, Zhijian, et al.
Pubblicazione: (2024)
MSCrackMamba: Leveraging Vision Mamba for Crack Detection in Fused Multispectral Imagery
di: Zhu, Qinfeng, et al.
Pubblicazione: (2024)
di: Zhu, Qinfeng, et al.
Pubblicazione: (2024)
Context-Enhanced Video Moment Retrieval with Large Language Models
di: Liu, Weijia, et al.
Pubblicazione: (2024)
di: Liu, Weijia, et al.
Pubblicazione: (2024)
OS-HGAdapter: Open Semantic Hypergraph Adapter for Large Language Models Assisted Entropy-Enhanced Image-Text Alignment
di: Chen, Rongjun, et al.
Pubblicazione: (2025)
di: Chen, Rongjun, et al.
Pubblicazione: (2025)
MOC-3D: Manifold-Order Consistency for Text-to-3D Generation
di: Fan, Chenyang, et al.
Pubblicazione: (2026)
di: Fan, Chenyang, et al.
Pubblicazione: (2026)
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
di: Wan, Siqi, et al.
Pubblicazione: (2025)
di: Wan, Siqi, et al.
Pubblicazione: (2025)
Improving Virtual Try-On with Garment-focused Diffusion Models
di: Wan, Siqi, et al.
Pubblicazione: (2024)
di: Wan, Siqi, et al.
Pubblicazione: (2024)
Proxy-Tuning: Tailoring Multimodal Autoregressive Models for Subject-Driven Image Generation
di: Wu, Yi, et al.
Pubblicazione: (2025)
di: Wu, Yi, et al.
Pubblicazione: (2025)
SceneDreamer360: Text-Driven 3D-Consistent Scene Generation with Panoramic Gaussian Splatting
di: Li, Wenrui, et al.
Pubblicazione: (2024)
di: Li, Wenrui, et al.
Pubblicazione: (2024)
STEAR: Layer-Aware Spatiotemporal Evidence Intervention for Hallucination Mitigation in Video Large Language Models
di: Fan, Linfeng, et al.
Pubblicazione: (2026)
di: Fan, Linfeng, et al.
Pubblicazione: (2026)
ReCorD: Reasoning and Correcting Diffusion for HOI Generation
di: Jiang-Lin, Jian-Yu, et al.
Pubblicazione: (2024)
di: Jiang-Lin, Jian-Yu, et al.
Pubblicazione: (2024)
VidCompress: Memory-Enhanced Temporal Compression for Video Understanding in Large Language Models
di: Lan, Xiaohan, et al.
Pubblicazione: (2024)
di: Lan, Xiaohan, et al.
Pubblicazione: (2024)
Spatio-Temporal Data Enhanced Vision-Language Model for Traffic Scene Understanding
di: Ma, Jingtian, et al.
Pubblicazione: (2025)
di: Ma, Jingtian, et al.
Pubblicazione: (2025)
Leveraging multimodal explanatory annotations for video interpretation with Modality Specific Dataset
di: Ancarani, Elisa, et al.
Pubblicazione: (2025)
di: Ancarani, Elisa, et al.
Pubblicazione: (2025)
TopoCode: Topologically Informed Error Detection and Correction in Communication Systems
di: Guo, Hongzhi
Pubblicazione: (2024)
di: Guo, Hongzhi
Pubblicazione: (2024)
OralGPT-Omni: A Versatile Dental Multimodal Large Language Model
di: Hao, Jing, et al.
Pubblicazione: (2025)
di: Hao, Jing, et al.
Pubblicazione: (2025)
Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models
di: Yu, Xiaomin, et al.
Pubblicazione: (2026)
di: Yu, Xiaomin, et al.
Pubblicazione: (2026)
Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework
di: Wang, Jing, et al.
Pubblicazione: (2025)
di: Wang, Jing, et al.
Pubblicazione: (2025)
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction
di: Li, Dong, et al.
Pubblicazione: (2025)
di: Li, Dong, et al.
Pubblicazione: (2025)
PRINTER:Deformation-Aware Adversarial Learning for Virtual IHC Staining with In Situ Fidelity
di: Yuan, Yizhe, et al.
Pubblicazione: (2025)
di: Yuan, Yizhe, et al.
Pubblicazione: (2025)
Leveraging Compressed Frame Sizes For Ultra-Fast Video Classification
di: Han, Yuxing, et al.
Pubblicazione: (2024)
di: Han, Yuxing, et al.
Pubblicazione: (2024)
Bridging the Gap: Sketch-Aware Interpolation Network for High-Quality Animation Sketch Inbetweening
di: Shen, Jiaming, et al.
Pubblicazione: (2023)
di: Shen, Jiaming, et al.
Pubblicazione: (2023)
Enhancing Interactive Image Retrieval With Query Rewriting Using Large Language Models and Vision Language Models
di: Zhu, Hongyi, et al.
Pubblicazione: (2024)
di: Zhu, Hongyi, et al.
Pubblicazione: (2024)
FeatDistill: A Feature Distillation Enhanced Multi-Expert Ensemble Framework for Robust AI-generated Image Detection
di: Tu, Zhilin, et al.
Pubblicazione: (2026)
di: Tu, Zhilin, et al.
Pubblicazione: (2026)
Towards Emotion Analysis in Short-form Videos: A Large-Scale Dataset and Baseline
di: Wu, Xuecheng, et al.
Pubblicazione: (2023)
di: Wu, Xuecheng, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Multimodal AI Teacher: Integrating Edge Computing and Reasoning Models for Enhanced Student Error Analysis
di: Tianlong Xu, et al.
Pubblicazione: (2025) -
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education
di: Zhang, Yi-Fan, et al.
Pubblicazione: (2025) -
CustomContrast: A Multilevel Contrastive Perspective For Subject-Driven Text-to-Image Customization
di: Chen, Nan, et al.
Pubblicazione: (2024) -
TraveLLaMA: A Multimodal Travel Assistant with Large-Scale Dataset and Structured Reasoning
di: Chu, Meng, et al.
Pubblicazione: (2025) -
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
di: Zhou, Tian-Yi, et al.
Pubblicazione: (2026)