Structured Prompting and Multi-Agent Knowledge Distillation for Traffic Video Interpretation and Risk Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yunxiang, Xu, Ningning, Yang, Jidong J. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Agent Visual-Language Reasoning for Comprehensive Highway Scene Understanding
by: Yang, Yunxiang, et al.
Published: (2025)
by: Yang, Yunxiang, et al.
Published: (2025)
Leveraging Scene Geometry and Depth Information for Robust Image Deraining
by: Xu, Ningning, et al.
Published: (2024)
by: Xu, Ningning, et al.
Published: (2024)
Enhancing autonomous vehicle safety in rain: a data-centric approach for clear vision
by: Seferian, Mark A., et al.
Published: (2024)
by: Seferian, Mark A., et al.
Published: (2024)
Enhancing Multimodal LLM for Detailed and Accurate Video Captioning using Multi-Round Preference Optimization
by: Tang, Changli, et al.
Published: (2024)
by: Tang, Changli, et al.
Published: (2024)
VisionUnite: A Vision-Language Foundation Model for Ophthalmology Enhanced with Clinical Knowledge
by: Li, Zihan, et al.
Published: (2024)
by: Li, Zihan, et al.
Published: (2024)
Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding
by: Guo, Weiyu, et al.
Published: (2025)
by: Guo, Weiyu, et al.
Published: (2025)
RankDVQA-mini: Knowledge Distillation-Driven Deep Video Quality Assessment
by: Feng, Chen, et al.
Published: (2023)
by: Feng, Chen, et al.
Published: (2023)
MTKD: Multi-Teacher Knowledge Distillation for Image Super-Resolution
by: Jiang, Yuxuan, et al.
Published: (2024)
by: Jiang, Yuxuan, et al.
Published: (2024)
SCKD: Semi-Supervised Cross-Modality Knowledge Distillation for 4D Radar Object Detection
by: Xu, Ruoyu, et al.
Published: (2024)
by: Xu, Ruoyu, et al.
Published: (2024)
REACT-KD: Region-Aware Cross-modal Topological Knowledge Distillation for Interpretable Medical Image Classification
by: Chen, Hongzhao, et al.
Published: (2025)
by: Chen, Hongzhao, et al.
Published: (2025)
Explainable Knowledge Distillation for Efficient Medical Image Classification
by: Mir, Aqib Nazir, et al.
Published: (2025)
by: Mir, Aqib Nazir, et al.
Published: (2025)
Efficient Knowledge Distillation of SAM for Medical Image Segmentation
by: Patil, Kunal Dasharath, et al.
Published: (2025)
by: Patil, Kunal Dasharath, et al.
Published: (2025)
STPNet: Scale-aware Text Prompt Network for Medical Image Segmentation
by: Shan, Dandan, et al.
Published: (2025)
by: Shan, Dandan, et al.
Published: (2025)
Edge AI-Enabled Chicken Health Detection Based on Enhanced FCOS-Lite and Knowledge Distillation
by: Tong, Qiang, et al.
Published: (2024)
by: Tong, Qiang, et al.
Published: (2024)
Target-Balanced Score Distillation
by: Xu, Zhou, et al.
Published: (2025)
by: Xu, Zhou, et al.
Published: (2025)
Multi-Task Multi-Scale Contrastive Knowledge Distillation for Efficient Medical Image Segmentation
by: Biswas, Risab
Published: (2024)
by: Biswas, Risab
Published: (2024)
Fundus2Video: Cross-Modal Angiography Video Generation from Static Fundus Photography with Clinical Knowledge Guidance
by: Zhang, Weiyi, et al.
Published: (2024)
by: Zhang, Weiyi, et al.
Published: (2024)
Comprehensive Evaluation of Multimodal AI Models in Medical Imaging Diagnosis: From Data Augmentation to Preference-Based Comparison
by: Ruan, Cailian, et al.
Published: (2024)
by: Ruan, Cailian, et al.
Published: (2024)
Uncovering Latent Pathological Signatures in Pulmonary CT via Cross-Window Knowledge Distillation
by: Peng, Bo, et al.
Published: (2026)
by: Peng, Bo, et al.
Published: (2026)
MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs
by: Shi, Baorong, et al.
Published: (2026)
by: Shi, Baorong, et al.
Published: (2026)
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models
by: Mehta, Deval, et al.
Published: (2025)
by: Mehta, Deval, et al.
Published: (2025)
A Transformer-Based Multi-Stream Approach for Isolated Iranian Sign Language Recognition
by: Ghadami, Ali, et al.
Published: (2024)
by: Ghadami, Ali, et al.
Published: (2024)
HiLight: Technical Report on the Motern AI Video Language Model
by: Wang, Zhiting, et al.
Published: (2024)
by: Wang, Zhiting, et al.
Published: (2024)
Beyond Textual Knowledge-Leveraging Multimodal Knowledge Bases for Enhancing Vision-and-Language Navigation
by: Yang, Dongsheng, et al.
Published: (2026)
by: Yang, Dongsheng, et al.
Published: (2026)
InternVQA: Advancing Compressed Video Quality Assessment with Distilling Large Foundation Model
by: Guan, Fengbin, et al.
Published: (2025)
by: Guan, Fengbin, et al.
Published: (2025)
Towards Interpretable Radiology Report Generation via Concept Bottlenecks using a Multi-Agentic RAG
by: Alam, Hasan Md Tusfiqur, et al.
Published: (2024)
by: Alam, Hasan Md Tusfiqur, et al.
Published: (2024)
How Does Diverse Interpretability of Textual Prompts Impact Medical Vision-Language Zero-Shot Tasks?
by: Wang, Sicheng, et al.
Published: (2024)
by: Wang, Sicheng, et al.
Published: (2024)
MvKSR: Multi-view Knowledge-guided Scene Recovery for Hazy and Rainy Degradation
by: Yang, Dong, et al.
Published: (2024)
by: Yang, Dong, et al.
Published: (2024)
Dual-Modal Lung Cancer AI: Interpretable Radiology and Microscopy with Clinical Risk Integration
by: Sukumal, Baramee, et al.
Published: (2026)
by: Sukumal, Baramee, et al.
Published: (2026)
VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis
by: Hoopes, Andrew, et al.
Published: (2024)
by: Hoopes, Andrew, et al.
Published: (2024)
Histo-Genomic Knowledge Distillation For Cancer Prognosis From Histopathology Whole Slide Images
by: Wang, Zhikang, et al.
Published: (2024)
by: Wang, Zhikang, et al.
Published: (2024)
Digital Staining with Knowledge Distillation: A Unified Framework for Unpaired and Paired-But-Misaligned Data
by: Xu, Ziwang, et al.
Published: (2025)
by: Xu, Ziwang, et al.
Published: (2025)
MoMA: Momentum Contrastive Learning with Multi-head Attention-based Knowledge Distillation for Histopathology Image Analysis
by: Vuong, Trinh Thi Le, et al.
Published: (2023)
by: Vuong, Trinh Thi Le, et al.
Published: (2023)
SFT-KD-Recon: Learning a Student-friendly Teacher for Knowledge Distillation in Magnetic Resonance Image Reconstruction
by: Gayathri, Matcha Naga, et al.
Published: (2023)
by: Gayathri, Matcha Naga, et al.
Published: (2023)
Involution and BSConv Multi-Depth Distillation Network for Lightweight Image Super-Resolution
by: Khatami-Rizi, Akram, et al.
Published: (2025)
by: Khatami-Rizi, Akram, et al.
Published: (2025)
PGAD: Prototype-Guided Adaptive Distillation for Multi-Modal Learning in AD Diagnosis
by: Li, Yanfei, et al.
Published: (2025)
by: Li, Yanfei, et al.
Published: (2025)
Scale-aware Adaptive Supervised Network with Limited Medical Annotations
by: Li, Zihan, et al.
Published: (2026)
by: Li, Zihan, et al.
Published: (2026)
Boosting Neural Representations for Videos with a Conditional Decoder
by: Zhang, Xinjie, et al.
Published: (2024)
by: Zhang, Xinjie, et al.
Published: (2024)
APSeg: Auto-Prompt Model with Acquired and Injected Knowledge for Nuclear Instance Segmentation and Classification
by: Xu, Liying, et al.
Published: (2025)
by: Xu, Liying, et al.
Published: (2025)
Face to Cartoon Incremental Super-Resolution using Knowledge Distillation
by: Devkatte, Trinetra, et al.
Published: (2024)
by: Devkatte, Trinetra, et al.
Published: (2024)
Similar Items
-
Multi-Agent Visual-Language Reasoning for Comprehensive Highway Scene Understanding
by: Yang, Yunxiang, et al.
Published: (2025) -
Leveraging Scene Geometry and Depth Information for Robust Image Deraining
by: Xu, Ningning, et al.
Published: (2024) -
Enhancing autonomous vehicle safety in rain: a data-centric approach for clear vision
by: Seferian, Mark A., et al.
Published: (2024) -
Enhancing Multimodal LLM for Detailed and Accurate Video Captioning using Multi-Round Preference Optimization
by: Tang, Changli, et al.
Published: (2024) -
VisionUnite: A Vision-Language Foundation Model for Ophthalmology Enhanced with Clinical Knowledge
by: Li, Zihan, et al.
Published: (2024)