Case-based reasoning approach for diagnostic screening of children with developmental delays
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Zichen, Li, Jiakang, Lai, Songning, Huang, Sitan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CCSRP: Robust Pruning of Spiking Neural Networks through Cooperative Coevolution
von: Song, Zichen, et al.
Veröffentlicht: (2024)
von: Song, Zichen, et al.
Veröffentlicht: (2024)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
von: Ling, Jun, et al.
Veröffentlicht: (2024)
von: Ling, Jun, et al.
Veröffentlicht: (2024)
A multi-purpose automatic editing system based on lecture semantics for remote education
von: Hu, Panwen, et al.
Veröffentlicht: (2024)
von: Hu, Panwen, et al.
Veröffentlicht: (2024)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
von: Cao, Pu, et al.
Veröffentlicht: (2023)
von: Cao, Pu, et al.
Veröffentlicht: (2023)
LPM 1.0: Video-based Character Performance Model
von: Zeng, Ailing, et al.
Veröffentlicht: (2026)
von: Zeng, Ailing, et al.
Veröffentlicht: (2026)
Robust Latent Representation Tuning for Image-text Classification
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
DualComp: End-to-End Learning of a Unified Dual-Modality Lossless Compressor
von: Zhao, Yan, et al.
Veröffentlicht: (2025)
von: Zhao, Yan, et al.
Veröffentlicht: (2025)
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
MM-Point: Multi-View Information-Enhanced Multi-Modal Self-Supervised 3D Point Cloud Understanding
von: Yu, Hai-Tao, et al.
Veröffentlicht: (2024)
von: Yu, Hai-Tao, et al.
Veröffentlicht: (2024)
Both Text and Images Leaked! A Systematic Analysis of Data Contamination in Multimodal LLM
von: Song, Dingjie, et al.
Veröffentlicht: (2024)
von: Song, Dingjie, et al.
Veröffentlicht: (2024)
Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment
von: Cai, Zhuoxuan, et al.
Veröffentlicht: (2025)
von: Cai, Zhuoxuan, et al.
Veröffentlicht: (2025)
TrueFake: A Real World Case Dataset of Last Generation Fake Images also Shared on Social Networks
von: Dell'Anna, Stefano, et al.
Veröffentlicht: (2025)
von: Dell'Anna, Stefano, et al.
Veröffentlicht: (2025)
Understanding Temporal Logic Consistency in Video-Language Models through Cross-Modal Attention Discriminability
von: Li, Chengzhi, et al.
Veröffentlicht: (2025)
von: Li, Chengzhi, et al.
Veröffentlicht: (2025)
Exploiting LMM-based knowledge for image classification tasks
von: Tzelepi, Maria, et al.
Veröffentlicht: (2024)
von: Tzelepi, Maria, et al.
Veröffentlicht: (2024)
DyRoNet: Dynamic Routing and Low-Rank Adapters for Autonomous Driving Streaming Perception
von: Huang, Xiang, et al.
Veröffentlicht: (2024)
von: Huang, Xiang, et al.
Veröffentlicht: (2024)
Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks
von: Li, Jie, et al.
Veröffentlicht: (2025)
von: Li, Jie, et al.
Veröffentlicht: (2025)
Webcam-based Pupil Diameter Prediction Benefits from Upscaling
von: Shah, Vijul, et al.
Veröffentlicht: (2024)
von: Shah, Vijul, et al.
Veröffentlicht: (2024)
Causal-Story: Local Causal Attention Utilizing Parameter-Efficient Tuning For Visual Story Synthesis
von: Song, Tianyi, et al.
Veröffentlicht: (2023)
von: Song, Tianyi, et al.
Veröffentlicht: (2023)
LLM-based Fusion of Multi-modal Features for Commercial Memorability Prediction
von: Pramov, Aleksandar
Veröffentlicht: (2025)
von: Pramov, Aleksandar
Veröffentlicht: (2025)
ReCorD: Reasoning and Correcting Diffusion for HOI Generation
von: Jiang-Lin, Jian-Yu, et al.
Veröffentlicht: (2024)
von: Jiang-Lin, Jian-Yu, et al.
Veröffentlicht: (2024)
PFB-Diff: Progressive Feature Blending Diffusion for Text-driven Image Editing
von: Huang, Wenjing, et al.
Veröffentlicht: (2023)
von: Huang, Wenjing, et al.
Veröffentlicht: (2023)
Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR
von: Zhang, Yulong, et al.
Veröffentlicht: (2025)
von: Zhang, Yulong, et al.
Veröffentlicht: (2025)
InstructFLIP: Exploring Unified Vision-Language Model for Face Anti-spoofing
von: Lin, Kun-Hsiang, et al.
Veröffentlicht: (2025)
von: Lin, Kun-Hsiang, et al.
Veröffentlicht: (2025)
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation
von: Huang, Feizhen, et al.
Veröffentlicht: (2025)
von: Huang, Feizhen, et al.
Veröffentlicht: (2025)
DASH: Dynamic Audio-Driven Semantic Chunking for Efficient Omnimodal Token Compression
von: Li, Bingzhou, et al.
Veröffentlicht: (2026)
von: Li, Bingzhou, et al.
Veröffentlicht: (2026)
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos
von: Yu, Jiashuo, et al.
Veröffentlicht: (2025)
von: Yu, Jiashuo, et al.
Veröffentlicht: (2025)
Can Large Language Models Grasp Event Signals? Exploring Pure Zero-Shot Event-based Recognition
von: Yu, Zongyou, et al.
Veröffentlicht: (2024)
von: Yu, Zongyou, et al.
Veröffentlicht: (2024)
Anti-Inpainting: A Proactive Defense Approach against Malicious Diffusion-based Inpainters under Unknown Conditions
von: Guo, Yimao, et al.
Veröffentlicht: (2025)
von: Guo, Yimao, et al.
Veröffentlicht: (2025)
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
von: Huang, Victor Shea-Jay, et al.
Veröffentlicht: (2025)
von: Huang, Victor Shea-Jay, et al.
Veröffentlicht: (2025)
CLASP: Cross-modal Salient Anchor-based Semantic Propagation for Weakly-supervised Dense Audio-Visual Event Localization
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025)
von: Zhou, Jinxing, et al.
Veröffentlicht: (2025)
MoPE-CLIP: Structured Pruning for Efficient Vision-Language Models with Module-wise Pruning Error Metric
von: Lin, Haokun, et al.
Veröffentlicht: (2024)
von: Lin, Haokun, et al.
Veröffentlicht: (2024)
AToken: A Unified Tokenizer for Vision
von: Lu, Jiasen, et al.
Veröffentlicht: (2025)
von: Lu, Jiasen, et al.
Veröffentlicht: (2025)
X-Adapter: Adding Universal Compatibility of Plugins for Upgraded Diffusion Model
von: Ran, Lingmin, et al.
Veröffentlicht: (2023)
von: Ran, Lingmin, et al.
Veröffentlicht: (2023)
BlobCtrl: Taming Controllable Blob for Element-level Image Editing
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
von: Yan, Xin, et al.
Veröffentlicht: (2024)
von: Yan, Xin, et al.
Veröffentlicht: (2024)
CLCR: Cross-Level Semantic Collaborative Representation for Multimodal Learning
von: Meng, Chunlei, et al.
Veröffentlicht: (2026)
von: Meng, Chunlei, et al.
Veröffentlicht: (2026)
Style-Preserving Lip Sync via Audio-Aware Style Reference
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
ASR-enhanced Multimodal Representation Learning for Cross-Domain Product Retrieval
von: Zhao, Ruixiang, et al.
Veröffentlicht: (2024)
von: Zhao, Ruixiang, et al.
Veröffentlicht: (2024)
Efficient Low-Resolution Face Recognition via Bridge Distillation
von: Ge, Shiming, et al.
Veröffentlicht: (2024)
von: Ge, Shiming, et al.
Veröffentlicht: (2024)
GiVE: Guiding Visual Encoder to Perceive Overlooked Information
von: Li, Junjie, et al.
Veröffentlicht: (2024)
von: Li, Junjie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CCSRP: Robust Pruning of Spiking Neural Networks through Cooperative Coevolution
von: Song, Zichen, et al.
Veröffentlicht: (2024) -
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
von: Ling, Jun, et al.
Veröffentlicht: (2024) -
A multi-purpose automatic editing system based on lecture semantics for remote education
von: Hu, Panwen, et al.
Veröffentlicht: (2024) -
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
von: Cao, Pu, et al.
Veröffentlicht: (2023) -
LPM 1.0: Video-based Character Performance Model
von: Zeng, Ailing, et al.
Veröffentlicht: (2026)