NutriScreener: Retrieval-Augmented Multi-Pose Graph Attention Network for Malnourishment Screening
Fuente:
arXiv
Guardado en:
| Autores principales: | Khan, Misaal, Vatsa, Mayank, Singh, Kuldeep, Singh, Richa |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Right Looks, Wrong Reasons: Compositional Fidelity in Text-to-Image Generation
por: Vatsa, Mayank, et al.
Publicado: (2025)
por: Vatsa, Mayank, et al.
Publicado: (2025)
TAIGen: Training-Free Adversarial Image Generation via Diffusion Models
por: Roy, Susim, et al.
Publicado: (2025)
por: Roy, Susim, et al.
Publicado: (2025)
Discerning the Chaos: Detecting Adversarial Perturbations while Disentangling Intentional from Unintentional Noises
por: Jain, Anubhooti, et al.
Publicado: (2024)
por: Jain, Anubhooti, et al.
Publicado: (2024)
Optimizing Skin Lesion Classification via Multimodal Data and Auxiliary Task Integration
por: Khurshid, Mahapara, et al.
Publicado: (2024)
por: Khurshid, Mahapara, et al.
Publicado: (2024)
Learn "No" to Say "Yes" Better: Improving Vision-Language Models via Negations
por: Singh, Jaisidh, et al.
Publicado: (2024)
por: Singh, Jaisidh, et al.
Publicado: (2024)
LitMAS: A Lightweight and Generalized Multi-Modal Anti-Spoofing Framework for Biometric Security
por: Gorthi, Nidheesh, et al.
Publicado: (2025)
por: Gorthi, Nidheesh, et al.
Publicado: (2025)
Low-Resolution Chest X-ray Classification via Knowledge Distillation and Multi-task Learning
por: Akhter, Yasmeena, et al.
Publicado: (2024)
por: Akhter, Yasmeena, et al.
Publicado: (2024)
RAVU: Retrieval Augmented Video Understanding with Compositional Reasoning over Graph
por: Malik, Sameer, et al.
Publicado: (2025)
por: Malik, Sameer, et al.
Publicado: (2025)
Unbiased Model Prediction Without Using Protected Attribute Information
por: Majumdar, Puspita, et al.
Publicado: (2026)
por: Majumdar, Puspita, et al.
Publicado: (2026)
Harmonizing Geometry and Uncertainty: Diffusion with Hyperspheres
por: Dosi, Muskan, et al.
Publicado: (2025)
por: Dosi, Muskan, et al.
Publicado: (2025)
Continual Unlearning for Foundational Text-to-Image Models without Generalization Erosion
por: Thakral, Kartik, et al.
Publicado: (2025)
por: Thakral, Kartik, et al.
Publicado: (2025)
HyperSpaceX: Radial and Angular Exploration of HyperSpherical Dimensions
por: Chiranjeev, Chiranjeev, et al.
Publicado: (2024)
por: Chiranjeev, Chiranjeev, et al.
Publicado: (2024)
Fine-Grained Erasure in Text-to-Image Diffusion-based Foundation Models
por: Thakral, Kartik, et al.
Publicado: (2025)
por: Thakral, Kartik, et al.
Publicado: (2025)
HYPERPOSE: Hyperbolic Kinematic Phase-Space Attention for 3D Human Pose Estimation
por: Thekkath, Vinduja, et al.
Publicado: (2026)
por: Thekkath, Vinduja, et al.
Publicado: (2026)
Poze: Sports Technique Feedback under Data Constraints
por: Singh, Agamdeep, et al.
Publicado: (2024)
por: Singh, Agamdeep, et al.
Publicado: (2024)
Navigating Text-to-Image Generative Bias across Indic Languages
por: Mittal, Surbhi, et al.
Publicado: (2024)
por: Mittal, Surbhi, et al.
Publicado: (2024)
3D WholeBody Pose Estimation based on Semantic Graph Attention Network and Distance Information
por: Wen, Sihan, et al.
Publicado: (2024)
por: Wen, Sihan, et al.
Publicado: (2024)
Augmenting End-to-End Steering Angle Prediction with CAN Bus Data
por: Singh, Amit
Publicado: (2023)
por: Singh, Amit
Publicado: (2023)
Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation
por: Shah, Arya, et al.
Publicado: (2026)
por: Shah, Arya, et al.
Publicado: (2026)
Pose as Clinical Prior: Learning Dual Representations for Scoliosis Screening
por: Zhou, Zirui, et al.
Publicado: (2025)
por: Zhou, Zirui, et al.
Publicado: (2025)
EPAM-Net: An Efficient Pose-driven Attention-guided Multimodal Network for Video Action Recognition
por: Abdelkawy, Ahmed, et al.
Publicado: (2024)
por: Abdelkawy, Ahmed, et al.
Publicado: (2024)
TruePose: Human-Parsing-guided Attention Diffusion for Full-ID Preserving Pose Transfer
por: Xu, Zhihong, et al.
Publicado: (2025)
por: Xu, Zhihong, et al.
Publicado: (2025)
TRISHUL: Towards Region Identification and Screen Hierarchy Understanding for Large VLM based GUI Agents
por: Singh, Kunal, et al.
Publicado: (2025)
por: Singh, Kunal, et al.
Publicado: (2025)
Auditing Facial Emotion Recognition Datasets for Posed Expressions and Racial Bias
por: Khan, Rina, et al.
Publicado: (2025)
por: Khan, Rina, et al.
Publicado: (2025)
PARIC: Probabilistic Attention Regularization for Language Guided Image Classification from Pre-trained Vison Language Models
por: Nautiyal, Mayank, et al.
Publicado: (2025)
por: Nautiyal, Mayank, et al.
Publicado: (2025)
When RAG Hurts: Diagnosing and Mitigating Attention Distraction in Retrieval-Augmented LVLMs
por: Zhao, Beidi, et al.
Publicado: (2026)
por: Zhao, Beidi, et al.
Publicado: (2026)
Remote Sensing Retrieval-Augmented Generation: Bridging Remote Sensing Imagery and Comprehensive Knowledge with a Multi-Modal Dataset and Retrieval-Augmented Generation Model
por: Wen, Congcong, et al.
Publicado: (2025)
por: Wen, Congcong, et al.
Publicado: (2025)
Interpretable Plant Leaf Disease Detection Using Attention-Enhanced CNN
por: Singh, Balram, et al.
Publicado: (2025)
por: Singh, Balram, et al.
Publicado: (2025)
MADTempo: An Interactive System for Multi-Event Temporal Video Retrieval with Query Augmentation
por: Vu, Huu-An, et al.
Publicado: (2025)
por: Vu, Huu-An, et al.
Publicado: (2025)
Delta-K: Boosting Multi-Instance Generation via Cross-Attention Augmentation
por: Wang, Zitong, et al.
Publicado: (2026)
por: Wang, Zitong, et al.
Publicado: (2026)
From No to Know: Taxonomy, Challenges, and Opportunities for Negation Understanding in Multimodal Foundation Models
por: Vatsa, Mayank, et al.
Publicado: (2025)
por: Vatsa, Mayank, et al.
Publicado: (2025)
GUNet: A Graph Convolutional Network United Diffusion Model for Stable and Diversity Pose Generation
por: Liang, Shuowen, et al.
Publicado: (2024)
por: Liang, Shuowen, et al.
Publicado: (2024)
mKG-RAG: Leveraging Multimodal Knowledge Graphs in Retrieval-Augmented Generation for Knowledge-intensive VQA
por: Yuan, Xu, et al.
Publicado: (2025)
por: Yuan, Xu, et al.
Publicado: (2025)
Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
por: Singh, Amit Kumar, et al.
Publicado: (2024)
por: Singh, Amit Kumar, et al.
Publicado: (2024)
Retrieval-Augmented Prompt for OOD Detection
por: Han, Ruisong, et al.
Publicado: (2025)
por: Han, Ruisong, et al.
Publicado: (2025)
Towards Scene Graph Anticipation
por: Peddi, Rohith, et al.
Publicado: (2024)
por: Peddi, Rohith, et al.
Publicado: (2024)
Keypoints as Dynamic Centroids for Unified Human Pose and Segmentation
por: Ahmad, Niaz, et al.
Publicado: (2025)
por: Ahmad, Niaz, et al.
Publicado: (2025)
Self-Attention Based Multi-Scale Graph Auto-Encoder Network of 3D Meshes
por: Nazir, Saqib, et al.
Publicado: (2025)
por: Nazir, Saqib, et al.
Publicado: (2025)
On Responsible Machine Learning Datasets with Fairness, Privacy, and Regulatory Norms
por: Mittal, Surbhi, et al.
Publicado: (2023)
por: Mittal, Surbhi, et al.
Publicado: (2023)
PoseMoE: Mixture-of-Experts Network for Monocular 3D Human Pose Estimation
por: Liu, Mengyuan, et al.
Publicado: (2025)
por: Liu, Mengyuan, et al.
Publicado: (2025)
Ejemplares similares
-
Right Looks, Wrong Reasons: Compositional Fidelity in Text-to-Image Generation
por: Vatsa, Mayank, et al.
Publicado: (2025) -
TAIGen: Training-Free Adversarial Image Generation via Diffusion Models
por: Roy, Susim, et al.
Publicado: (2025) -
Discerning the Chaos: Detecting Adversarial Perturbations while Disentangling Intentional from Unintentional Noises
por: Jain, Anubhooti, et al.
Publicado: (2024) -
Optimizing Skin Lesion Classification via Multimodal Data and Auxiliary Task Integration
por: Khurshid, Mahapara, et al.
Publicado: (2024) -
Learn "No" to Say "Yes" Better: Improving Vision-Language Models via Negations
por: Singh, Jaisidh, et al.
Publicado: (2024)