AURORA: Adaptive Unified Representation for Robust Ultrasound Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Khan, Ufaq, Teja, L. D. M. S. Sai, Shakiru, Ayuba, Shaaban, Mai A., Xie, Yutong, Bilal, Muhammad, Khan, Muhammad Haris |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
by: Khan, Adnan, et al.
Published: (2024)
by: Khan, Adnan, et al.
Published: (2024)
MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
by: Khan, Ufaq, et al.
Published: (2026)
by: Khan, Ufaq, et al.
Published: (2026)
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
by: Khan, Ufaq, et al.
Published: (2025)
by: Khan, Ufaq, et al.
Published: (2025)
Robust and Label-Efficient Deep Waste Detection
by: Abid, Hassan, et al.
Published: (2025)
by: Abid, Hassan, et al.
Published: (2025)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
by: Ding, Bonan, et al.
Published: (2026)
by: Ding, Bonan, et al.
Published: (2026)
CountZES: Counting via Zero-Shot Exemplar Selection
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock
by: Nawaz, Umair, et al.
Published: (2025)
by: Nawaz, Umair, et al.
Published: (2025)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
by: Danish, Muhammad Sohail, et al.
Published: (2025)
by: Danish, Muhammad Sohail, et al.
Published: (2025)
Depth Attention for Robust RGB Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
by: Ali, Eman, et al.
Published: (2023)
by: Ali, Eman, et al.
Published: (2023)
OpenEarthAgent: A Unified Framework for Tool-Augmented Geospatial Agents
by: Shabbir, Akashah, et al.
Published: (2026)
by: Shabbir, Akashah, et al.
Published: (2026)
Divergent Domains, Convergent Grading: Enhancing Generalization in Diabetic Retinopathy Grading
by: Chokuwa, Sharon, et al.
Published: (2024)
by: Chokuwa, Sharon, et al.
Published: (2024)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
by: Sharifdeen, Ashshak, et al.
Published: (2025)
by: Sharifdeen, Ashshak, et al.
Published: (2025)
Realistic and Efficient Face Swapping: A Unified Approach with Diffusion Models
by: Baliah, Sanoojan, et al.
Published: (2024)
by: Baliah, Sanoojan, et al.
Published: (2024)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
by: Gunduboina, Hariseetharam, et al.
Published: (2025)
by: Gunduboina, Hariseetharam, et al.
Published: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
by: Ali, Eman, et al.
Published: (2024)
by: Ali, Eman, et al.
Published: (2024)
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
RAPTOR+: A Visually Grounded Vision-Language Framework to Improve Clinical Trust and Auditability in Automated Cancer Referral Processing
by: Abioye, Sofiat, et al.
Published: (2026)
by: Abioye, Sofiat, et al.
Published: (2026)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
Agentic AI for Remote Sensing: Technical Challenges and Research Directions
by: Munir, Muhammad Akhtar, et al.
Published: (2026)
by: Munir, Muhammad Akhtar, et al.
Published: (2026)
A Spatiotemporal Approach to Tri-Perspective Representation for 3D Semantic Occupancy Prediction
by: Silva, Sathira, et al.
Published: (2024)
by: Silva, Sathira, et al.
Published: (2024)
MedPromptX: Grounded Multimodal Prompting for Chest X-ray Diagnosis
by: Shaaban, Mai A., et al.
Published: (2024)
by: Shaaban, Mai A., et al.
Published: (2024)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Towards PerSense++: Advancing Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
ReConText3D: Replay-based Continual Text-to-3D Generation
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
by: Baliah, Sanoojan, et al.
Published: (2026)
by: Baliah, Sanoojan, et al.
Published: (2026)
Judging from Support-set: A New Way to Utilize Few-Shot Segmentation for Segmentation Refinement Process
by: Moon, Seonghyeon, et al.
Published: (2024)
by: Moon, Seonghyeon, et al.
Published: (2024)
Pose-Guided Self-Training with Two-Stage Clustering for Unsupervised Landmark Discovery
by: Tourani, Siddharth, et al.
Published: (2024)
by: Tourani, Siddharth, et al.
Published: (2024)
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
by: Kuckreja, Kartik, et al.
Published: (2026)
by: Kuckreja, Kartik, et al.
Published: (2026)
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
by: Ali, Eman, et al.
Published: (2025)
by: Ali, Eman, et al.
Published: (2025)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
by: Danish, Muhammad Sohail, et al.
Published: (2024)
by: Danish, Muhammad Sohail, et al.
Published: (2024)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
ThinkGeo: Evaluating Tool-Augmented Agents for Remote Sensing Tasks
by: Shabbir, Akashah, et al.
Published: (2025)
by: Shabbir, Akashah, et al.
Published: (2025)
CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders
by: Ahamed, Shihab Aaqil, et al.
Published: (2025)
by: Ahamed, Shihab Aaqil, et al.
Published: (2025)
Towards Generalizing to Unseen Domains with Few Labels
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)
TLAC: Two-stage LMM Augmented CLIP for Zero-Shot Classification
by: Munir, Ans, et al.
Published: (2025)
by: Munir, Ans, et al.
Published: (2025)
Compositional Zero-Shot Learning: A Survey
by: Munir, Ans, et al.
Published: (2025)
by: Munir, Ans, et al.
Published: (2025)
TactileNet: Bridging the Accessibility Gap with AI-Generated Tactile Graphics for Individuals with Vision Impairment
by: Khan, Adnan, et al.
Published: (2025)
by: Khan, Adnan, et al.
Published: (2025)
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
by: Silva, Sathira, et al.
Published: (2025)
by: Silva, Sathira, et al.
Published: (2025)
Unsupervised Deep Graph Matching Based on Cycle Consistency
by: Tourani, Siddharth, et al.
Published: (2023)
by: Tourani, Siddharth, et al.
Published: (2023)
Similar Items
-
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
by: Khan, Adnan, et al.
Published: (2024) -
MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
by: Khan, Ufaq, et al.
Published: (2026) -
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
by: Khan, Ufaq, et al.
Published: (2025) -
Robust and Label-Efficient Deep Waste Detection
by: Abid, Hassan, et al.
Published: (2025) -
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
by: Ding, Bonan, et al.
Published: (2026)