Gespeichert in:
| Hauptverfasser: | Nasiri-Sarvi, Ali, Nguyen, Anh Tien, Rivaz, Hassan, Samaras, Dimitris, Hosseini, Mahdi S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.12403 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SPARC: Concept-Aligned Sparse Autoencoders for Cross-Model and Cross-Modal Interpretability
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2025)
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2025)
Vision Mamba for Classification of Breast Ultrasound Images
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024)
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024)
Vim4Path: Self-Supervised Vision Mamba for Histopathology Images
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024)
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024)
Ultrasound Image Generation using Latent Diffusion Models
von: Freiche, Benoit, et al.
Veröffentlicht: (2025)
von: Freiche, Benoit, et al.
Veröffentlicht: (2025)
2DMamba: Efficient State Space Model for Image Representation with Applications on Giga-Pixel Whole Slide Image Classification
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
LBMamba: Locally Bi-directional Mamba
von: Zhang, Jingwei, et al.
Veröffentlicht: (2025)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2025)
CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values
von: Koleilat, Taha, et al.
Veröffentlicht: (2025)
von: Koleilat, Taha, et al.
Veröffentlicht: (2025)
Self-supervised co-salient object detection via feature correspondence at multiple scales
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
Vision Transformer for Classification of Breast Ultrasound Images
von: Gheflati, Behnaz, et al.
Veröffentlicht: (2021)
von: Gheflati, Behnaz, et al.
Veröffentlicht: (2021)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
Comparative Analysis of Diffusion Generative Models in Computational Pathology
von: Thakkar, Denisha, et al.
Veröffentlicht: (2024)
von: Thakkar, Denisha, et al.
Veröffentlicht: (2024)
TopoDiffusionNet: A Topology-aware Diffusion Model
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Assessing Sample Quality via the Latent Space of Generative Models
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
Evi-Steer: Learning to Steer Biomedical Vision-Language Models through Efficient and Generalizable Evidential Tuning
von: Koleilat, Taha, et al.
Veröffentlicht: (2026)
von: Koleilat, Taha, et al.
Veröffentlicht: (2026)
MI-NeRF: Learning a Single Face NeRF from Multiple Identities
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
MIGS: Multi-Identity Gaussian Splatting via Tensor Decomposition
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
VLEER: Vision and Language Embeddings for Explainable Whole Slide Image Representation
von: Nguyen, Anh Tien, et al.
Veröffentlicht: (2025)
von: Nguyen, Anh Tien, et al.
Veröffentlicht: (2025)
Phrase-Instance Alignment for Generalized Referring Segmentation
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024)
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2024)
Fast constrained sampling in pre-trained diffusion models
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
Grounding DINO-US-SAM: Text-Prompted Multi-Organ Segmentation in Ultrasound with LoRA-Tuned Vision-Language Models
von: Rasaee, Hamza, et al.
Veröffentlicht: (2025)
von: Rasaee, Hamza, et al.
Veröffentlicht: (2025)
Learning 3D Reconstruction with Priors in Test Time
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Importance-Based Token Merging for Efficient Image and Video Generation
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2026)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2026)
Improving Contrastive Learning for Referring Expression Counting
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
CORA: Consistency-Guided Semi-Supervised Framework for Reasoning Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2025)
von: Howlader, Prantik, et al.
Veröffentlicht: (2025)
Reliability of deep learning models for anatomical landmark detection: The role of inter-rater variability
von: Salari, Soorena, et al.
Veröffentlicht: (2024)
von: Salari, Soorena, et al.
Veröffentlicht: (2024)
Rig3DGS: Creating Controllable Portraits from Casual Monocular Videos
von: Rivero, Alfredo, et al.
Veröffentlicht: (2024)
von: Rivero, Alfredo, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
PathSegDiff: Pathology Segmentation using Diffusion model representations
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
MedCLIP-SAMv2: Towards Universal Text-Driven Medical Image Segmentation
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
BiomedCoOp: Learning to Prompt for Biomedical Vision-Language Models
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
Training Deep Visual Networks Beyond Loss and Accuracy Through a Dynamical Systems Approach
von: La Quang, Hai, et al.
Veröffentlicht: (2026)
von: La Quang, Hai, et al.
Veröffentlicht: (2026)
Explainable AI and susceptibility to adversarial attacks: a case study in classification of breast ultrasound images
von: Rasaee, Hamza, et al.
Veröffentlicht: (2021)
von: Rasaee, Hamza, et al.
Veröffentlicht: (2021)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Efficient INT8 Single-Image Super-Resolution via Deployment-Aware Quantization and Teacher-Guided Training
von: Nguyen, Pham Phuong Nam, et al.
Veröffentlicht: (2026)
von: Nguyen, Pham Phuong Nam, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SPARC: Concept-Aligned Sparse Autoencoders for Cross-Model and Cross-Modal Interpretability
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2025) -
Vision Mamba for Classification of Breast Ultrasound Images
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024) -
Vim4Path: Self-Supervised Vision Mamba for Histopathology Images
von: Nasiri-Sarvi, Ali, et al.
Veröffentlicht: (2024) -
Ultrasound Image Generation using Latent Diffusion Models
von: Freiche, Benoit, et al.
Veröffentlicht: (2025) -
2DMamba: Efficient State Space Model for Image Representation with Applications on Giga-Pixel Whole Slide Image Classification
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)