Distilling Specialized Orders for Visual Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pramanik, Rishav, Sghaier, Amin, Aminbeidokhti, Masih, Rodriguez, Juan A., Poupon, Antoine, Vazquez, David, Pal, Christopher, Yin, Zhaozheng, Pedersoli, Marco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Source-Free Domain Adaptation for YOLO Object Detection
von: Varailhon, Simon, et al.
Veröffentlicht: (2024)
von: Varailhon, Simon, et al.
Veröffentlicht: (2024)
Revisiting Mixout: An Overlooked Path to Robust Finetuning
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
High-Rate Mixout: Revisiting Mixout for Robust Domain Generalization
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
LT-Soups: Bridging Head and Tail Classes via Subsampled Model Soups
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
von: Muralidharan, Srikanth, et al.
Veröffentlicht: (2025)
von: Muralidharan, Srikanth, et al.
Veröffentlicht: (2025)
WiSE-OD: Benchmarking Robustness in Infrared Object Detection
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2025)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2025)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
Domain Generalization by Rejecting Extreme Augmentations
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2023)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2023)
Modality Translation for Object Detection Adaptation Without Forgetting Prior Knowledge
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2024)
WASH: Train your Ensemble with Communication-Efficient Weight Shuffling, then Average
von: Fournier, Louis, et al.
Veröffentlicht: (2024)
von: Fournier, Louis, et al.
Veröffentlicht: (2024)
HalluciDet: Hallucinating RGB Modality for Person Detection Through Privileged Information
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2023)
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2023)
SANEval: Open-Vocabulary Compositional Benchmarks with Failure-mode Diagnosis
von: Pramanik, Rishav, et al.
Veröffentlicht: (2026)
von: Pramanik, Rishav, et al.
Veröffentlicht: (2026)
VectorGym: A Multitask Benchmark for SVG Code Generation, Sketching, and Editing
von: Rodriguez, Juan, et al.
Veröffentlicht: (2026)
von: Rodriguez, Juan, et al.
Veröffentlicht: (2026)
Rendering-Aware Reinforcement Learning for Vector Graphics Generation
von: Rodriguez, Juan A., et al.
Veröffentlicht: (2025)
von: Rodriguez, Juan A., et al.
Veröffentlicht: (2025)
Alignment, Mining and Fusion: Representation Alignment with Hard Negative Mining and Selective Knowledge Fusion for Medical Visual Question Answering
von: Zou, Yuanhao, et al.
Veröffentlicht: (2025)
von: Zou, Yuanhao, et al.
Veröffentlicht: (2025)
How (Mis)calibrated is Your Federated CLIP and What To Do About It?
von: Singha, Mainak, et al.
Veröffentlicht: (2025)
von: Singha, Mainak, et al.
Veröffentlicht: (2025)
StarVector: Generating Scalable Vector Graphics Code from Images and Text
von: Rodriguez, Juan A., et al.
Veröffentlicht: (2023)
von: Rodriguez, Juan A., et al.
Veröffentlicht: (2023)
Attention-Enhanced Co-Interactive Fusion Network (AECIF-Net) for Automated Structural Condition Assessment in Visual Inspection
von: Zhang, Chenyu, et al.
Veröffentlicht: (2023)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2023)
CTA: Cross-Task Alignment for Better Test Time Training
von: Barbeau, Samuel, et al.
Veröffentlicht: (2025)
von: Barbeau, Samuel, et al.
Veröffentlicht: (2025)
SemiDAViL: Semi-supervised Domain Adaptation with Vision-Language Guidance for Semantic Segmentation
von: Basak, Hritam, et al.
Veröffentlicht: (2025)
von: Basak, Hritam, et al.
Veröffentlicht: (2025)
CCDNet: Learning to Detect Camouflage against Distractors in Infrared Small Target Detection
von: Liao, Zikai, et al.
Veröffentlicht: (2026)
von: Liao, Zikai, et al.
Veröffentlicht: (2026)
Multi Teacher Privileged Knowledge Distillation for Multimodal Expression Recognition
von: Aslam, Muhammad Haseeb, et al.
Veröffentlicht: (2024)
von: Aslam, Muhammad Haseeb, et al.
Veröffentlicht: (2024)
TeD-Loc: Text Distillation for Weakly Supervised Object Localization
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2025)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2025)
Unsupervised Object Discovery: A Comprehensive Survey and Unified Taxonomy
von: Villa-Vásquez, José-Fabian, et al.
Veröffentlicht: (2024)
von: Villa-Vásquez, José-Fabian, et al.
Veröffentlicht: (2024)
MiPa: Mixed Patch Infrared-Visible Modality Agnostic Object Detection
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
Weakly-Supervised Semantic Segmentation with Image-Level Labels: from Traditional Models to Foundation Models
von: Chen, Zhaozheng, et al.
Veröffentlicht: (2023)
von: Chen, Zhaozheng, et al.
Veröffentlicht: (2023)
Neural Architecture Search by Learning a Hierarchical Search Space
von: Roshtkhari, Mehraveh Javan, et al.
Veröffentlicht: (2025)
von: Roshtkhari, Mehraveh Javan, et al.
Veröffentlicht: (2025)
Distilling Privileged Multimodal Information for Expression Recognition using Optimal Transport
von: Aslam, Muhammad Haseeb, et al.
Veröffentlicht: (2024)
von: Aslam, Muhammad Haseeb, et al.
Veröffentlicht: (2024)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
von: Mishra, Shambhavi, et al.
Veröffentlicht: (2024)
von: Mishra, Shambhavi, et al.
Veröffentlicht: (2024)
VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors
von: Belal, Atif, et al.
Veröffentlicht: (2025)
von: Belal, Atif, et al.
Veröffentlicht: (2025)
Learning De-Biased Representations for Remote-Sensing Imagery
von: Tian, Zichen, et al.
Veröffentlicht: (2024)
von: Tian, Zichen, et al.
Veröffentlicht: (2024)
A Realistic Protocol for Evaluation of Weakly Supervised Object Localization
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
Advancements in Repetitive Action Counting: Joint-Based PoseRAC Model With Improved Performance
von: Chen, Haodong, et al.
Veröffentlicht: (2023)
von: Chen, Haodong, et al.
Veröffentlicht: (2023)
Generative Animations: A Multi-Model Pipeline for Prompt-Driven Motion Synthesis
von: Khurana, Mannat, et al.
Veröffentlicht: (2026)
von: Khurana, Mannat, et al.
Veröffentlicht: (2026)
Data-free Knowledge Distillation for Fine-grained Visual Categorization
von: Shao, Renrong, et al.
Veröffentlicht: (2024)
von: Shao, Renrong, et al.
Veröffentlicht: (2024)
ASBA: A-line State Space Model and B-line Attention for Sparse Optical Doppler Tomography Reconstruction
von: Li, Zhenghong, et al.
Veröffentlicht: (2026)
von: Li, Zhenghong, et al.
Veröffentlicht: (2026)
Low-Rank Expert Merging for Multi-Source Domain Adaptation in Person Re-Identification
von: Nehdi, Taha Mustapha, et al.
Veröffentlicht: (2025)
von: Nehdi, Taha Mustapha, et al.
Veröffentlicht: (2025)
Guided Interpretable Facial Expression Recognition via Spatial Action Unit Cues
von: Belharbi, Soufiane, et al.
Veröffentlicht: (2024)
von: Belharbi, Soufiane, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Source-Free Domain Adaptation for YOLO Object Detection
von: Varailhon, Simon, et al.
Veröffentlicht: (2024) -
Revisiting Mixout: An Overlooked Path to Robust Finetuning
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025) -
High-Rate Mixout: Revisiting Mixout for Robust Domain Generalization
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025) -
LT-Soups: Bridging Head and Tail Classes via Subsampled Model Soups
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025) -
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
von: Muralidharan, Srikanth, et al.
Veröffentlicht: (2025)