Salvato in:
| Autori principali: | Banerjee, Soumya, Verma, Vinay K., Mukherjee, Avideep, Gupta, Deepak, Namboodiri, Vinay P., Rai, Piyush |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2309.08227 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RISSOLE: Parameter-efficient Diffusion Models via Block-wise Generation and Retrieval-Guidance
di: Mukherjee, Avideep, et al.
Pubblicazione: (2024)
di: Mukherjee, Avideep, et al.
Pubblicazione: (2024)
NOVO: Unlearning-Compliant Vision Transformers
di: Roy, Soumya, et al.
Pubblicazione: (2025)
di: Roy, Soumya, et al.
Pubblicazione: (2025)
ADROIT: A Self-Supervised Framework for Learning Robust Representations for Active Learning
di: Banerjee, Soumya, et al.
Pubblicazione: (2025)
di: Banerjee, Soumya, et al.
Pubblicazione: (2025)
StyleYourSmile: Cross-Domain Face Retargeting Without Paired Multi-Style Data
di: Dey, Avirup, et al.
Pubblicazione: (2025)
di: Dey, Avirup, et al.
Pubblicazione: (2025)
CLIP Adaptation by Intra-modal Overlap Reduction
di: Kravets, Alexey, et al.
Pubblicazione: (2024)
di: Kravets, Alexey, et al.
Pubblicazione: (2024)
TalkLoRA: Low-Rank Adaptation for Speech-Driven Animation
di: Saunders, Jack, et al.
Pubblicazione: (2024)
di: Saunders, Jack, et al.
Pubblicazione: (2024)
Trusting Semantic Segmentation Networks
di: Some, Samik, et al.
Pubblicazione: (2024)
di: Some, Samik, et al.
Pubblicazione: (2024)
Can Unsupervised Segmentation Reduce Annotation Costs for Video Semantic Segmentation?
di: Some, Samik, et al.
Pubblicazione: (2026)
di: Some, Samik, et al.
Pubblicazione: (2026)
Dubbing for Everyone: Data-Efficient Visual Dubbing using Neural Rendering Priors
di: Saunders, Jack, et al.
Pubblicazione: (2024)
di: Saunders, Jack, et al.
Pubblicazione: (2024)
Determinantal Point Process as an alternative to NMS
di: Some, Samik, et al.
Pubblicazione: (2020)
di: Some, Samik, et al.
Pubblicazione: (2020)
MedFocusCLIP : Improving few shot classification in medical datasets using pixel wise attention
di: Arora, Aadya, et al.
Pubblicazione: (2025)
di: Arora, Aadya, et al.
Pubblicazione: (2025)
Self-supervised Representation Learning for Cell Event Recognition through Time Arrow Prediction
di: Chen, Cangxiong, et al.
Pubblicazione: (2024)
di: Chen, Cangxiong, et al.
Pubblicazione: (2024)
Rethinking Few Shot CLIP Benchmarks: A Critical Analysis in the Inductive Setting
di: Kravets, Alexey, et al.
Pubblicazione: (2025)
di: Kravets, Alexey, et al.
Pubblicazione: (2025)
EIDT-V: Exploiting Intersections in Diffusion Trajectories for Model-Agnostic, Zero-Shot, Training-Free Text-to-Video Generation
di: Jagpal, Diljeet, et al.
Pubblicazione: (2025)
di: Jagpal, Diljeet, et al.
Pubblicazione: (2025)
RAW: Robust Avatar Watermarking -- Benchmarking and Baseline
di: Parry, Jack, et al.
Pubblicazione: (2026)
di: Parry, Jack, et al.
Pubblicazione: (2026)
Interpretability Transfer from Language to Vision via Sparse Autoencoders
di: Kravets, Alexey, et al.
Pubblicazione: (2026)
di: Kravets, Alexey, et al.
Pubblicazione: (2026)
Single Stage Warped Cloth Learning and Semantic-Contextual Attention Feature Fusion for Virtual TryOn
di: Pathak, Sanhita, et al.
Pubblicazione: (2023)
di: Pathak, Sanhita, et al.
Pubblicazione: (2023)
Resource Efficient Perception for Vision Systems
di: Subramanyam, A V, et al.
Pubblicazione: (2024)
di: Subramanyam, A V, et al.
Pubblicazione: (2024)
Towards Accurate Lip-to-Speech Synthesis in-the-Wild
di: Hegde, Sindhu, et al.
Pubblicazione: (2024)
di: Hegde, Sindhu, et al.
Pubblicazione: (2024)
Convolutional Prompting meets Language Models for Continual Learning
di: Roy, Anurag, et al.
Pubblicazione: (2024)
di: Roy, Anurag, et al.
Pubblicazione: (2024)
GraVITON: Graph based garment warping with attention guided inversion for Virtual-tryon
di: Pathak, Sanhita, et al.
Pubblicazione: (2024)
di: Pathak, Sanhita, et al.
Pubblicazione: (2024)
Heracles: A Hybrid SSM-Transformer Model for High-Resolution Image and Time-Series Analysis
di: Patro, Badri N., et al.
Pubblicazione: (2024)
di: Patro, Badri N., et al.
Pubblicazione: (2024)
Federated Learning with Uncertainty and Personalization via Efficient Second-order Optimization
di: Pal, Shivam, et al.
Pubblicazione: (2024)
di: Pal, Shivam, et al.
Pubblicazione: (2024)
FOCUS: Forcing In-Context Object Localization through Visual Support Constraints and Policy Optimization
di: Karim, Mohammed Asad, et al.
Pubblicazione: (2026)
di: Karim, Mohammed Asad, et al.
Pubblicazione: (2026)
Uncertainty-Aware Vision-Language Segmentation for Medical Imaging
di: Das, Aryan, et al.
Pubblicazione: (2026)
di: Das, Aryan, et al.
Pubblicazione: (2026)
A Novel Cloud-Based Diffusion-Guided Hybrid Model for High-Accuracy Accident Detection in Intelligent Transportation Systems
di: Sai, Siva, et al.
Pubblicazione: (2025)
di: Sai, Siva, et al.
Pubblicazione: (2025)
Rethinking Test Time Scaling for Flow-Matching Generative Models
di: Yu, Qingtao, et al.
Pubblicazione: (2025)
di: Yu, Qingtao, et al.
Pubblicazione: (2025)
Illumination-Aware Contactless Fingerprint Spoof Detection via Paired Flash-Non-Flash Imaging
di: Sahoo, Roja, et al.
Pubblicazione: (2026)
di: Sahoo, Roja, et al.
Pubblicazione: (2026)
ShortCheck: Checkworthiness Detection of Multilingual Short-Form Videos
di: Vatndal, Henrik, et al.
Pubblicazione: (2025)
di: Vatndal, Henrik, et al.
Pubblicazione: (2025)
Streaming Drag-Oriented Interactive Video Manipulation: Drag Anything, Anytime!
di: Zhou, Junbao, et al.
Pubblicazione: (2025)
di: Zhou, Junbao, et al.
Pubblicazione: (2025)
Reliable or Deceptive? Investigating Gated Features for Smooth Visual Explanations in CNNs
di: Mitra, Soham, et al.
Pubblicazione: (2024)
di: Mitra, Soham, et al.
Pubblicazione: (2024)
Efficient Text-Guided Convolutional Adapter for the Diffusion Model
di: Das, Aryan, et al.
Pubblicazione: (2026)
di: Das, Aryan, et al.
Pubblicazione: (2026)
DiffSTR: Controlled Diffusion Models for Scene Text Removal
di: Pathak, Sanhita, et al.
Pubblicazione: (2024)
di: Pathak, Sanhita, et al.
Pubblicazione: (2024)
Chirality in Action: Time-Aware Video Representation Learning by Latent Straightening
di: Bagad, Piyush, et al.
Pubblicazione: (2025)
di: Bagad, Piyush, et al.
Pubblicazione: (2025)
Shortcut Invariance: Targeted Jacobian Regularization in Disentangled Latent Space
di: Pal, Shivam, et al.
Pubblicazione: (2025)
di: Pal, Shivam, et al.
Pubblicazione: (2025)
MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts
di: Edula, Vinay, et al.
Pubblicazione: (2026)
di: Edula, Vinay, et al.
Pubblicazione: (2026)
RefDiffNet: Learning to Expose Subtle PCB Defects Before Detection
di: Edula, Vinay, et al.
Pubblicazione: (2026)
di: Edula, Vinay, et al.
Pubblicazione: (2026)
Roommate Compatibility Detection Through Machine Learning Techniques
di: Lamba, Mansha, et al.
Pubblicazione: (2020)
di: Lamba, Mansha, et al.
Pubblicazione: (2020)
HAD: Heterogeneity-Aware Distillation for Lifelong Heterogeneous Learning
di: Zhang, Xuerui, et al.
Pubblicazione: (2026)
di: Zhang, Xuerui, et al.
Pubblicazione: (2026)
Comparative Evaluation of Traditional Methods and Deep Learning for Brain Glioma Imaging. Review Paper
di: Janardhan, Kiranmayee, et al.
Pubblicazione: (2026)
di: Janardhan, Kiranmayee, et al.
Pubblicazione: (2026)
Documenti analoghi
-
RISSOLE: Parameter-efficient Diffusion Models via Block-wise Generation and Retrieval-Guidance
di: Mukherjee, Avideep, et al.
Pubblicazione: (2024) -
NOVO: Unlearning-Compliant Vision Transformers
di: Roy, Soumya, et al.
Pubblicazione: (2025) -
ADROIT: A Self-Supervised Framework for Learning Robust Representations for Active Learning
di: Banerjee, Soumya, et al.
Pubblicazione: (2025) -
StyleYourSmile: Cross-Domain Face Retargeting Without Paired Multi-Style Data
di: Dey, Avirup, et al.
Pubblicazione: (2025) -
CLIP Adaptation by Intra-modal Overlap Reduction
di: Kravets, Alexey, et al.
Pubblicazione: (2024)