Unified Multi-Dataset Training for TBPS
Fuente:
arXiv
Saved in:
| Main Authors: | Chatterjee, Nilanjana, Garg, Sidharatha, Subramanyam, A V, Lall, Brejesh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging band diversity for feature selection in EO data
by: Hussain, Sadia, et al.
Published: (2025)
by: Hussain, Sadia, et al.
Published: (2025)
GraVITON: Graph based garment warping with attention guided inversion for Virtual-tryon
by: Pathak, Sanhita, et al.
Published: (2024)
by: Pathak, Sanhita, et al.
Published: (2024)
Single Stage Warped Cloth Learning and Semantic-Contextual Attention Feature Fusion for Virtual TryOn
by: Pathak, Sanhita, et al.
Published: (2023)
by: Pathak, Sanhita, et al.
Published: (2023)
DiffSTR: Controlled Diffusion Models for Scene Text Removal
by: Pathak, Sanhita, et al.
Published: (2024)
by: Pathak, Sanhita, et al.
Published: (2024)
Boosting Weak Positives for Text Based Person Search
by: Modi, Akshay, et al.
Published: (2025)
by: Modi, Akshay, et al.
Published: (2025)
Knowledge Distillation in Vision Transformers: A Critical Review
by: Habib, Gousia, et al.
Published: (2023)
by: Habib, Gousia, et al.
Published: (2023)
Stride-Net: Fairness-Aware Disentangled Representation Learning for Chest X-Ray Diagnosis
by: Rashid, Darakshan, et al.
Published: (2026)
by: Rashid, Darakshan, et al.
Published: (2026)
Optimizing Vision Transformers with Data-Free Knowledge Transfer
by: Habib, Gousia, et al.
Published: (2024)
by: Habib, Gousia, et al.
Published: (2024)
Continual Segmentation under Joint Nonstationarity
by: Pandey, Prashant, et al.
Published: (2026)
by: Pandey, Prashant, et al.
Published: (2026)
Dual Thinking and Logical Processing -- Are Multi-modal Large Language Models Closing the Gap with Human Vision ?
by: Dayanandan, Kailas, et al.
Published: (2024)
by: Dayanandan, Kailas, et al.
Published: (2024)
AD-Relight: Training-Free Banner Relighting via Illumination Translation with Diffusion Priors
by: Mishra, Rameshwar, et al.
Published: (2026)
by: Mishra, Rameshwar, et al.
Published: (2026)
A Comprehensive Review of Knowledge Distillation in Computer Vision
by: Habib, Gousia, et al.
Published: (2024)
by: Habib, Gousia, et al.
Published: (2024)
LIB-KD: Teaching Inductive Bias for Efficient Vision Transformer Distillation and Compression
by: Habib, Gousia, et al.
Published: (2023)
by: Habib, Gousia, et al.
Published: (2023)
A Comprehensive Survey on Synthetic Infrared Image synthesis
by: Upadhyay, Avinash, et al.
Published: (2024)
by: Upadhyay, Avinash, et al.
Published: (2024)
Image Synthesis with Graph Conditioning: CLIP-Guided Diffusion Models for Scene Graphs
by: Mishra, Rameshwar, et al.
Published: (2024)
by: Mishra, Rameshwar, et al.
Published: (2024)
Keypoint Aware Masked Image Modelling
by: Krishna, Madhava, et al.
Published: (2024)
by: Krishna, Madhava, et al.
Published: (2024)
Resource Efficient Perception for Vision Systems
by: Subramanyam, A V, et al.
Published: (2024)
by: Subramanyam, A V, et al.
Published: (2024)
Novel View Synthesis using DDIM Inversion
by: Singh, Sehajdeep, et al.
Published: (2025)
by: Singh, Sehajdeep, et al.
Published: (2025)
Online Pseudo-Label Unified Object Detection for Multiple Datasets Training
by: Tang, XiaoJun, et al.
Published: (2024)
by: Tang, XiaoJun, et al.
Published: (2024)
Conditional Consistency Guided Image Translation and Enhancement
by: Bhagat, Amil, et al.
Published: (2025)
by: Bhagat, Amil, et al.
Published: (2025)
The Deepfake Detective: Interpreting Neural Forensics Through Sparse Features and Manifolds
by: Sahoo, Subramanyam, et al.
Published: (2025)
by: Sahoo, Subramanyam, et al.
Published: (2025)
Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
by: Lall, Vishakha, et al.
Published: (2025)
by: Lall, Vishakha, et al.
Published: (2025)
Benchmarking Multi-dimensional AIGC Video Quality Assessment: A Dataset and Unified Model
by: Zhang, Zhichao, et al.
Published: (2024)
by: Zhang, Zhichao, et al.
Published: (2024)
UKnow: A Unified Knowledge Protocol with Multimodal Knowledge Graph Datasets for Reasoning and Vision-Language Pre-Training
by: Gong, Biao, et al.
Published: (2023)
by: Gong, Biao, et al.
Published: (2023)
Multi-Dataset Cross-Domain Knowledge Distillation for Unified Medical Image Segmentation, Classification, and Detection
by: Ciprian-Mihai, Ceausescu, et al.
Published: (2026)
by: Ciprian-Mihai, Ceausescu, et al.
Published: (2026)
PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation
by: Jeong, Uyoung, et al.
Published: (2025)
by: Jeong, Uyoung, et al.
Published: (2025)
Estimating Human Poses Across Datasets: A Unified Skeleton and Multi-Teacher Distillation Approach
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
Towards Unified Benchmark and Models for Multi-Modal Perceptual Metrics
by: Ghazanfari, Sara, et al.
Published: (2024)
by: Ghazanfari, Sara, et al.
Published: (2024)
When Big Models Train Small Ones: Label-Free Model Parity Alignment for Efficient Visual Question Answering using Small VLMs
by: Penamakuri, Abhirama Subramanyam, et al.
Published: (2025)
by: Penamakuri, Abhirama Subramanyam, et al.
Published: (2025)
DECIDER: Leveraging Foundation Model Priors for Improved Model Failure Detection and Explanation
by: Subramanyam, Rakshith, et al.
Published: (2024)
by: Subramanyam, Rakshith, et al.
Published: (2024)
UniRain: Unified Image Deraining with RAG-based Dataset Distillation and Multi-objective Reweighted Optimization
by: Yang, Qianfeng, et al.
Published: (2026)
by: Yang, Qianfeng, et al.
Published: (2026)
EEmo-Logic: A Unified Dataset and Multi-Stage Framework for Comprehensive Image-Evoked Emotion Assessment
by: Gao, Lancheng, et al.
Published: (2026)
by: Gao, Lancheng, et al.
Published: (2026)
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition
by: He, Yu, et al.
Published: (2026)
by: He, Yu, et al.
Published: (2026)
Unified Generative and Discriminative Training for Multi-modal Large Language Models
by: Chow, Wei, et al.
Published: (2024)
by: Chow, Wei, et al.
Published: (2024)
FSCA-Net: Feature-Separated Cross-Attention Network for Robust Multi-Dataset Training
by: Chen, Yuehai
Published: (2026)
by: Chen, Yuehai
Published: (2026)
PicoEyes: Unified Gaze Estimation Framework for Mixed Reality with a Large-Scale Multi-View Dataset
by: Duan, Fuxin, et al.
Published: (2026)
by: Duan, Fuxin, et al.
Published: (2026)
Dataset Distillation by Automatic Training Trajectories
by: Liu, Dai, et al.
Published: (2024)
by: Liu, Dai, et al.
Published: (2024)
Explainable Gait Abnormality Detection Using Dual-Dataset CNN-LSTM Models
by: Agarwal, Parth, et al.
Published: (2025)
by: Agarwal, Parth, et al.
Published: (2025)
ImgEdit: A Unified Image Editing Dataset and Benchmark
by: Ye, Yang, et al.
Published: (2025)
by: Ye, Yang, et al.
Published: (2025)
Flash-Unified: A Training-Free and Task-Aware Acceleration Framework for Native Unified Models
by: Ke, Junlong, et al.
Published: (2026)
by: Ke, Junlong, et al.
Published: (2026)
Similar Items
-
Leveraging band diversity for feature selection in EO data
by: Hussain, Sadia, et al.
Published: (2025) -
GraVITON: Graph based garment warping with attention guided inversion for Virtual-tryon
by: Pathak, Sanhita, et al.
Published: (2024) -
Single Stage Warped Cloth Learning and Semantic-Contextual Attention Feature Fusion for Virtual TryOn
by: Pathak, Sanhita, et al.
Published: (2023) -
DiffSTR: Controlled Diffusion Models for Scene Text Removal
by: Pathak, Sanhita, et al.
Published: (2024) -
Boosting Weak Positives for Text Based Person Search
by: Modi, Akshay, et al.
Published: (2025)