Multi-scale Attention Guided Pose Transfer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roy, Prasun, Bhattacharya, Saumik, Ghosh, Subhankar, Pal, Umapada |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TIPS: Text-Induced Pose Synthesis
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
STEFANN: Scene Text Editor using Font Adaptive Neural Network
von: Roy, Prasun, et al.
Veröffentlicht: (2019)
von: Roy, Prasun, et al.
Veröffentlicht: (2019)
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
Scene Aware Person Image Generation through Global Contextual Conditioning
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
Semantically Consistent Person Image Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
von: Das, Alloy, et al.
Veröffentlicht: (2023)
von: Das, Alloy, et al.
Veröffentlicht: (2023)
Effects of Degradations on Deep Neural Network Architectures
von: Roy, Prasun, et al.
Veröffentlicht: (2018)
von: Roy, Prasun, et al.
Veröffentlicht: (2018)
DRG-Font: Dynamic Reference-Guided Few-shot Font Generation via Contrastive Style-Content Disentanglement
von: Chakraborty, Rejoy, et al.
Veröffentlicht: (2026)
von: Chakraborty, Rejoy, et al.
Veröffentlicht: (2026)
A CNN Based Framework for Unistroke Numeral Recognition in Air-Writing
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
Position and Rotation Invariant Sign Language Recognition from 3D Kinect Data with Recurrent Neural Networks
von: Roy, Prasun, et al.
Veröffentlicht: (2020)
von: Roy, Prasun, et al.
Veröffentlicht: (2020)
Correlation Weighted Prototype-based Self-Supervised One-Shot Segmentation of Medical Images
von: Manna, Siladittya, et al.
Veröffentlicht: (2024)
von: Manna, Siladittya, et al.
Veröffentlicht: (2024)
MIO : Mutual Information Optimization using Self-Supervised Binary Contrastive Learning
von: Manna, Siladittya, et al.
Veröffentlicht: (2021)
von: Manna, Siladittya, et al.
Veröffentlicht: (2021)
Follow-Your-MultiPose: Tuning-Free Multi-Character Text-to-Video Generation via Pose Guidance
von: Zhang, Beiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Beiyuan, et al.
Veröffentlicht: (2024)
SAGE-GAN: Towards Realistic and Robust Segmentation of Spatially Ordered Nanoparticles via Attention-Guided GANs
von: Pal, Anindya, et al.
Veröffentlicht: (2026)
von: Pal, Anindya, et al.
Veröffentlicht: (2026)
Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
Decorrelation-based Self-Supervised Visual Representation Learning for Writer Identification
von: Maitra, Arkadip, et al.
Veröffentlicht: (2024)
von: Maitra, Arkadip, et al.
Veröffentlicht: (2024)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
von: Das, Alloy, et al.
Veröffentlicht: (2024)
von: Das, Alloy, et al.
Veröffentlicht: (2024)
Multi-scale Bottleneck Transformer for Weakly Supervised Multimodal Violence Detection
von: Sun, Shengyang, et al.
Veröffentlicht: (2024)
von: Sun, Shengyang, et al.
Veröffentlicht: (2024)
Spatiotemporal Graph Guided Multi-modal Network for Livestreaming Product Retrieval
von: Hu, Xiaowan, et al.
Veröffentlicht: (2024)
von: Hu, Xiaowan, et al.
Veröffentlicht: (2024)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations
von: Gupta, Parul, et al.
Veröffentlicht: (2025)
von: Gupta, Parul, et al.
Veröffentlicht: (2025)
Multi-scale Activation, Refinement, and Aggregation: Exploring Diverse Cues for Fine-Grained Bird Recognition
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction
von: Li, Dong, et al.
Veröffentlicht: (2025)
von: Li, Dong, et al.
Veröffentlicht: (2025)
Detection and Recovery of Adversarial Slow-Pose Drift in Offloaded Visual-Inertial Odometry
von: Saha, Soruya, et al.
Veröffentlicht: (2025)
von: Saha, Soruya, et al.
Veröffentlicht: (2025)
Dynamically Scaled Temperature in Self-Supervised Contrastive Learning
von: Manna, Siladittya, et al.
Veröffentlicht: (2023)
von: Manna, Siladittya, et al.
Veröffentlicht: (2023)
Bridging the Pose-Semantic Gap: A Cascade Framework for Text-Based Person Anomaly Search
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
von: Xie, Zequn, et al.
Veröffentlicht: (2026)
Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
Harmonizing Attention: Training-free Texture-aware Geometry Transfer
von: Ikuta, Eito, et al.
Veröffentlicht: (2024)
von: Ikuta, Eito, et al.
Veröffentlicht: (2024)
Detection of Cyberbullying in GIF using AI
von: Dave, Pal, et al.
Veröffentlicht: (2025)
von: Dave, Pal, et al.
Veröffentlicht: (2025)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
HDiffTG: A Lightweight Hybrid Diffusion-Transformer-GCN Architecture for 3D Human Pose Estimation
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
Word-level Sign Language Recognition with Multi-stream Neural Networks Focusing on Local Regions and Skeletal Information
von: Maruyama, Mizuki, et al.
Veröffentlicht: (2021)
von: Maruyama, Mizuki, et al.
Veröffentlicht: (2021)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
von: Wei, Fangda, et al.
Veröffentlicht: (2026)
von: Wei, Fangda, et al.
Veröffentlicht: (2026)
Joint Flow And Feature Refinement Using Attention For Video Restoration
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
Embedded Heterogeneous Attention Transformer for Cross-lingual Image Captioning
von: Song, Zijie, et al.
Veröffentlicht: (2023)
von: Song, Zijie, et al.
Veröffentlicht: (2023)
DiffuseST: Unleashing the Capability of the Diffusion Model for Style Transfer
von: Hu, Ying, et al.
Veröffentlicht: (2024)
von: Hu, Ying, et al.
Veröffentlicht: (2024)
Group-based Distinctive Image Captioning with Memory Difference Encoding and Attention
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025)
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
von: Xie, Liping, et al.
Veröffentlicht: (2025)
von: Xie, Liping, et al.
Veröffentlicht: (2025)
Reference-Guided Identity Preserving Face Restoration
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TIPS: Text-Induced Pose Synthesis
von: Roy, Prasun, et al.
Veröffentlicht: (2022) -
STEFANN: Scene Text Editor using Font Adaptive Neural Network
von: Roy, Prasun, et al.
Veröffentlicht: (2019) -
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2025) -
Scene Aware Person Image Generation through Global Contextual Conditioning
von: Roy, Prasun, et al.
Veröffentlicht: (2022) -
Semantically Consistent Person Image Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2023)