STEFANN: Scene Text Editor using Font Adaptive Neural Network
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roy, Prasun, Bhattacharya, Saumik, Ghosh, Subhankar, Pal, Umapada |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2019
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-scale Attention Guided Pose Transfer
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
von: Das, Alloy, et al.
Veröffentlicht: (2023)
von: Das, Alloy, et al.
Veröffentlicht: (2023)
Scene Aware Person Image Generation through Global Contextual Conditioning
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
TIPS: Text-Induced Pose Synthesis
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
von: Roy, Prasun, et al.
Veröffentlicht: (2022)
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
Semantically Consistent Person Image Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
Effects of Degradations on Deep Neural Network Architectures
von: Roy, Prasun, et al.
Veröffentlicht: (2018)
von: Roy, Prasun, et al.
Veröffentlicht: (2018)
DRG-Font: Dynamic Reference-Guided Few-shot Font Generation via Contrastive Style-Content Disentanglement
von: Chakraborty, Rejoy, et al.
Veröffentlicht: (2026)
von: Chakraborty, Rejoy, et al.
Veröffentlicht: (2026)
Position and Rotation Invariant Sign Language Recognition from 3D Kinect Data with Recurrent Neural Networks
von: Roy, Prasun, et al.
Veröffentlicht: (2020)
von: Roy, Prasun, et al.
Veröffentlicht: (2020)
A CNN Based Framework for Unistroke Numeral Recognition in Air-Writing
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
von: Roy, Prasun, et al.
Veröffentlicht: (2023)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
von: Das, Alloy, et al.
Veröffentlicht: (2024)
von: Das, Alloy, et al.
Veröffentlicht: (2024)
MIO : Mutual Information Optimization using Self-Supervised Binary Contrastive Learning
von: Manna, Siladittya, et al.
Veröffentlicht: (2021)
von: Manna, Siladittya, et al.
Veröffentlicht: (2021)
Correlation Weighted Prototype-based Self-Supervised One-Shot Segmentation of Medical Images
von: Manna, Siladittya, et al.
Veröffentlicht: (2024)
von: Manna, Siladittya, et al.
Veröffentlicht: (2024)
ControlText: Unlocking Controllable Fonts in Multilingual Text Rendering without Font Annotations
von: Jiang, Bowen, et al.
Veröffentlicht: (2025)
von: Jiang, Bowen, et al.
Veröffentlicht: (2025)
Scene-Text Grounding for Text-Based Video Question Answering
von: Zhou, Sheng, et al.
Veröffentlicht: (2024)
von: Zhou, Sheng, et al.
Veröffentlicht: (2024)
MorphText: Deep Morphology Regularized Arbitrary-shape Scene Text Detection
von: Xu, Chengpei, et al.
Veröffentlicht: (2024)
von: Xu, Chengpei, et al.
Veröffentlicht: (2024)
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering
von: Zhou, Sheng, et al.
Veröffentlicht: (2025)
von: Zhou, Sheng, et al.
Veröffentlicht: (2025)
SceneDreamer360: Text-Driven 3D-Consistent Scene Generation with Panoramic Gaussian Splatting
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation
von: Cai, Haonan, et al.
Veröffentlicht: (2026)
von: Cai, Haonan, et al.
Veröffentlicht: (2026)
Detection of Cyberbullying in GIF using AI
von: Dave, Pal, et al.
Veröffentlicht: (2025)
von: Dave, Pal, et al.
Veröffentlicht: (2025)
Decorrelation-based Self-Supervised Visual Representation Learning for Writer Identification
von: Maitra, Arkadip, et al.
Veröffentlicht: (2024)
von: Maitra, Arkadip, et al.
Veröffentlicht: (2024)
Word-level Sign Language Recognition with Multi-stream Neural Networks Focusing on Local Regions and Skeletal Information
von: Maruyama, Mizuki, et al.
Veröffentlicht: (2021)
von: Maruyama, Mizuki, et al.
Veröffentlicht: (2021)
T2VParser: Adaptive Decomposition Tokens for Partial Alignment in Text to Video Retrieval
von: Li, Yili, et al.
Veröffentlicht: (2025)
von: Li, Yili, et al.
Veröffentlicht: (2025)
AGSP-DSA: An Adaptive Graph Signal Processing Framework for Robust Multimodal Fusion with Dynamic Semantic Alignment
von: Karthikeya, KV, et al.
Veröffentlicht: (2026)
von: Karthikeya, KV, et al.
Veröffentlicht: (2026)
CAMeL: Cross-modality Adaptive Meta-Learning for Text-based Person Retrieval
von: Yu, Hang, et al.
Veröffentlicht: (2025)
von: Yu, Hang, et al.
Veröffentlicht: (2025)
4D Gaussian Splatting with Scale-aware Residual Field and Adaptive Optimization for Real-time Rendering of Temporally Complex Dynamic Scenes
von: Yan, Jinbo, et al.
Veröffentlicht: (2024)
von: Yan, Jinbo, et al.
Veröffentlicht: (2024)
Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
Show Me the World in My Language: Establishing the First Baseline for Scene-Text to Scene-Text Translation
von: Vaidya, Shreyas, et al.
Veröffentlicht: (2023)
von: Vaidya, Shreyas, et al.
Veröffentlicht: (2023)
Single Image Dehazing Using Scene Depth Ordering
von: Ling, Pengyang, et al.
Veröffentlicht: (2024)
von: Ling, Pengyang, et al.
Veröffentlicht: (2024)
Dynamically Scaled Temperature in Self-Supervised Contrastive Learning
von: Manna, Siladittya, et al.
Veröffentlicht: (2023)
von: Manna, Siladittya, et al.
Veröffentlicht: (2023)
SNP-S3: Shared Network Pre-training and Significant Semantic Strengthening for Various Video-Text Tasks
von: Dong, Xingning, et al.
Veröffentlicht: (2024)
von: Dong, Xingning, et al.
Veröffentlicht: (2024)
Scene Graph Generation with Role-Playing Large Language Models
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
Modality-Aware Shot Relating and Comparing for Video Scene Detection
von: Tan, Jiawei, et al.
Veröffentlicht: (2024)
von: Tan, Jiawei, et al.
Veröffentlicht: (2024)
CPSL: Representing Volumetric Video via Content-Promoted Scene Layers
von: Hu, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Hu, Kaiyuan, et al.
Veröffentlicht: (2025)
SCENEFORGE: Enhancing 3D-text alignment with Structured Scene Compositions
von: Sbrolli, Cristian, et al.
Veröffentlicht: (2025)
von: Sbrolli, Cristian, et al.
Veröffentlicht: (2025)
Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion
von: Wang, Xinghan, et al.
Veröffentlicht: (2024)
von: Wang, Xinghan, et al.
Veröffentlicht: (2024)
DBDH: A Dual-Branch Dual-Head Neural Network for Invisible Embedded Regions Localization
von: Zhao, Chengxin, et al.
Veröffentlicht: (2024)
von: Zhao, Chengxin, et al.
Veröffentlicht: (2024)
Unbiased Video Scene Graph Generation via Visual and Semantic Dual Debiasing
von: Li, Yanjun, et al.
Veröffentlicht: (2025)
von: Li, Yanjun, et al.
Veröffentlicht: (2025)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
von: Li, Zeju, et al.
Veröffentlicht: (2024)
von: Li, Zeju, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multi-scale Attention Guided Pose Transfer
von: Roy, Prasun, et al.
Veröffentlicht: (2022) -
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
von: Das, Alloy, et al.
Veröffentlicht: (2023) -
Scene Aware Person Image Generation through Global Contextual Conditioning
von: Roy, Prasun, et al.
Veröffentlicht: (2022) -
TIPS: Text-Induced Pose Synthesis
von: Roy, Prasun, et al.
Veröffentlicht: (2022) -
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2025)