One-Shot Multilingual Font Generation Via ViT
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Zhiheng, Liu, Jiarui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RepViT: Revisiting Mobile CNN From ViT Perspective
por: Wang, Ao, et al.
Publicado: (2023)
por: Wang, Ao, et al.
Publicado: (2023)
Deeper Inside Deep ViT
por: Hong, Sungrae
Publicado: (2025)
por: Hong, Sungrae
Publicado: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
por: Zhong, Yunshan, et al.
Publicado: (2023)
por: Zhong, Yunshan, et al.
Publicado: (2023)
Applying ViT in Generalized Few-shot Semantic Segmentation
por: Geng, Liyuan, et al.
Publicado: (2024)
por: Geng, Liyuan, et al.
Publicado: (2024)
ViT-5: Vision Transformers for The Mid-2020s
por: Wang, Feng, et al.
Publicado: (2026)
por: Wang, Feng, et al.
Publicado: (2026)
TFS-ViT: Token-Level Feature Stylization for Domain Generalization
por: Noori, Mehrdad, et al.
Publicado: (2023)
por: Noori, Mehrdad, et al.
Publicado: (2023)
Let ViT Speak: Generative Language-Image Pre-training
por: Fang, Yan, et al.
Publicado: (2026)
por: Fang, Yan, et al.
Publicado: (2026)
Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
por: Liu, Jiani, et al.
Publicado: (2025)
por: Liu, Jiani, et al.
Publicado: (2025)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
por: Zhu, Chen, et al.
Publicado: (2025)
por: Zhu, Chen, et al.
Publicado: (2025)
DA-Font: Few-Shot Font Generation via Dual-Attention Hybrid Integration
por: Chen, Weiran, et al.
Publicado: (2025)
por: Chen, Weiran, et al.
Publicado: (2025)
Few-Part-Shot Font Generation
por: Akiba, Masaki, et al.
Publicado: (2025)
por: Akiba, Masaki, et al.
Publicado: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
por: Khoramdel, Javad, et al.
Publicado: (2024)
por: Khoramdel, Javad, et al.
Publicado: (2024)
Rethinking Random Masking in Self-Distillation on ViT
por: Seong, Jihyeon, et al.
Publicado: (2025)
por: Seong, Jihyeon, et al.
Publicado: (2025)
Your ViT is Secretly an Image Segmentation Model
por: Kerssies, Tommie, et al.
Publicado: (2025)
por: Kerssies, Tommie, et al.
Publicado: (2025)
ViT$^3$: Unlocking Test-Time Training in Vision
por: Han, Dongchen, et al.
Publicado: (2025)
por: Han, Dongchen, et al.
Publicado: (2025)
Few-Shot Class-Incremental Model Attribution Using Learnable Representation From CLIP-ViT Features
por: Lee, Hanbyul, et al.
Publicado: (2025)
por: Lee, Hanbyul, et al.
Publicado: (2025)
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection
por: Zhang, Jiangning, et al.
Publicado: (2023)
por: Zhang, Jiangning, et al.
Publicado: (2023)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
por: Chattopadhyay, Nandish, et al.
Publicado: (2026)
por: Chattopadhyay, Nandish, et al.
Publicado: (2026)
Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation
por: Zhao, Wangbo, et al.
Publicado: (2024)
por: Zhao, Wangbo, et al.
Publicado: (2024)
ViTCAE: ViT-based Class-conditioned Autoencoder
por: Jebraeeli, Vahid, et al.
Publicado: (2025)
por: Jebraeeli, Vahid, et al.
Publicado: (2025)
U-REPA: Aligning Diffusion U-Nets to ViTs
por: Tian, Yuchuan, et al.
Publicado: (2025)
por: Tian, Yuchuan, et al.
Publicado: (2025)
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers
por: Ibtehaz, Nabil, et al.
Publicado: (2024)
por: Ibtehaz, Nabil, et al.
Publicado: (2024)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
por: Puy, Gilles, et al.
Publicado: (2026)
por: Puy, Gilles, et al.
Publicado: (2026)
Mobile U-ViT: Revisiting large kernel and U-shaped ViT for efficient medical image segmentation
por: Tang, Fenghe, et al.
Publicado: (2025)
por: Tang, Fenghe, et al.
Publicado: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
por: Shah, Arya, et al.
Publicado: (2025)
por: Shah, Arya, et al.
Publicado: (2025)
SFMViT: SlowFast Meet ViT in Chaotic World
por: Lin, Jiaying, et al.
Publicado: (2024)
por: Lin, Jiaying, et al.
Publicado: (2024)
Improve Contrastive Clustering Performance by Multiple Fusing-Augmenting ViT Blocks
por: Wang, Cheng, et al.
Publicado: (2025)
por: Wang, Cheng, et al.
Publicado: (2025)
Elastic ViTs from Pretrained Models without Retraining
por: Simoncini, Walter, et al.
Publicado: (2025)
por: Simoncini, Walter, et al.
Publicado: (2025)
ViT-AdaLA: Adapting Vision Transformers with Linear Attention
por: Li, Yifan, et al.
Publicado: (2026)
por: Li, Yifan, et al.
Publicado: (2026)
Pretrained ViTs Yield Versatile Representations For Medical Images
por: Matsoukas, Christos, et al.
Publicado: (2023)
por: Matsoukas, Christos, et al.
Publicado: (2023)
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
por: Ma, Xiaochen, et al.
Publicado: (2023)
por: Ma, Xiaochen, et al.
Publicado: (2023)
ViT-1.58b: Mobile Vision Transformers in the 1-bit Era
por: Yuan, Zhengqing, et al.
Publicado: (2024)
por: Yuan, Zhengqing, et al.
Publicado: (2024)
How to train your ViT for OOD Detection
por: Mueller, Maximilian, et al.
Publicado: (2024)
por: Mueller, Maximilian, et al.
Publicado: (2024)
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
por: Ma, Yunsheng, et al.
Publicado: (2022)
por: Ma, Yunsheng, et al.
Publicado: (2022)
ViT-Lens: Towards Omni-modal Representations
por: Lei, Weixian, et al.
Publicado: (2023)
por: Lei, Weixian, et al.
Publicado: (2023)
PRANCE: Joint Token-Optimization and Structural Channel-Pruning for Adaptive ViT Inference
por: Li, Ye, et al.
Publicado: (2024)
por: Li, Ye, et al.
Publicado: (2024)
EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation
por: Liu, Longfei, et al.
Publicado: (2026)
por: Liu, Longfei, et al.
Publicado: (2026)
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer
por: Tai, Yu-Shan, et al.
Publicado: (2024)
por: Tai, Yu-Shan, et al.
Publicado: (2024)
Charm: The Missing Piece in ViT fine-tuning for Image Aesthetic Assessment
por: Behrad, Fatemeh, et al.
Publicado: (2025)
por: Behrad, Fatemeh, et al.
Publicado: (2025)
ToaSt: Token Channel Selection and Structured Pruning for Efficient ViT
por: Moon, Hyunchan, et al.
Publicado: (2026)
por: Moon, Hyunchan, et al.
Publicado: (2026)
Ejemplares similares
-
RepViT: Revisiting Mobile CNN From ViT Perspective
por: Wang, Ao, et al.
Publicado: (2023) -
Deeper Inside Deep ViT
por: Hong, Sungrae
Publicado: (2025) -
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
por: Zhong, Yunshan, et al.
Publicado: (2023) -
Applying ViT in Generalized Few-shot Semantic Segmentation
por: Geng, Liyuan, et al.
Publicado: (2024) -
ViT-5: Vision Transformers for The Mid-2020s
por: Wang, Feng, et al.
Publicado: (2026)