PyramidStyler: Transformer-Based Neural Style Transfer with Pyramidal Positional Encoding and Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Durairaju, Raahul Krishna, Saruladha, K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
D2Styler: Advancing Arbitrary Style Transfer with Discrete Diffusion Methods
by: Susladkar, Onkar, et al.
Published: (2024)
by: Susladkar, Onkar, et al.
Published: (2024)
Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding
by: Chen, Zhanpeng, et al.
Published: (2025)
by: Chen, Zhanpeng, et al.
Published: (2025)
Global Feature Pyramid Network
by: Xiao, Weilin, et al.
Published: (2023)
by: Xiao, Weilin, et al.
Published: (2023)
A Unified Framework for Microscopy Defocus Deblur with Multi-Pyramid Transformer and Contrastive Learning
by: Zhang, Yuelin, et al.
Published: (2024)
by: Zhang, Yuelin, et al.
Published: (2024)
BatStyler: Advancing Multi-category Style Generation for Source-free Domain Generalization
by: Xu, Xiusheng, et al.
Published: (2025)
by: Xu, Xiusheng, et al.
Published: (2025)
Fast-iTPN: Integrally Pre-Trained Transformer Pyramid Network with Token Migration
by: Tian, Yunjie, et al.
Published: (2022)
by: Tian, Yunjie, et al.
Published: (2022)
Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation
by: Liu, Hanbing, et al.
Published: (2023)
by: Liu, Hanbing, et al.
Published: (2023)
Deep Learning-Based Fatigue Cracks Detection in Bridge Girders using Feature Pyramid Networks
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Pyramid Self-contrastive Learning Framework for Test-time Ultrasound Image Denoising
by: Zhang, Jiajing, et al.
Published: (2026)
by: Zhang, Jiajing, et al.
Published: (2026)
Power of Boundary and Reflection: Semantic Transparent Object Segmentation using Pyramid Vision Transformer with Transparent Cues
by: Vu, Tuan-Anh, et al.
Published: (2025)
by: Vu, Tuan-Anh, et al.
Published: (2025)
PyraTok: Language-Aligned Pyramidal Tokenizer for Video Understanding and Generation
by: Susladkar, Onkar, et al.
Published: (2026)
by: Susladkar, Onkar, et al.
Published: (2026)
Pyramid Coder: Hierarchical Code Generator for Compositional Visual Question Answering
by: Shen, Ruoyue, et al.
Published: (2024)
by: Shen, Ruoyue, et al.
Published: (2024)
An Enhanced Pyramid Feature Network Based on Long-Range Dependencies for Multi-Organ Medical Image Segmentation
by: Tan, Dayu, et al.
Published: (2025)
by: Tan, Dayu, et al.
Published: (2025)
An Innovative Framework for Breast Cancer Detection Using Pyramid Adaptive Atrous Convolution, Transformer Integration, and Multi-Scale Feature Fusion
by: Pour, Ehsan Sadeghi, et al.
Published: (2026)
by: Pour, Ehsan Sadeghi, et al.
Published: (2026)
Weierstrass Positional Encoding for Vision Transformers
by: Xin, Zhihang, et al.
Published: (2026)
by: Xin, Zhihang, et al.
Published: (2026)
SPFFNet: Strip Perception and Feature Fusion Spatial Pyramid Pooling for Fabric Defect Detection
by: Zhao, Peizhe, et al.
Published: (2025)
by: Zhao, Peizhe, et al.
Published: (2025)
iPEAR: Iterative Pyramid Estimation with Attention and Residuals for Deformable Medical Image Registration
by: Wu, Heming, et al.
Published: (2025)
by: Wu, Heming, et al.
Published: (2025)
Imbalance-Aware Culvert-Sewer Defect Segmentation Using an Enhanced Feature Pyramid Network
by: Alshawi, Rasha, et al.
Published: (2024)
by: Alshawi, Rasha, et al.
Published: (2024)
SHARP-Net: A Refined Pyramid Network for Deficiency Segmentation in Culverts and Sewer Pipes
by: Alshawi, Rasha, et al.
Published: (2024)
by: Alshawi, Rasha, et al.
Published: (2024)
Training Frozen Feature Pyramid DINOv2 for Eyelid Measurements with Infinite Encoding and Orthogonal Regularization
by: Chen, Chun-Hung
Published: (2025)
by: Chen, Chun-Hung
Published: (2025)
UniVid: Pyramid Diffusion Model for High Quality Video Generation
by: Xiao, Xinyu, et al.
Published: (2026)
by: Xiao, Xinyu, et al.
Published: (2026)
PVTAdpNet: Polyp Segmentation using Pyramid vision transformer with a novel Adapter block
by: Nezhad, Arshia Yousefi, et al.
Published: (2025)
by: Nezhad, Arshia Yousefi, et al.
Published: (2025)
PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation
by: Li, Xiaolong, et al.
Published: (2025)
by: Li, Xiaolong, et al.
Published: (2025)
Benchmarking Hierarchical Image Pyramid Transformer for the classification of colon biopsies and polyps in histopathology images
by: Contreras, Nohemi Sofia Leon, et al.
Published: (2024)
by: Contreras, Nohemi Sofia Leon, et al.
Published: (2024)
Pyramid Hierarchical Masked Diffusion Model for Imaging Synthesis
by: Xiao, Xiaojiao, et al.
Published: (2025)
by: Xiao, Xiaojiao, et al.
Published: (2025)
Deformable Image Registration with Multi-scale Feature Fusion from Shared Encoder, Auxiliary and Pyramid Decoders
by: Zhou, Hongchao, et al.
Published: (2024)
by: Zhou, Hongchao, et al.
Published: (2024)
S$^2$-FPN: Scale-ware Strip Attention Guided Feature Pyramid Network for Real-time Semantic Segmentation
by: Elhassan, Mohammed A. M., et al.
Published: (2022)
by: Elhassan, Mohammed A. M., et al.
Published: (2022)
Curve-based Neural Style Transfer
by: Chen, Yu-hsuan, et al.
Published: (2023)
by: Chen, Yu-hsuan, et al.
Published: (2023)
PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents
by: Zhang, Yuqun, et al.
Published: (2025)
by: Zhang, Yuqun, et al.
Published: (2025)
PHPQ: Pyramid Hybrid Pooling Quantization for Efficient Fine-Grained Image Retrieval
by: Zeng, Ziyun, et al.
Published: (2021)
by: Zeng, Ziyun, et al.
Published: (2021)
Face Pyramid Vision Transformer
by: Islam, Khawar, et al.
Published: (2022)
by: Islam, Khawar, et al.
Published: (2022)
SPARO: Selective Attention for Robust and Compositional Transformer Encodings for Vision
by: Vani, Ankit, et al.
Published: (2024)
by: Vani, Ankit, et al.
Published: (2024)
Spatiotemporal Pyramid Flow Matching for Climate Emulation
by: Irvin, Jeremy Andrew, et al.
Published: (2025)
by: Irvin, Jeremy Andrew, et al.
Published: (2025)
A 2D Semantic-Aware Position Encoding for Vision Transformers
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
AesFA: An Aesthetic Feature-Aware Arbitrary Neural Style Transfer
by: Kwon, Joonwoo, et al.
Published: (2023)
by: Kwon, Joonwoo, et al.
Published: (2023)
A3-FPN: Asymptotic Content-Aware Pyramid Attention Network for Dense Visual Prediction
by: Qin, Meng'en, et al.
Published: (2026)
by: Qin, Meng'en, et al.
Published: (2026)
Cameras as Relative Positional Encoding
by: Li, Ruilong, et al.
Published: (2025)
by: Li, Ruilong, et al.
Published: (2025)
RLMiniStyler: Light-weight RL Style Agent for Arbitrary Sequential Neural Style Generation
by: Hu, Jing, et al.
Published: (2025)
by: Hu, Jing, et al.
Published: (2025)
StyleRF-VolVis: Style Transfer of Neural Radiance Fields for Expressive Volume Visualization
by: Tang, Kaiyuan, et al.
Published: (2024)
by: Tang, Kaiyuan, et al.
Published: (2024)
AnoStyler: Text-Driven Localized Anomaly Generation via Lightweight Style Transfer
by: So, Yulim, et al.
Published: (2025)
by: So, Yulim, et al.
Published: (2025)
Similar Items
-
D2Styler: Advancing Arbitrary Style Transfer with Discrete Diffusion Methods
by: Susladkar, Onkar, et al.
Published: (2024) -
Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding
by: Chen, Zhanpeng, et al.
Published: (2025) -
Global Feature Pyramid Network
by: Xiao, Weilin, et al.
Published: (2023) -
A Unified Framework for Microscopy Defocus Deblur with Multi-Pyramid Transformer and Contrastive Learning
by: Zhang, Yuelin, et al.
Published: (2024) -
BatStyler: Advancing Multi-category Style Generation for Source-free Domain Generalization
by: Xu, Xiusheng, et al.
Published: (2025)