WIPES: Wavelet-based Visual Primitives
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Wenhao, Zhu, Hao, Wu, Delong, Kang, Di, Bao, Linchao, Cao, Xun, Ma, Zhan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural Poisson Solver: A Universal and Continuous Framework for Natural Signal Blending
by: Wu, Delong, et al.
Published: (2024)
by: Wu, Delong, et al.
Published: (2024)
Neural Point-based Volumetric Avatar: Surface-guided Neural Points for Efficient and Photorealistic Volumetric Head Avatar
by: Wang, Cong, et al.
Published: (2023)
by: Wang, Cong, et al.
Published: (2023)
MC-Bench: A Benchmark for Multi-Context Visual Grounding in the Era of MLLMs
by: Xu, Yunqiu, et al.
Published: (2024)
by: Xu, Yunqiu, et al.
Published: (2024)
DAGSM: Disentangled Avatar Generation with GS-enhanced Mesh
by: Zhuang, Jingyu, et al.
Published: (2024)
by: Zhuang, Jingyu, et al.
Published: (2024)
Mitigating Ambiguities in 3D Classification with Gaussian Splatting
by: Zhang, Ruiqi, et al.
Published: (2025)
by: Zhang, Ruiqi, et al.
Published: (2025)
PIFu for the Real World: A Self-supervised Framework to Reconstruct Dressed Human from Single-view Images
by: Xiong, Zhangyang, et al.
Published: (2022)
by: Xiong, Zhangyang, et al.
Published: (2022)
MTC-VAE: Multi-Level Temporal Compression with Content Awareness
by: Dong, Yubo, et al.
Published: (2026)
by: Dong, Yubo, et al.
Published: (2026)
GPD: Guided Progressive Distillation for Fast and High-Quality Video Generation
by: Liang, Xiao, et al.
Published: (2026)
by: Liang, Xiao, et al.
Published: (2026)
Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval
by: Suo, Yucheng, et al.
Published: (2024)
by: Suo, Yucheng, et al.
Published: (2024)
3DID: Direct 3D Inverse Design for Aerodynamics with Physics-Aware Optimization
by: Hao, Yuze, et al.
Published: (2025)
by: Hao, Yuze, et al.
Published: (2025)
FINER++: Building a Family of Variable-periodic Functions for Activating Implicit Neural Representation
by: Zhu, Hao, et al.
Published: (2024)
by: Zhu, Hao, et al.
Published: (2024)
Point2Primitive: CAD Reconstruction from Point Cloud by Direct Primitive Prediction
by: Ma, Xinzhu, et al.
Published: (2025)
by: Ma, Xinzhu, et al.
Published: (2025)
H3R: Hybrid Multi-view Correspondence for Generalizable 3D Reconstruction
by: Jia, Heng, et al.
Published: (2025)
by: Jia, Heng, et al.
Published: (2025)
MeGA: Hybrid Mesh-Gaussian Head Avatar for High-Fidelity Rendering and Head Editing
by: Wang, Cong, et al.
Published: (2024)
by: Wang, Cong, et al.
Published: (2024)
Collaborative Group: Composed Image Retrieval via Consensus Learning from Noisy Annotations
by: Zhang, Xu, et al.
Published: (2023)
by: Zhang, Xu, et al.
Published: (2023)
Oscillating Dispersion for Maximal Light-throughput Spectral Imaging
by: Zhang, Jiuyun, et al.
Published: (2026)
by: Zhang, Jiuyun, et al.
Published: (2026)
WMamba: Wavelet-based Mamba for Face Forgery Detection
by: Peng, Siran, et al.
Published: (2025)
by: Peng, Siran, et al.
Published: (2025)
Gaussian Primitives for Deformable Image Registration
by: Li, Jihe, et al.
Published: (2024)
by: Li, Jihe, et al.
Published: (2024)
From Trial to Triumph: Advancing Long Video Understanding via Visual Context Sample Scaling and Self-reward Alignment
by: Suo, Yucheng, et al.
Published: (2025)
by: Suo, Yucheng, et al.
Published: (2025)
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
by: Yang, Xiangpeng, et al.
Published: (2025)
by: Yang, Xiangpeng, et al.
Published: (2025)
EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing
by: Yang, Xiangpeng, et al.
Published: (2024)
by: Yang, Xiangpeng, et al.
Published: (2024)
Combating Label Noise With A General Surrogate Model For Sample Selection
by: Liang, Chao, et al.
Published: (2023)
by: Liang, Chao, et al.
Published: (2023)
Slimmable Networks for Contrastive Self-supervised Learning
by: Zhao, Shuai, et al.
Published: (2022)
by: Zhao, Shuai, et al.
Published: (2022)
DGL: Dynamic Global-Local Prompt Tuning for Text-Video Retrieval
by: Yang, Xiangpeng, et al.
Published: (2024)
by: Yang, Xiangpeng, et al.
Published: (2024)
CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model
by: Zhao, Shuai, et al.
Published: (2023)
by: Zhao, Shuai, et al.
Published: (2023)
FreeLong: Training-Free Long Video Generation with SpectralBlend Temporal Attention
by: Lu, Yu, et al.
Published: (2024)
by: Lu, Yu, et al.
Published: (2024)
RW-Net: Enhancing Few-Shot Point Cloud Classification with a Wavelet Transform Projection-based Network
by: Zhang, Haosheng, et al.
Published: (2025)
by: Zhang, Haosheng, et al.
Published: (2025)
Event Transformer
by: Jiang, Bin, et al.
Published: (2022)
by: Jiang, Bin, et al.
Published: (2022)
Efficient Visual Representation Learning with Heat Conduction Equation
by: Zhang, Zhemin, et al.
Published: (2024)
by: Zhang, Zhemin, et al.
Published: (2024)
Wavelet based inpainting detection
by: Adrian-Alin, Barglazan, et al.
Published: (2024)
by: Adrian-Alin, Barglazan, et al.
Published: (2024)
TEXTRIX: Latent Attribute Grid for Native Texture Generation and Beyond
by: Zeng, Yifei, et al.
Published: (2025)
by: Zeng, Yifei, et al.
Published: (2025)
Domain Generalization through Spatial Relation Induction over Visual Primitives
by: Nguyen, Dat, et al.
Published: (2026)
by: Nguyen, Dat, et al.
Published: (2026)
PACE: Post-Causal Entropy Modeling for Learned LiDAR Point Cloud Compression
by: Zhu, Jiahao, et al.
Published: (2026)
by: Zhu, Jiahao, et al.
Published: (2026)
SuperPrimitive: Scene Reconstruction at a Primitive Level
by: Mazur, Kirill, et al.
Published: (2023)
by: Mazur, Kirill, et al.
Published: (2023)
Unveiling the Power of Wavelets: A Wavelet-based Kolmogorov-Arnold Network for Hyperspectral Image Classification
by: Seydi, Seyd Teymoor, et al.
Published: (2024)
by: Seydi, Seyd Teymoor, et al.
Published: (2024)
Unified Primitive Proxies for Structured Shape Completion
by: Chen, Zhaiyu, et al.
Published: (2026)
by: Chen, Zhaiyu, et al.
Published: (2026)
Text-DiFuse: An Interactive Multi-Modal Image Fusion Framework based on Text-modulated Diffusion Model
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
WaveletFormerNet: A Transformer-based Wavelet Network for Real-world Non-homogeneous and Dense Fog Removal
by: Zhang, Shengli, et al.
Published: (2024)
by: Zhang, Shengli, et al.
Published: (2024)
Micro-macro Wavelet-based Gaussian Splatting for 3D Reconstruction from Unconstrained Images
by: Li, Yihui, et al.
Published: (2025)
by: Li, Yihui, et al.
Published: (2025)
Test-Time Adaptation with CLIP Reward for Zero-Shot Generalization in Vision-Language Models
by: Zhao, Shuai, et al.
Published: (2023)
by: Zhao, Shuai, et al.
Published: (2023)
Similar Items
-
Neural Poisson Solver: A Universal and Continuous Framework for Natural Signal Blending
by: Wu, Delong, et al.
Published: (2024) -
Neural Point-based Volumetric Avatar: Surface-guided Neural Points for Efficient and Photorealistic Volumetric Head Avatar
by: Wang, Cong, et al.
Published: (2023) -
MC-Bench: A Benchmark for Multi-Context Visual Grounding in the Era of MLLMs
by: Xu, Yunqiu, et al.
Published: (2024) -
DAGSM: Disentangled Avatar Generation with GS-enhanced Mesh
by: Zhuang, Jingyu, et al.
Published: (2024) -
Mitigating Ambiguities in 3D Classification with Gaussian Splatting
by: Zhang, Ruiqi, et al.
Published: (2025)