PixelGen: Improving Pixel Diffusion with Perceptual Supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Zehong, Xu, Ruihan, Zhang, Shiliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PixelGen: Rethinking Embedded Camera Systems
von: Li, Kunjun, et al.
Veröffentlicht: (2024)
von: Li, Kunjun, et al.
Veröffentlicht: (2024)
DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025)
von: Liu, Ye, et al.
Veröffentlicht: (2025)
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
von: Jiang, Liyao, et al.
Veröffentlicht: (2024)
von: Jiang, Liyao, et al.
Veröffentlicht: (2024)
PixelArena: A benchmark for Pixel-Precision Visual Intelligence
von: Liang, Feng, et al.
Veröffentlicht: (2025)
von: Liang, Feng, et al.
Veröffentlicht: (2025)
HybridStitch: Pixel and Timestep Level Model Stitching for Diffusion Acceleration
von: Sun, Desen, et al.
Veröffentlicht: (2026)
von: Sun, Desen, et al.
Veröffentlicht: (2026)
The Thinking Pixel: Recursive Sparse Reasoning in Multimodal Diffusion Latents
von: Sun, Yuwei, et al.
Veröffentlicht: (2026)
von: Sun, Yuwei, et al.
Veröffentlicht: (2026)
PixelWeb: The First Web GUI Dataset with Pixel-Wise Labels
von: Yang, Qi, et al.
Veröffentlicht: (2025)
von: Yang, Qi, et al.
Veröffentlicht: (2025)
Pixel-Space Post-Training of Latent Diffusion Models
von: Zhang, Christina, et al.
Veröffentlicht: (2024)
von: Zhang, Christina, et al.
Veröffentlicht: (2024)
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy Labels
von: Mu, Chenyu, et al.
Veröffentlicht: (2025)
von: Mu, Chenyu, et al.
Veröffentlicht: (2025)
PixelSmile: Toward Fine-Grained Facial Expression Editing
von: Hua, Jiabin, et al.
Veröffentlicht: (2026)
von: Hua, Jiabin, et al.
Veröffentlicht: (2026)
Beyond Pixels: Vector-to-Graph Transformation for Reliable Schematic Auditing
von: Ma, Chengwei, et al.
Veröffentlicht: (2026)
von: Ma, Chengwei, et al.
Veröffentlicht: (2026)
Single-Frame Point-Pixel Registration via Supervised Cross-Modal Feature Matching
von: Han, Yu, et al.
Veröffentlicht: (2025)
von: Han, Yu, et al.
Veröffentlicht: (2025)
L2P: Unlocking Latent Potential for Pixel Generation
von: Chen, Zhennan, et al.
Veröffentlicht: (2026)
von: Chen, Zhennan, et al.
Veröffentlicht: (2026)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
Cycle Pixel Difference Network for Crisp Edge Detection
von: Liu, Changsong, et al.
Veröffentlicht: (2024)
von: Liu, Changsong, et al.
Veröffentlicht: (2024)
FlattenGPT: Depth Compression for Transformer with Layer Flattening
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding
von: Wang, Nan, et al.
Veröffentlicht: (2026)
von: Wang, Nan, et al.
Veröffentlicht: (2026)
RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations
von: Ge, Yanhao, et al.
Veröffentlicht: (2026)
von: Ge, Yanhao, et al.
Veröffentlicht: (2026)
Pixel is a Barrier: Diffusion Models Are More Adversarially Robust Than We Think
von: Xue, Haotian, et al.
Veröffentlicht: (2024)
von: Xue, Haotian, et al.
Veröffentlicht: (2024)
ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models
von: Zhou, Qin, et al.
Veröffentlicht: (2025)
von: Zhou, Qin, et al.
Veröffentlicht: (2025)
Tracing Copied Pixels and Regularizing Patch Affinity in Copy Detection
von: Lu, Yichen, et al.
Veröffentlicht: (2026)
von: Lu, Yichen, et al.
Veröffentlicht: (2026)
From Pampas to Pixels: Fine-Tuning Diffusion Models for Gaúcho Heritage
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024)
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024)
Pixelis: Reasoning in Pixels, from Seeing to Acting
von: Zhou, Yunpeng
Veröffentlicht: (2026)
von: Zhou, Yunpeng
Veröffentlicht: (2026)
From Pixels to Predicates Structuring urban perception with scene graphs
von: Liu, Yunlong, et al.
Veröffentlicht: (2025)
von: Liu, Yunlong, et al.
Veröffentlicht: (2025)
Beyond Pixels: Visual Metaphor Transfer via Schema-Driven Agentic Reasoning
von: Xu, Yu, et al.
Veröffentlicht: (2026)
von: Xu, Yu, et al.
Veröffentlicht: (2026)
MedReasoner: Reinforcement Learning Drives Reasoning Grounding from Clinical Thought to Pixel-Level Precision
von: Yan, Zhonghao, et al.
Veröffentlicht: (2025)
von: Yan, Zhonghao, et al.
Veröffentlicht: (2025)
Pixel-Grounded Retrieval for Knowledgeable Large Multimodal Models
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2026)
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2026)
PEAR: Pixel-aligned Expressive humAn mesh Recovery
von: Wu, Jiahao, et al.
Veröffentlicht: (2026)
von: Wu, Jiahao, et al.
Veröffentlicht: (2026)
GLaMM: Pixel Grounding Large Multimodal Model
von: Rasheed, Hanoona, et al.
Veröffentlicht: (2023)
von: Rasheed, Hanoona, et al.
Veröffentlicht: (2023)
PixelBytes: Catching Unified Representation for Multimodal Generation
von: Furfaro, Fabien
Veröffentlicht: (2024)
von: Furfaro, Fabien
Veröffentlicht: (2024)
Optimal Video Compression using Pixel Shift Tracking
von: Panneerselvam, Hitesh Saai Mananchery, et al.
Veröffentlicht: (2024)
von: Panneerselvam, Hitesh Saai Mananchery, et al.
Veröffentlicht: (2024)
PixelBytes: Catching Unified Embedding for Multimodal Generation
von: Furfaro, Fabien
Veröffentlicht: (2024)
von: Furfaro, Fabien
Veröffentlicht: (2024)
Pixels, Patterns, but No Poetry: To See The World like Humans
von: Gao, Hongcheng, et al.
Veröffentlicht: (2025)
von: Gao, Hongcheng, et al.
Veröffentlicht: (2025)
Exploring the Limits of Semantic Image Compression at Micro-bits per Pixel
von: Dotzel, Jordan, et al.
Veröffentlicht: (2024)
von: Dotzel, Jordan, et al.
Veröffentlicht: (2024)
Lattice Boltzmann Model for Learning Real-World Pixel Dynamicity
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
Imbalanced Medical Image Segmentation with Pixel-dependent Noisy Labels
von: Guo, Erjian, et al.
Veröffentlicht: (2025)
von: Guo, Erjian, et al.
Veröffentlicht: (2025)
PAT: Pixel-wise Adaptive Training for Long-tailed Segmentation
von: Do, Khoi, et al.
Veröffentlicht: (2024)
von: Do, Khoi, et al.
Veröffentlicht: (2024)
Pixel-Aligned Multi-View Generation with Depth Guided Decoder
von: Tang, Zhenggang, et al.
Veröffentlicht: (2024)
von: Tang, Zhenggang, et al.
Veröffentlicht: (2024)
Towards Efficient Pixel Labeling for Industrial Anomaly Detection and Localization
von: Wu, Jingqi, et al.
Veröffentlicht: (2025)
von: Wu, Jingqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PixelGen: Rethinking Embedded Camera Systems
von: Li, Kunjun, et al.
Veröffentlicht: (2024) -
DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation
von: Ma, Zehong, et al.
Veröffentlicht: (2025) -
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025) -
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
von: Jiang, Liyao, et al.
Veröffentlicht: (2024) -
PixelArena: A benchmark for Pixel-Precision Visual Intelligence
von: Liang, Feng, et al.
Veröffentlicht: (2025)