Neural Network Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Kai, Tang, Dongwen, Zeng, Boya, Yin, Yida, Xu, Zhaopan, Zhou, Yukun, Zang, Zelin, Darrell, Trevor, Liu, Zhuang, You, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding Bias in Large-Scale Visual Datasets
by: Zeng, Boya, et al.
Published: (2024)
by: Zeng, Boya, et al.
Published: (2024)
Generative Modeling of Weights: Generalization or Memorization?
by: Zeng, Boya, et al.
Published: (2025)
by: Zeng, Boya, et al.
Published: (2025)
LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models
by: Lian, Long, et al.
Published: (2023)
by: Lian, Long, et al.
Published: (2023)
Segment Anything without Supervision
by: Wang, XuDong, et al.
Published: (2024)
by: Wang, XuDong, et al.
Published: (2024)
Memorization in 3D Shape Generation: An Empirical Study
by: Pu, Shu, et al.
Published: (2025)
by: Pu, Shu, et al.
Published: (2025)
Readout Guidance: Learning Control from Diffusion Features
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
Can World Simulators Reason? Gen-ViRe: A Generative Visual Reasoning Benchmark
by: Liu, Xinxin, et al.
Published: (2025)
by: Liu, Xinxin, et al.
Published: (2025)
Diffusion Hyperfeatures: Searching Through Time and Space for Semantic Correspondence
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
DAM: Dual Active Learning with Multimodal Foundation Model for Source-Free Domain Adaptation
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
RAPID^3: Tri-Level Reinforced Acceleration Policies for Diffusion Transformer
by: Zhao, Wangbo, et al.
Published: (2025)
by: Zhao, Wangbo, et al.
Published: (2025)
Vector Quantized Feature Fields for Fast 3D Semantic Lifting
by: Tang, George, et al.
Published: (2025)
by: Tang, George, et al.
Published: (2025)
Fast Image-based Neural Relighting with Translucency-Reflection Modeling
by: Zhu, Shizhan, et al.
Published: (2023)
by: Zhu, Shizhan, et al.
Published: (2023)
DiffAug: Enhance Unsupervised Contrastive Learning with Domain-Knowledge-Free Diffusion-based Data Augmentation
by: Zang, Zelin, et al.
Published: (2023)
by: Zang, Zelin, et al.
Published: (2023)
PEBench: A Fictitious Dataset to Benchmark Machine Unlearning for Multimodal Large Language Models
by: Xu, Zhaopan, et al.
Published: (2025)
by: Xu, Zhaopan, et al.
Published: (2025)
SCEESR: Semantic-Control Edge Enhancement for Diffusion-Based Super-Resolution
by: Zhuang, Yun Kai
Published: (2025)
by: Zhuang, Yun Kai
Published: (2025)
UEval: A Benchmark for Unified Multimodal Generation
by: Li, Bo, et al.
Published: (2026)
by: Li, Bo, et al.
Published: (2026)
Multi-source Domain Adaptation for Panoramic Semantic Segmentation
by: Jiang, Jing, et al.
Published: (2024)
by: Jiang, Jing, et al.
Published: (2024)
LLM-grounded Video Diffusion Models
by: Lian, Long, et al.
Published: (2023)
by: Lian, Long, et al.
Published: (2023)
UnSAMv2: Self-Supervised Learning Enables Segment Anything at Any Granularity
by: Yu, Junwei, et al.
Published: (2025)
by: Yu, Junwei, et al.
Published: (2025)
ORAL: Prompting Your Large-Scale LoRAs via Conditional Recurrent Diffusion
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models
by: Harrington, Anne, et al.
Published: (2025)
by: Harrington, Anne, et al.
Published: (2025)
Hierarchical Augmentation and Distillation for Class Incremental Audio-Visual Video Recognition
by: Zuo, Yukun, et al.
Published: (2024)
by: Zuo, Yukun, et al.
Published: (2024)
DiffFNO: Diffusion Fourier Neural Operator
by: Liu, Xiaoyi, et al.
Published: (2024)
by: Liu, Xiaoyi, et al.
Published: (2024)
The Comparison of Individual Cat Recognition Using Neural Networks
by: Li, Mingxuan, et al.
Published: (2024)
by: Li, Mingxuan, et al.
Published: (2024)
Human-Aligned Bench: Fine-Grained Assessment of Reasoning Ability in MLLMs vs. Humans
by: Qiu, Yansheng, et al.
Published: (2025)
by: Qiu, Yansheng, et al.
Published: (2025)
Finding Visual Task Vectors
by: Hojel, Alberto, et al.
Published: (2024)
by: Hojel, Alberto, et al.
Published: (2024)
When Do We Not Need Larger Vision Models?
by: Shi, Baifeng, et al.
Published: (2024)
by: Shi, Baifeng, et al.
Published: (2024)
VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images
by: Zhou, Guanyu, et al.
Published: (2026)
by: Zhou, Guanyu, et al.
Published: (2026)
Dynamic Diffusion Transformer
by: Zhao, Wangbo, et al.
Published: (2024)
by: Zhao, Wangbo, et al.
Published: (2024)
InstanceDiffusion: Instance-level Control for Image Generation
by: Wang, Xudong, et al.
Published: (2024)
by: Wang, Xudong, et al.
Published: (2024)
Part-aware Shape Generation with Latent 3D Diffusion of Neural Voxel Fields
by: Huang, Yuhang, et al.
Published: (2024)
by: Huang, Yuhang, et al.
Published: (2024)
Vision-Language Models Create Cross-Modal Task Representations
by: Luo, Grace, et al.
Published: (2024)
by: Luo, Grace, et al.
Published: (2024)
Revisiting Cross-Attention Mechanisms: Leveraging Beneficial Noise for Domain-Adaptive Learning
by: Zang, Zelin, et al.
Published: (2026)
by: Zang, Zelin, et al.
Published: (2026)
Visual Lexicon: Rich Image Features in Language Space
by: Wang, XuDong, et al.
Published: (2024)
by: Wang, XuDong, et al.
Published: (2024)
Diff-PCC: Diffusion-based Neural Compression for 3D Point Clouds
by: Liu, Kai, et al.
Published: (2024)
by: Liu, Kai, et al.
Published: (2024)
Hierarchical Prompts for Rehearsal-free Continual Learning
by: Zuo, Yukun, et al.
Published: (2024)
by: Zuo, Yukun, et al.
Published: (2024)
Spiking Neural Networks Need High Frequency Information
by: Fang, Yuetong, et al.
Published: (2025)
by: Fang, Yuetong, et al.
Published: (2025)
Bridge then Begin Anew: Generating Target-relevant Intermediate Model for Source-free Visual Emotion Adaptation
by: Zhu, Jiankun, et al.
Published: (2024)
by: Zhu, Jiankun, et al.
Published: (2024)
Reconstruction Alignment Improves Unified Multimodal Models
by: Xie, Ji, et al.
Published: (2025)
by: Xie, Ji, et al.
Published: (2025)
MPBench: A Comprehensive Multimodal Reasoning Benchmark for Process Errors Identification
by: Xu, Zhaopan, et al.
Published: (2025)
by: Xu, Zhaopan, et al.
Published: (2025)
Similar Items
-
Understanding Bias in Large-Scale Visual Datasets
by: Zeng, Boya, et al.
Published: (2024) -
Generative Modeling of Weights: Generalization or Memorization?
by: Zeng, Boya, et al.
Published: (2025) -
LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models
by: Lian, Long, et al.
Published: (2023) -
Segment Anything without Supervision
by: Wang, XuDong, et al.
Published: (2024) -
Memorization in 3D Shape Generation: An Empirical Study
by: Pu, Shu, et al.
Published: (2025)