Salvato in:
| Autori principali: | Jia, Zexi, Luo, Pengcheng, Zhong, Yijia, Zhang, Jinchao, Zhou, Jie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2603.08064 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Manifold-Optimal Guidance: A Unified Riemannian Control View of Diffusion Guidance
di: Jia, Zexi, et al.
Pubblicazione: (2026)
di: Jia, Zexi, et al.
Pubblicazione: (2026)
Too Vivid to Be Real? Benchmarking and Calibrating Generative Color Fidelity
di: Fang, Zhengyao, et al.
Pubblicazione: (2026)
di: Fang, Zhengyao, et al.
Pubblicazione: (2026)
StyleDecoupler: Generalizable Artistic Style Disentanglement
di: Jia, Zexi, et al.
Pubblicazione: (2026)
di: Jia, Zexi, et al.
Pubblicazione: (2026)
CoDA: Color Distribution Probing for Efficient and Generalizable AI-Generated Image Detection
di: Jia, Zexi, et al.
Pubblicazione: (2026)
di: Jia, Zexi, et al.
Pubblicazione: (2026)
Exploring Specular Reflection Inconsistency for Generalizable Face Forgery Detection
di: Fei, Hongyan, et al.
Pubblicazione: (2026)
di: Fei, Hongyan, et al.
Pubblicazione: (2026)
RDTF: Resource-efficient Dual-mask Training Framework for Multi-frame Animated Sticker Generation
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025)
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
di: Jia, Zexi, et al.
Pubblicazione: (2025)
di: Jia, Zexi, et al.
Pubblicazione: (2025)
Semantic to Structure: Learning Structural Representations for Infringement Detection
di: Huang, Chuanwei, et al.
Pubblicazione: (2025)
di: Huang, Chuanwei, et al.
Pubblicazione: (2025)
A Visual Leap in CLIP Compositionality Reasoning through Generation of Counterfactual Sets
di: Jia, Zexi, et al.
Pubblicazione: (2025)
di: Jia, Zexi, et al.
Pubblicazione: (2025)
WalkVLM:Aid Visually Impaired People Walking by Vision Language Model
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
F2RVLM: Boosting Fine-grained Fragment Retrieval for Multi-Modal Long-form Dialogue with Vision Language Model
di: Bi, Hanbo, et al.
Pubblicazione: (2025)
di: Bi, Hanbo, et al.
Pubblicazione: (2025)
From Imitation to Innovation: The Emergence of AI Unique Artistic Styles and the Challenge of Copyright Protection
di: Jia, Zexi, et al.
Pubblicazione: (2025)
di: Jia, Zexi, et al.
Pubblicazione: (2025)
Mobius: A High Efficient Spatial-Temporal Parallel Training Paradigm for Text-to-Video Generation Task
di: Yang, Yiran, et al.
Pubblicazione: (2024)
di: Yang, Yiran, et al.
Pubblicazione: (2024)
A Synthetic-to-Real Dehazing Method based on Domain Unification
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025)
Generative Video Compression with One-Dimensional Latent Representation
di: Zheng, Zihan, et al.
Pubblicazione: (2026)
di: Zheng, Zihan, et al.
Pubblicazione: (2026)
Real2Code: Reconstruct Articulated Objects via Code Generation
di: Mandi, Zhao, et al.
Pubblicazione: (2024)
di: Mandi, Zhao, et al.
Pubblicazione: (2024)
Scalable Geometric Fracture Assembly via Co-creation Space among Assemblers
di: Zhang, Ruiyuan, et al.
Pubblicazione: (2023)
di: Zhang, Ruiyuan, et al.
Pubblicazione: (2023)
Video2LoRA: Unified Semantic-Controlled Video Generation via Per-Reference-Video LoRA
di: Wu, Zexi, et al.
Pubblicazione: (2026)
di: Wu, Zexi, et al.
Pubblicazione: (2026)
Flying Bird Object Detection Algorithm in Surveillance Video Based on Motion Information
di: Sun, Ziwei, et al.
Pubblicazione: (2023)
di: Sun, Ziwei, et al.
Pubblicazione: (2023)
Semantic One-Dimensional Tokenizer for Image Reconstruction and Generation
di: Qu, Yunpeng, et al.
Pubblicazione: (2026)
di: Qu, Yunpeng, et al.
Pubblicazione: (2026)
Generative Semantic Coding for Ultra-Low Bitrate Visual Communication and Analysis
di: Chen, Weiming, et al.
Pubblicazione: (2025)
di: Chen, Weiming, et al.
Pubblicazione: (2025)
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models
di: Hu, Nanxing, et al.
Pubblicazione: (2025)
di: Hu, Nanxing, et al.
Pubblicazione: (2025)
HIR-ALIGN: Enhancing Hyperspectral Image Restoration via Diffusion-Based Data Generation
di: Pang, Li, et al.
Pubblicazione: (2026)
di: Pang, Li, et al.
Pubblicazione: (2026)
Improving the Generalization of Segmentation Foundation Model under Distribution Shift via Weakly Supervised Adaptation
di: Zhang, Haojie, et al.
Pubblicazione: (2023)
di: Zhang, Haojie, et al.
Pubblicazione: (2023)
Why Settle for One? Text-to-ImageSet Generation and Evaluation
di: Jia, Chengyou, et al.
Pubblicazione: (2025)
di: Jia, Chengyou, et al.
Pubblicazione: (2025)
Unleashing Video Language Models for Fine-grained HRCT Report Generation
di: Fang, Yingying, et al.
Pubblicazione: (2026)
di: Fang, Yingying, et al.
Pubblicazione: (2026)
ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
di: Zhang, Ting, et al.
Pubblicazione: (2024)
di: Zhang, Ting, et al.
Pubblicazione: (2024)
Understanding the Implicit User Intention via Reasoning with Large Language Model for Image Editing
di: Wang, Yijia, et al.
Pubblicazione: (2025)
di: Wang, Yijia, et al.
Pubblicazione: (2025)
MemGS: Memory-Efficient Gaussian Splatting for Real-Time SLAM
di: Bai, Yinlong, et al.
Pubblicazione: (2025)
di: Bai, Yinlong, et al.
Pubblicazione: (2025)
Low-Dimensional Gradient Helps Out-of-Distribution Detection
di: Wu, Yingwen, et al.
Pubblicazione: (2023)
di: Wu, Yingwen, et al.
Pubblicazione: (2023)
Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens
di: Kim, Dongwon, et al.
Pubblicazione: (2025)
di: Kim, Dongwon, et al.
Pubblicazione: (2025)
Diffusion Models Learn Low-Dimensional Distributions via Subspace Clustering
di: Wang, Peng, et al.
Pubblicazione: (2024)
di: Wang, Peng, et al.
Pubblicazione: (2024)
Perceptual Video Coding for Machines via Satisfied Machine Ratio Modeling
di: Zhang, Qi, et al.
Pubblicazione: (2022)
di: Zhang, Qi, et al.
Pubblicazione: (2022)
Towards One-step Causal Video Generation via Adversarial Self-Distillation
di: Yang, Yongqi, et al.
Pubblicazione: (2025)
di: Yang, Yongqi, et al.
Pubblicazione: (2025)
Multi-concept Model Immunization through Differentiable Model Merging
di: Zheng, Amber Yijia, et al.
Pubblicazione: (2024)
di: Zheng, Amber Yijia, et al.
Pubblicazione: (2024)
OneViewAll: Semantic Prior Guided One-View 6D Pose Estimation for Novel Objects
di: Luo, Yang, et al.
Pubblicazione: (2026)
di: Luo, Yang, et al.
Pubblicazione: (2026)
Boosting Semi-Supervised Medical Image Segmentation via Masked Image Consistency and Discrepancy Learning
di: Zhou, Pengcheng, et al.
Pubblicazione: (2025)
di: Zhou, Pengcheng, et al.
Pubblicazione: (2025)
CoDoL: Conditional Domain Prompt Learning for Out-of-Distribution Generalization
di: Zhang, Min, et al.
Pubblicazione: (2025)
di: Zhang, Min, et al.
Pubblicazione: (2025)
HGP-Mamba: Integrating Histology and Generated Protein Features for Mamba-based Multimodal Survival Risk Prediction
di: Dai, Jing, et al.
Pubblicazione: (2026)
di: Dai, Jing, et al.
Pubblicazione: (2026)
Collaborative Low-Rank Adaptation for Pre-Trained Vision Transformers
di: Liu, Zheng, et al.
Pubblicazione: (2025)
di: Liu, Zheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Manifold-Optimal Guidance: A Unified Riemannian Control View of Diffusion Guidance
di: Jia, Zexi, et al.
Pubblicazione: (2026) -
Too Vivid to Be Real? Benchmarking and Calibrating Generative Color Fidelity
di: Fang, Zhengyao, et al.
Pubblicazione: (2026) -
StyleDecoupler: Generalizable Artistic Style Disentanglement
di: Jia, Zexi, et al.
Pubblicazione: (2026) -
CoDA: Color Distribution Probing for Efficient and Generalizable AI-Generated Image Detection
di: Jia, Zexi, et al.
Pubblicazione: (2026) -
Exploring Specular Reflection Inconsistency for Generalizable Face Forgery Detection
di: Fei, Hongyan, et al.
Pubblicazione: (2026)