Saved in:
| Main Authors: | Li, Xinghan, Yu, Yue, Song, Xue, Shan, Haijun, Chen, Jingjing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.09314 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection
by: Li, Xinghan, et al.
Published: (2026)
by: Li, Xinghan, et al.
Published: (2026)
Biomechanics-Guided Residual Approach to Generalizable Human Motion Generation and Estimation
by: Kang, Zixi, et al.
Published: (2025)
by: Kang, Zixi, et al.
Published: (2025)
The Silent Assistant: NoiseQuery as Implicit Guidance for Goal-Driven Image Generation
by: Wang, Ruoyu, et al.
Published: (2024)
by: Wang, Ruoyu, et al.
Published: (2024)
Doubly Abductive Counterfactual Inference for Text-based Image Editing
by: Song, Xue, et al.
Published: (2024)
by: Song, Xue, et al.
Published: (2024)
Learning an Implicit Physics Model for Image-based Fluid Simulation
by: Jia, Emily Yue-Ting, et al.
Published: (2025)
by: Jia, Emily Yue-Ting, et al.
Published: (2025)
Generating Attribute-Aware Human Motions from Textual Prompt
by: Wang, Xinghan, et al.
Published: (2025)
by: Wang, Xinghan, et al.
Published: (2025)
REL-SF4PASS: Panoramic Semantic Segmentation with REL Depth Representation and Spherical Fusion
by: Li, Xuewei, et al.
Published: (2026)
by: Li, Xuewei, et al.
Published: (2026)
Revealing Temporal Label Noise in Multimodal Hateful Video Classification
by: Yang, Shuonan, et al.
Published: (2025)
by: Yang, Shuonan, et al.
Published: (2025)
LoRA of Change: Learning to Generate LoRA for the Editing Instruction from A Single Before-After Image Pair
by: Song, Xue, et al.
Published: (2024)
by: Song, Xue, et al.
Published: (2024)
Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion
by: Wang, Xinghan, et al.
Published: (2024)
by: Wang, Xinghan, et al.
Published: (2024)
DBINDS -- Can Initial Noise from Diffusion Model Inversion Help Reveal AI-Generated Videos?
by: Wu, Yanlin, et al.
Published: (2025)
by: Wu, Yanlin, et al.
Published: (2025)
Noise Adaption Network for Morse Code Image Classification
by: Wang, Xiaxia, et al.
Published: (2024)
by: Wang, Xiaxia, et al.
Published: (2024)
Consistent Diffusion: Denoising Diffusion Model with Data-Consistent Training for Image Restoration
by: Cheng, Xinlong, et al.
Published: (2024)
by: Cheng, Xinlong, et al.
Published: (2024)
IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning
by: Sun, Zhichao, et al.
Published: (2026)
by: Sun, Zhichao, et al.
Published: (2026)
Generalizing Vision-Language Models with Dedicated Prompt Guidance
by: Li, Xinyao, et al.
Published: (2025)
by: Li, Xinyao, et al.
Published: (2025)
Consistent and Optimal Solution to Camera Motion Estimation
by: Zeng, Guangyang, et al.
Published: (2024)
by: Zeng, Guangyang, et al.
Published: (2024)
From Synthetic to Real: Toward Identity-Consistent Makeup Transfer with Synthetic and Real Data
by: Yu, Yue, et al.
Published: (2026)
by: Yu, Yue, et al.
Published: (2026)
Semantic Generative Tuning for Unified Multimodal Models
by: Yu, Songsong, et al.
Published: (2026)
by: Yu, Songsong, et al.
Published: (2026)
Dive3D: Diverse Distillation-based Text-to-3D Generation via Score Implicit Matching
by: Bai, Weimin, et al.
Published: (2025)
by: Bai, Weimin, et al.
Published: (2025)
Boosting Self-Supervised Tracking with Contextual Prompts and Noise Learning
by: Zheng, Yaozong, et al.
Published: (2026)
by: Zheng, Yaozong, et al.
Published: (2026)
Generalizable Implicit Motion Modeling for Video Frame Interpolation
by: Guo, Zujin, et al.
Published: (2024)
by: Guo, Zujin, et al.
Published: (2024)
Gait Recognition via Collaborating Discriminative and Generative Diffusion Models
by: Xiong, Haijun, et al.
Published: (2025)
by: Xiong, Haijun, et al.
Published: (2025)
Interacted Planes Reveal 3D Line Mapping
by: Ke, Zeran, et al.
Published: (2026)
by: Ke, Zeran, et al.
Published: (2026)
Position: Towards Implicit Prompt For Text-To-Image Models
by: Yang, Yue, et al.
Published: (2024)
by: Yang, Yue, et al.
Published: (2024)
Implicit Concept Removal of Diffusion Models
by: Liu, Zhili, et al.
Published: (2023)
by: Liu, Zhili, et al.
Published: (2023)
GrOCE:Graph-Guided Online Concept Erasure for Text-to-Image Diffusion Models
by: Han, Ning, et al.
Published: (2025)
by: Han, Ning, et al.
Published: (2025)
Visual Implicit Autoregressive Modeling
by: Jiang, Pengfei, et al.
Published: (2026)
by: Jiang, Pengfei, et al.
Published: (2026)
I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
NDM: A Noise-driven Detection and Mitigation Framework against Implicit Sexual Intentions in Text-to-Image Generation
by: Sun, Yitong, et al.
Published: (2025)
by: Sun, Yitong, et al.
Published: (2025)
DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models
by: Zhou, Yue, et al.
Published: (2026)
by: Zhou, Yue, et al.
Published: (2026)
MotionFlow:Learning Implicit Motion Flow for Complex Camera Trajectory Control in Video Generation
by: Lei, Guojun, et al.
Published: (2025)
by: Lei, Guojun, et al.
Published: (2025)
SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing
by: Li, Ruihuang, et al.
Published: (2024)
by: Li, Ruihuang, et al.
Published: (2024)
Q2A: Querying Implicit Fully Continuous Feature Pyramid to Align Features for Medical Image Segmentation
by: Yu, Jiahao, et al.
Published: (2024)
by: Yu, Jiahao, et al.
Published: (2024)
OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks
by: Wang, Jiayu, et al.
Published: (2025)
by: Wang, Jiayu, et al.
Published: (2025)
NegVSR: Augmenting Negatives for Generalized Noise Modeling in Real-World Video Super-Resolution
by: Song, Yexing, et al.
Published: (2023)
by: Song, Yexing, et al.
Published: (2023)
Noise Fusion-based Distillation Learning for Anomaly Detection in Complex Industrial Environments
by: Yu, Jiawen, et al.
Published: (2025)
by: Yu, Jiawen, et al.
Published: (2025)
GIFS: Neural Implicit Function for General Shape Representation
by: Ye, Jianglong, et al.
Published: (2022)
by: Ye, Jianglong, et al.
Published: (2022)
CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation
by: Huang, Kaiyi, et al.
Published: (2026)
by: Huang, Kaiyi, et al.
Published: (2026)
Boosting of Implicit Neural Representation-based Image Denoiser
by: Yan, Zipei, et al.
Published: (2024)
by: Yan, Zipei, et al.
Published: (2024)
Disentangling Hardness from Noise: An Uncertainty-Driven Model-Agnostic Framework for Long-Tailed Remote Sensing Classification
by: Ding, Chi, et al.
Published: (2026)
by: Ding, Chi, et al.
Published: (2026)
Similar Items
-
VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection
by: Li, Xinghan, et al.
Published: (2026) -
Biomechanics-Guided Residual Approach to Generalizable Human Motion Generation and Estimation
by: Kang, Zixi, et al.
Published: (2025) -
The Silent Assistant: NoiseQuery as Implicit Guidance for Goal-Driven Image Generation
by: Wang, Ruoyu, et al.
Published: (2024) -
Doubly Abductive Counterfactual Inference for Text-based Image Editing
by: Song, Xue, et al.
Published: (2024) -
Learning an Implicit Physics Model for Image-based Fluid Simulation
by: Jia, Emily Yue-Ting, et al.
Published: (2025)