Saved in:
| Main Authors: | Zhou, Hang, Cai, Jiale, Ye, Yuteng, Feng, Yonghui, Gao, Chenxing, Yu, Junqing, Song, Zikai, Yang, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.09026 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Progressive Text-to-Image Diffusion with Soft Latent Direction
by: Ye, YuTeng, et al.
Published: (2023)
by: Ye, YuTeng, et al.
Published: (2023)
CA-Diff: Collaborative Anatomy Diffusion for Brain Tissue Segmentation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
Attacking Transformers with Feature Diversity Adversarial Perturbation
by: Gao, Chenxing, et al.
Published: (2024)
by: Gao, Chenxing, et al.
Published: (2024)
Coupled Mamba: Enhanced Multi-modal Fusion with Coupled State Space Model
by: Li, Wenbing, et al.
Published: (2024)
by: Li, Wenbing, et al.
Published: (2024)
CurEvo: Curriculum-Guided Self-Evolution for Video Understanding
by: Zeng, Guiyi, et al.
Published: (2026)
by: Zeng, Guiyi, et al.
Published: (2026)
Optimized View and Geometry Distillation from Multi-view Diffuser
by: Zhang, Youjia, et al.
Published: (2023)
by: Zhang, Youjia, et al.
Published: (2023)
LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing
by: Li, Wenbing, et al.
Published: (2025)
by: Li, Wenbing, et al.
Published: (2025)
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
by: Zhan, Yu-Wei, et al.
Published: (2025)
by: Zhan, Yu-Wei, et al.
Published: (2025)
HumanSAM: Classifying Human-centric Forgery Videos in Human Spatial, Appearance, and Motion Anomaly
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
MVP: Winning Solution to SMP Challenge 2025 Video Track
by: Ye, Liliang, et al.
Published: (2025)
by: Ye, Liliang, et al.
Published: (2025)
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
Correction to Parallel Symmetric Appearance‐Motion Framework with Diffusion and Refinement Blocks for Video Anomaly Detection System
Published: (2025)
Published: (2025)
SF2T: Self-supervised Fragment Finetuning of Video-LLMs for Fine-Grained Understanding
by: Hu, Yangliu, et al.
Published: (2025)
by: Hu, Yangliu, et al.
Published: (2025)
GA-S$^3$: Comprehensive Social Network Simulation with Group Agents
by: Zhang, Yunyao, et al.
Published: (2025)
by: Zhang, Yunyao, et al.
Published: (2025)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
by: Zhai, Yuanhao, et al.
Published: (2024)
by: Zhai, Yuanhao, et al.
Published: (2024)
Motion Semantics Guided Normalizing Flow for Privacy-Preserving Video Anomaly Detection
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Cross-Modality Masked Learning for Survival Prediction in ICI Treated NSCLC Patients
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
by: Liu, Huijie, et al.
Published: (2025)
by: Liu, Huijie, et al.
Published: (2025)
DAM-VSR: Disentanglement of Appearance and Motion for Video Super-Resolution
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
Hypergraph-State Collaborative Reasoning for Multi-Object Tracking
by: Song, Zikai, et al.
Published: (2026)
by: Song, Zikai, et al.
Published: (2026)
Appearance Blur-driven AutoEncoder and Motion-guided Memory Module for Video Anomaly Detection
by: Lyu, Jiahao, et al.
Published: (2024)
by: Lyu, Jiahao, et al.
Published: (2024)
OmniTrend: Content-Context Modeling for Scalable Social Popularity Prediction
by: Ye, Liliang, et al.
Published: (2026)
by: Ye, Liliang, et al.
Published: (2026)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
by: Zhang, Xinglang, et al.
Published: (2026)
by: Zhang, Xinglang, et al.
Published: (2026)
HyperFusion: Hierarchical Multimodal Ensemble Learning for Social Media Popularity Prediction
by: Ye, Liliang, et al.
Published: (2025)
by: Ye, Liliang, et al.
Published: (2025)
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
by: Wang, Dali, et al.
Published: (2026)
by: Wang, Dali, et al.
Published: (2026)
DiffusionTrack: Diffusion Model For Multi-Object Tracking
by: Luo, Run, et al.
Published: (2023)
by: Luo, Run, et al.
Published: (2023)
Ref-GS: Directional Factorization for 2D Gaussian Splatting
by: Zhang, Youjia, et al.
Published: (2024)
by: Zhang, Youjia, et al.
Published: (2024)
Coupling Macro Dynamics and Micro States for Long-Horizon Social Simulation
by: Zhang, Yunyao, et al.
Published: (2026)
by: Zhang, Yunyao, et al.
Published: (2026)
HotComment: A Benchmark for Evaluating Popularity of Online Comments
by: Wu, Yafeng, et al.
Published: (2026)
by: Wu, Yafeng, et al.
Published: (2026)
Tracking the Unstable: Appearance-Guided Motion Modeling for Robust Multi-Object Tracking in UAV-Captured Videos
by: Ma, Jianbo, et al.
Published: (2025)
by: Ma, Jianbo, et al.
Published: (2025)
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
by: Chefer, Hila, et al.
Published: (2025)
by: Chefer, Hila, et al.
Published: (2025)
PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
by: Gao, Mingju, et al.
Published: (2026)
by: Gao, Mingju, et al.
Published: (2026)
Artemis: Articulated Neural Pets with Appearance and Motion synthesis
by: Luo, Haimin, et al.
Published: (2022)
by: Luo, Haimin, et al.
Published: (2022)
Dual Conditioned Motion Diffusion for Pose-Based Video Anomaly Detection
by: Wang, Hongsong, et al.
Published: (2024)
by: Wang, Hongsong, et al.
Published: (2024)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Uncertainty-Guided Appearance-Motion Association Network for Out-of-Distribution Action Detection
by: Fang, Xiang, et al.
Published: (2024)
by: Fang, Xiang, et al.
Published: (2024)
Jigsaw3D: Disentangled 3D Style Transfer via Patch Shuffling and Masking
by: Ye, Yuteng, et al.
Published: (2025)
by: Ye, Yuteng, et al.
Published: (2025)
IntervenSim: Intervention-Aware Social Network Simulation for Opinion Dynamics
by: Zhang, Yunyao, et al.
Published: (2026)
by: Zhang, Yunyao, et al.
Published: (2026)
Autogenic Language Embedding for Coherent Point Tracking
by: Song, Zikai, et al.
Published: (2024)
by: Song, Zikai, et al.
Published: (2024)
GateMOT: Q-Gated Attention for Dense Object Tracking
by: Lv, Mingjin, et al.
Published: (2026)
by: Lv, Mingjin, et al.
Published: (2026)
Similar Items
-
Progressive Text-to-Image Diffusion with Soft Latent Direction
by: Ye, YuTeng, et al.
Published: (2023) -
CA-Diff: Collaborative Anatomy Diffusion for Brain Tissue Segmentation
by: Xing, Qilong, et al.
Published: (2025) -
Attacking Transformers with Feature Diversity Adversarial Perturbation
by: Gao, Chenxing, et al.
Published: (2024) -
Coupled Mamba: Enhanced Multi-modal Fusion with Coupled State Space Model
by: Li, Wenbing, et al.
Published: (2024) -
CurEvo: Curriculum-Guided Self-Evolution for Video Understanding
by: Zeng, Guiyi, et al.
Published: (2026)