Is this Generated Person Existed in Real-world? Fine-grained Detecting and Calibrating Abnormal Human-body
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zeqing, Ma, Qingyang, Wan, Wentao, Li, Haojie, Wang, Keze, Tian, Yonghong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
A Stepwise Distillation Learning Strategy for Non-differentiable Visual Programming Frameworks on Visual Reasoning Tasks
von: Wan, Wentao, et al.
Veröffentlicht: (2023)
von: Wan, Wentao, et al.
Veröffentlicht: (2023)
Enhancing Visual Programming for Visual Reasoning via Probabilistic Graphs
von: Wan, Wentao, et al.
Veröffentlicht: (2025)
von: Wan, Wentao, et al.
Veröffentlicht: (2025)
Towards Top-Down Reasoning: An Explainable Multi-Agent Approach for Visual Question Answering
von: Wang, Zeqing, et al.
Veröffentlicht: (2023)
von: Wang, Zeqing, et al.
Veröffentlicht: (2023)
Adversarial Reconstruction Feedback for Robust Fine-grained Generalization
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
Fine-grained Abnormality Prompt Learning for Zero-shot Anomaly Detection
von: Zhu, Jiawen, et al.
Veröffentlicht: (2024)
von: Zhu, Jiawen, et al.
Veröffentlicht: (2024)
3DAlign-DAER: Dynamic Attention Policy and Efficient Retrieval Strategy for Fine-grained 3D-Text Alignment at Scale
von: Fan, Yijia, et al.
Veröffentlicht: (2025)
von: Fan, Yijia, et al.
Veröffentlicht: (2025)
VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
Bridge to Non-Barrier Communication: Gloss-Prompted Fine-grained Cued Speech Gesture Generation with Diffusion Model
von: Lei, Wentao, et al.
Veröffentlicht: (2024)
von: Lei, Wentao, et al.
Veröffentlicht: (2024)
Adaptive-VoCo: Complexity-Aware Visual Token Compression for Vision-Language Models
von: Guo, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Guo, Xiaoyang, et al.
Veröffentlicht: (2025)
Weather-R1: Logically Consistent Reinforcement Fine-Tuning for Multimodal Reasoning in Meteorology
von: Wu, Kaiyu, et al.
Veröffentlicht: (2026)
von: Wu, Kaiyu, et al.
Veröffentlicht: (2026)
Fine-grained Dynamic Network for Generic Event Boundary Detection
von: Zheng, Ziwei, et al.
Veröffentlicht: (2024)
von: Zheng, Ziwei, et al.
Veröffentlicht: (2024)
Aligning Human Motion Generation with Human Perceptions
von: Wang, Haoru, et al.
Veröffentlicht: (2024)
von: Wang, Haoru, et al.
Veröffentlicht: (2024)
MFVLR: Multi-domain Fine-grained Vision-Language Reconstruction for Generalizable Diffusion Face Forgery Detection and Localization
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
Learning to Align Generative Appearance Priors for Fine-grained Image Retrieval
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
FineParser: A Fine-grained Spatio-temporal Action Parser for Human-centric Action Quality Assessment
von: Xu, Jinglin, et al.
Veröffentlicht: (2024)
von: Xu, Jinglin, et al.
Veröffentlicht: (2024)
FinePhys: Fine-grained Human Action Generation by Explicitly Incorporating Physical Laws for Effective Skeletal Guidance
von: Shao, Dian, et al.
Veröffentlicht: (2025)
von: Shao, Dian, et al.
Veröffentlicht: (2025)
Language-driven Fine-grained Retrieval
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
FinPercep-RM: A Fine-grained Reward Model and Co-evolutionary Curriculum for RL-based Real-world Super-Resolution
von: Liu, Yidi, et al.
Veröffentlicht: (2025)
von: Liu, Yidi, et al.
Veröffentlicht: (2025)
TACR-YOLO: A Real-time Detection Framework for Abnormal Human Behaviors Enhanced with Coordinate and Task-Aware Representations
von: Yin, Xinyi, et al.
Veröffentlicht: (2025)
von: Yin, Xinyi, et al.
Veröffentlicht: (2025)
Towards an Effective Action-Region Tracking Framework for Fine-grained Video Action Recognition
von: Sun, Baoli, et al.
Veröffentlicht: (2025)
von: Sun, Baoli, et al.
Veröffentlicht: (2025)
LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face Editing
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
STORM: Search-Guided Generative World Models for Robotic Manipulation
von: Lin, Wenjun, et al.
Veröffentlicht: (2025)
von: Lin, Wenjun, et al.
Veröffentlicht: (2025)
Automated Detection of Abnormalities in Zebrafish Development
von: Sivaprasad, Sarath, et al.
Veröffentlicht: (2026)
von: Sivaprasad, Sarath, et al.
Veröffentlicht: (2026)
From Dataset to Real-world: General 3D Object Detection via Generalized Cross-domain Few-shot Learning
von: Li, Shuangzhi, et al.
Veröffentlicht: (2025)
von: Li, Shuangzhi, et al.
Veröffentlicht: (2025)
FineDiffusion: Scaling up Diffusion Models for Fine-grained Image Generation with 10,000 Classes
von: Pan, Ziying, et al.
Veröffentlicht: (2024)
von: Pan, Ziying, et al.
Veröffentlicht: (2024)
Modeling Thousands of Human Annotators for Generalizable Text-to-Image Person Re-identification
von: Jiang, Jiayu, et al.
Veröffentlicht: (2025)
von: Jiang, Jiayu, et al.
Veröffentlicht: (2025)
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
von: Wang, Kun, et al.
Veröffentlicht: (2025)
von: Wang, Kun, et al.
Veröffentlicht: (2025)
TriLoRA: Integrating SVD for Advanced Style Personalization in Text-to-Image Generation
von: Feng, Chengcheng, et al.
Veröffentlicht: (2024)
von: Feng, Chengcheng, et al.
Veröffentlicht: (2024)
Unleashing Video Language Models for Fine-grained HRCT Report Generation
von: Fang, Yingying, et al.
Veröffentlicht: (2026)
von: Fang, Yingying, et al.
Veröffentlicht: (2026)
Leveraging Structure Knowledge and Deep Models for the Detection of Abnormal Handwritten Text
von: Wang, Zi-Rui
Veröffentlicht: (2024)
von: Wang, Zi-Rui
Veröffentlicht: (2024)
When Every Millisecond Counts: Real-Time Anomaly Detection via the Multimodal Asynchronous Hybrid Network
von: Xiao, Dong, et al.
Veröffentlicht: (2025)
von: Xiao, Dong, et al.
Veröffentlicht: (2025)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
Occlusion-Aware 3D Motion Interpretation for Abnormal Behavior Detection
von: Li, Su, et al.
Veröffentlicht: (2024)
von: Li, Su, et al.
Veröffentlicht: (2024)
Attention Calibration for Disentangled Text-to-Image Personalization
von: Zhang, Yanbing, et al.
Veröffentlicht: (2024)
von: Zhang, Yanbing, et al.
Veröffentlicht: (2024)
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
von: Fang, Xueji, et al.
Veröffentlicht: (2026)
von: Fang, Xueji, et al.
Veröffentlicht: (2026)
Object Detection using Event Camera: A MoE Heat Conduction based Detector and A New Benchmark Dataset
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Every Subtlety Counts: Fine-grained Person Independence Micro-Action Recognition via Distributionally Robust Optimization
von: Cui, Feng-Qi, et al.
Veröffentlicht: (2025)
von: Cui, Feng-Qi, et al.
Veröffentlicht: (2025)
Task-oriented Embedding Counts: Heuristic Clustering-driven Feature Fine-tuning for Whole Slide Image Classification
von: Wang, Xuenian, et al.
Veröffentlicht: (2024)
von: Wang, Xuenian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models
von: Wang, Zeqing, et al.
Veröffentlicht: (2025) -
A Stepwise Distillation Learning Strategy for Non-differentiable Visual Programming Frameworks on Visual Reasoning Tasks
von: Wan, Wentao, et al.
Veröffentlicht: (2023) -
Enhancing Visual Programming for Visual Reasoning via Probabilistic Graphs
von: Wan, Wentao, et al.
Veröffentlicht: (2025) -
Towards Top-Down Reasoning: An Explainable Multi-Agent Approach for Visual Question Answering
von: Wang, Zeqing, et al.
Veröffentlicht: (2023) -
Adversarial Reconstruction Feedback for Robust Fine-grained Generalization
von: Wang, Shijie, et al.
Veröffentlicht: (2025)