PersonificationNet: Making customized subject act like a person
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Tianchu, Li, Pengyu, Wang, Biao, Hua, Xiansheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Animalbooth: multimodal feature enhancement for animal subject personalization
von: Liu, Chen, et al.
Veröffentlicht: (2025)
von: Liu, Chen, et al.
Veröffentlicht: (2025)
Intrinsic Gradient Suppression for Label-Noise Prompt Tuning in Vision-Language Models
von: Li, Jiayu, et al.
Veröffentlicht: (2026)
von: Li, Jiayu, et al.
Veröffentlicht: (2026)
Efficient Token Compression for Vision Transformer with Spatial Information Preserved
von: Mao, Junzhu, et al.
Veröffentlicht: (2025)
von: Mao, Junzhu, et al.
Veröffentlicht: (2025)
PromptLNet: Region-Adaptive Aesthetic Enhancement via Prompt Guidance in Low-Light Enhancement Net
von: Yin, Jun, et al.
Veröffentlicht: (2025)
von: Yin, Jun, et al.
Veröffentlicht: (2025)
Semi-supervised Semantic Segmentation with Multi-Constraint Consistency Learning
von: Yin, Jianjian, et al.
Veröffentlicht: (2025)
von: Yin, Jianjian, et al.
Veröffentlicht: (2025)
EZIGen: Enhancing zero-shot personalized image generation with precise subject encoding and decoupled guidance
von: Duan, Zicheng, et al.
Veröffentlicht: (2024)
von: Duan, Zicheng, et al.
Veröffentlicht: (2024)
ParkingTwin: Training-Free Streaming 3D Reconstruction for Parking-Lot Digital Twins
von: Liu, Xinhao, et al.
Veröffentlicht: (2026)
von: Liu, Xinhao, et al.
Veröffentlicht: (2026)
MRIo3DS-Net: A Mutually Reinforcing Images to 3D Surface RNN-like framework for model-adaptation indoor 3D reconstruction
von: Li, Chang, et al.
Veröffentlicht: (2024)
von: Li, Chang, et al.
Veröffentlicht: (2024)
GeoLink: A 3D-Aware Framework Towards Better Generalization in Cross-View Geo-Localization
von: Zhang, Hongyang, et al.
Veröffentlicht: (2026)
von: Zhang, Hongyang, et al.
Veröffentlicht: (2026)
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
MMAD: Multi-label Micro-Action Detection in Videos
von: Li, Kun, et al.
Veröffentlicht: (2024)
von: Li, Kun, et al.
Veröffentlicht: (2024)
R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning
von: Wang, Biao, et al.
Veröffentlicht: (2025)
von: Wang, Biao, et al.
Veröffentlicht: (2025)
CrownGen: Patient-customized Crown Generation via Point Diffusion Model
von: Bae, Juyoung, et al.
Veröffentlicht: (2025)
von: Bae, Juyoung, et al.
Veröffentlicht: (2025)
ParameterNet: Parameters Are All You Need
von: Han, Kai, et al.
Veröffentlicht: (2023)
von: Han, Kai, et al.
Veröffentlicht: (2023)
Unit Region Encoding: A Unified and Compact Geometry-aware Representation for Floorplan Applications
von: Zhang, Huichao, et al.
Veröffentlicht: (2025)
von: Zhang, Huichao, et al.
Veröffentlicht: (2025)
DRPCA-Net: Make Robust PCA Great Again for Infrared Small Target Detection
von: Xiong, Zihao, et al.
Veröffentlicht: (2025)
von: Xiong, Zihao, et al.
Veröffentlicht: (2025)
Visual-Semantic Graph Matching Net for Zero-Shot Learning
von: Duan, Bowen, et al.
Veröffentlicht: (2024)
von: Duan, Bowen, et al.
Veröffentlicht: (2024)
FG-TreeSeg: Flow-Guided Tree Crown Segmentation without Instance Annotations
von: Chen, Pengyu, et al.
Veröffentlicht: (2026)
von: Chen, Pengyu, et al.
Veröffentlicht: (2026)
MoE-DiffIR: Task-customized Diffusion Priors for Universal Compressed Image Restoration
von: Ren, Yulin, et al.
Veröffentlicht: (2024)
von: Ren, Yulin, et al.
Veröffentlicht: (2024)
ArchShapeNet:An Interpretable 3D-CNN Framework for Evaluating Architectural Shapes
von: Yin, Jun, et al.
Veröffentlicht: (2025)
von: Yin, Jun, et al.
Veröffentlicht: (2025)
A Survey on fMRI-based Brain Decoding for Reconstructing Multimodal Stimuli
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
Probing Deep into Temporal Profile Makes the Infrared Small Target Detector Much Better
von: Li, Ruojing, et al.
Veröffentlicht: (2025)
von: Li, Ruojing, et al.
Veröffentlicht: (2025)
WS-DETR: Robust Water Surface Object Detection through Vision-Radar Fusion with Detection Transformer
von: Yin, Huilin, et al.
Veröffentlicht: (2025)
von: Yin, Huilin, et al.
Veröffentlicht: (2025)
RoNet: Rotation-oriented Continuous Image Translation
von: Li, Yi, et al.
Veröffentlicht: (2024)
von: Li, Yi, et al.
Veröffentlicht: (2024)
A Collaborative Jade Recognition System for Mobile Devices Based on Lightweight and Large Models
von: Wang, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2025)
What Makes ImageNet Look Unlike LAION
von: Shirali, Ali, et al.
Veröffentlicht: (2023)
von: Shirali, Ali, et al.
Veröffentlicht: (2023)
GMFL-Net: A Global Multi-geometric Feature Learning Network for Repetitive Action Counting
von: Li, Jun, et al.
Veröffentlicht: (2024)
von: Li, Jun, et al.
Veröffentlicht: (2024)
EoS-FM: Can an Ensemble of Specialist Models act as a Generalist Feature Extractor?
von: Adorni, Pierre, et al.
Veröffentlicht: (2025)
von: Adorni, Pierre, et al.
Veröffentlicht: (2025)
Sketch2PoseNet: Efficient and Generalized Sketch to 3D Human Pose Prediction
von: Wang, Li, et al.
Veröffentlicht: (2025)
von: Wang, Li, et al.
Veröffentlicht: (2025)
Cross-Modal Urban Sensing: Evaluating Sound-Vision Alignment Across Street-Level and Aerial Imagery
von: Chen, Pengyu, et al.
Veröffentlicht: (2025)
von: Chen, Pengyu, et al.
Veröffentlicht: (2025)
Micro-gesture Online Recognition using Learnable Query Points
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
EMDFNet: Efficient Multi-scale and Diverse Feature Network for Traffic Sign Detection
von: Li, Pengyu, et al.
Veröffentlicht: (2024)
von: Li, Pengyu, et al.
Veröffentlicht: (2024)
VLRM: Vision-Language Models act as Reward Models for Image Captioning
von: Dzabraev, Maksim, et al.
Veröffentlicht: (2024)
von: Dzabraev, Maksim, et al.
Veröffentlicht: (2024)
RepControlNet: ControlNet Reparameterization
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
Technical Report for ActivityNet Challenge 2022 -- Temporal Action Localization
von: Chen, Shimin, et al.
Veröffentlicht: (2024)
von: Chen, Shimin, et al.
Veröffentlicht: (2024)
Glissando-Net: Deep sinGLe vIew category level poSe eStimation ANd 3D recOnstruction
von: Sun, Bo, et al.
Veröffentlicht: (2025)
von: Sun, Bo, et al.
Veröffentlicht: (2025)
OpenMAP-BrainAge: Generalizable and Interpretable Brain Age Predictor
von: Kan, Pengyu, et al.
Veröffentlicht: (2025)
von: Kan, Pengyu, et al.
Veröffentlicht: (2025)
HCF-Net: Hierarchical Context Fusion Network for Infrared Small Object Detection
von: Xu, Shibiao, et al.
Veröffentlicht: (2024)
von: Xu, Shibiao, et al.
Veröffentlicht: (2024)
SOFTooth: Semantics-Enhanced Order-Aware Fusion for Tooth Instance Segmentation
von: Li, Xiaolan, et al.
Veröffentlicht: (2025)
von: Li, Xiaolan, et al.
Veröffentlicht: (2025)
BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video Deflickering
von: Qiu, Xinmin, et al.
Veröffentlicht: (2024)
von: Qiu, Xinmin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Animalbooth: multimodal feature enhancement for animal subject personalization
von: Liu, Chen, et al.
Veröffentlicht: (2025) -
Intrinsic Gradient Suppression for Label-Noise Prompt Tuning in Vision-Language Models
von: Li, Jiayu, et al.
Veröffentlicht: (2026) -
Efficient Token Compression for Vision Transformer with Spatial Information Preserved
von: Mao, Junzhu, et al.
Veröffentlicht: (2025) -
PromptLNet: Region-Adaptive Aesthetic Enhancement via Prompt Guidance in Low-Light Enhancement Net
von: Yin, Jun, et al.
Veröffentlicht: (2025) -
Semi-supervised Semantic Segmentation with Multi-Constraint Consistency Learning
von: Yin, Jianjian, et al.
Veröffentlicht: (2025)