S3Editor: A Sparse Semantic-Disentangled Self-Training Framework for Face Video Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Guangzhi, Chen, Tianyi, Ghasedi, Kamran, Wu, HsiangTao, Ding, Tianyu, Nuesmeyer, Chris, Zharkov, Ilya, Kankanhalli, Mohan, Liang, Luming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
von: Zhu, Haidong, et al.
Veröffentlicht: (2023)
von: Zhu, Haidong, et al.
Veröffentlicht: (2023)
FORA: Fast-Forward Caching in Diffusion Transformer Acceleration
von: Selvaraju, Pratheba, et al.
Veröffentlicht: (2024)
von: Selvaraju, Pratheba, et al.
Veröffentlicht: (2024)
Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
AdaContour: Adaptive Contour Descriptor with Hierarchical Representation
von: Ding, Tianyu, et al.
Veröffentlicht: (2024)
von: Ding, Tianyu, et al.
Veröffentlicht: (2024)
The Efficiency Spectrum of Large Language Models: An Algorithmic Survey
von: Ding, Tianyu, et al.
Veröffentlicht: (2023)
von: Ding, Tianyu, et al.
Veröffentlicht: (2023)
DREAM: Diffusion Rectification and Estimation-Adaptive Models
von: Zhou, Jinxin, et al.
Veröffentlicht: (2023)
von: Zhou, Jinxin, et al.
Veröffentlicht: (2023)
ProCrop: Learning Aesthetic Image Cropping from Professional Compositions
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
DistiLLM-2: A Contrastive Approach Boosts the Distillation of LLMs
von: Ko, Jongwoo, et al.
Veröffentlicht: (2025)
von: Ko, Jongwoo, et al.
Veröffentlicht: (2025)
Motion Graph Unleashed: A Novel Approach to Video Prediction
von: Zhong, Yiqi, et al.
Veröffentlicht: (2024)
von: Zhong, Yiqi, et al.
Veröffentlicht: (2024)
HESSO: Towards Automatic Efficient and User Friendly Any Neural Network Training and Pruning
von: Chen, Tianyi, et al.
Veröffentlicht: (2024)
von: Chen, Tianyi, et al.
Veröffentlicht: (2024)
Cat-AIR: Content and Task-Aware All-in-One Image Restoration
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
OFER: Occluded Face Expression Reconstruction
von: Selvaraju, Pratheba, et al.
Veröffentlicht: (2024)
von: Selvaraju, Pratheba, et al.
Veröffentlicht: (2024)
Automatic Joint Structured Pruning and Quantization for Efficient Neural Network Training and Compression
von: Qu, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Qu, Xiaoyi, et al.
Veröffentlicht: (2025)
UnStar: Unlearning with Self-Taught Anti-Sample Reasoning for LLMs
von: Sinha, Yash, et al.
Veröffentlicht: (2024)
von: Sinha, Yash, et al.
Veröffentlicht: (2024)
SCAN: Bootstrapping Contrastive Pre-training for Data Efficiency
von: Guo, Yangyang, et al.
Veröffentlicht: (2024)
von: Guo, Yangyang, et al.
Veröffentlicht: (2024)
FractalForensics: Proactive Deepfake Detection and Localization via Fractal Watermarks
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)
Translution: Unifying Self-attention and Convolution for Adaptive and Relative Modeling
von: Fan, Hehe, et al.
Veröffentlicht: (2025)
von: Fan, Hehe, et al.
Veröffentlicht: (2025)
Word-Anchored Temporal Forgery Localization
von: Wang, Tianyi, et al.
Veröffentlicht: (2026)
von: Wang, Tianyi, et al.
Veröffentlicht: (2026)
Training-Free Disentangled Text-Guided Image Editing via Sparse Latent Constraints
von: Shabrina, Mutiara, et al.
Veröffentlicht: (2025)
von: Shabrina, Mutiara, et al.
Veröffentlicht: (2025)
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
Diffusion Facial Forgery Detection
von: Cheng, Harry, et al.
Veröffentlicht: (2024)
von: Cheng, Harry, et al.
Veröffentlicht: (2024)
Object-Centric Framework for Video Moment Retrieval
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
VideoAutoArena: An Automated Arena for Evaluating Large Multimodal Models in Video Analysis through User Simulation
von: Luo, Ziyang, et al.
Veröffentlicht: (2024)
von: Luo, Ziyang, et al.
Veröffentlicht: (2024)
Image-to-Image Translation with Disentangled Latent Vectors for Face Editing
von: Dalva, Yusuf, et al.
Veröffentlicht: (2023)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2023)
Finetuning Text-to-Image Diffusion Models for Fairness
von: Shen, Xudong, et al.
Veröffentlicht: (2023)
von: Shen, Xudong, et al.
Veröffentlicht: (2023)
Joint Vision-Language Social Bias Removal for CLIP
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
Bullying the Machine: How Personas Increase LLM Vulnerability
von: Xu, Ziwei, et al.
Veröffentlicht: (2025)
von: Xu, Ziwei, et al.
Veröffentlicht: (2025)
DVI: Disentangling Semantic and Visual Identity for Training-Free Personalized Generation
von: Li, Guandong, et al.
Veröffentlicht: (2025)
von: Li, Guandong, et al.
Veröffentlicht: (2025)
Context-aware Sparse Spatiotemporal Learning for Event-based Vision
von: Wang, Shenqi, et al.
Veröffentlicht: (2025)
von: Wang, Shenqi, et al.
Veröffentlicht: (2025)
Hallucination is Inevitable: An Innate Limitation of Large Language Models
von: Xu, Ziwei, et al.
Veröffentlicht: (2024)
von: Xu, Ziwei, et al.
Veröffentlicht: (2024)
SemFaceEdit: Semantic Face Editing on Generative Radiance Manifolds
von: Verma, Shashikant, et al.
Veröffentlicht: (2025)
von: Verma, Shashikant, et al.
Veröffentlicht: (2025)
Reasoning LLMs are Wandering Solution Explorers
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
VidHal: Benchmarking Temporal Hallucinations in Vision LLMs
von: Choong, Wey Yeh, et al.
Veröffentlicht: (2024)
von: Choong, Wey Yeh, et al.
Veröffentlicht: (2024)
FluxSpace: Disentangled Semantic Editing in Rectified Flow Transformers
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
Nearest Neighbor Projection Removal Adversarial Training
von: Singh, Himanshu, et al.
Veröffentlicht: (2025)
von: Singh, Himanshu, et al.
Veröffentlicht: (2025)
Detecting Deepfakes via Hamiltonian Dynamics
von: Cheng, Harry, et al.
Veröffentlicht: (2026)
von: Cheng, Harry, et al.
Veröffentlicht: (2026)
Fair Deepfake Detectors Can Generalize
von: Cheng, Harry, et al.
Veröffentlicht: (2025)
von: Cheng, Harry, et al.
Veröffentlicht: (2025)
RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild
von: Xu, Danni, et al.
Veröffentlicht: (2025)
von: Xu, Danni, et al.
Veröffentlicht: (2025)
Aggregating Diverse Cue Experts for AI-Generated Image Detection
von: Tan, Lei, et al.
Veröffentlicht: (2026)
von: Tan, Lei, et al.
Veröffentlicht: (2026)
Unified Editing of Panorama, 3D Scenes, and Videos Through Disentangled Self-Attention Injection
von: Kwon, Gihyun, et al.
Veröffentlicht: (2024)
von: Kwon, Gihyun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
von: Zhu, Haidong, et al.
Veröffentlicht: (2023) -
FORA: Fast-Forward Caching in Diffusion Transformer Acceleration
von: Selvaraju, Pratheba, et al.
Veröffentlicht: (2024) -
Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025) -
AdaContour: Adaptive Contour Descriptor with Hierarchical Representation
von: Ding, Tianyu, et al.
Veröffentlicht: (2024) -
The Efficiency Spectrum of Large Language Models: An Algorithmic Survey
von: Ding, Tianyu, et al.
Veröffentlicht: (2023)