CigTime: Corrective Instruction Generation Through Inverse Motion Editing
Fuente:
arXiv
Guardado en:
| Autores principales: | Fang, Qihang, Tang, Chengcheng, Tekin, Bugra, Yang, Yanchao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DEAL-300K: Diffusion-based Editing Area Localization with a 300K-Scale Dataset and Frequency-Prompted Baseline
por: Zhang, Rui, et al.
Publicado: (2025)
por: Zhang, Rui, et al.
Publicado: (2025)
BID: Boundary-Interior Decoding for Unsupervised Temporal Action Localization Pre-Trainin
por: Fang, Qihang, et al.
Publicado: (2024)
por: Fang, Qihang, et al.
Publicado: (2024)
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
por: Egorov, Konstantin, et al.
Publicado: (2025)
por: Egorov, Konstantin, et al.
Publicado: (2025)
Efficient Vision-based Vehicle Speed Estimation
por: Macko, Andrej, et al.
Publicado: (2025)
por: Macko, Andrej, et al.
Publicado: (2025)
N-DriverMotion: Driver motion learning and prediction using an event-based camera and directly trained spiking neural networks on Loihi 2
por: Chung, Hyo Jong, et al.
Publicado: (2024)
por: Chung, Hyo Jong, et al.
Publicado: (2024)
Stylized Face Sketch Extraction via Generative Prior with Limited Data
por: Yun, Kwan, et al.
Publicado: (2024)
por: Yun, Kwan, et al.
Publicado: (2024)
LeGO: Leveraging a Surface Deformation Network for Animatable Stylized Face Generation with One Example
por: Yoon, Soyeon, et al.
Publicado: (2024)
por: Yoon, Soyeon, et al.
Publicado: (2024)
Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting
por: Maheshkar, Vaishali, et al.
Publicado: (2025)
por: Maheshkar, Vaishali, et al.
Publicado: (2025)
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
por: Lu, Wanglong, et al.
Publicado: (2024)
por: Lu, Wanglong, et al.
Publicado: (2024)
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
por: Choi, Seonghwa, et al.
Publicado: (2025)
por: Choi, Seonghwa, et al.
Publicado: (2025)
SymFace: Additional Facial Symmetry Loss for Deep Face Recognition
por: Prakash, Pritesh, et al.
Publicado: (2024)
por: Prakash, Pritesh, et al.
Publicado: (2024)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
por: Zafeiri, Maria, et al.
Publicado: (2024)
por: Zafeiri, Maria, et al.
Publicado: (2024)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
por: Sartinas, Evangelos G., et al.
Publicado: (2023)
por: Sartinas, Evangelos G., et al.
Publicado: (2023)
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
por: Khan, Md Ashik, et al.
Publicado: (2025)
por: Khan, Md Ashik, et al.
Publicado: (2025)
Distilling foundation models for robust and efficient models in digital pathology
por: Filiot, Alexandre, et al.
Publicado: (2025)
por: Filiot, Alexandre, et al.
Publicado: (2025)
Harnessing Deep Learning and Satellite Imagery for Post-Buyout Land Cover Mapping
por: Otal, Hakan T., et al.
Publicado: (2024)
por: Otal, Hakan T., et al.
Publicado: (2024)
HuMoCon: Concept Discovery for Human Motion Understanding
por: Fang, Qihang, et al.
Publicado: (2025)
por: Fang, Qihang, et al.
Publicado: (2025)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
por: Ömercikoğlu, Ahmet Can, et al.
Publicado: (2025)
por: Ömercikoğlu, Ahmet Can, et al.
Publicado: (2025)
SelectiveKD: A semi-supervised framework for cancer detection in DBT through Knowledge Distillation and Pseudo-labeling
por: Dillard, Laurent, et al.
Publicado: (2024)
por: Dillard, Laurent, et al.
Publicado: (2024)
Eleven Primitives and Three Gates: The Universal Structure of Computational Imaging
por: Yang, Chengshuai, et al.
Publicado: (2026)
por: Yang, Chengshuai, et al.
Publicado: (2026)
Tri-Plane Mamba: Efficiently Adapting Segment Anything Model for 3D Medical Images
por: Wang, Hualiang, et al.
Publicado: (2024)
por: Wang, Hualiang, et al.
Publicado: (2024)
DeepFusionNet: Autoencoder-Based Low-Light Image Enhancement and Super-Resolution
por: Çalışkan, Halil Hüseyin, et al.
Publicado: (2025)
por: Çalışkan, Halil Hüseyin, et al.
Publicado: (2025)
HieraEdgeNet: A Multi-Scale Edge-Enhanced Framework for Automated Pollen Recognition
por: Long, Yuchong, et al.
Publicado: (2025)
por: Long, Yuchong, et al.
Publicado: (2025)
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
por: Cai, Changpeng, et al.
Publicado: (2024)
por: Cai, Changpeng, et al.
Publicado: (2024)
EditP23: 3D Editing via Propagation of Image Prompts to Multi-View
por: Bar-On, Roi, et al.
Publicado: (2025)
por: Bar-On, Roi, et al.
Publicado: (2025)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
por: Lu, Wanglong, et al.
Publicado: (2024)
por: Lu, Wanglong, et al.
Publicado: (2024)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
por: Cao, Bryan Bo, et al.
Publicado: (2024)
por: Cao, Bryan Bo, et al.
Publicado: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
por: Merugu, Ranjith, et al.
Publicado: (2025)
por: Merugu, Ranjith, et al.
Publicado: (2025)
FastFit: Accelerating Multi-Reference Virtual Try-On via Cacheable Diffusion Models
por: Chong, Zheng, et al.
Publicado: (2025)
por: Chong, Zheng, et al.
Publicado: (2025)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
por: Lu, Wanglong, et al.
Publicado: (2025)
por: Lu, Wanglong, et al.
Publicado: (2025)
PolyGlotFake: A Novel Multilingual and Multimodal DeepFake Dataset
por: Hou, Yang, et al.
Publicado: (2024)
por: Hou, Yang, et al.
Publicado: (2024)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
por: Zhang, Jian, et al.
Publicado: (2024)
por: Zhang, Jian, et al.
Publicado: (2024)
VitalLens 2.0: High-Fidelity rPPG for Heart Rate Variability Estimation from Face Video
por: Rouast, Philipp V.
Publicado: (2025)
por: Rouast, Philipp V.
Publicado: (2025)
Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis
por: Narimani, Mohammadreza, et al.
Publicado: (2026)
por: Narimani, Mohammadreza, et al.
Publicado: (2026)
Satellite-Net: Automatic Extraction of Land Cover Indicators from Satellite Imagery by Deep Learning
por: Bernasconi, Eleonora, et al.
Publicado: (2019)
por: Bernasconi, Eleonora, et al.
Publicado: (2019)
S3Simulator: A benchmarking Side Scan Sonar Simulator dataset for Underwater Image Analysis
por: S, Kamal Basha, et al.
Publicado: (2024)
por: S, Kamal Basha, et al.
Publicado: (2024)
Balanced conic rectified flow
por: Kim, Shin Seong, et al.
Publicado: (2025)
por: Kim, Shin Seong, et al.
Publicado: (2025)
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
por: Gerats, Beerend G. A., et al.
Publicado: (2024)
por: Gerats, Beerend G. A., et al.
Publicado: (2024)
MotionAnymesh: Physics-Grounded Articulation for Simulation-Ready Digital Twins
por: Xu, WenBo, et al.
Publicado: (2026)
por: Xu, WenBo, et al.
Publicado: (2026)
ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection
por: Samson, Hema Hariharan
Publicado: (2026)
por: Samson, Hema Hariharan
Publicado: (2026)
Ejemplares similares
-
DEAL-300K: Diffusion-based Editing Area Localization with a 300K-Scale Dataset and Frequency-Prompted Baseline
por: Zhang, Rui, et al.
Publicado: (2025) -
BID: Boundary-Interior Decoding for Unsupervised Temporal Action Localization Pre-Trainin
por: Fang, Qihang, et al.
Publicado: (2024) -
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
por: Egorov, Konstantin, et al.
Publicado: (2025) -
Efficient Vision-based Vehicle Speed Estimation
por: Macko, Andrej, et al.
Publicado: (2025) -
N-DriverMotion: Driver motion learning and prediction using an event-based camera and directly trained spiking neural networks on Loihi 2
por: Chung, Hyo Jong, et al.
Publicado: (2024)