Optical-Flow Guided Prompt Optimization for Coherent Video Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nam, Hyelin, Kim, Jaemin, Lee, Dohun, Ye, Jong Chul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UNICORN: Ultrasound Nakagami Imaging via Score Matching and Adaptation for Assessing Hepatic Steatosis
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2026)
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2026)
Multiscale Color Guided Attention Ensemble Classifier for Age-Related Macular Degeneration using Concurrent Fundus and Optical Coherence Tomography Images
von: Gupta, Pragya, et al.
Veröffentlicht: (2024)
von: Gupta, Pragya, et al.
Veröffentlicht: (2024)
Multimodal Fusion and Coherence Modeling for Video Topic Segmentation
von: Yu, Hai, et al.
Veröffentlicht: (2024)
von: Yu, Hai, et al.
Veröffentlicht: (2024)
Review and Recommendations for using Artificial Intelligence in Intracoronary Optical Coherence Tomography Analysis
von: Chen, Xu, et al.
Veröffentlicht: (2025)
von: Chen, Xu, et al.
Veröffentlicht: (2025)
IQNet: Image Quality Assessment Guided Just Noticeable Difference Prefiltering For Versatile Video Coding
von: Sun, Yu-Han, et al.
Veröffentlicht: (2023)
von: Sun, Yu-Han, et al.
Veröffentlicht: (2023)
EchoLVFM: One-Step Video Generation via Latent Flow Matching for Echocardiogram Synthesis
von: Oladokun, Emmanuel, et al.
Veröffentlicht: (2026)
von: Oladokun, Emmanuel, et al.
Veröffentlicht: (2026)
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models
von: Mehta, Deval, et al.
Veröffentlicht: (2025)
von: Mehta, Deval, et al.
Veröffentlicht: (2025)
HazeMatching: Dehazing Light Microscopy Images with Guided Conditional Flow Matching
von: Ray, Anirban, et al.
Veröffentlicht: (2025)
von: Ray, Anirban, et al.
Veröffentlicht: (2025)
EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence
von: Li, Yahui, et al.
Veröffentlicht: (2026)
von: Li, Yahui, et al.
Veröffentlicht: (2026)
Conditional Brownian Bridge Diffusion Model for VHR SAR to Optical Image Translation
von: Kim, Seon-Hoon, et al.
Veröffentlicht: (2024)
von: Kim, Seon-Hoon, et al.
Veröffentlicht: (2024)
GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2025)
von: Kim, Gwanghyun, et al.
Veröffentlicht: (2025)
CV-VAE: A Compatible Video VAE for Latent Generative Video Models
von: Zhao, Sijie, et al.
Veröffentlicht: (2024)
von: Zhao, Sijie, et al.
Veröffentlicht: (2024)
IPG: Incremental Patch Generation for Generalized Adversarial Patch Training
von: Lee, Wonho, et al.
Veröffentlicht: (2025)
von: Lee, Wonho, et al.
Veröffentlicht: (2025)
Fundus2Video: Cross-Modal Angiography Video Generation from Static Fundus Photography with Clinical Knowledge Guidance
von: Zhang, Weiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Weiyi, et al.
Veröffentlicht: (2024)
Explainable and Controllable Motion Curve Guided Cardiac Ultrasound Video Generation
von: Yu, Junxuan, et al.
Veröffentlicht: (2024)
von: Yu, Junxuan, et al.
Veröffentlicht: (2024)
Spatially Covariant Image Registration with Text Prompts
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
AI-Based Stroke Rehabilitation Domiciliary Assessment System with ST_GCN Attention
von: Lim, Suhyeon, et al.
Veröffentlicht: (2025)
von: Lim, Suhyeon, et al.
Veröffentlicht: (2025)
HyperspectralMAE: The Hyperspectral Imagery Classification Model using Fourier-Encoded Dual-Branch Masked Autoencoder
von: Jeong, Wooyoung, et al.
Veröffentlicht: (2025)
von: Jeong, Wooyoung, et al.
Veröffentlicht: (2025)
Prompt Mechanisms in Medical Imaging: A Comprehensive Survey
von: Yang, Hao, et al.
Veröffentlicht: (2025)
von: Yang, Hao, et al.
Veröffentlicht: (2025)
Beyond Calibration: Confounding Pathology Limits Foundation Model Specificity in Abdominal Trauma CT
von: Raythatha, Jineel H, et al.
Veröffentlicht: (2026)
von: Raythatha, Jineel H, et al.
Veröffentlicht: (2026)
MedSAM2: Segment Anything in 3D Medical Images and Videos
von: Ma, Jun, et al.
Veröffentlicht: (2025)
von: Ma, Jun, et al.
Veröffentlicht: (2025)
LUMINA-Net: Low-light Upgrade through Multi-stage Illumination and Noise Adaptation Network for Image Enhancement
von: Siddiqua, Namrah, et al.
Veröffentlicht: (2025)
von: Siddiqua, Namrah, et al.
Veröffentlicht: (2025)
Loss-Aware Automatic Selection of Structured Pruning Criteria for Deep Neural Network Acceleration
von: Ghimire, Deepak, et al.
Veröffentlicht: (2025)
von: Ghimire, Deepak, et al.
Veröffentlicht: (2025)
Optimizing Breast Cancer Detection in Mammograms: A Comprehensive Study of Transfer Learning, Resolution Reduction, and Multi-View Classification
von: Petrini, Daniel G. P., et al.
Veröffentlicht: (2025)
von: Petrini, Daniel G. P., et al.
Veröffentlicht: (2025)
VideoQA-SC: Adaptive Semantic Communication for Video Question Answering
von: Guo, Jiangyuan, et al.
Veröffentlicht: (2024)
von: Guo, Jiangyuan, et al.
Veröffentlicht: (2024)
Guided Lensless Polarization Imaging
von: Kraicer, Noa, et al.
Veröffentlicht: (2026)
von: Kraicer, Noa, et al.
Veröffentlicht: (2026)
Replace-then-Perturb: Targeted Adversarial Attacks With Visual Reasoning for Vision-Language Models
von: Jang, Jonggyu, et al.
Veröffentlicht: (2024)
von: Jang, Jonggyu, et al.
Veröffentlicht: (2024)
Efficient Deep Learning Approaches for Processing Ultra-Widefield Retinal Imaging
von: Kim, Siwon, et al.
Veröffentlicht: (2025)
von: Kim, Siwon, et al.
Veröffentlicht: (2025)
Prediction of Frozen Region Growth in Kidney Cryoablation Intervention Using a 3D Flow-Matching Model
von: Yoon, Siyeop, et al.
Veröffentlicht: (2025)
von: Yoon, Siyeop, et al.
Veröffentlicht: (2025)
VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis
von: Hoopes, Andrew, et al.
Veröffentlicht: (2024)
von: Hoopes, Andrew, et al.
Veröffentlicht: (2024)
SpurBreast: A Curated Dataset for Investigating Spurious Correlations in Real-world Breast MRI Classification
von: Won, Jong Bum, et al.
Veröffentlicht: (2025)
von: Won, Jong Bum, et al.
Veröffentlicht: (2025)
Generative Artificial Intelligence in Medical Imaging: Foundations, Progress, and Clinical Translation
von: Zhou, Xuanru, et al.
Veröffentlicht: (2025)
von: Zhou, Xuanru, et al.
Veröffentlicht: (2025)
Active Prompt Tuning Enables Gpt-40 To Do Efficient Classification Of Microscopy Images
von: Kandiyana, Abhiram, et al.
Veröffentlicht: (2024)
von: Kandiyana, Abhiram, et al.
Veröffentlicht: (2024)
AI Guided Early Screening of Cervical Cancer
von: S I, Dharanidharan, et al.
Veröffentlicht: (2024)
von: S I, Dharanidharan, et al.
Veröffentlicht: (2024)
AMRG: Extend Vision Language Models for Automatic Mammography Report Generation
von: Sung, Nak-Jun, et al.
Veröffentlicht: (2025)
von: Sung, Nak-Jun, et al.
Veröffentlicht: (2025)
Efficient Flow Matching for Sparse-View CT Reconstruction
von: Shi, Jiayang, et al.
Veröffentlicht: (2026)
von: Shi, Jiayang, et al.
Veröffentlicht: (2026)
Characterizing Motion Encoding in Video Diffusion Timesteps
von: Baherwani, Vatsal, et al.
Veröffentlicht: (2025)
von: Baherwani, Vatsal, et al.
Veröffentlicht: (2025)
Deep Blind Super-Resolution for Satellite Video
von: Xiao, Yi, et al.
Veröffentlicht: (2024)
von: Xiao, Yi, et al.
Veröffentlicht: (2024)
EVAN: Evolutional Video Streaming Adaptation via Neural Representation
von: Liu, Mufan, et al.
Veröffentlicht: (2024)
von: Liu, Mufan, et al.
Veröffentlicht: (2024)
Fundus Image-based Visual Acuity Assessment with PAC-Guarantees
von: Jang, Sooyong, et al.
Veröffentlicht: (2024)
von: Jang, Sooyong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
UNICORN: Ultrasound Nakagami Imaging via Score Matching and Adaptation for Assessing Hepatic Steatosis
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2026) -
Multiscale Color Guided Attention Ensemble Classifier for Age-Related Macular Degeneration using Concurrent Fundus and Optical Coherence Tomography Images
von: Gupta, Pragya, et al.
Veröffentlicht: (2024) -
Multimodal Fusion and Coherence Modeling for Video Topic Segmentation
von: Yu, Hai, et al.
Veröffentlicht: (2024) -
Review and Recommendations for using Artificial Intelligence in Intracoronary Optical Coherence Tomography Analysis
von: Chen, Xu, et al.
Veröffentlicht: (2025) -
IQNet: Image Quality Assessment Guided Just Noticeable Difference Prefiltering For Versatile Video Coding
von: Sun, Yu-Han, et al.
Veröffentlicht: (2023)