Gespeichert in:
| Hauptverfasser: | Cheng, Weijin, Liu, Jianzhi, Deng, Jiawen, Ren, Fuji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2401.01128 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SSP-SAM: SAM with Semantic-Spatial Prompt for Referring Expression Segmentation
von: Tang, Wei, et al.
Veröffentlicht: (2026)
von: Tang, Wei, et al.
Veröffentlicht: (2026)
F3-Pruning: A Training-Free and Generalized Pruning Strategy towards Faster and Finer Text-to-Video Synthesis
von: Su, Sitong, et al.
Veröffentlicht: (2023)
von: Su, Sitong, et al.
Veröffentlicht: (2023)
Road Rage Reasoning with Vision-language Models (VLMs): Task Definition and Evaluation Dataset
von: Weng, Yibing, et al.
Veröffentlicht: (2025)
von: Weng, Yibing, et al.
Veröffentlicht: (2025)
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
von: Jing, Zonglei, et al.
Veröffentlicht: (2025)
von: Jing, Zonglei, et al.
Veröffentlicht: (2025)
SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration
von: Zhang, Yuhong, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhong, et al.
Veröffentlicht: (2024)
Cross-head mutual Mean-Teaching for semi-supervised medical image segmentation
von: Li, Wei, et al.
Veröffentlicht: (2023)
von: Li, Wei, et al.
Veröffentlicht: (2023)
MSSTNet: A Multi-Scale Spatio-Temporal CNN-Transformer Network for Dynamic Facial Expression Recognition
von: Wang, Linhuang, et al.
Veröffentlicht: (2024)
von: Wang, Linhuang, et al.
Veröffentlicht: (2024)
PromptEnhancer: A Simple Approach to Enhance Text-to-Image Models via Chain-of-Thought Prompt Rewriting
von: Wang, Linqing, et al.
Veröffentlicht: (2025)
von: Wang, Linqing, et al.
Veröffentlicht: (2025)
Prompt2Fashion: An automatically generated fashion dataset
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
von: Argyrou, Georgia, et al.
Veröffentlicht: (2024)
SSG-Dit: A Spatial Signal Guided Framework for Controllable Video Generation
von: Hu, Peng, et al.
Veröffentlicht: (2025)
von: Hu, Peng, et al.
Veröffentlicht: (2025)
StyleSSP: Sampling StartPoint Enhancement for Training-free Diffusion-based Method for Style Transfer
von: Xu, Ruojun, et al.
Veröffentlicht: (2025)
von: Xu, Ruojun, et al.
Veröffentlicht: (2025)
Toward Generalist Anomaly Detection via In-context Residual Learning with Few-shot Sample Prompts
von: Zhu, Jiawen, et al.
Veröffentlicht: (2024)
von: Zhu, Jiawen, et al.
Veröffentlicht: (2024)
TransUNext: towards a more advanced U-shaped framework for automatic vessel segmentation in the fundus image
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Multi-view learning for automatic classification of multi-wavelength auroral images
von: Yang, Qiuju, et al.
Veröffentlicht: (2023)
von: Yang, Qiuju, et al.
Veröffentlicht: (2023)
Physics-informed simulation framework for realistic sonar image generation and statistical validation
von: S, Kamal Basha, et al.
Veröffentlicht: (2026)
von: S, Kamal Basha, et al.
Veröffentlicht: (2026)
OCCO: LVM-guided Infrared and Visible Image Fusion Framework based on Object-aware and Contextual COntrastive Learning
von: Li, Hui, et al.
Veröffentlicht: (2025)
von: Li, Hui, et al.
Veröffentlicht: (2025)
SSP-GNN: Learning to Track via Bilevel Optimization
von: Golias, Griffin, et al.
Veröffentlicht: (2024)
von: Golias, Griffin, et al.
Veröffentlicht: (2024)
AICL: Action In-Context Learning for Video Diffusion Model
von: Liu, Jianzhi, et al.
Veröffentlicht: (2024)
von: Liu, Jianzhi, et al.
Veröffentlicht: (2024)
ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models
von: Song, Zirui, et al.
Veröffentlicht: (2025)
von: Song, Zirui, et al.
Veröffentlicht: (2025)
Exploring Kolmogorov-Arnold networks for realistic image sharpness assessment
von: Yu, Shaode, et al.
Veröffentlicht: (2024)
von: Yu, Shaode, et al.
Veröffentlicht: (2024)
Knowledge-Guided Prompt Learning for Deepfake Facial Image Detection
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
Fine-grained Abnormality Prompt Learning for Zero-shot Anomaly Detection
von: Zhu, Jiawen, et al.
Veröffentlicht: (2024)
von: Zhu, Jiawen, et al.
Veröffentlicht: (2024)
Toward distortion-aware change detection in realistic scenarios
von: Zhao, Yitao, et al.
Veröffentlicht: (2024)
von: Zhao, Yitao, et al.
Veröffentlicht: (2024)
Safe-SD: Safe and Traceable Stable Diffusion with Text Prompt Trigger for Invisible Generative Watermarking
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
Improving image synthesis with diffusion-negative sampling
von: Desai, Alakh, et al.
Veröffentlicht: (2024)
von: Desai, Alakh, et al.
Veröffentlicht: (2024)
SUTrack: Towards Simple and Unified Single Object Tracking
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
FACT: A Simple and Efficient Framework for Active Finetuning
von: Xu, Wenshuai, et al.
Veröffentlicht: (2026)
von: Xu, Wenshuai, et al.
Veröffentlicht: (2026)
SimpleProc: Fully Procedural Synthetic Data from Simple Rules for Multi-View Stereo
von: Ma, Zeyu, et al.
Veröffentlicht: (2026)
von: Ma, Zeyu, et al.
Veröffentlicht: (2026)
Universal Prompt Optimizer for Safe Text-to-Image Generation
von: Wu, Zongyu, et al.
Veröffentlicht: (2024)
von: Wu, Zongyu, et al.
Veröffentlicht: (2024)
SSP-RACL: Classification of Noisy Fundus Images with Self-Supervised Pretraining and Robust Adaptive Credal Loss
von: Ye, Mengwen, et al.
Veröffentlicht: (2024)
von: Ye, Mengwen, et al.
Veröffentlicht: (2024)
A Simple and Effective Point-based Network for Event Camera 6-DOFs Pose Relocalization
von: Ren, Hongwei, et al.
Veröffentlicht: (2024)
von: Ren, Hongwei, et al.
Veröffentlicht: (2024)
DCPT: Darkness Clue-Prompted Tracking in Nighttime UAVs
von: Zhu, Jiawen, et al.
Veröffentlicht: (2023)
von: Zhu, Jiawen, et al.
Veröffentlicht: (2023)
Leveraging a realistic synthetic database to learn Shape-from-Shading for estimating the colon depth in colonoscopy images
von: Ruano, Josué, et al.
Veröffentlicht: (2023)
von: Ruano, Josué, et al.
Veröffentlicht: (2023)
Towards synthetic generation of realistic wooden logs
von: Zolotarev, Fedor, et al.
Veröffentlicht: (2025)
von: Zolotarev, Fedor, et al.
Veröffentlicht: (2025)
SimpleFusion: A Simple Fusion Framework for Infrared and Visible Images
von: Chen, Ming, et al.
Veröffentlicht: (2024)
von: Chen, Ming, et al.
Veröffentlicht: (2024)
NSFW-Classifier Guided Prompt Sanitization for Safe Text-to-Image Generation
von: Xie, Yu, et al.
Veröffentlicht: (2025)
von: Xie, Yu, et al.
Veröffentlicht: (2025)
Generative Technology for Human Emotion Recognition: A Scope Review
von: Ma, Fei, et al.
Veröffentlicht: (2024)
von: Ma, Fei, et al.
Veröffentlicht: (2024)
ArtisanGS: Interactive Tools for Gaussian Splat Selection with AI and Human in the Loop
von: Tsang, Clement Fuji, et al.
Veröffentlicht: (2026)
von: Tsang, Clement Fuji, et al.
Veröffentlicht: (2026)
Segmentation by registration-enabled SAM prompt engineering using five reference images
von: Chen, Yaxi, et al.
Veröffentlicht: (2024)
von: Chen, Yaxi, et al.
Veröffentlicht: (2024)
TIPS Over Tricks: Simple Prompts for Effective Zero-shot Anomaly Detection
von: Salehi, Alireza, et al.
Veröffentlicht: (2026)
von: Salehi, Alireza, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SSP-SAM: SAM with Semantic-Spatial Prompt for Referring Expression Segmentation
von: Tang, Wei, et al.
Veröffentlicht: (2026) -
F3-Pruning: A Training-Free and Generalized Pruning Strategy towards Faster and Finer Text-to-Video Synthesis
von: Su, Sitong, et al.
Veröffentlicht: (2023) -
Road Rage Reasoning with Vision-language Models (VLMs): Task Definition and Evaluation Dataset
von: Weng, Yibing, et al.
Veröffentlicht: (2025) -
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
von: Jing, Zonglei, et al.
Veröffentlicht: (2025) -
SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration
von: Zhang, Yuhong, et al.
Veröffentlicht: (2024)