Learning Gaussian Data Augmentation in Feature Space for One-shot Object Detection in Manga
Fuente:
arXiv
Guardado en:
| Autores principales: | Taniguchi, Takara, Furuta, Ryosuke |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MangaUB: A Manga Understanding Benchmark for Large Multimodal Models
por: Ikuta, Hikaru, et al.
Publicado: (2024)
por: Ikuta, Hikaru, et al.
Publicado: (2024)
EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching
por: Taniguchi, Takara, et al.
Publicado: (2026)
por: Taniguchi, Takara, et al.
Publicado: (2026)
Inference-time Trajectory Optimization for Manga Image Editing
por: Furuta, Ryosuke
Publicado: (2026)
por: Furuta, Ryosuke
Publicado: (2026)
SFFNet: Synergistic Feature Fusion Network With Dual-Domain Edge Enhancement for UAV Image Object Detection
por: Zhang, Wenfeng, et al.
Publicado: (2026)
por: Zhang, Wenfeng, et al.
Publicado: (2026)
OneHOI: Unifying Human-Object Interaction Generation and Editing
por: Hoe, Jiun Tian, et al.
Publicado: (2026)
por: Hoe, Jiun Tian, et al.
Publicado: (2026)
OT-DETECTOR: Delving into Optimal Transport for Zero-shot Out-of-Distribution Detection
por: Liu, Yu, et al.
Publicado: (2025)
por: Liu, Yu, et al.
Publicado: (2025)
Do LLMs Understand Visual Anomalies? Uncovering LLM's Capabilities in Zero-shot Anomaly Detection
por: Zhu, Jiaqi, et al.
Publicado: (2024)
por: Zhu, Jiaqi, et al.
Publicado: (2024)
Dual Mutual Learning Network with Global-local Awareness for RGB-D Salient Object Detection
por: Yi, Kang, et al.
Publicado: (2025)
por: Yi, Kang, et al.
Publicado: (2025)
Spatial-Temporal Human-Object Interaction Detection
por: Sun, Xu, et al.
Publicado: (2025)
por: Sun, Xu, et al.
Publicado: (2025)
DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection
por: Guo, Junjie, et al.
Publicado: (2024)
por: Guo, Junjie, et al.
Publicado: (2024)
SmartSplat: Feature-Smart Gaussians for Scalable Compression of Ultra-High-Resolution Images
por: Li, Linfei, et al.
Publicado: (2025)
por: Li, Linfei, et al.
Publicado: (2025)
Compression of 3D Gaussian Splatting with Optimized Feature Planes and Standard Video Codecs
por: Lee, Soonbin, et al.
Publicado: (2025)
por: Lee, Soonbin, et al.
Publicado: (2025)
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection
por: Cheung, Tsun-Hin, et al.
Publicado: (2024)
por: Cheung, Tsun-Hin, et al.
Publicado: (2024)
Other Tokens Matter: Exploring Global and Local Features of Vision Transformers for Object Re-Identification
por: Wang, Yingquan, et al.
Publicado: (2024)
por: Wang, Yingquan, et al.
Publicado: (2024)
On the Robustness of Human-Object Interaction Detection against Distribution Shift
por: Xie, Chi, et al.
Publicado: (2025)
por: Xie, Chi, et al.
Publicado: (2025)
REArtGS: Reconstructing and Generating Articulated Objects via 3D Gaussian Splatting with Geometric and Motion Constraints
por: Wu, Di, et al.
Publicado: (2025)
por: Wu, Di, et al.
Publicado: (2025)
GAIA: Zero-shot Talking Avatar Generation
por: He, Tianyu, et al.
Publicado: (2023)
por: He, Tianyu, et al.
Publicado: (2023)
Modularized Zero-shot VQA with Pre-trained Models
por: Cao, Rui, et al.
Publicado: (2023)
por: Cao, Rui, et al.
Publicado: (2023)
Accelerated Event-Based Feature Detection and Compression for Surveillance Video Systems
por: Freeman, Andrew C., et al.
Publicado: (2023)
por: Freeman, Andrew C., et al.
Publicado: (2023)
Divide-and-Conquer: Confluent Triple-Flow Network for RGB-T Salient Object Detection
por: Tang, Hao, et al.
Publicado: (2024)
por: Tang, Hao, et al.
Publicado: (2024)
Multi-Scale and Detail-Enhanced Segment Anything Model for Salient Object Detection
por: Gao, Shixuan, et al.
Publicado: (2024)
por: Gao, Shixuan, et al.
Publicado: (2024)
HMPE:HeatMap Embedding for Efficient Transformer-Based Small Object Detection
por: Zeng, YangChen
Publicado: (2025)
por: Zeng, YangChen
Publicado: (2025)
WaveMamba: Wavelet-Driven Mamba Fusion for RGB-Infrared Object Detection
por: Zhu, Haodong, et al.
Publicado: (2025)
por: Zhu, Haodong, et al.
Publicado: (2025)
Selective Vision-Language Subspace Projection for Few-shot CLIP
por: Zhu, Xingyu, et al.
Publicado: (2024)
por: Zhu, Xingyu, et al.
Publicado: (2024)
Pseudo-triplet Guided Few-shot Composed Image Retrieval
por: Hou, Bohan, et al.
Publicado: (2024)
por: Hou, Bohan, et al.
Publicado: (2024)
GMFVAD: Using Grained Multi-modal Feature to Improve Video Anomaly Detection
por: Dai, Guangyu, et al.
Publicado: (2025)
por: Dai, Guangyu, et al.
Publicado: (2025)
$\mathbf{C}^2$Former: Calibrated and Complementary Transformer for RGB-Infrared Object Detection
por: Yuan, Maoxun, et al.
Publicado: (2023)
por: Yuan, Maoxun, et al.
Publicado: (2023)
Efficiently Collecting Training Dataset for 2D Object Detection by Online Visual Feedback
por: Kiyokawa, Takuya, et al.
Publicado: (2023)
por: Kiyokawa, Takuya, et al.
Publicado: (2023)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
por: Zhang, Zhenxing, et al.
Publicado: (2024)
por: Zhang, Zhenxing, et al.
Publicado: (2024)
MU-MAE: Multimodal Masked Autoencoders-Based One-Shot Learning
por: Liu, Rex, et al.
Publicado: (2024)
por: Liu, Rex, et al.
Publicado: (2024)
SM3Det: A Unified Model for Multi-Modal Remote Sensing Object Detection
por: Li, Yuxuan, et al.
Publicado: (2024)
por: Li, Yuxuan, et al.
Publicado: (2024)
Segmentation-Based Attention Entropy: Detecting and Mitigating Object Hallucinations in Large Vision-Language Models
por: Song, Jiale, et al.
Publicado: (2026)
por: Song, Jiale, et al.
Publicado: (2026)
A Simple Task-aware Contrastive Local Descriptor Selection Strategy for Few-shot Learning between inter class and intra class
por: Qiao, Qian, et al.
Publicado: (2024)
por: Qiao, Qian, et al.
Publicado: (2024)
Efficient Object-centric Representation Learning with Pre-trained Geometric Prior
por: Khac, Phúc H. Le, et al.
Publicado: (2024)
por: Khac, Phúc H. Le, et al.
Publicado: (2024)
KAN-SAM: Kolmogorov-Arnold Network Guided Segment Anything Model for RGB-T Salient Object Detection
por: Li, Xingyuan, et al.
Publicado: (2025)
por: Li, Xingyuan, et al.
Publicado: (2025)
Interpretable Zero-shot Referring Expression Comprehension with Query-driven Scene Graphs
por: Wu, Yike, et al.
Publicado: (2026)
por: Wu, Yike, et al.
Publicado: (2026)
Augment Before Copy-Paste: Data and Memory Efficiency-Oriented Instance Segmentation Framework for Sport-scenes
por: Hsu, Chih-Chung, et al.
Publicado: (2024)
por: Hsu, Chih-Chung, et al.
Publicado: (2024)
Multi-Modal Image Fusion via Intervention-Stable Feature Learning
por: Wang, Xue, et al.
Publicado: (2026)
por: Wang, Xue, et al.
Publicado: (2026)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
por: Lian, Zheng, et al.
Publicado: (2023)
por: Lian, Zheng, et al.
Publicado: (2023)
MultiColor: Image Colorization by Learning from Multiple Color Spaces
por: Du, Xiangcheng, et al.
Publicado: (2024)
por: Du, Xiangcheng, et al.
Publicado: (2024)
Ejemplares similares
-
MangaUB: A Manga Understanding Benchmark for Large Multimodal Models
por: Ikuta, Hikaru, et al.
Publicado: (2024) -
EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching
por: Taniguchi, Takara, et al.
Publicado: (2026) -
Inference-time Trajectory Optimization for Manga Image Editing
por: Furuta, Ryosuke
Publicado: (2026) -
SFFNet: Synergistic Feature Fusion Network With Dual-Domain Edge Enhancement for UAV Image Object Detection
por: Zhang, Wenfeng, et al.
Publicado: (2026) -
OneHOI: Unifying Human-Object Interaction Generation and Editing
por: Hoe, Jiun Tian, et al.
Publicado: (2026)