Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Pei-Chi, Yao, Yi, Hsu, Chan-Feng, Xie, HongXia, Chen, Hung-Jen, Shuai, Hong-Han, Cheng, Wen-Huang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025)
por: Raoufi, Behnam, et al.
Publicado: (2025)
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
por: Deng, Pei, et al.
Publicado: (2025)
por: Deng, Pei, et al.
Publicado: (2025)
Semi supervised GAN for smart microscopy, fast and data efficient cell cycle classification
por: Manick, Rajeev, et al.
Publicado: (2026)
por: Manick, Rajeev, et al.
Publicado: (2026)
EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis
por: Guo, Yijie, et al.
Publicado: (2025)
por: Guo, Yijie, et al.
Publicado: (2025)
CoMatcher: Multi-View Collaborative Feature Matching
por: Zhang, Jintao, et al.
Publicado: (2025)
por: Zhang, Jintao, et al.
Publicado: (2025)
Visible Iris Area as a Quality Metric for Reliable Iris Recognition Under Pupil Dilation and Eyelid Occlusion
por: Pessaud, Jack, et al.
Publicado: (2025)
por: Pessaud, Jack, et al.
Publicado: (2025)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
por: Durrani, Hamza Ahmed, et al.
Publicado: (2026)
por: Durrani, Hamza Ahmed, et al.
Publicado: (2026)
From eye to AI: studying rodent social behavior in the era of machine Learning
por: Chindemi, Giuseppe, et al.
Publicado: (2025)
por: Chindemi, Giuseppe, et al.
Publicado: (2025)
Single-Shot Metric Depth from Focused Plenoptic Cameras
por: Lasheras-Hernandez, Blanca, et al.
Publicado: (2024)
por: Lasheras-Hernandez, Blanca, et al.
Publicado: (2024)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
por: Duguay, Simon-Olivier, et al.
Publicado: (2026)
por: Duguay, Simon-Olivier, et al.
Publicado: (2026)
Optimizing the image correction pipeline for pedestrian detection in the thermal-infrared domain
por: Karam, Christophe, et al.
Publicado: (2024)
por: Karam, Christophe, et al.
Publicado: (2024)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
por: Wang, Yiming, et al.
Publicado: (2026)
por: Wang, Yiming, et al.
Publicado: (2026)
UnCageNet: Tracking and Pose Estimation of Caged Animal
por: Dutta, Sayak, et al.
Publicado: (2025)
por: Dutta, Sayak, et al.
Publicado: (2025)
A Light Perspective for 3D Object Detection
por: Pederiva, Marcelo Eduardo, et al.
Publicado: (2025)
por: Pederiva, Marcelo Eduardo, et al.
Publicado: (2025)
Mistake Attribution: Fine-Grained Mistake Understanding in Egocentric Videos
por: Li, Yayuan, et al.
Publicado: (2025)
por: Li, Yayuan, et al.
Publicado: (2025)
Combining Absolute and Semi-Generalized Relative Poses for Visual Localization
por: Panek, Vojtech, et al.
Publicado: (2024)
por: Panek, Vojtech, et al.
Publicado: (2024)
A Guide to Structureless Visual Localization
por: Panek, Vojtech, et al.
Publicado: (2025)
por: Panek, Vojtech, et al.
Publicado: (2025)
SelvaBox: A high-resolution dataset for tropical tree crown detection
por: Baudchon, Hugo, et al.
Publicado: (2025)
por: Baudchon, Hugo, et al.
Publicado: (2025)
Facial Attribute Based Text Guided Face Anonymization
por: Muştu, Mustafa İzzet, et al.
Publicado: (2025)
por: Muştu, Mustafa İzzet, et al.
Publicado: (2025)
SPMamba-YOLO: An Underwater Object Detection Network Based on Multi-Scale Feature Enhancement and Global Context Modeling
por: Liao, Guanghao, et al.
Publicado: (2026)
por: Liao, Guanghao, et al.
Publicado: (2026)
Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion
por: Zhu, Yu, et al.
Publicado: (2025)
por: Zhu, Yu, et al.
Publicado: (2025)
Image-Based Leopard Seal Recognition: Approaches and Challenges in Current Automated Systems
por: Salazar, Jorge Yero, et al.
Publicado: (2024)
por: Salazar, Jorge Yero, et al.
Publicado: (2024)
Dense Motion Captioning
por: Xu, Shiyao, et al.
Publicado: (2025)
por: Xu, Shiyao, et al.
Publicado: (2025)
TD3Net: A temporal densely connected multi-dilated convolutional network for lipreading
por: Lee, Byung Hoon, et al.
Publicado: (2025)
por: Lee, Byung Hoon, et al.
Publicado: (2025)
MoDE: Mixture of Diffusion Experts for Any Occluded Face Recognition
por: Fan, Qiannan, et al.
Publicado: (2025)
por: Fan, Qiannan, et al.
Publicado: (2025)
NeuroGaze-Distill: Brain-informed Distillation and Depression-Inspired Geometric Priors for Robust Facial Emotion Recognition
por: Li, Zilin, et al.
Publicado: (2025)
por: Li, Zilin, et al.
Publicado: (2025)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
por: Panek, Vojtech, et al.
Publicado: (2026)
por: Panek, Vojtech, et al.
Publicado: (2026)
Prompt Sensitivity in Vision-Language Grounding: How Small Changes in Wording Affect Object Detection
por: Deka, Dawar Jyoti, et al.
Publicado: (2026)
por: Deka, Dawar Jyoti, et al.
Publicado: (2026)
Dynamic Arthroscopic Navigation System for Anterior Cruciate Ligament Reconstruction Based on Multi-level Memory Architecture
por: Wang, Shuo, et al.
Publicado: (2025)
por: Wang, Shuo, et al.
Publicado: (2025)
OmniAcc: Personalized Accessibility Assistant Using Generative AI
por: Karki, Siddhant, et al.
Publicado: (2025)
por: Karki, Siddhant, et al.
Publicado: (2025)
Towards Accurate and Efficient Waste Image Classification: A Hybrid Deep Learning and Machine Learning Approach
por: Nguyen, Ngoc-Bao-Quang, et al.
Publicado: (2025)
por: Nguyen, Ngoc-Bao-Quang, et al.
Publicado: (2025)
A Vision-Language Model for Focal Liver Lesion Classification
por: Jian, Song, et al.
Publicado: (2025)
por: Jian, Song, et al.
Publicado: (2025)
PhysVideoGenerator: Towards Physically Aware Video Generation via Latent Physics Guidance
por: Satish, Siddarth Nilol Kundur, et al.
Publicado: (2026)
por: Satish, Siddarth Nilol Kundur, et al.
Publicado: (2026)
Deep Learning-based Depth Estimation Methods from Monocular Image and Videos: A Comprehensive Survey
por: Rajapaksha, Uchitha, et al.
Publicado: (2024)
por: Rajapaksha, Uchitha, et al.
Publicado: (2024)
SimWorld: A Unified Benchmark for Simulator-Conditioned Scene Generation via World Model
por: Li, Xinqing, et al.
Publicado: (2025)
por: Li, Xinqing, et al.
Publicado: (2025)
Exploring Surround-View Fisheye Camera 3D Object Detection
por: Li, Changcai, et al.
Publicado: (2025)
por: Li, Changcai, et al.
Publicado: (2025)
High-Frequency Semantics and Geometric Priors for End-to-End Detection Transformers in Challenging UAV Imagery
por: Peng, Hongxing, et al.
Publicado: (2025)
por: Peng, Hongxing, et al.
Publicado: (2025)
VLM-VPI: A Vision-Language Reasoning Framework for Improving Automated Vehicle-Pedestrian Interactions
por: Pu, Qingwen, et al.
Publicado: (2026)
por: Pu, Qingwen, et al.
Publicado: (2026)
Capacity Constraint Analysis Using Object Detection for Smart Manufacturing
por: Ahmad, Hafiz Mughees, et al.
Publicado: (2024)
por: Ahmad, Hafiz Mughees, et al.
Publicado: (2024)
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
por: Ahmad, Hafiz Mughees, et al.
Publicado: (2024)
por: Ahmad, Hafiz Mughees, et al.
Publicado: (2024)
Ejemplares similares
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025) -
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
por: Deng, Pei, et al.
Publicado: (2025) -
Semi supervised GAN for smart microscopy, fast and data efficient cell cycle classification
por: Manick, Rajeev, et al.
Publicado: (2026) -
EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis
por: Guo, Yijie, et al.
Publicado: (2025) -
CoMatcher: Multi-View Collaborative Feature Matching
por: Zhang, Jintao, et al.
Publicado: (2025)