Noise-Robust Tiny Object Localization with Flows
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Huixin, Yang, Linlin, Chen, Ronyu, Gu, Kerui, Zhang, Baochang, Yao, Angela, Cao, Xianbin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Uncertainty-Aware Gradient Stabilization for Small Object Detection
by: Sun, Huixin, et al.
Published: (2023)
by: Sun, Huixin, et al.
Published: (2023)
P4Q: Learning to Prompt for Quantization in Visual-language Models
by: Sun, Huixin, et al.
Published: (2024)
by: Sun, Huixin, et al.
Published: (2024)
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
by: Wang, Ziming, et al.
Published: (2023)
by: Wang, Ziming, et al.
Published: (2023)
Online Test-time Adaptation for 3D Human Pose Estimation: A Practical Perspective with Estimated 2D Poses
by: Lin, Qiuxia, et al.
Published: (2025)
by: Lin, Qiuxia, et al.
Published: (2025)
TinyFormer: Preserving Tiny Objects in YOLO-DETR Hybrid Real-time Detectors
by: Hsieh, Jun-Wei, et al.
Published: (2026)
by: Hsieh, Jun-Wei, et al.
Published: (2026)
AMLRIS: Alignment-aware Masked Learning for Referring Image Segmentation
by: Chen, Tongfei, et al.
Published: (2026)
by: Chen, Tongfei, et al.
Published: (2026)
Scale-Aware Relay and Scale-Adaptive Loss for Tiny Object Detection in Aerial Images
by: Li, Jinfu, et al.
Published: (2025)
by: Li, Jinfu, et al.
Published: (2025)
Heterogeneous Graph Transformer for Multiple Tiny Object Tracking in RGB-T Videos
by: Xu, Qingyu, et al.
Published: (2024)
by: Xu, Qingyu, et al.
Published: (2024)
Localization, balance and affinity: a stronger multifaceted collaborative salient object detector in remote sensing images
by: Xie, Yakun, et al.
Published: (2024)
by: Xie, Yakun, et al.
Published: (2024)
Fusion-Mamba for Cross-modality Object Detection
by: Dong, Wenhao, et al.
Published: (2024)
by: Dong, Wenhao, et al.
Published: (2024)
TinyDrop: Tiny Model Guided Token Dropping for Vision Transformers
by: Wang, Guoxin, et al.
Published: (2025)
by: Wang, Guoxin, et al.
Published: (2025)
T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models
by: Yang, Jindong, et al.
Published: (2025)
by: Yang, Jindong, et al.
Published: (2025)
Learning domain-invariant features through channel-level sparsification for Out-Of Distribution Generalization
by: Pei, Haoran, et al.
Published: (2026)
by: Pei, Haoran, et al.
Published: (2026)
AnchorDS: Anchoring Dynamic Sources for Semantically Consistent Text-to-3D Generation
by: Zhu, Jiayin, et al.
Published: (2025)
by: Zhu, Jiayin, et al.
Published: (2025)
Can Vision-Language Models be a Good Guesser? Exploring VLMs for Times and Location Reasoning
by: Zhang, Gengyuan, et al.
Published: (2023)
by: Zhang, Gengyuan, et al.
Published: (2023)
KITRO: Refining Human Mesh by 2D Clues and Kinematic-tree Rotation
by: Yang, Fengyuan, et al.
Published: (2024)
by: Yang, Fengyuan, et al.
Published: (2024)
Robust Tiny Object Detection in Aerial Images amidst Label Noise
by: Zhu, Haoran, et al.
Published: (2024)
by: Zhu, Haoran, et al.
Published: (2024)
COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection
by: Peng, Peiran, et al.
Published: (2025)
by: Peng, Peiran, et al.
Published: (2025)
A Channel-ensemble Approach: Unbiased and Low-variance Pseudo-labels is Critical for Semi-supervised Classification
by: Wu, Jiaqi, et al.
Published: (2024)
by: Wu, Jiaqi, et al.
Published: (2024)
Image Forgery Localization via Guided Noise and Multi-Scale Feature Aggregation
by: Niu, Yakun, et al.
Published: (2024)
by: Niu, Yakun, et al.
Published: (2024)
APLA: Additional Perturbation for Latent Noise with Adversarial Training Enables Consistency
by: Yao, Yupu, et al.
Published: (2023)
by: Yao, Yupu, et al.
Published: (2023)
StreamTinyNet: video streaming analysis with spatial-temporal TinyML
by: Shalby, Hazem Hesham Yousef, et al.
Published: (2024)
by: Shalby, Hazem Hesham Yousef, et al.
Published: (2024)
RELO: Reinforcement Learning to Localize for Visual Object Tracking
by: Chen, Xin, et al.
Published: (2026)
by: Chen, Xin, et al.
Published: (2026)
Agentic Surgical AI: Surgeon Style Fingerprinting and Privacy Risk Quantification via Discrete Diffusion in a Vision-Language-Action Framework
by: Zhan, Huixin, et al.
Published: (2025)
by: Zhan, Huixin, et al.
Published: (2025)
T2T-VICL: Unlocking the Boundaries of Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs
by: Xia, Shao-Jun, et al.
Published: (2025)
by: Xia, Shao-Jun, et al.
Published: (2025)
TinySAM 2: Extreme Memory Compression for Efficient Track Anything Model
by: Ding, Zhaoyuan, et al.
Published: (2026)
by: Ding, Zhaoyuan, et al.
Published: (2026)
TinyVLM: Zero-Shot Object Detection on Microcontrollers via Vision-Language Distillation with Matryoshka Embeddings
by: Wilson, Bibin
Published: (2026)
by: Wilson, Bibin
Published: (2026)
Prompt as Knowledge Bank: Boost Vision-language model via Structural Representation for zero-shot medical detection
by: Yang, Yuguang, et al.
Published: (2025)
by: Yang, Yuguang, et al.
Published: (2025)
An Empirical Study on the Robustness of YOLO Models for Underwater Object Detection
by: Nabahirwa, Edwine, et al.
Published: (2025)
by: Nabahirwa, Edwine, et al.
Published: (2025)
Is Your Text-to-Image Model Robust to Caption Noise?
by: Yu, Weichen, et al.
Published: (2024)
by: Yu, Weichen, et al.
Published: (2024)
Object-fabrication Targeted Attack for Object Detection
by: Zhang, Xuchong, et al.
Published: (2022)
by: Zhang, Xuchong, et al.
Published: (2022)
Decom--CAM: Tell Me What You See, In Details! Feature-Level Interpretation via Decomposition Class Activation Map
by: Yang, Yuguang, et al.
Published: (2023)
by: Yang, Yuguang, et al.
Published: (2023)
SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection
by: Tan, Wanying, et al.
Published: (2026)
by: Tan, Wanying, et al.
Published: (2026)
HRGS: Hierarchical Gaussian Splatting for Memory-Efficient High-Resolution 3D Reconstruction
by: Li, Changbai, et al.
Published: (2025)
by: Li, Changbai, et al.
Published: (2025)
Robust Domain Generalization for Multi-modal Object Recognition
by: Qiao, Yuxin, et al.
Published: (2024)
by: Qiao, Yuxin, et al.
Published: (2024)
FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation
by: Zeng, Huajian, et al.
Published: (2026)
by: Zeng, Huajian, et al.
Published: (2026)
From Noisy Labels to Intrinsic Structure: A Geometric-Structural Dual-Guided Framework for Noise-Robust Medical Image Segmentation
by: Wang, Tao, et al.
Published: (2025)
by: Wang, Tao, et al.
Published: (2025)
NDM: A Noise-driven Detection and Mitigation Framework against Implicit Sexual Intentions in Text-to-Image Generation
by: Sun, Yitong, et al.
Published: (2025)
by: Sun, Yitong, et al.
Published: (2025)
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
by: Li, Wenxuan, et al.
Published: (2026)
by: Li, Wenxuan, et al.
Published: (2026)
Improved Noise Schedule for Diffusion Training
by: Hang, Tiankai, et al.
Published: (2024)
by: Hang, Tiankai, et al.
Published: (2024)
Similar Items
-
Uncertainty-Aware Gradient Stabilization for Small Object Detection
by: Sun, Huixin, et al.
Published: (2023) -
P4Q: Learning to Prompt for Quantization in Visual-language Models
by: Sun, Huixin, et al.
Published: (2024) -
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
by: Wang, Ziming, et al.
Published: (2023) -
Online Test-time Adaptation for 3D Human Pose Estimation: A Practical Perspective with Estimated 2D Poses
by: Lin, Qiuxia, et al.
Published: (2025) -
TinyFormer: Preserving Tiny Objects in YOLO-DETR Hybrid Real-time Detectors
by: Hsieh, Jun-Wei, et al.
Published: (2026)