Advanced Feature Manipulation for Enhanced Change Detection Leveraging Natural Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Zhenglin, Huang, Yangchen, Zhu, Mengran, Zhang, Jingyu, Chang, JingHao, Liu, Houze |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Research on the Application of Computer Vision Based on Deep Learning in Autonomous Driving Technology
by: Zhang, Jingyu, et al.
Published: (2024)
by: Zhang, Jingyu, et al.
Published: (2024)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
A Weak-Signal-Aware Framework for Subsurface Defect Detection: Mechanisms for Enhancing Low-SCR Hyperbolic Signatures
by: Zhang, Wenbo, et al.
Published: (2026)
by: Zhang, Wenbo, et al.
Published: (2026)
Change3D: Revisiting Change Detection and Captioning from A Video Modeling Perspective
by: Zhu, Duowang, et al.
Published: (2025)
by: Zhu, Duowang, et al.
Published: (2025)
Group Critical-token Policy Optimization for Autoregressive Image Generation
by: Zhang, Guohui, et al.
Published: (2025)
by: Zhang, Guohui, et al.
Published: (2025)
Uni-RCM: Unified Reference-guided Cross-modal Mapping for Multi-Class Anomaly Detection
by: Wu, Yangchen, et al.
Published: (2026)
by: Wu, Yangchen, et al.
Published: (2026)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
by: Trusca, Maria Mihaela, et al.
Published: (2024)
by: Trusca, Maria Mihaela, et al.
Published: (2024)
Leveraging Large Language Models for Multimodal Search
by: Barbany, Oriol, et al.
Published: (2024)
by: Barbany, Oriol, et al.
Published: (2024)
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism
by: Zhang, Congzhi, et al.
Published: (2025)
by: Zhang, Congzhi, et al.
Published: (2025)
Enhancing Breast Cancer Detection with Vision Transformers and Graph Neural Networks
by: Cai, Yeming, et al.
Published: (2025)
by: Cai, Yeming, et al.
Published: (2025)
Leveraging Geometric Priors for Unaligned Scene Change Detection
by: Liu, Ziling, et al.
Published: (2025)
by: Liu, Ziling, et al.
Published: (2025)
Skeleton-OOD: An End-to-End Skeleton-Based Model for Robust Out-of-Distribution Human Action Detection
by: Xu, Jing, et al.
Published: (2024)
by: Xu, Jing, et al.
Published: (2024)
Integrated Dynamic Phenological Feature for Remote Sensing Image Land Cover Change Detection
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
Sakuga-42M Dataset: Scaling Up Cartoon Research
by: Pan, Zhenglin
Published: (2024)
by: Pan, Zhenglin
Published: (2024)
LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation
by: Yue, Tongtian, et al.
Published: (2025)
by: Yue, Tongtian, et al.
Published: (2025)
LG-CD: Enhancing Language-Guided Change Detection through SAM2 Adaptation
by: Liu, Yixiao, et al.
Published: (2025)
by: Liu, Yixiao, et al.
Published: (2025)
Research on Deep Learning Model of Feature Extraction Based on Convolutional Neural Network
by: Liu, Houze, et al.
Published: (2024)
by: Liu, Houze, et al.
Published: (2024)
SO-DETR: Leveraging Dual-Domain Features and Knowledge Distillation for Small Object Detection
by: Zhang, Huaxiang, et al.
Published: (2025)
by: Zhang, Huaxiang, et al.
Published: (2025)
CADRef: Robust Out-of-Distribution Detection via Class-Aware Decoupled Relative Feature Leveraging
by: Ling, Zhiwei, et al.
Published: (2025)
by: Ling, Zhiwei, et al.
Published: (2025)
CFPFormer: Feature-pyramid like Transformer Decoder for Segmentation and Detection
by: Cai, Hongyi, et al.
Published: (2024)
by: Cai, Hongyi, et al.
Published: (2024)
ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding
by: Sun, Dongwei, et al.
Published: (2026)
by: Sun, Dongwei, et al.
Published: (2026)
A New Benchmark and Model for Challenging Image Manipulation Detection
by: Zhang, Zhenfei, et al.
Published: (2023)
by: Zhang, Zhenfei, et al.
Published: (2023)
Exploiting Modality-Specific Features For Multi-Modal Manipulation Detection And Grounding
by: Wang, Jiazhen, et al.
Published: (2023)
by: Wang, Jiazhen, et al.
Published: (2023)
Spatial-DISE: A Unified Benchmark for Evaluating Spatial Reasoning in Vision-Language Models
by: Huang, Xinmiao, et al.
Published: (2025)
by: Huang, Xinmiao, et al.
Published: (2025)
Leveraging Fine-Grained Information and Noise Decoupling for Remote Sensing Change Detection
by: Du, Qiangang, et al.
Published: (2024)
by: Du, Qiangang, et al.
Published: (2024)
PhyUnfold-Net: Advancing Remote Sensing Change Detection with Physics-Guided Deep Unfolding
by: Lei, Zelin, et al.
Published: (2026)
by: Lei, Zelin, et al.
Published: (2026)
ChangeViT: Unleashing Plain Vision Transformers for Change Detection
by: Zhu, Duowang, et al.
Published: (2024)
by: Zhu, Duowang, et al.
Published: (2024)
Bring Remote Sensing Object Detect Into Nature Language Model: Using SFT Method
by: Wang, Fei, et al.
Published: (2025)
by: Wang, Fei, et al.
Published: (2025)
BusterX++: Towards Unified Cross-Modal AI-Generated Content Detection and Explanation with MLLM
by: Wen, Haiquan, et al.
Published: (2025)
by: Wen, Haiquan, et al.
Published: (2025)
Glass Surface Detection: Leveraging Reflection Dynamics in Flash/No-flash Imagery
by: Yan, Tao, et al.
Published: (2025)
by: Yan, Tao, et al.
Published: (2025)
Towards Explainable Bilingual Multimodal Misinformation Detection and Localization
by: He, Yiwei, et al.
Published: (2025)
by: He, Yiwei, et al.
Published: (2025)
Enhancing Medical Image Segmentation with Deep Learning and Diffusion Models
by: Liu, Houze, et al.
Published: (2024)
by: Liu, Houze, et al.
Published: (2024)
LASFNet: A Lightweight Attention-Guided Self-Modulation Feature Fusion Network for Multimodal Object Detection
by: Hao, Lei, et al.
Published: (2025)
by: Hao, Lei, et al.
Published: (2025)
Leveraging Semantic Cues from Foundation Vision Models for Enhanced Local Feature Correspondence
by: Cadar, Felipe, et al.
Published: (2024)
by: Cadar, Felipe, et al.
Published: (2024)
Advancing Video Self-Supervised Learning via Image Foundation Models
by: Wu, Jingwei, et al.
Published: (2025)
by: Wu, Jingwei, et al.
Published: (2025)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
by: Hu, Zhuozhao, et al.
Published: (2025)
by: Hu, Zhuozhao, et al.
Published: (2025)
Detecting Text Manipulation in Images using Vision Language Models
by: Vidit, Vidit, et al.
Published: (2025)
by: Vidit, Vidit, et al.
Published: (2025)
Geometric Features Enhanced Human-Object Interaction Detection
by: Zhu, Manli, et al.
Published: (2024)
by: Zhu, Manli, et al.
Published: (2024)
Dome-DETR: DETR with Density-Oriented Feature-Query Manipulation for Efficient Tiny Object Detection
by: Hu, Zhangchi, et al.
Published: (2025)
by: Hu, Zhangchi, et al.
Published: (2025)
Similar Items
-
Research on the Application of Computer Vision Based on Deep Learning in Autonomous Driving Technology
by: Zhang, Jingyu, et al.
Published: (2024) -
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024) -
A Weak-Signal-Aware Framework for Subsurface Defect Detection: Mechanisms for Enhancing Low-SCR Hyperbolic Signatures
by: Zhang, Wenbo, et al.
Published: (2026) -
Change3D: Revisiting Change Detection and Captioning from A Video Modeling Perspective
by: Zhu, Duowang, et al.
Published: (2025) -
Group Critical-token Policy Optimization for Autoregressive Image Generation
by: Zhang, Guohui, et al.
Published: (2025)