A Visually Attentive Splice Localization Network with Multi-Domain Feature Extractor and Multi-Receptive Field Upsampler
Fuente:
arXiv
Saved in:
| Main Authors: | Yadav, Ankit, Vishwakarma, Dinesh Kumar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Effective Image Forensics via A Novel Computationally Efficient Framework and A New Image Splice Dataset
by: Yadav, Ankit, et al.
Published: (2024)
by: Yadav, Ankit, et al.
Published: (2024)
Datasets, Clues and State-of-the-Arts for Multimedia Forensics: An Extensive Review
by: Yadav, Ankit, et al.
Published: (2024)
by: Yadav, Ankit, et al.
Published: (2024)
VyAnG-Net: A Novel Multi-Modal Sarcasm Recognition Model by Uncovering Visual, Acoustic and Glossary Features
by: Pandey, Ananya, et al.
Published: (2024)
by: Pandey, Ananya, et al.
Published: (2024)
A Noise and Edge extraction-based dual-branch method for Shallowfake and Deepfake Localization
by: Dagar, Deepak, et al.
Published: (2024)
by: Dagar, Deepak, et al.
Published: (2024)
Target-Dependent Multimodal Sentiment Analysis Via Employing Visual-to Emotional-Caption Translation Network using Visual-Caption Pairs
by: Pandey, Ananya, et al.
Published: (2024)
by: Pandey, Ananya, et al.
Published: (2024)
Gait Recognition with Temporal Kolmogorov-Arnold Networks
by: Asad, Mohammed, et al.
Published: (2026)
by: Asad, Mohammed, et al.
Published: (2026)
Contrastive Learning-based Multi Modal Architecture for Emoticon Prediction by Employing Image-Text Pairs
by: Pandey, Ananya, et al.
Published: (2024)
by: Pandey, Ananya, et al.
Published: (2024)
Modelling Visual Semantics via Image Captioning to extract Enhanced Multi-Level Cross-Modal Semantic Incongruity Representation with Attention for Multimodal Sarcasm Detection
by: Aggarwal, Sajal, et al.
Published: (2024)
by: Aggarwal, Sajal, et al.
Published: (2024)
MobileMamba: Lightweight Multi-Receptive Visual Mamba Network
by: He, Haoyang, et al.
Published: (2024)
by: He, Haoyang, et al.
Published: (2024)
ARFC-WAHNet: Adaptive Receptive Field Convolution and Wavelet-Attentive Hierarchical Network for Infrared Small Target Detection
by: Cui, Xingye, et al.
Published: (2025)
by: Cui, Xingye, et al.
Published: (2025)
Spectral Probing of Feature Upsamplers in 2D-to-3D Scene Reconstruction
by: Xiao, Ling, et al.
Published: (2026)
by: Xiao, Ling, et al.
Published: (2026)
Adversarial Patch for 3D Local Feature Extractor
by: Pao, Yu Wen, et al.
Published: (2024)
by: Pao, Yu Wen, et al.
Published: (2024)
PV-SSD: A Multi-Modal Point Cloud Feature Fusion Method for Projection Features and Variable Receptive Field Voxel Features
by: Shao, Yongxin, et al.
Published: (2023)
by: Shao, Yongxin, et al.
Published: (2023)
Tex-ViT: A Generalizable, Robust, Texture-based dual-branch cross-attention deepfake detector
by: Dagar, Deepak, et al.
Published: (2024)
by: Dagar, Deepak, et al.
Published: (2024)
A Refreshed Similarity-based Upsampler for Direct High-Ratio Feature Upsampling
by: Zhou, Minghao, et al.
Published: (2024)
by: Zhou, Minghao, et al.
Published: (2024)
Multi-Scale Cross-Fusion and Edge-Supervision Network for Image Splicing Localization
by: Niu, Yakun, et al.
Published: (2024)
by: Niu, Yakun, et al.
Published: (2024)
Landmark Guided Visual Feature Extractor for Visual Speech Recognition with Limited Resource
by: Yang, Lei, et al.
Published: (2025)
by: Yang, Lei, et al.
Published: (2025)
AFRDA: Attentive Feature Refinement for Domain Adaptive Semantic Segmentation
by: Khan, Md. Al-Masrur, et al.
Published: (2025)
by: Khan, Md. Al-Masrur, et al.
Published: (2025)
Analyzing the Feature Extractor Networks for Face Image Synthesis
by: Sarıtaş, Erdi, et al.
Published: (2024)
by: Sarıtaş, Erdi, et al.
Published: (2024)
GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild
by: Peng, Guozhen, et al.
Published: (2024)
by: Peng, Guozhen, et al.
Published: (2024)
Progressive Classifier and Feature Extractor Adaptation for Unsupervised Domain Adaptation on Point Clouds
by: Wang, Zicheng, et al.
Published: (2023)
by: Wang, Zicheng, et al.
Published: (2023)
HST-MRF: Heterogeneous Swin Transformer with Multi-Receptive Field for Medical Image Segmentation
by: Huang, Xiaofei, et al.
Published: (2023)
by: Huang, Xiaofei, et al.
Published: (2023)
LAHNet: Local Attentive Hashing Network for Point Cloud Registration
by: Qu, Wentao, et al.
Published: (2025)
by: Qu, Wentao, et al.
Published: (2025)
Generative Fields: Uncovering Hierarchical Feature Control for StyleGAN via Inverted Receptive Fields
by: He, Zhuo, et al.
Published: (2025)
by: He, Zhuo, et al.
Published: (2025)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
Adaptive Deep Iris Feature Extractor at Arbitrary Resolutions
by: Shoji, Yuho, et al.
Published: (2024)
by: Shoji, Yuho, et al.
Published: (2024)
COSMo: CLIP Talks on Open-Set Multi-Target Domain Adaptation
by: Monga, Munish, et al.
Published: (2024)
by: Monga, Munish, et al.
Published: (2024)
Knee Osteoarthritis Severity Prediction using an Attentive Multi-Scale Deep Convolutional Neural Network
by: Jain, Rohit Kumar, et al.
Published: (2021)
by: Jain, Rohit Kumar, et al.
Published: (2021)
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models
by: Yadav, Ankit, et al.
Published: (2025)
by: Yadav, Ankit, et al.
Published: (2025)
Attentive Illumination Decomposition Model for Multi-Illuminant White Balancing
by: Kim, Dongyoung, et al.
Published: (2024)
by: Kim, Dongyoung, et al.
Published: (2024)
Multi-Receptive Field Ensemble with Cross-Entropy Masking for Class Imbalance in Remote Sensing Change Detection
by: Naveed, Humza, et al.
Published: (2025)
by: Naveed, Humza, et al.
Published: (2025)
RFAConv: Receptive-Field Attention Convolution for Improving Convolutional Neural Networks
by: Zhang, Xin, et al.
Published: (2023)
by: Zhang, Xin, et al.
Published: (2023)
Wavelet Convolutions for Large Receptive Fields
by: Finder, Shahaf E., et al.
Published: (2024)
by: Finder, Shahaf E., et al.
Published: (2024)
Restricted Receptive Fields for Face Verification
by: Ozturk, Kagan, et al.
Published: (2025)
by: Ozturk, Kagan, et al.
Published: (2025)
A Pairwise DomMix Attentive Adversarial Network for Unsupervised Domain Adaptive Object Detection
by: Shao, Jie, et al.
Published: (2024)
by: Shao, Jie, et al.
Published: (2024)
Efficient Pretraining Model based on Multi-Scale Local Visual Field Feature Reconstruction for PCB CT Image Element Segmentation
by: Chen, Chen, et al.
Published: (2024)
by: Chen, Chen, et al.
Published: (2024)
BoxingVI: A Multi-Modal Benchmark for Boxing Action Recognition and Localization
by: Kumar, Rahul, et al.
Published: (2025)
by: Kumar, Rahul, et al.
Published: (2025)
MarsSeg: Mars Surface Semantic Segmentation with Multi-level Extractor and Connector
by: Li, Junbo, et al.
Published: (2024)
by: Li, Junbo, et al.
Published: (2024)
A Trainable Feature Extractor Module for Deep Neural Networks and Scanpath Classification
by: Fuhl, Wolfgang
Published: (2024)
by: Fuhl, Wolfgang
Published: (2024)
Gaussian Splatting Feature Fields for Privacy-Preserving Visual Localization
by: Pietrantoni, Maxime, et al.
Published: (2025)
by: Pietrantoni, Maxime, et al.
Published: (2025)
Similar Items
-
Towards Effective Image Forensics via A Novel Computationally Efficient Framework and A New Image Splice Dataset
by: Yadav, Ankit, et al.
Published: (2024) -
Datasets, Clues and State-of-the-Arts for Multimedia Forensics: An Extensive Review
by: Yadav, Ankit, et al.
Published: (2024) -
VyAnG-Net: A Novel Multi-Modal Sarcasm Recognition Model by Uncovering Visual, Acoustic and Glossary Features
by: Pandey, Ananya, et al.
Published: (2024) -
A Noise and Edge extraction-based dual-branch method for Shallowfake and Deepfake Localization
by: Dagar, Deepak, et al.
Published: (2024) -
Target-Dependent Multimodal Sentiment Analysis Via Employing Visual-to Emotional-Caption Translation Network using Visual-Caption Pairs
by: Pandey, Ananya, et al.
Published: (2024)