Guardado en:
| Autores principales: | Pokala, Praveen Kumar, Patibandla, Jaya Sai Kiran, Pandey, Naveen Kumar, Pailla, Balakrishna Reddy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2402.00918 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AQUALLM: Audio Question Answering Data Generation Using Large Language Models
por: Behera, Swarup Ranjan, et al.
Publicado: (2023)
por: Behera, Swarup Ranjan, et al.
Publicado: (2023)
SNIFR : Boosting Fine-Grained Child Harmful Content Detection Through Audio-Visual Alignment with Cascaded Cross-Transformer
por: Phukan, Orchid Chetia, et al.
Publicado: (2025)
por: Phukan, Orchid Chetia, et al.
Publicado: (2025)
Capsule Endoscopy Multi-classification via Gated Attention and Wavelet Transformations
por: Panchananam, Lakshmi Srinivas, et al.
Publicado: (2024)
por: Panchananam, Lakshmi Srinivas, et al.
Publicado: (2024)
Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model
por: Mazna, Romaric, et al.
Publicado: (2026)
por: Mazna, Romaric, et al.
Publicado: (2026)
LTCA: Long-range Temporal Context Attention for Referring Video Object Segmentation
por: Yan, Cilin, et al.
Publicado: (2025)
por: Yan, Cilin, et al.
Publicado: (2025)
Diabetic Retinopathy Lesion Segmentation through Attention Mechanisms
por: Jithesh, Aruna, et al.
Publicado: (2026)
por: Jithesh, Aruna, et al.
Publicado: (2026)
Multi-Context Temporal Consistent Modeling for Referring Video Object Segmentation
por: Choi, Sun-Hyuk, et al.
Publicado: (2025)
por: Choi, Sun-Hyuk, et al.
Publicado: (2025)
Multi-Scale Foreground-Background Confidence for Out-of-Distribution Segmentation
por: Marschall, Samuel, et al.
Publicado: (2024)
por: Marschall, Samuel, et al.
Publicado: (2024)
WB LUTs: Contrastive Learning for White Balancing Lookup Tables
por: Manne, Sai Kumar Reddy, et al.
Publicado: (2024)
por: Manne, Sai Kumar Reddy, et al.
Publicado: (2024)
Curvature Informed Furthest Point Sampling
por: Bhardwaj, Shubham, et al.
Publicado: (2024)
por: Bhardwaj, Shubham, et al.
Publicado: (2024)
Enhancing Pneumonia Diagnosis and Severity Assessment through Deep Learning: A Comprehensive Approach Integrating CNN Classification and Infection Segmentation
por: Mallidi, S Kumar Reddy
Publicado: (2025)
por: Mallidi, S Kumar Reddy
Publicado: (2025)
AADNet: Attention aware Demoiréing Network
por: Reddy, M Rakesh, et al.
Publicado: (2024)
por: Reddy, M Rakesh, et al.
Publicado: (2024)
Turn-by-Turn Indoor Navigation for the Visually Impaired
por: Srinivasaiah, Santosh, et al.
Publicado: (2024)
por: Srinivasaiah, Santosh, et al.
Publicado: (2024)
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
por: Truong, Quang-Trung, et al.
Publicado: (2024)
por: Truong, Quang-Trung, et al.
Publicado: (2024)
Spatio-Temporal Attention for Consistent Video Semantic Segmentation in Automated Driving
por: Varghese, Serin, et al.
Publicado: (2026)
por: Varghese, Serin, et al.
Publicado: (2026)
Biomechanical-phase based Temporal Segmentation in Sports Videos: a Demonstration on Javelin-Throw
por: Badatya, Bikash Kumar, et al.
Publicado: (2025)
por: Badatya, Bikash Kumar, et al.
Publicado: (2025)
Overcoming Small Data Limitations in Video-Based Infant Respiration Estimation
por: Song, Liyang, et al.
Publicado: (2025)
por: Song, Liyang, et al.
Publicado: (2025)
Chitranuvad: Adapting Multi-Lingual LLMs for Multimodal Translation
por: Khan, Shaharukh, et al.
Publicado: (2025)
por: Khan, Shaharukh, et al.
Publicado: (2025)
Learning to Refocus with Video Diffusion Models
por: Tedla, SaiKiran, et al.
Publicado: (2025)
por: Tedla, SaiKiran, et al.
Publicado: (2025)
Estimating Vehicle Speed on Roadways Using RNNs and Transformers: A Video-based Approach
por: Mareddy, Sai Krishna Reddy, et al.
Publicado: (2025)
por: Mareddy, Sai Krishna Reddy, et al.
Publicado: (2025)
FOCUS: Towards Universal Foreground Segmentation
por: You, Zuyao, et al.
Publicado: (2025)
por: You, Zuyao, et al.
Publicado: (2025)
Language-Guided Temporal Token Pruning for Efficient VideoLLM Processing
por: Kumar, Yogesh
Publicado: (2025)
por: Kumar, Yogesh
Publicado: (2025)
Learning Local and Global Temporal Contexts for Video Semantic Segmentation
por: Sun, Guolei, et al.
Publicado: (2022)
por: Sun, Guolei, et al.
Publicado: (2022)
Towards Inclusive Face Recognition Through Synthetic Ethnicity Alteration
por: Chandaliya, Praveen Kumar, et al.
Publicado: (2024)
por: Chandaliya, Praveen Kumar, et al.
Publicado: (2024)
TexTAR : Textual Attribute Recognition in Multi-domain and Multi-lingual Document Images
por: Kumar, Rohan, et al.
Publicado: (2025)
por: Kumar, Rohan, et al.
Publicado: (2025)
UnSegGNet: Unsupervised Image Segmentation using Graph Neural Networks
por: Reddy, Kovvuri Sai Gopal, et al.
Publicado: (2024)
por: Reddy, Kovvuri Sai Gopal, et al.
Publicado: (2024)
UnSeGArmaNet: Unsupervised Image Segmentation using Graph Neural Networks with Convolutional ARMA Filters
por: Reddy, Kovvuri Sai Gopal, et al.
Publicado: (2024)
por: Reddy, Kovvuri Sai Gopal, et al.
Publicado: (2024)
Embodiment: Self-Supervised Depth Estimation Based on Camera Models
por: Zhang, Jinchang, et al.
Publicado: (2024)
por: Zhang, Jinchang, et al.
Publicado: (2024)
Multi-Stain Multi-Level Convolutional Network for Multi-Tissue Breast Cancer Image Segmentation
por: Modi, Akash, et al.
Publicado: (2024)
por: Modi, Akash, et al.
Publicado: (2024)
Joint Flow And Feature Refinement Using Attention For Video Restoration
por: Merugu, Ranjith, et al.
Publicado: (2025)
por: Merugu, Ranjith, et al.
Publicado: (2025)
Point Tracking as a Temporal Cue for Robust Myocardial Segmentation in Echocardiography Videos
por: Khodabakhshian, Bahar, et al.
Publicado: (2026)
por: Khodabakhshian, Bahar, et al.
Publicado: (2026)
Foreground-Covering Prototype Generation and Matching for SAM-Aided Few-Shot Segmentation
por: Park, Suho, et al.
Publicado: (2025)
por: Park, Suho, et al.
Publicado: (2025)
Chitrarth: Bridging Vision and Language for a Billion People
por: Khan, Shaharukh, et al.
Publicado: (2025)
por: Khan, Shaharukh, et al.
Publicado: (2025)
Multi-dimension Transformer with Attention-based Filtering for Medical Image Segmentation
por: Wang, Wentao, et al.
Publicado: (2024)
por: Wang, Wentao, et al.
Publicado: (2024)
SegMAN: Omni-scale Context Modeling with State Space Models and Local Attention for Semantic Segmentation
por: Fu, Yunxiang, et al.
Publicado: (2024)
por: Fu, Yunxiang, et al.
Publicado: (2024)
Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion
por: Xue, Yu, et al.
Publicado: (2026)
por: Xue, Yu, et al.
Publicado: (2026)
CEM-FBGTinyDet: Context-Enhanced Foreground Balance with Gradient Tuning for tiny Objects
por: Liu, Tao, et al.
Publicado: (2025)
por: Liu, Tao, et al.
Publicado: (2025)
Certified vs. Empirical Adversarial Robust-ness via Hybrid Convolutions with Attention Stochasticity
por: Dhar, Joy, et al.
Publicado: (2026)
por: Dhar, Joy, et al.
Publicado: (2026)
Modelling Visual Semantics via Image Captioning to extract Enhanced Multi-Level Cross-Modal Semantic Incongruity Representation with Attention for Multimodal Sarcasm Detection
por: Aggarwal, Sajal, et al.
Publicado: (2024)
por: Aggarwal, Sajal, et al.
Publicado: (2024)
Robust Foreground-Background Separation for Severely-Degraded Videos Using Convolutional Sparse Representation Modeling
por: Naganuma, Kazuki, et al.
Publicado: (2025)
por: Naganuma, Kazuki, et al.
Publicado: (2025)
Ejemplares similares
-
AQUALLM: Audio Question Answering Data Generation Using Large Language Models
por: Behera, Swarup Ranjan, et al.
Publicado: (2023) -
SNIFR : Boosting Fine-Grained Child Harmful Content Detection Through Audio-Visual Alignment with Cascaded Cross-Transformer
por: Phukan, Orchid Chetia, et al.
Publicado: (2025) -
Capsule Endoscopy Multi-classification via Gated Attention and Wavelet Transformations
por: Panchananam, Lakshmi Srinivas, et al.
Publicado: (2024) -
Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model
por: Mazna, Romaric, et al.
Publicado: (2026) -
LTCA: Long-range Temporal Context Attention for Referring Video Object Segmentation
por: Yan, Cilin, et al.
Publicado: (2025)