Gespeichert in:
| Hauptverfasser: | Menghani, Gaurav, Kumar, Ravi, Kumar, Sanjiv |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2411.07501 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
B-DENSE: Branching For Dense Ensemble Network Supervision Efficiency
von: Puniani, Cherish, et al.
Veröffentlicht: (2026)
von: Puniani, Cherish, et al.
Veröffentlicht: (2026)
Semantically Controllable Augmentations for Generalizable Robot Learning
von: Chen, Zoey, et al.
Veröffentlicht: (2024)
von: Chen, Zoey, et al.
Veröffentlicht: (2024)
Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
von: Singh, Amit Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Amit Kumar, et al.
Veröffentlicht: (2024)
APTx: better activation function than MISH, SWISH, and ReLU's variants used in deep learning
von: Kumar, Ravin
Veröffentlicht: (2022)
von: Kumar, Ravin
Veröffentlicht: (2022)
D-CODA: Diffusion for Coordinated Dual-Arm Data Augmentation
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2025)
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2025)
Residual SODAP: Residual Self-Organizing Domain-Adaptive Prompting with Structural Knowledge Preservation for Continual Learning
von: Oh, Gyutae, et al.
Veröffentlicht: (2026)
von: Oh, Gyutae, et al.
Veröffentlicht: (2026)
RS-CA-HSICT: A Residual and Spatial Channel Augmented CNN Transformer Framework for Monkeypox Detection
von: Iqbal, Rashid, et al.
Veröffentlicht: (2025)
von: Iqbal, Rashid, et al.
Veröffentlicht: (2025)
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
von: Chen, Jason, et al.
Veröffentlicht: (2025)
von: Chen, Jason, et al.
Veröffentlicht: (2025)
Residual Kolmogorov-Arnold Network for Enhanced Deep Learning
von: Yu, Ray Congrui, et al.
Veröffentlicht: (2024)
von: Yu, Ray Congrui, et al.
Veröffentlicht: (2024)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Improved Cotton Leaf Disease Classification Using Parameter-Efficient Deep Learning Framework
von: Patra, Aswini Kumar, et al.
Veröffentlicht: (2024)
von: Patra, Aswini Kumar, et al.
Veröffentlicht: (2024)
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
von: Shrivastava, Gaurav, et al.
Veröffentlicht: (2024)
von: Shrivastava, Gaurav, et al.
Veröffentlicht: (2024)
Generative AI in Vision: A Survey on Models, Metrics and Applications
von: Raut, Gaurav, et al.
Veröffentlicht: (2024)
von: Raut, Gaurav, et al.
Veröffentlicht: (2024)
Diabetic Retinopathy Classification using Downscaling Algorithms and Deep Learning
von: Doshi, Nishi, et al.
Veröffentlicht: (2026)
von: Doshi, Nishi, et al.
Veröffentlicht: (2026)
Table Detection with Active Learning
von: Gautam, Somraj, et al.
Veröffentlicht: (2025)
von: Gautam, Somraj, et al.
Veröffentlicht: (2025)
IDAL: Improved Domain Adaptive Learning for Natural Images Dataset
von: Gupta, Ravi Kant, et al.
Veröffentlicht: (2025)
von: Gupta, Ravi Kant, et al.
Veröffentlicht: (2025)
Improved Classification of Nitrogen Stress Severity in Plants Under Combined Stress Conditions Using Spatio-Temporal Deep Learning Framework
von: Patra, Aswini Kumar, et al.
Veröffentlicht: (2025)
von: Patra, Aswini Kumar, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning for Urban Air Quality Management: Multi-Objective Optimization of Pollution Mitigation Booth Placement in Metropolitan Environments
von: Rajesh, Kirtan, et al.
Veröffentlicht: (2025)
von: Rajesh, Kirtan, et al.
Veröffentlicht: (2025)
What Can We Learn from Inter-Annotator Variability in Skin Lesion Segmentation?
von: Abhishek, Kumar, et al.
Veröffentlicht: (2025)
von: Abhishek, Kumar, et al.
Veröffentlicht: (2025)
Learning Regional Monsoon Patterns with a Multimodal Attention U-Net
von: Mazumder, Swaib Ilias, et al.
Veröffentlicht: (2025)
von: Mazumder, Swaib Ilias, et al.
Veröffentlicht: (2025)
A Systematic Survey on Deep Learning Architectures for Point Cloud Classification and Segmentation
von: Kamal, Minhas, et al.
Veröffentlicht: (2026)
von: Kamal, Minhas, et al.
Veröffentlicht: (2026)
Generation of Indian Sign Language Letters, Numbers, and Words
von: Yadav, Ajeet Kumar, et al.
Veröffentlicht: (2025)
von: Yadav, Ajeet Kumar, et al.
Veröffentlicht: (2025)
Vision-Language Models Provide Promptable Representations for Reinforcement Learning
von: Chen, William, et al.
Veröffentlicht: (2024)
von: Chen, William, et al.
Veröffentlicht: (2024)
ReCoRe: Regularized Contrastive Representation Learning of World Model
von: Poudel, Rudra P. K., et al.
Veröffentlicht: (2023)
von: Poudel, Rudra P. K., et al.
Veröffentlicht: (2023)
Cooperative Meta-Learning with Gradient Augmentation
von: Shin, Jongyun, et al.
Veröffentlicht: (2024)
von: Shin, Jongyun, et al.
Veröffentlicht: (2024)
Self-supervised Learning for Hyperspectral Images of Trees
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025)
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025)
REGEN: Learning Compact Video Embedding with (Re-)Generative Decoder
von: Zhang, Yitian, et al.
Veröffentlicht: (2025)
von: Zhang, Yitian, et al.
Veröffentlicht: (2025)
Towards Real-Time 2D Mapping: Harnessing Drones, AI, and Computer Vision for Advanced Insights
von: Agnur, Bharath Kumar
Veröffentlicht: (2024)
von: Agnur, Bharath Kumar
Veröffentlicht: (2024)
When Better Eyes Lead to Blindness: A Diagnostic Study of the Information Bottleneck in CNN-LSTM Image Captioning Models
von: Gupta, Hitesh Kumar
Veröffentlicht: (2025)
von: Gupta, Hitesh Kumar
Veröffentlicht: (2025)
Revisiting Data Augmentation in Deep Reinforcement Learning
von: Hu, Jianshu, et al.
Veröffentlicht: (2024)
von: Hu, Jianshu, et al.
Veröffentlicht: (2024)
Learning to Balance: Diverse Normalization for Cloth-Changing Person Re-Identification
von: Wang, Hongjun, et al.
Veröffentlicht: (2024)
von: Wang, Hongjun, et al.
Veröffentlicht: (2024)
MTCNET: Multi-task Learning Paradigm for Crowd Count Estimation
von: Kumar, Abhay, et al.
Veröffentlicht: (2019)
von: Kumar, Abhay, et al.
Veröffentlicht: (2019)
Residual-SwinCA-Net: A Channel-Aware Integrated Residual CNN-Swin Transformer for Malignant Lesion Segmentation in BUSI
von: Naz, Saeeda, et al.
Veröffentlicht: (2025)
von: Naz, Saeeda, et al.
Veröffentlicht: (2025)
A Novel Defense Against Poisoning Attacks on Federated Learning: LayerCAM Augmented with Autoencoder
von: Zheng, Jingjing, et al.
Veröffentlicht: (2024)
von: Zheng, Jingjing, et al.
Veröffentlicht: (2024)
Mechanisms of Non-Monotonic Scaling in Vision Transformers
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
Measuring the (Un)Faithfulness of Concept-Based Explanations
von: Kumar, Shubham, et al.
Veröffentlicht: (2025)
von: Kumar, Shubham, et al.
Veröffentlicht: (2025)
Parameter Reduction Improves Vision Transformers: A Comparative Study of Sharing and Width Reduction
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
Successes and Limitations of Object-centric Models at Compositional Generalisation
von: Montero, Milton L., et al.
Veröffentlicht: (2024)
von: Montero, Milton L., et al.
Veröffentlicht: (2024)
SupReMix: Supervised Contrastive Learning for Medical Imaging Regression with Mixup
von: Wu, Yilei, et al.
Veröffentlicht: (2023)
von: Wu, Yilei, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
von: Batra, Sumeet, et al.
Veröffentlicht: (2024) -
B-DENSE: Branching For Dense Ensemble Network Supervision Efficiency
von: Puniani, Cherish, et al.
Veröffentlicht: (2026) -
Semantically Controllable Augmentations for Generalizable Robot Learning
von: Chen, Zoey, et al.
Veröffentlicht: (2024) -
Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
von: Singh, Amit Kumar, et al.
Veröffentlicht: (2024) -
APTx: better activation function than MISH, SWISH, and ReLU's variants used in deep learning
von: Kumar, Ravin
Veröffentlicht: (2022)