Automatic Scene Generation: State-of-the-Art Techniques, Models, Datasets, Challenges, and Future Prospects
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fime, Awal Ahmed, Mahmud, Saifuddin, Das, Arpita, Islam, Md. Sunzidul, Kim, Hong-Hoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distributed LLMs and Multimodal Large Language Models: A Survey on Advances, Challenges, and Future Directions
von: Amini, Hadi, et al.
Veröffentlicht: (2025)
von: Amini, Hadi, et al.
Veröffentlicht: (2025)
Bangla Sign Language Translation: Dataset Creation Challenges, Benchmarking and Prospects
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025)
A Survey on Wi-Fi Sensing Generalizability: Taxonomy, Techniques, Datasets, and Future Research Prospects
von: Wang, Fei, et al.
Veröffentlicht: (2025)
von: Wang, Fei, et al.
Veröffentlicht: (2025)
Deep Learning Approach for Enhancing Oral Squamous Cell Carcinoma with LIME Explainable AI Technique
von: Islam, Samiha, et al.
Veröffentlicht: (2024)
von: Islam, Samiha, et al.
Veröffentlicht: (2024)
AI Art Neural Constellation: Revealing the Collective and Contrastive State of AI-Generated and Human Art
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2024)
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2024)
Investigating Prompting Techniques for Zero- and Few-Shot Visual Question Answering
von: Awal, Rabiul, et al.
Veröffentlicht: (2023)
von: Awal, Rabiul, et al.
Veröffentlicht: (2023)
Deep Learning for Ophthalmology: The State-of-the-Art and Future Trends
von: Nguyen, Duy M. H., et al.
Veröffentlicht: (2025)
von: Nguyen, Duy M. H., et al.
Veröffentlicht: (2025)
MMIS: Multimodal Dataset for Interior Scene Visual Generation and Recognition
von: Kassab, Hozaifa, et al.
Veröffentlicht: (2024)
von: Kassab, Hozaifa, et al.
Veröffentlicht: (2024)
Leveraging Pre-trained CNNs for Efficient Feature Extraction in Rice Leaf Disease Classification
von: Sobuj, Md. Shohanur Islam, et al.
Veröffentlicht: (2024)
von: Sobuj, Md. Shohanur Islam, et al.
Veröffentlicht: (2024)
A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects
von: Xie, Guohuan, et al.
Veröffentlicht: (2025)
von: Xie, Guohuan, et al.
Veröffentlicht: (2025)
DEEGITS: Deep Learning based Framework for Measuring Heterogenous Traffic State in Challenging Traffic Scenarios
von: Islam, Muttahirul, et al.
Veröffentlicht: (2024)
von: Islam, Muttahirul, et al.
Veröffentlicht: (2024)
Instruction-Guided Lesion Segmentation for Chest X-rays with Automatically Generated Large-Scale Dataset
von: Choi, Geon, et al.
Veröffentlicht: (2025)
von: Choi, Geon, et al.
Veröffentlicht: (2025)
A study on Deep Convolutional Neural Networks, transfer learning, and Mnet model for Cervical Cancer Detection
von: Sagor, Saifuddin, et al.
Veröffentlicht: (2025)
von: Sagor, Saifuddin, et al.
Veröffentlicht: (2025)
Deepfake Media Forensics: State of the Art and Challenges Ahead
von: Amerini, Irene, et al.
Veröffentlicht: (2024)
von: Amerini, Irene, et al.
Veröffentlicht: (2024)
BdSLW60: A Word-Level Bangla Sign Language Dataset
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2024)
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2024)
NeuroAPS-Net: Neuro-Anatomically Aware Point Cloud Representation for Efficient Alzheimer's Disease Classification
von: Islam, Towhidul, et al.
Veröffentlicht: (2026)
von: Islam, Towhidul, et al.
Veröffentlicht: (2026)
DeepfakeArt Challenge: A Benchmark Dataset for Generative AI Art Forgery and Data Poisoning Detection
von: Aboutalebi, Hossein, et al.
Veröffentlicht: (2023)
von: Aboutalebi, Hossein, et al.
Veröffentlicht: (2023)
A Critical Analysis on Machine Learning Techniques for Video-based Human Activity Recognition of Surveillance Systems: A Review
von: Jahan, Shahriar, et al.
Veröffentlicht: (2024)
von: Jahan, Shahriar, et al.
Veröffentlicht: (2024)
An Image Dataset of Common Skin Diseases of Bangladesh and Benchmarking Performance with Machine Learning Models
von: Hossain, Sazzad, et al.
Veröffentlicht: (2026)
von: Hossain, Sazzad, et al.
Veröffentlicht: (2026)
AquaFuse: Waterbody Fusion for Physics Guided View Synthesis of Underwater Scenes
von: Siddique, Md Abu Bakr, et al.
Veröffentlicht: (2024)
von: Siddique, Md Abu Bakr, et al.
Veröffentlicht: (2024)
Reducing Label Dependency for Underwater Scene Understanding: A Survey of Datasets, Techniques and Applications
von: Raine, Scarlett, et al.
Veröffentlicht: (2024)
von: Raine, Scarlett, et al.
Veröffentlicht: (2024)
Skin Cancer Images Classification using Transfer Learning Techniques
von: Islam, Md Sirajul, et al.
Veröffentlicht: (2024)
von: Islam, Md Sirajul, et al.
Veröffentlicht: (2024)
ChartZero: Synthetic Priors Enable Zero Shot Chart Data Extraction
von: Islam, Md Touhidul, et al.
Veröffentlicht: (2026)
von: Islam, Md Touhidul, et al.
Veröffentlicht: (2026)
WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection
von: Hong, Yan, et al.
Veröffentlicht: (2024)
von: Hong, Yan, et al.
Veröffentlicht: (2024)
State-of-the-Art Transformer Models for Image Super-Resolution: Techniques, Challenges, and Applications
von: Dutta, Debasish, et al.
Veröffentlicht: (2025)
von: Dutta, Debasish, et al.
Veröffentlicht: (2025)
Towards Automatic Power Battery Detection: New Challenge, Benchmark Dataset and Baseline
von: Zhao, Xiaoqi, et al.
Veröffentlicht: (2023)
von: Zhao, Xiaoqi, et al.
Veröffentlicht: (2023)
A Survey on Deep Learning for Polyp Segmentation: Techniques, Challenges and Future Trends
von: Mei, Jiaxin, et al.
Veröffentlicht: (2023)
von: Mei, Jiaxin, et al.
Veröffentlicht: (2023)
Involution-Infused DenseNet with Two-Step Compression for Resource-Efficient Plant Disease Classification
von: Ahmed, T., et al.
Veröffentlicht: (2025)
von: Ahmed, T., et al.
Veröffentlicht: (2025)
Efficient Leaf Disease Classification and Segmentation using Midpoint Normalization Technique and Attention Mechanism
von: Taufik, Enam Ahmed, et al.
Veröffentlicht: (2025)
von: Taufik, Enam Ahmed, et al.
Veröffentlicht: (2025)
MOSEv2: A More Challenging Dataset for Video Object Segmentation in Complex Scenes
von: Ding, Henghui, et al.
Veröffentlicht: (2025)
von: Ding, Henghui, et al.
Veröffentlicht: (2025)
VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites
von: Islam, Md. Adnanul, et al.
Veröffentlicht: (2025)
von: Islam, Md. Adnanul, et al.
Veröffentlicht: (2025)
EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
3D Scene Prompting for Scene-Consistent Camera-Controllable Video Generation
von: Lee, JoungBin, et al.
Veröffentlicht: (2025)
von: Lee, JoungBin, et al.
Veröffentlicht: (2025)
Evaluation of State-of-the-Art Deep Learning Techniques for Plant Disease and Pest Detection
von: Banerjee, Saptarshi, et al.
Veröffentlicht: (2025)
von: Banerjee, Saptarshi, et al.
Veröffentlicht: (2025)
Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos
von: Mahmud, Syed Mumtahin, et al.
Veröffentlicht: (2025)
von: Mahmud, Syed Mumtahin, et al.
Veröffentlicht: (2025)
A Survey on Long Video Generation: Challenges, Methods, and Prospects
von: Li, Chengxuan, et al.
Veröffentlicht: (2024)
von: Li, Chengxuan, et al.
Veröffentlicht: (2024)
Attentive Dilated Convolution for Automatic Sleep Staging using Force-directed Layout
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
von: Jobayer, Md, et al.
Veröffentlicht: (2024)
Pre-training for Action Recognition with Automatically Generated Fractal Datasets
von: Svyezhentsev, Davyd, et al.
Veröffentlicht: (2024)
von: Svyezhentsev, Davyd, et al.
Veröffentlicht: (2024)
State-of-the-Art Fails in the Art of Damage Detection
von: Ivanova, Daniela, et al.
Veröffentlicht: (2024)
von: Ivanova, Daniela, et al.
Veröffentlicht: (2024)
DENSER: 3D Gaussians Splatting for Scene Reconstruction of Dynamic Urban Environments
von: Mohamad, Mahmud A., et al.
Veröffentlicht: (2024)
von: Mohamad, Mahmud A., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Distributed LLMs and Multimodal Large Language Models: A Survey on Advances, Challenges, and Future Directions
von: Amini, Hadi, et al.
Veröffentlicht: (2025) -
Bangla Sign Language Translation: Dataset Creation Challenges, Benchmarking and Prospects
von: Rubaiyeat, Husne Ara, et al.
Veröffentlicht: (2025) -
A Survey on Wi-Fi Sensing Generalizability: Taxonomy, Techniques, Datasets, and Future Research Prospects
von: Wang, Fei, et al.
Veröffentlicht: (2025) -
Deep Learning Approach for Enhancing Oral Squamous Cell Carcinoma with LIME Explainable AI Technique
von: Islam, Samiha, et al.
Veröffentlicht: (2024) -
AI Art Neural Constellation: Revealing the Collective and Contrastive State of AI-Generated and Human Art
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2024)