Learning to Find Missing Video Frames with Synthetic Data Augmentation: A General Framework and Application in Generating Thermal Images Using RGB Cameras
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Andersen, Mathias Viborg, Greer, Ross, Møgelmose, Andreas, Trivedi, Mohan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
The Why, When, and How to Use Active Learning in Large-Data-Driven 3D Object Detection for Safe Autonomous Driving: An Empirical Exploration
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Language-Driven Active Learning for Diverse Open-Set 3D Object Detection
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Towards Explainable, Safe Autonomous Driving with Language Embeddings for Novelty Identification and Active Learning: Framework and Experimental Analysis with Real-World Data Sets
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
von: Gopalkrishnan, Akshay, et al.
Veröffentlicht: (2024)
von: Gopalkrishnan, Akshay, et al.
Veröffentlicht: (2024)
Perception Without Vision for Trajectory Prediction: Ego Vehicle Dynamics as Scene Representation for Efficient Active Learning in Autonomous Driving
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
ActiveAnno3D -- An Active Learning Framework for Multi-Modal 3D Object Detection
von: Ghita, Ahmed, et al.
Veröffentlicht: (2024)
von: Ghita, Ahmed, et al.
Veröffentlicht: (2024)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
von: Keskar, Maitrayee, et al.
Veröffentlicht: (2025)
von: Keskar, Maitrayee, et al.
Veröffentlicht: (2025)
Finding the Reflection Point: Unpadding Images to Remove Data Augmentation Artifacts in Large Open Source Image Datasets for Machine Learning
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
von: Bossen, Tonko E. W., et al.
Veröffentlicht: (2025)
von: Bossen, Tonko E. W., et al.
Veröffentlicht: (2025)
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
von: Chen, Jason, et al.
Veröffentlicht: (2025)
von: Chen, Jason, et al.
Veröffentlicht: (2025)
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
von: Shriram, Shashank, et al.
Veröffentlicht: (2025)
von: Shriram, Shashank, et al.
Veröffentlicht: (2025)
A New Perspective On AI Safety Through Control Theory Methodologies
von: Ullrich, Lars, et al.
Veröffentlicht: (2025)
von: Ullrich, Lars, et al.
Veröffentlicht: (2025)
Vision and Language: Novel Representations and Artificial intelligence for Driving Scene Safety Assessment and Autonomous Vehicle Planning
von: Greer, Ross, et al.
Veröffentlicht: (2026)
von: Greer, Ross, et al.
Veröffentlicht: (2026)
Data-Driven Feature Tracking for Event Cameras With and Without Frames
von: Messikommer, Nico, et al.
Veröffentlicht: (2022)
von: Messikommer, Nico, et al.
Veröffentlicht: (2022)
Beyond General Prompts: Automated Prompt Refinement using Contrastive Class Alignment Scores for Disambiguating Objects in Vision-Language Models
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
von: Choi, Lucas, et al.
Veröffentlicht: (2025)
Creativity and Visual Communication from Machine to Musician: Sharing a Score through a Robotic Camera
von: Greer, Ross, et al.
Veröffentlicht: (2024)
von: Greer, Ross, et al.
Veröffentlicht: (2024)
Automated Data Curation Using GPS & NLP to Generate Instruction-Action Pairs for Autonomous Vehicle Vision-Language Navigation Datasets
von: Roque, Guillermo, et al.
Veröffentlicht: (2025)
von: Roque, Guillermo, et al.
Veröffentlicht: (2025)
SVAD: From Single Image to 3D Avatar via Synthetic Data Generation with Video Diffusion and Data Augmentation
von: Choi, Yonwoo
Veröffentlicht: (2025)
von: Choi, Yonwoo
Veröffentlicht: (2025)
Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation
von: Fei, Yuanchen, et al.
Veröffentlicht: (2026)
von: Fei, Yuanchen, et al.
Veröffentlicht: (2026)
Synthetic Data Augmentation for Table Detection: Re-evaluating TableNet's Performance with Automatically Generated Document Images
von: Sahukara, Krishna, et al.
Veröffentlicht: (2025)
von: Sahukara, Krishna, et al.
Veröffentlicht: (2025)
A Conditional Generative Framework for Synthetic Data Augmentation in Segmenting Thin and Elongated Structures in Biological Images
von: Liu, Yi, et al.
Veröffentlicht: (2025)
von: Liu, Yi, et al.
Veröffentlicht: (2025)
TAEGAN: Generating Synthetic Tabular Data For Data Augmentation
von: Li, Jiayu, et al.
Veröffentlicht: (2024)
von: Li, Jiayu, et al.
Veröffentlicht: (2024)
MemCam: Memory-Augmented Camera Control for Consistent Video Generation
von: Gao, Xinhang, et al.
Veröffentlicht: (2026)
von: Gao, Xinhang, et al.
Veröffentlicht: (2026)
Synthetic Data Generation for Augmenting Small Samples
von: Liu, Dan, et al.
Veröffentlicht: (2025)
von: Liu, Dan, et al.
Veröffentlicht: (2025)
Masked Clinical Modelling: A Framework for Synthetic and Augmented Survival Data Generation
von: Kuo, Nicholas I-Hsien, et al.
Veröffentlicht: (2024)
von: Kuo, Nicholas I-Hsien, et al.
Veröffentlicht: (2024)
Optimal Multispectral Imaging using RGB Cameras
von: Matulić, Tomislav, et al.
Veröffentlicht: (2026)
von: Matulić, Tomislav, et al.
Veröffentlicht: (2026)
Synthetic ECG Generation for Data Augmentation and Transfer Learning in Arrhythmia Classification
von: Núñez, José Fernando, et al.
Veröffentlicht: (2024)
von: Núñez, José Fernando, et al.
Veröffentlicht: (2024)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
von: Kirchner, Sven, et al.
Veröffentlicht: (2025)
Frame-Guided Synthetic Claim Generation for Automatic Fact-Checking Using High-Volume Tabular Data
von: Devasier, Jacob, et al.
Veröffentlicht: (2026)
von: Devasier, Jacob, et al.
Veröffentlicht: (2026)
Illuminant Estimation Using RGB Camera Image and Ambient Light Sensor Signal
von: Yuyang Liu, et al.
Veröffentlicht: (2025)
von: Yuyang Liu, et al.
Veröffentlicht: (2025)
ThermalGen: Style-Disentangled Flow-Based Generative Models for RGB-to-Thermal Image Translation
von: Xiao, Jiuhong, et al.
Veröffentlicht: (2025)
von: Xiao, Jiuhong, et al.
Veröffentlicht: (2025)
Increasing the Diversity in RGB-to-Thermal Image Translation for Automotive Applications
von: Wang, Kaili, et al.
Veröffentlicht: (2025)
von: Wang, Kaili, et al.
Veröffentlicht: (2025)
Enhancing Highway Safety: Accident Detection on the A9 Test Stretch Using Roadside Sensors
von: Zimmer, Walter, et al.
Veröffentlicht: (2025)
von: Zimmer, Walter, et al.
Veröffentlicht: (2025)
Synthetic Thermal and RGB Videos for Automatic Pain Assessment utilizing a Vision-MLP Architecture
von: Gkikas, Stefanos, et al.
Veröffentlicht: (2024)
von: Gkikas, Stefanos, et al.
Veröffentlicht: (2024)
FRAG: Frame Selection Augmented Generation for Long Video and Long Document Understanding
von: Huang, De-An, et al.
Veröffentlicht: (2025)
von: Huang, De-An, et al.
Veröffentlicht: (2025)
Data Augmentation for Generating Synthetic Electrogastrogram Time Series
von: Miljković, Nadica, et al.
Veröffentlicht: (2023)
von: Miljković, Nadica, et al.
Veröffentlicht: (2023)
MAG-V: A Multi-Agent Framework for Synthetic Data Generation and Verification
von: Sengupta, Saptarshi, et al.
Veröffentlicht: (2024)
von: Sengupta, Saptarshi, et al.
Veröffentlicht: (2024)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
von: Wang, Boyang, et al.
Veröffentlicht: (2025)
von: Wang, Boyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
The Why, When, and How to Use Active Learning in Large-Data-Driven 3D Object Detection for Safe Autonomous Driving: An Empirical Exploration
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
Language-Driven Active Learning for Diverse Open-Set 3D Object Detection
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
Towards Explainable, Safe Autonomous Driving with Language Embeddings for Novelty Identification and Active Learning: Framework and Experimental Analysis with Real-World Data Sets
von: Greer, Ross, et al.
Veröffentlicht: (2024) -
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
von: Gopalkrishnan, Akshay, et al.
Veröffentlicht: (2024)