OpenViGA: Video Generation for Automotive Driving Scenes by Streamlining and Fine-Tuning Open Source Models with Public Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Möller, Björn, Li, Zhengyang, Stelzer, Malte, Graave, Thomas, Bettels, Fabian, Ataya, Muaaz, Fingscheidt, Tim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Noise-Robust AV-ASR Using Visual Features Both in the Whisper Encoder and Decoder
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)
Improving Block-Wise LLM Quantization by 4-bit Block-Wise Optimal Float (BOF4): Analysis and Variations
von: Blumenberg, Patrick, et al.
Veröffentlicht: (2025)
von: Blumenberg, Patrick, et al.
Veröffentlicht: (2025)
Engineering of Hallucination in Generative AI: It's not a Bug, it's a Feature
von: Fingscheidt, Tim, et al.
Veröffentlicht: (2026)
von: Fingscheidt, Tim, et al.
Veröffentlicht: (2026)
A Lightweight Image Super-Resolution Transformer Trained on Low-Resolution Images Only
von: Möller, Björn, et al.
Veröffentlicht: (2025)
von: Möller, Björn, et al.
Veröffentlicht: (2025)
Foundation Models for Amodal Video Instance Segmentation in Automated Driving
von: Breitenstein, Jasmin, et al.
Veröffentlicht: (2024)
von: Breitenstein, Jasmin, et al.
Veröffentlicht: (2024)
MRDrive: An Open Source Mixed Reality Driving Simulator for Automotive User Research
von: Ebel, Patrick, et al.
Veröffentlicht: (2026)
von: Ebel, Patrick, et al.
Veröffentlicht: (2026)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
Efficient Multi-View 3D Object Detection by Dynamic Token Selection and Fine-Tuning
von: Nazir, Danish, et al.
Veröffentlicht: (2026)
von: Nazir, Danish, et al.
Veröffentlicht: (2026)
DisContSE: Single-Step Diffusion Speech Enhancement Based on Joint Discrete and Continuous Embeddings
von: Fu, Yihui, et al.
Veröffentlicht: (2026)
von: Fu, Yihui, et al.
Veröffentlicht: (2026)
RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration
von: Ridder, Fabian, et al.
Veröffentlicht: (2026)
von: Ridder, Fabian, et al.
Veröffentlicht: (2026)
Prompting and Fine-Tuning Open-Sourced Large Language Models for Stance Classification
von: Cruickshank, Iain J., et al.
Veröffentlicht: (2023)
von: Cruickshank, Iain J., et al.
Veröffentlicht: (2023)
Fine-Tuning and Evaluating Open-Source Large Language Models for the Army Domain
von: Ruiz, Daniel C., et al.
Veröffentlicht: (2024)
von: Ruiz, Daniel C., et al.
Veröffentlicht: (2024)
BiDexHand: Design and Evaluation of an Open-Source 16-DoF Biomimetic Dexterous Hand
von: Weng, Zhengyang Kris
Veröffentlicht: (2025)
von: Weng, Zhengyang Kris
Veröffentlicht: (2025)
Teachers' Experiences with Implementing Open-ended Labs in High School Physics Classe
von: Talafian, Hamideh, et al.
Veröffentlicht: (2024)
von: Talafian, Hamideh, et al.
Veröffentlicht: (2024)
LiGuard: A Streamlined Open-Source Framework for Rapid & Interactive Lidar Research
von: Shahbaz, Muhammad, et al.
Veröffentlicht: (2025)
von: Shahbaz, Muhammad, et al.
Veröffentlicht: (2025)
Physical Design of UET-RVMCU: A Streamlined Open-Source RISC-V Microcontroller
von: Azhar, Abdullah, et al.
Veröffentlicht: (2026)
von: Azhar, Abdullah, et al.
Veröffentlicht: (2026)
Disengagement Analysis and Field Tests of a Prototypical Open-Source Level 4 Autonomous Driving System
von: Seegert, Marvin, et al.
Veröffentlicht: (2026)
von: Seegert, Marvin, et al.
Veröffentlicht: (2026)
NEFMind: Parameter-Efficient Fine-Tuning of Open-Source LLMs for Telecom APIs Automation
von: Khan, Zainab, et al.
Veröffentlicht: (2025)
von: Khan, Zainab, et al.
Veröffentlicht: (2025)
UniVA: Universal Video Agent towards Open-Source Next-Generation Video Generalist
von: Liang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Liang, Zhengyang, et al.
Veröffentlicht: (2025)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
TEACar: An Open-Source Autonomous Driving Platform
von: Zhang, Zhongzheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongzheng, et al.
Veröffentlicht: (2026)
Distributed Semantic Segmentation with Efficient Joint Source and Task Decoding
von: Nazir, Danish, et al.
Veröffentlicht: (2024)
von: Nazir, Danish, et al.
Veröffentlicht: (2024)
The Value and Values of Openness: Open Source, AI Openness, and the Future of Digital Public Goods
von: Barth, Jennifer
Veröffentlicht: (2025)
von: Barth, Jennifer
Veröffentlicht: (2025)
OpenEMMA: Open-Source Multimodal Model for End-to-End Autonomous Driving
von: Xing, Shuo, et al.
Veröffentlicht: (2024)
von: Xing, Shuo, et al.
Veröffentlicht: (2024)
Chirality Detection in Scanning Tunneling Microscopy Data Using Artificial Intelligence
von: Tim J. Seifert, et al.
Veröffentlicht: (2024)
von: Tim J. Seifert, et al.
Veröffentlicht: (2024)
An Integrated Approach to Importance Sampling and Machine Learning for Efficient Monte Carlo Estimation of Distortion Risk Measures in Black Box Models
von: Bettels, Sören, et al.
Veröffentlicht: (2024)
von: Bettels, Sören, et al.
Veröffentlicht: (2024)
Open‐Source Tools for the Fabrication and Characterization of Organic Electronics
von: Julian F. Butscher, et al.
Veröffentlicht: (2024)
von: Julian F. Butscher, et al.
Veröffentlicht: (2024)
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning
von: Alizadeh, Meysam, et al.
Veröffentlicht: (2023)
von: Alizadeh, Meysam, et al.
Veröffentlicht: (2023)
Streamlined Open-Vocabulary Human-Object Interaction Detection
von: Sun, Chang, et al.
Veröffentlicht: (2026)
von: Sun, Chang, et al.
Veröffentlicht: (2026)
OpenTwinMap: An Open-Source Digital Twin Generator for Urban Autonomous Driving
von: Richardson, Alex, et al.
Veröffentlicht: (2025)
von: Richardson, Alex, et al.
Veröffentlicht: (2025)
Open-Sora Plan: Open-Source Large Video Generation Model
von: Lin, Bin, et al.
Veröffentlicht: (2024)
von: Lin, Bin, et al.
Veröffentlicht: (2024)
Efficient High-Performance Bark-Scale Neural Network for Residual Echo and Noise Suppression
von: Seidel, Ernst, et al.
Veröffentlicht: (2024)
von: Seidel, Ernst, et al.
Veröffentlicht: (2024)
ViSoLex: An Open-Source Repository for Vietnamese Social Media Lexical Normalization
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2025)
Video-Browser: Towards Agentic Open-web Video Browsing
von: Liang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Liang, Zhengyang, et al.
Veröffentlicht: (2025)
Fine-Tuning Open Video Generators for Cinematic Scene Synthesis: A Small-Data Pipeline with LoRA and Wan2.1 I2V
von: Akarsu, Meftun, et al.
Veröffentlicht: (2025)
von: Akarsu, Meftun, et al.
Veröffentlicht: (2025)
Collaborative Dynamic 3D Scene Graphs for Open-Vocabulary Urban Scene Understanding
von: Steinke, Tim, et al.
Veröffentlicht: (2025)
von: Steinke, Tim, et al.
Veröffentlicht: (2025)
Application-Driven Pedagogical Knowledge Optimization of Open-Source LLMs via Reinforcement Learning and Supervised Fine-Tuning
von: Singh, Navan Preet, et al.
Veröffentlicht: (2026)
von: Singh, Navan Preet, et al.
Veröffentlicht: (2026)
Rethinking Scale: The Efficacy of Fine-Tuned Open-Source LLMs in Large-Scale Reproducible Social Science Research
von: Carammia, Marcello, et al.
Veröffentlicht: (2024)
von: Carammia, Marcello, et al.
Veröffentlicht: (2024)
Open Source, Open Threats? Investigating Security Challenges in Open-Source Software
von: Akhavani, Seyed Ali, et al.
Veröffentlicht: (2025)
von: Akhavani, Seyed Ali, et al.
Veröffentlicht: (2025)
Articulated 3D Scene Graphs for Open-World Mobile Manipulation
von: Büchner, Martin, et al.
Veröffentlicht: (2026)
von: Büchner, Martin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Noise-Robust AV-ASR Using Visual Features Both in the Whisper Encoder and Decoder
von: Li, Zhengyang, et al.
Veröffentlicht: (2026) -
Improving Block-Wise LLM Quantization by 4-bit Block-Wise Optimal Float (BOF4): Analysis and Variations
von: Blumenberg, Patrick, et al.
Veröffentlicht: (2025) -
Engineering of Hallucination in Generative AI: It's not a Bug, it's a Feature
von: Fingscheidt, Tim, et al.
Veröffentlicht: (2026) -
A Lightweight Image Super-Resolution Transformer Trained on Low-Resolution Images Only
von: Möller, Björn, et al.
Veröffentlicht: (2025) -
Foundation Models for Amodal Video Instance Segmentation in Automated Driving
von: Breitenstein, Jasmin, et al.
Veröffentlicht: (2024)