GPUDrive: Data-driven, multi-agent driving simulation at 1 million FPS
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kazemkhani, Saman, Pandya, Aarav, Cornelisse, Daphne, Shacklett, Brennan, Vinitsky, Eugene |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Megakernel vs Wavefront GPU Path Tracing
von: Padilla, Rafael, et al.
Veröffentlicht: (2026)
von: Padilla, Rafael, et al.
Veröffentlicht: (2026)
A 129FPS Full HD Real-Time Accelerator for 3D Gaussian Splatting
von: Chang, Fang-Chi, et al.
Veröffentlicht: (2026)
von: Chang, Fang-Chi, et al.
Veröffentlicht: (2026)
Building reliable sim driving agents by scaling self-play
von: Cornelisse, Daphne, et al.
Veröffentlicht: (2025)
von: Cornelisse, Daphne, et al.
Veröffentlicht: (2025)
GauRast: Enhancing GPU Triangle Rasterizers to Accelerate 3D Gaussian Splatting
von: Li, Sixu, et al.
Veröffentlicht: (2025)
von: Li, Sixu, et al.
Veröffentlicht: (2025)
SCALE-Sim v3: A modular cycle-accurate systolic accelerator simulator for end-to-end system analysis
von: Raj, Ritik, et al.
Veröffentlicht: (2025)
von: Raj, Ritik, et al.
Veröffentlicht: (2025)
Enhancing software-hardware co-design for HEP by low-overhead profiling of single- and multi-threaded programs on diverse architectures with Adaptyst
von: Graczyk, Maksymilian, et al.
Veröffentlicht: (2025)
von: Graczyk, Maksymilian, et al.
Veröffentlicht: (2025)
Assessing Tenstorrent's RISC-V MatMul Acceleration Capabilities
von: Cavagna, Hiari Pizzini, et al.
Veröffentlicht: (2025)
von: Cavagna, Hiari Pizzini, et al.
Veröffentlicht: (2025)
HiKonv: Maximizing the Throughput of Quantized Convolution With Novel Bit-wise Management and Computation
von: Chen, Yao, et al.
Veröffentlicht: (2022)
von: Chen, Yao, et al.
Veröffentlicht: (2022)
Karatsuba Matrix Multiplication and its Efficient Custom Hardware Implementations
von: Pogue, Trevor E., et al.
Veröffentlicht: (2025)
von: Pogue, Trevor E., et al.
Veröffentlicht: (2025)
Silicon Showdown: Performance, Efficiency, and Ecosystem Barriers in Consumer-Grade LLM Inference
von: Javat, Abdurrahman, et al.
Veröffentlicht: (2026)
von: Javat, Abdurrahman, et al.
Veröffentlicht: (2026)
Strassen Multisystolic Array Hardware Architectures
von: Pogue, Trevor E., et al.
Veröffentlicht: (2025)
von: Pogue, Trevor E., et al.
Veröffentlicht: (2025)
Accelerating LLM Inference via Dynamic KV Cache Placement in Heterogeneous Memory System
von: Fang, Yunhua, et al.
Veröffentlicht: (2025)
von: Fang, Yunhua, et al.
Veröffentlicht: (2025)
KWT-Tiny: RISC-V Accelerated, Embedded Keyword Spotting Transformer
von: Al-Qawlaq, Aness, et al.
Veröffentlicht: (2024)
von: Al-Qawlaq, Aness, et al.
Veröffentlicht: (2024)
LLM-Driven Design Space Exploration of FPGA-based Accelerators
von: Sharma, Vinamra, et al.
Veröffentlicht: (2026)
von: Sharma, Vinamra, et al.
Veröffentlicht: (2026)
Data-Driven Power Modeling and Monitoring via Hardware Performance Counter Tracking
von: Mazzola, Sergio, et al.
Veröffentlicht: (2025)
von: Mazzola, Sergio, et al.
Veröffentlicht: (2025)
WaSP: Warp Scheduling to Mimic Prefetching in Graphics Workloads
von: Joseph, Diya, et al.
Veröffentlicht: (2024)
von: Joseph, Diya, et al.
Veröffentlicht: (2024)
Gaussian Blending Unit: An Edge GPU Plug-in for Real-Time Gaussian-Based Rendering in AR/VR
von: Ye, Zhifan, et al.
Veröffentlicht: (2025)
von: Ye, Zhifan, et al.
Veröffentlicht: (2025)
Cicero: Addressing Algorithmic and Architectural Bottlenecks in Neural Rendering by Radiance Warping and Memory Optimizations
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
Minimizing Ray Tracing Memory Traffic through Quantized Structures and Ray Stream Tracing
von: Grauer, Moritz, et al.
Veröffentlicht: (2025)
von: Grauer, Moritz, et al.
Veröffentlicht: (2025)
Vorion: A RISC-V GPU with Hardware-Accelerated 3D Gaussian Rendering and Training
von: Wang, Yipeng, et al.
Veröffentlicht: (2025)
von: Wang, Yipeng, et al.
Veröffentlicht: (2025)
Potamoi: Accelerating Neural Rendering via a Unified Streaming Architecture
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
GEMM-GS: Accelerating 3D Gaussian Splatting on Tensor Cores with GEMM-Compatible Blending
von: Li, Haomin, et al.
Veröffentlicht: (2026)
von: Li, Haomin, et al.
Veröffentlicht: (2026)
A Quantitative Analysis and Guidelines of Data Streaming Accelerator in Modern Intel Xeon Scalable Processors
von: Kuper, Reese, et al.
Veröffentlicht: (2023)
von: Kuper, Reese, et al.
Veröffentlicht: (2023)
Characterizing VLA Models: Identifying the Action Generation Bottleneck for Edge AI Architectures
von: Vishwanathan, Manoj, et al.
Veröffentlicht: (2026)
von: Vishwanathan, Manoj, et al.
Veröffentlicht: (2026)
Heterogeneous Memory Benchmarking Toolkit
von: Ghaemi, Golsana, et al.
Veröffentlicht: (2025)
von: Ghaemi, Golsana, et al.
Veröffentlicht: (2025)
Simulation-Driven Evaluation of Chiplet-Based Architectures Using VisualSim
von: Ali, Wajid, et al.
Veröffentlicht: (2025)
von: Ali, Wajid, et al.
Veröffentlicht: (2025)
Enhancing Instruction Prefetching via Cache and TLB Management
von: Jamet, Alexandre Valentin, et al.
Veröffentlicht: (2026)
von: Jamet, Alexandre Valentin, et al.
Veröffentlicht: (2026)
ETM2: Empowering Traditional Memory Bandwidth Regulation using ETM
von: Zuepke, Alexander, et al.
Veröffentlicht: (2026)
von: Zuepke, Alexander, et al.
Veröffentlicht: (2026)
Towards CPU Performance Prediction: New Challenge Benchmark Dataset and Novel Approach
von: Liu, Xiaoman
Veröffentlicht: (2024)
von: Liu, Xiaoman
Veröffentlicht: (2024)
Adaptive Cache Pollution Control for Large Language Model Inference Workloads Using Temporal CNN-Based Prediction and Priority-Aware Replacement
von: Liu, Songze, et al.
Veröffentlicht: (2025)
von: Liu, Songze, et al.
Veröffentlicht: (2025)
Recurrent CircuitSAT Sampling for Sequential Circuits
von: Ardakani, Arash, et al.
Veröffentlicht: (2025)
von: Ardakani, Arash, et al.
Veröffentlicht: (2025)
Introducing the Arm-membench Throughput Benchmark
von: Burth, Cyrill, et al.
Veröffentlicht: (2025)
von: Burth, Cyrill, et al.
Veröffentlicht: (2025)
SPEC CPU2026: Characterization, Representativeness, and Cross-Suite Comparison
von: Li, Ruihao, et al.
Veröffentlicht: (2026)
von: Li, Ruihao, et al.
Veröffentlicht: (2026)
Single 32-bit Sub-Channel DDR5 DIMMs: Architecture, Performance Bounds, and Standardisation
von: Ke, Chih-Hua
Veröffentlicht: (2026)
von: Ke, Chih-Hua
Veröffentlicht: (2026)
Regular-Dead on Arrival: Characterizing and Protecting Against Dead-Entry TLB Misses in GPU Microarchitectures
von: Anik, Shafayat Mowla, et al.
Veröffentlicht: (2026)
von: Anik, Shafayat Mowla, et al.
Veröffentlicht: (2026)
Makinote: An FPGA-Based HW/SW Platform for Pre-Silicon Emulation of RISC-V Designs
von: Perdomo, Elias, et al.
Veröffentlicht: (2024)
von: Perdomo, Elias, et al.
Veröffentlicht: (2024)
AI Load Dynamics--A Power Electronics Perspective
von: Li, Yuzhuo, et al.
Veröffentlicht: (2025)
von: Li, Yuzhuo, et al.
Veröffentlicht: (2025)
SAHM: State-Aware Heterogeneous Multicore for Single-Thread Performance
von: Wadle, Shayne, et al.
Veröffentlicht: (2025)
von: Wadle, Shayne, et al.
Veröffentlicht: (2025)
ONNXim: A Fast, Cycle-level Multi-core NPU Simulator
von: Ham, Hyungkyu, et al.
Veröffentlicht: (2024)
von: Ham, Hyungkyu, et al.
Veröffentlicht: (2024)
LightningSimV2: Faster and Scalable Simulation for High-Level Synthesis via Graph Compilation and Optimization
von: Sarkar, Rishov, et al.
Veröffentlicht: (2024)
von: Sarkar, Rishov, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Megakernel vs Wavefront GPU Path Tracing
von: Padilla, Rafael, et al.
Veröffentlicht: (2026) -
A 129FPS Full HD Real-Time Accelerator for 3D Gaussian Splatting
von: Chang, Fang-Chi, et al.
Veröffentlicht: (2026) -
Building reliable sim driving agents by scaling self-play
von: Cornelisse, Daphne, et al.
Veröffentlicht: (2025) -
GauRast: Enhancing GPU Triangle Rasterizers to Accelerate 3D Gaussian Splatting
von: Li, Sixu, et al.
Veröffentlicht: (2025) -
SCALE-Sim v3: A modular cycle-accurate systolic accelerator simulator for end-to-end system analysis
von: Raj, Ritik, et al.
Veröffentlicht: (2025)