CUROCKET: Optimizing ROCKET for GPU
Fuente:
arXiv
Saved in:
| Main Authors: | Stüven, Ole, Moenck, Keno, Schüppstuhl, Thorsten |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HIT-ROCKET: Hadamard-vector Inner-product Transformer for ROCKET
by: Hao, Wang, et al.
Published: (2025)
by: Hao, Wang, et al.
Published: (2025)
Open-vocabulary 3D scene perception in industrial environments
by: Moenck, Keno, et al.
Published: (2026)
by: Moenck, Keno, et al.
Published: (2026)
Industrial Language-Image Dataset (ILID): Adapting Vision Foundation Models for Industrial Settings
by: Moenck, Keno, et al.
Published: (2024)
by: Moenck, Keno, et al.
Published: (2024)
ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression
by: Ali, Ammar, et al.
Published: (2026)
by: Ali, Ammar, et al.
Published: (2026)
Decentralized Time Series Classification with ROCKET Features
by: Casella, Bruno, et al.
Published: (2025)
by: Casella, Bruno, et al.
Published: (2025)
ROCKET-2: Steering Visuomotor Policy via Cross-View Goal Alignment
by: Cai, Shaofei, et al.
Published: (2025)
by: Cai, Shaofei, et al.
Published: (2025)
Wear Classification of Abrasive Flap Wheels using a Hierarchical Deep Learning Approach
by: Kähler, Falko, et al.
Published: (2026)
by: Kähler, Falko, et al.
Published: (2026)
GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization
by: Khan, Zaid, et al.
Published: (2026)
by: Khan, Zaid, et al.
Published: (2026)
A GPU-accelerated Large-scale Simulator for Transportation System Optimization Benchmarking
by: Zhang, Jun, et al.
Published: (2024)
by: Zhang, Jun, et al.
Published: (2024)
Tri-Accel: Curvature-Aware Precision-Adaptive and Memory-Elastic Optimization for Efficient GPU Usage
by: Sheibanian, Mohsen, et al.
Published: (2025)
by: Sheibanian, Mohsen, et al.
Published: (2025)
A layered architecture for log analysis in complex IT systems
by: Wittkopp, Thorsten
Published: (2025)
by: Wittkopp, Thorsten
Published: (2025)
Improving Efficiency of GPU Kernel Optimization Agents using a Domain-Specific Language and Speed-of-Light Guidance
by: Hari, Siva Kumar Sastry, et al.
Published: (2026)
by: Hari, Siva Kumar Sastry, et al.
Published: (2026)
GPU Memory Requirement Prediction for Deep Learning Task Based on Bidirectional Gated Recurrent Unit Optimization Transformer
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
KernelSkill: A Multi-Agent Framework for GPU Kernel Optimization
by: Sun, Qitong, et al.
Published: (2026)
by: Sun, Qitong, et al.
Published: (2026)
Omniwise: Predicting GPU Kernels Performance with LLMs
by: Wang, Zixian, et al.
Published: (2025)
by: Wang, Zixian, et al.
Published: (2025)
AgentKernelArena: Generalization-Aware Benchmarking of GPU Kernel Optimization Agents
by: Younesian, Sharareh, et al.
Published: (2026)
by: Younesian, Sharareh, et al.
Published: (2026)
Scaling On-Device GPU Inference for Large Generative Models
by: Tang, Jiuqiang, et al.
Published: (2025)
by: Tang, Jiuqiang, et al.
Published: (2025)
Prompt Optimization with Logged Bandit Data
by: Kiyohara, Haruka, et al.
Published: (2025)
by: Kiyohara, Haruka, et al.
Published: (2025)
Low-distortion and GPU-compatible Tree Embeddings in Hyperbolic Space
by: van Spengler, Max, et al.
Published: (2025)
by: van Spengler, Max, et al.
Published: (2025)
GPU Kernel Scientist: An LLM-Driven Framework for Iterative Kernel Optimization
by: Andrews, Martin, et al.
Published: (2025)
by: Andrews, Martin, et al.
Published: (2025)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
by: Nishimori, Soichiro, et al.
Published: (2026)
by: Nishimori, Soichiro, et al.
Published: (2026)
JaxWildfire: A GPU-Accelerated Wildfire Simulator for Reinforcement Learning
by: Çakır, Ufuk, et al.
Published: (2025)
by: Çakır, Ufuk, et al.
Published: (2025)
Implementation and Analysis of GPU Algorithms for Vecchia Approximation
by: James, Zachary, et al.
Published: (2024)
by: James, Zachary, et al.
Published: (2024)
AI for the prediction of early stages of Alzheimer's disease from neuroimaging biomarkers -- A narrative review of a growing field
by: Rudroff, Thorsten, et al.
Published: (2024)
by: Rudroff, Thorsten, et al.
Published: (2024)
Novel GPU Boruta algorithms for feature selection from high-dimensional data
by: Li, Xurui, et al.
Published: (2026)
by: Li, Xurui, et al.
Published: (2026)
GPU-Accelerated Deep Learning for Heatwave Prediction and Urban Heat Risk Assessment
by: Alihodžić, Adis
Published: (2026)
by: Alihodžić, Adis
Published: (2026)
Towards Self-Supervised Foundation Models for Critical Care Time Series
by: Jagd, Katja Naasunnguaq, et al.
Published: (2025)
by: Jagd, Katja Naasunnguaq, et al.
Published: (2025)
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
by: Zhang, Tuo, et al.
Published: (2025)
by: Zhang, Tuo, et al.
Published: (2025)
An Optimized Toolbox for Advanced Image Processing with Tsetlin Machine Composites
by: Grønningsæter, Ylva, et al.
Published: (2024)
by: Grønningsæter, Ylva, et al.
Published: (2024)
GPU-accelerated simulated annealing based on p-bits with real-world device-variability modeling
by: Onizawa, Naoya, et al.
Published: (2026)
by: Onizawa, Naoya, et al.
Published: (2026)
StructPrune: Structured Global Pruning asymptotics with $\mathcal{O}(\sqrt{N})$ GPU Memory
by: Song, Xinyuan, et al.
Published: (2025)
by: Song, Xinyuan, et al.
Published: (2025)
Exploring State Space and Reasoning by Elimination in Tsetlin Machines
by: Kadhim, Ahmed K., et al.
Published: (2024)
by: Kadhim, Ahmed K., et al.
Published: (2024)
SOL-ExecBench: Speed-of-Light Benchmarking for Real-World GPU Kernels Against Hardware Limits
by: Lin, Edward, et al.
Published: (2026)
by: Lin, Edward, et al.
Published: (2026)
Fisher-Bingham-like normalizing flows on the sphere
by: Glüsenkamp, Thorsten
Published: (2025)
by: Glüsenkamp, Thorsten
Published: (2025)
Agentic retrieval-augmented reasoning reshapes collective reliability under model variability in radiology question answering
by: Farajiamiri, Mina, et al.
Published: (2026)
by: Farajiamiri, Mina, et al.
Published: (2026)
An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference
by: Yao, Feiyu, et al.
Published: (2026)
by: Yao, Feiyu, et al.
Published: (2026)
FastKernels: Benchmarking GPU Kernel Generation in Production
by: Oliaro, Gabriele, et al.
Published: (2026)
by: Oliaro, Gabriele, et al.
Published: (2026)
Attention on the Sphere
by: Bonev, Boris, et al.
Published: (2025)
by: Bonev, Boris, et al.
Published: (2025)
Hybrid Learning and Optimization-Based Dynamic Scheduling for DL Workloads on Heterogeneous GPU Clusters
by: Dongare, Shruti, et al.
Published: (2025)
by: Dongare, Shruti, et al.
Published: (2025)
Online GPU Energy Optimization with Switching-Aware Bandits
by: Xu, Xiongxiao, et al.
Published: (2024)
by: Xu, Xiongxiao, et al.
Published: (2024)
Similar Items
-
HIT-ROCKET: Hadamard-vector Inner-product Transformer for ROCKET
by: Hao, Wang, et al.
Published: (2025) -
Open-vocabulary 3D scene perception in industrial environments
by: Moenck, Keno, et al.
Published: (2026) -
Industrial Language-Image Dataset (ILID): Adapting Vision Foundation Models for Industrial Settings
by: Moenck, Keno, et al.
Published: (2024) -
ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression
by: Ali, Ammar, et al.
Published: (2026) -
Decentralized Time Series Classification with ROCKET Features
by: Casella, Bruno, et al.
Published: (2025)