Hidden costs for inference with deep network on embedded system devices
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Chankyu, Choi, Woohyun, Park, Sangwook |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Hardness of Learning One Hidden Layer Neural Networks
von: Li, Shuchen, et al.
Veröffentlicht: (2024)
von: Li, Shuchen, et al.
Veröffentlicht: (2024)
From Pseudorandomness to Multi-Group Fairness and Back
von: Dwork, Cynthia, et al.
Veröffentlicht: (2023)
von: Dwork, Cynthia, et al.
Veröffentlicht: (2023)
A Quantitative Definition of Intelligence
von: Choi, Kang-Sin
Veröffentlicht: (2026)
von: Choi, Kang-Sin
Veröffentlicht: (2026)
Is uniform expressivity too restrictive? Towards efficient expressivity of graph neural networks
von: Khalife, Sammy, et al.
Veröffentlicht: (2024)
von: Khalife, Sammy, et al.
Veröffentlicht: (2024)
Noise-tolerant learnability of shallow quantum circuits from statistics and the cost of quantum pseudorandomness
von: Wadhwa, Chirag, et al.
Veröffentlicht: (2024)
von: Wadhwa, Chirag, et al.
Veröffentlicht: (2024)
How Hard Is Continuous Clustering? Lower Bounds from the Existential Theory of the Reals
von: Majumdar, Angshul
Veröffentlicht: (2026)
von: Majumdar, Angshul
Veröffentlicht: (2026)
Spiky Rank and Its Applications to Rigidity and Circuits
von: Hambardzumyan, Lianna, et al.
Veröffentlicht: (2026)
von: Hambardzumyan, Lianna, et al.
Veröffentlicht: (2026)
Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete
von: Li, Qian, et al.
Veröffentlicht: (2026)
von: Li, Qian, et al.
Veröffentlicht: (2026)
Decision Tree Learning on Product Spaces
von: Moakahr, Arshia Soltani, et al.
Veröffentlicht: (2026)
von: Moakahr, Arshia Soltani, et al.
Veröffentlicht: (2026)
Optimizing Computational-Statistical Runtime for Wasserstein Distance Estimation
von: Jacobs, Peter Matthew, et al.
Veröffentlicht: (2026)
von: Jacobs, Peter Matthew, et al.
Veröffentlicht: (2026)
On the Computational Hardness of Transformers
von: Saha, Barna, et al.
Veröffentlicht: (2026)
von: Saha, Barna, et al.
Veröffentlicht: (2026)
Certifiable Boolean Reasoning Is Universal
von: Li, Wenhao, et al.
Veröffentlicht: (2026)
von: Li, Wenhao, et al.
Veröffentlicht: (2026)
Sandwiching Polynomials for Geometric Concepts with Low Intrinsic Dimension
von: Klivans, Adam R., et al.
Veröffentlicht: (2026)
von: Klivans, Adam R., et al.
Veröffentlicht: (2026)
Polyhedral Instability Governs Regret in Online Learning
von: Li, Yuetai, et al.
Veröffentlicht: (2026)
von: Li, Yuetai, et al.
Veröffentlicht: (2026)
On the Hardness of Learning Regular Expressions
von: Attias, Idan, et al.
Veröffentlicht: (2025)
von: Attias, Idan, et al.
Veröffentlicht: (2025)
Learnability of Parameter-Bounded Bayes Nets
von: Bhattacharyya, Arnab, et al.
Veröffentlicht: (2024)
von: Bhattacharyya, Arnab, et al.
Veröffentlicht: (2024)
A Little Depth Goes a Long Way: The Expressive Power of Log-Depth Transformers
von: Merrill, William, et al.
Veröffentlicht: (2025)
von: Merrill, William, et al.
Veröffentlicht: (2025)
Low-Rank Matrix Approximation for Neural Network Compression
von: Cherukuri, Kalyan, et al.
Veröffentlicht: (2025)
von: Cherukuri, Kalyan, et al.
Veröffentlicht: (2025)
Proximity to Losslessly Compressible Parameters
von: Farrugia-Roberts, Matthew
Veröffentlicht: (2023)
von: Farrugia-Roberts, Matthew
Veröffentlicht: (2023)
Lower Bounds for Chain-of-Thought Reasoning in Hard-Attention Transformers
von: Amiri, Alireza, et al.
Veröffentlicht: (2025)
von: Amiri, Alireza, et al.
Veröffentlicht: (2025)
Statistical and Computational Guarantees of Kernel Max-Sliced Wasserstein Distances
von: Wang, Jie, et al.
Veröffentlicht: (2024)
von: Wang, Jie, et al.
Veröffentlicht: (2024)
Fundamental Limits of Crystalline Equivariant Graph Neural Networks: A Circuit Complexity Perspective
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
A Logic for Expressing Log-Precision Transformers
von: Merrill, William, et al.
Veröffentlicht: (2022)
von: Merrill, William, et al.
Veröffentlicht: (2022)
Smoothed Analysis for Learning Concepts with Low Intrinsic Dimension
von: Chandrasekaran, Gautam, et al.
Veröffentlicht: (2024)
von: Chandrasekaran, Gautam, et al.
Veröffentlicht: (2024)
Distribution-Specific Agnostic Conditional Classification With Halfspaces
von: Huang, Jizhou, et al.
Veröffentlicht: (2025)
von: Huang, Jizhou, et al.
Veröffentlicht: (2025)
Chain of Thought Empowers Transformers to Solve Inherently Serial Problems
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
Ask, and it shall be given: On the Turing completeness of prompting
von: Qiu, Ruizhong, et al.
Veröffentlicht: (2024)
von: Qiu, Ruizhong, et al.
Veröffentlicht: (2024)
Necessary and Sufficient Oracles: Toward a Computational Taxonomy For Reinforcement Learning
von: Rohatgi, Dhruv, et al.
Veröffentlicht: (2025)
von: Rohatgi, Dhruv, et al.
Veröffentlicht: (2025)
How Global Calibration Strengthens Multiaccuracy
von: Casacuberta, Sílvia, et al.
Veröffentlicht: (2025)
von: Casacuberta, Sílvia, et al.
Veröffentlicht: (2025)
Diffusion Language Models are Provably Optimal Parallel Samplers
von: Jiang, Haozhe, et al.
Veröffentlicht: (2025)
von: Jiang, Haozhe, et al.
Veröffentlicht: (2025)
Additive Models Explained: A Computational Complexity Approach
von: Bassan, Shahaf, et al.
Veröffentlicht: (2025)
von: Bassan, Shahaf, et al.
Veröffentlicht: (2025)
Large Language Models on Small Resource-Constrained Systems: Performance Characterization, Analysis and Trade-offs
von: Seymour, Liam, et al.
Veröffentlicht: (2024)
von: Seymour, Liam, et al.
Veröffentlicht: (2024)
Data Debugging is NP-hard for Classifiers Trained with SGD
von: Guo, Zizheng, et al.
Veröffentlicht: (2024)
von: Guo, Zizheng, et al.
Veröffentlicht: (2024)
Constant Bit-size Transformers Are Turing Complete
von: Li, Qian, et al.
Veröffentlicht: (2025)
von: Li, Qian, et al.
Veröffentlicht: (2025)
New Hardness Results for Low-Rank Matrix Completion
von: Chawin, Dror, et al.
Veröffentlicht: (2025)
von: Chawin, Dror, et al.
Veröffentlicht: (2025)
Reachability In Simple Neural Networks
von: Sälzer, Marco, et al.
Veröffentlicht: (2022)
von: Sälzer, Marco, et al.
Veröffentlicht: (2022)
Smoothed Agnostic Learning of Halfspaces over the Hypercube
von: Kou, Yiwen, et al.
Veröffentlicht: (2025)
von: Kou, Yiwen, et al.
Veröffentlicht: (2025)
Deep Learning as a Convex Paradigm of Computation: Minimizing Circuit Size with ResNets
von: Jacot, Arthur
Veröffentlicht: (2025)
von: Jacot, Arthur
Veröffentlicht: (2025)
fruit-SALAD: A Style Aligned Artwork Dataset to reveal similarity perception in image embeddings
von: Ohm, Tillmann, et al.
Veröffentlicht: (2024)
von: Ohm, Tillmann, et al.
Veröffentlicht: (2024)
When does Metropolized Hamiltonian Monte Carlo provably outperform Metropolis-adjusted Langevin algorithm?
von: Chen, Yuansi, et al.
Veröffentlicht: (2023)
von: Chen, Yuansi, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
On the Hardness of Learning One Hidden Layer Neural Networks
von: Li, Shuchen, et al.
Veröffentlicht: (2024) -
From Pseudorandomness to Multi-Group Fairness and Back
von: Dwork, Cynthia, et al.
Veröffentlicht: (2023) -
A Quantitative Definition of Intelligence
von: Choi, Kang-Sin
Veröffentlicht: (2026) -
Is uniform expressivity too restrictive? Towards efficient expressivity of graph neural networks
von: Khalife, Sammy, et al.
Veröffentlicht: (2024) -
Noise-tolerant learnability of shallow quantum circuits from statistics and the cost of quantum pseudorandomness
von: Wadhwa, Chirag, et al.
Veröffentlicht: (2024)