Carbon Intensity-Aware Adaptive Inference of DNNs
Fuente:
arXiv
Saved in:
| Main Author: | Jung, Jiwan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HAWX: A Hardware-Aware FrameWork for Fast and Scalable ApproXimation of DNNs
by: Nazari, Samira, et al.
Published: (2026)
by: Nazari, Samira, et al.
Published: (2026)
Adaptive Graph Rewiring to Mitigate Over-Squashing in Mesh-Based GNNs for Fluid Dynamics Simulations
by: Seo, Sangwoo, et al.
Published: (2025)
by: Seo, Sangwoo, et al.
Published: (2025)
Arithmetic-Intensity-Aware Quantization
by: Singh, Taig, et al.
Published: (2025)
by: Singh, Taig, et al.
Published: (2025)
CEAR: Certified Ensemble Adversarial Robustness in DNNs
by: Sadig, Daniel, et al.
Published: (2026)
by: Sadig, Daniel, et al.
Published: (2026)
DART: Input-Difficulty-AwaRe Adaptive Threshold for Early-Exit DNNs
by: Patne, Parth, et al.
Published: (2026)
by: Patne, Parth, et al.
Published: (2026)
How DNNs break the Curse of Dimensionality: Compositionality and Symmetry Learning
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Role-Aware Conditional Inference for Spatiotemporal Ecosystem Carbon Flux Prediction
by: Sun, Yiming, et al.
Published: (2026)
by: Sun, Yiming, et al.
Published: (2026)
Rapid Deployment of DNNs for Edge Computing via Structured Pruning at Initialization
by: Eccles, Bailey J., et al.
Published: (2024)
by: Eccles, Bailey J., et al.
Published: (2024)
Efficient Noise Mitigation for Enhancing Inference Accuracy in DNNs on Mixed-Signal Accelerators
by: Azizi, Seyedarmin, et al.
Published: (2024)
by: Azizi, Seyedarmin, et al.
Published: (2024)
ECQ$^{\text{x}}$: Explainability-Driven Quantization for Low-Bit and Sparse DNNs
by: Becking, Daniel, et al.
Published: (2021)
by: Becking, Daniel, et al.
Published: (2021)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
Efficient Triple Modular Redundancy for Reliability Enhancement of DNNs Using Explainable AI
by: Soroush, Kimia, et al.
Published: (2025)
by: Soroush, Kimia, et al.
Published: (2025)
Relationship between Uncertainty in DNNs and Adversarial Attacks
by: Ogonna, Mabel, et al.
Published: (2024)
by: Ogonna, Mabel, et al.
Published: (2024)
Random weights of DNNs and emergence of fixed points
by: Berlyand, L., et al.
Published: (2025)
by: Berlyand, L., et al.
Published: (2025)
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
by: Lu, Guanxi, et al.
Published: (2025)
by: Lu, Guanxi, et al.
Published: (2025)
Bayesian Critique-Tune-Based Reinforcement Learning with Adaptive Pressure for Multi-Intersection Traffic Signal Control
by: Duan, Wenchang, et al.
Published: (2024)
by: Duan, Wenchang, et al.
Published: (2024)
Instance-Adaptive Parametrization for Amortized Variational Inference
by: Pollastro, Andrea, et al.
Published: (2026)
by: Pollastro, Andrea, et al.
Published: (2026)
RAP: Runtime Adaptive Pruning for LLM Inference
by: Liu, Huanrong, et al.
Published: (2025)
by: Liu, Huanrong, et al.
Published: (2025)
TRINE: A Token-Aware, Runtime-Adaptive FPGA Inference Engine for Multimodal AI
by: Oh, Hyunwoo, et al.
Published: (2026)
by: Oh, Hyunwoo, et al.
Published: (2026)
JADAI: Jointly Amortizing Adaptive Design and Bayesian Inference
by: Bracher, Niels, et al.
Published: (2025)
by: Bracher, Niels, et al.
Published: (2025)
Defining and Extracting generalizable interaction primitives from DNNs
by: Chen, Lu, et al.
Published: (2024)
by: Chen, Lu, et al.
Published: (2024)
Hardware/Software Co-Design of RISC-V Extensions for Accelerating Sparse DNNs on FPGAs
by: Sabih, Muhammad, et al.
Published: (2025)
by: Sabih, Muhammad, et al.
Published: (2025)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
Multi-LLM Adaptive Conformal Inference for Reliable LLM Responses
by: Noh, Kangjun, et al.
Published: (2026)
by: Noh, Kangjun, et al.
Published: (2026)
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
by: Yan, Jiaming, et al.
Published: (2025)
by: Yan, Jiaming, et al.
Published: (2025)
Improving Day-Ahead Grid Carbon Intensity Forecasting by Joint Modeling of Local-Temporal and Cross-Variable Dependencies Across Different Frequencies
by: Zhang, Bowen, et al.
Published: (2026)
by: Zhang, Bowen, et al.
Published: (2026)
3D Optimization for AI Inference Scaling: Balancing Accuracy, Cost, and Latency
by: Jung, Minseok, et al.
Published: (2025)
by: Jung, Minseok, et al.
Published: (2025)
TIDE: Temporal Incremental Draft Engine for Self-Improving LLM Inference
by: Park, Jiyoung, et al.
Published: (2026)
by: Park, Jiyoung, et al.
Published: (2026)
Magic for the Age of Quantized DNNs
by: Sawada, Yoshihide, et al.
Published: (2024)
by: Sawada, Yoshihide, et al.
Published: (2024)
Adaptive Edge Learning for Density-Aware Graph Generation
by: Razavi, Seyedeh Ava Razi, et al.
Published: (2026)
by: Razavi, Seyedeh Ava Razi, et al.
Published: (2026)
FAAR: Format-Aware Adaptive Rounding for NVFP4
by: Li, Hanglin, et al.
Published: (2026)
by: Li, Hanglin, et al.
Published: (2026)
ADWIN: Adaptive Windows for Horizon-Aware On-Policy Distillation
by: Liang, Kun, et al.
Published: (2026)
by: Liang, Kun, et al.
Published: (2026)
Avoiding Premature Collapse: Adaptive Annealing for Entropy-Regularized Structural Inference
by: Liu, Yizhi
Published: (2026)
by: Liu, Yizhi
Published: (2026)
Bounding-Box Inference for Error-Aware Model-Based Reinforcement Learning
by: Talvitie, Erin J., et al.
Published: (2024)
by: Talvitie, Erin J., et al.
Published: (2024)
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
by: Tan, Qitao, et al.
Published: (2025)
by: Tan, Qitao, et al.
Published: (2025)
Efficient Post-Training Augmentation for Adaptive Inference in Heterogeneous and Distributed IoT Environments
by: Sponner, Max, et al.
Published: (2024)
by: Sponner, Max, et al.
Published: (2024)
RAMP: Reinforcement Adaptive Mixed Precision Quantization for Efficient On Device LLM Inference
by: Gautam, Arpit Singh, et al.
Published: (2026)
by: Gautam, Arpit Singh, et al.
Published: (2026)
CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration
by: Han, Yuning, et al.
Published: (2026)
by: Han, Yuning, et al.
Published: (2026)
Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs
by: Rottoli, Michael, et al.
Published: (2026)
by: Rottoli, Michael, et al.
Published: (2026)
AdaQAT: Adaptive Bit-Width Quantization-Aware Training
by: Gernigon, Cédric, et al.
Published: (2024)
by: Gernigon, Cédric, et al.
Published: (2024)
Similar Items
-
HAWX: A Hardware-Aware FrameWork for Fast and Scalable ApproXimation of DNNs
by: Nazari, Samira, et al.
Published: (2026) -
Adaptive Graph Rewiring to Mitigate Over-Squashing in Mesh-Based GNNs for Fluid Dynamics Simulations
by: Seo, Sangwoo, et al.
Published: (2025) -
Arithmetic-Intensity-Aware Quantization
by: Singh, Taig, et al.
Published: (2025) -
CEAR: Certified Ensemble Adversarial Robustness in DNNs
by: Sadig, Daniel, et al.
Published: (2026) -
DART: Input-Difficulty-AwaRe Adaptive Threshold for Early-Exit DNNs
by: Patne, Parth, et al.
Published: (2026)