When Servers Meet Species: A Fab-to-Grave Lens on Computing's Biodiversity Impact
Fuente:
arXiv
Guardado en:
| Autores principales: | Shi, Tianyao, Kumar, Ritbik, Hua, Inez, Ding, Yi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Not All Water Consumption Is Equal: A Water Stress Weighted Metric for Sustainable Computing
por: Wu, Yanran, et al.
Publicado: (2025)
por: Wu, Yanran, et al.
Publicado: (2025)
GreenLLM: Disaggregating Large Language Model Serving on Heterogeneous GPUs for Lower Carbon Emissions
por: Shi, Tianyao, et al.
Publicado: (2024)
por: Shi, Tianyao, et al.
Publicado: (2024)
Cache Your Prompt When It's Green: Carbon-Aware Caching for Large Language Model Serving
por: Tian, Yuyang, et al.
Publicado: (2025)
por: Tian, Yuyang, et al.
Publicado: (2025)
Systematic Characterization of LLM Quantization: A Performance, Energy, and Quality Perspective
por: Shi, Tianyao, et al.
Publicado: (2025)
por: Shi, Tianyao, et al.
Publicado: (2025)
Revisiting Computational Storage for Data Integrity and Security
por: Shi, Chao, et al.
Publicado: (2025)
por: Shi, Chao, et al.
Publicado: (2025)
CCSS: Hardware-Accelerated RTL Simulation with Fast Combinational Logic Computing and Sequential Logic Synchronization
por: Feng, Weigang, et al.
Publicado: (2025)
por: Feng, Weigang, et al.
Publicado: (2025)
Memory-Centric Computing: Solving Computing's Memory Problem
por: Mutlu, Onur, et al.
Publicado: (2025)
por: Mutlu, Onur, et al.
Publicado: (2025)
Towards Compute-Aware In-Switch Computing for LLMs Tensor-Parallelism on Multi-GPU Systems
por: Zhang, Chen, et al.
Publicado: (2026)
por: Zhang, Chen, et al.
Publicado: (2026)
Optimizing Task Scheduling in Fog Computing with Deadline Awareness
por: Sirjani, Mohammad Sadegh, et al.
Publicado: (2025)
por: Sirjani, Mohammad Sadegh, et al.
Publicado: (2025)
Memory-Centric Computing: Recent Advances in Processing-in-DRAM
por: Mutlu, Onur, et al.
Publicado: (2024)
por: Mutlu, Onur, et al.
Publicado: (2024)
Optimizing Distributed ML Communication with Fused Computation-Collective Operations
por: Punniyamurthy, Kishore, et al.
Publicado: (2023)
por: Punniyamurthy, Kishore, et al.
Publicado: (2023)
Optimizing ML Concurrent Computation and Communication with GPU DMA Engines
por: Agrawal, Anirudha, et al.
Publicado: (2024)
por: Agrawal, Anirudha, et al.
Publicado: (2024)
Accelerating MoE with Dynamic In-Switch Computing on Multi-GPUs
por: Zhang, Qijun, et al.
Publicado: (2026)
por: Zhang, Qijun, et al.
Publicado: (2026)
Knowledge-Guided Attention-Inspired Learning for Task Offloading in Vehicle Edge Computing
por: Ma, Ke, et al.
Publicado: (2025)
por: Ma, Ke, et al.
Publicado: (2025)
BlockAMC: Scalable In-Memory Analog Matrix Computing for Solving Linear Systems
por: Pan, Lunshuai, et al.
Publicado: (2024)
por: Pan, Lunshuai, et al.
Publicado: (2024)
Deep Learning and Machine Learning with GPGPU and CUDA: Unlocking the Power of Parallel Computing
por: Li, Ming, et al.
Publicado: (2024)
por: Li, Ming, et al.
Publicado: (2024)
How Fast Can Graph Computations Go on Fine-grained Parallel Architectures
por: Wang, Yuqing, et al.
Publicado: (2025)
por: Wang, Yuqing, et al.
Publicado: (2025)
UniFormer: Unified and Efficient Transformer for Reasoning Across General and Custom Computing
por: Ran, Zhuoheng, et al.
Publicado: (2025)
por: Ran, Zhuoheng, et al.
Publicado: (2025)
RapidOMS: FPGA-based Open Modification Spectral Library Searching with HD Computing
por: Pinge, Sumukh, et al.
Publicado: (2024)
por: Pinge, Sumukh, et al.
Publicado: (2024)
FpgaHub: Fpga-centric Hyper-heterogeneous Computing Platform for Big Data Analytics
por: Wang, Zeke, et al.
Publicado: (2025)
por: Wang, Zeke, et al.
Publicado: (2025)
PULSAR: Simultaneous Many-Row Activation for Reliable and High-Performance Computing in Off-the-Shelf DRAM Chips
por: Yuksel, Ismail Emir, et al.
Publicado: (2023)
por: Yuksel, Ismail Emir, et al.
Publicado: (2023)
Next-generation Probabilistic Computing Hardware with 3D MOSAICs, Illusion Scale-up, and Co-design
por: Srimani, Tathagata, et al.
Publicado: (2024)
por: Srimani, Tathagata, et al.
Publicado: (2024)
Conduit: Programmer-Transparent Near-Data Processing Using Multiple Compute-Capable Resources in Solid State Drives
por: Nadig, Rakesh, et al.
Publicado: (2026)
por: Nadig, Rakesh, et al.
Publicado: (2026)
WWW: What, When, Where to Compute-in-Memory
por: Sharma, Tanvi, et al.
Publicado: (2023)
por: Sharma, Tanvi, et al.
Publicado: (2023)
Torrent: A Distributed DMA for Efficient and Flexible Point-to-Multipoint Data Movement
por: Deng, Yunhao, et al.
Publicado: (2025)
por: Deng, Yunhao, et al.
Publicado: (2025)
XDMA: A Distributed, Extensible DMA Architecture for Layout-Flexible Data Movements in Heterogeneous Multi-Accelerator SoCs
por: Kong, Fanchen, et al.
Publicado: (2025)
por: Kong, Fanchen, et al.
Publicado: (2025)
EDEA: Efficient Dual-Engine Accelerator for Depthwise Separable Convolution with Direct Data Transfer
por: Chen, Yi, et al.
Publicado: (2025)
por: Chen, Yi, et al.
Publicado: (2025)
FlexVector: A SpMM Vector Processor with Flexible VRF for GCNs on Varying-Sparsity Graphs
por: Li, Bohan, et al.
Publicado: (2026)
por: Li, Bohan, et al.
Publicado: (2026)
Data-aware Dynamic Execution of Irregular Workloads on Heterogeneous Systems
por: Bai, Zhenyu, et al.
Publicado: (2025)
por: Bai, Zhenyu, et al.
Publicado: (2025)
MLDSE: Scaling Design Space Exploration Infrastructure for Multi-Level Hardware
por: Qu, Huanyu, et al.
Publicado: (2025)
por: Qu, Huanyu, et al.
Publicado: (2025)
CMDS: Cross-layer Dataflow Optimization for DNN Accelerators Exploiting Multi-bank Memories
por: Shi, Man, et al.
Publicado: (2024)
por: Shi, Man, et al.
Publicado: (2024)
Exploring energy consumption of AI frameworks on a 64-core RV64 Server CPU
por: Malenza, Giulio, et al.
Publicado: (2025)
por: Malenza, Giulio, et al.
Publicado: (2025)
Pooling Engram Conditional Memory in Large Language Models using CXL
por: Ma, Ruiyang, et al.
Publicado: (2026)
por: Ma, Ruiyang, et al.
Publicado: (2026)
Adaptive Multi-Objective Tiered Storage Configuration for KV Cache in LLM Service
por: Zheng, Xianzhe, et al.
Publicado: (2026)
por: Zheng, Xianzhe, et al.
Publicado: (2026)
TT-Edge: A Hardware-Software Co-Design for Energy-Efficient Tensor-Train Decomposition on Edge AI
por: Kwak, Hyunseok, et al.
Publicado: (2025)
por: Kwak, Hyunseok, et al.
Publicado: (2025)
The DEEP-ER project: I/O and resiliency extensions for the Cluster-Booster architecture
por: Kreuzer, Anke, et al.
Publicado: (2019)
por: Kreuzer, Anke, et al.
Publicado: (2019)
An Evaluation and Comparison of GPU Hardware and Solver Libraries for Accelerating the OPM Flow Reservoir Simulator
por: Qiu, Tong Dong, et al.
Publicado: (2023)
por: Qiu, Tong Dong, et al.
Publicado: (2023)
SLIM: A Heterogeneous Accelerator for Edge Inference of Sparse Large Language Model via Adaptive Thresholding
por: Xu, Weihong, et al.
Publicado: (2025)
por: Xu, Weihong, et al.
Publicado: (2025)
COMET: A Framework for Modeling Compound Operation Dataflows with Explicit Collectives
por: Negi, Shubham, et al.
Publicado: (2025)
por: Negi, Shubham, et al.
Publicado: (2025)
MVDRAM: Enabling GeMV Execution in Unmodified DRAM for Low-Bit LLM Acceleration
por: Kubo, Tatsuya, et al.
Publicado: (2025)
por: Kubo, Tatsuya, et al.
Publicado: (2025)
Ejemplares similares
-
Not All Water Consumption Is Equal: A Water Stress Weighted Metric for Sustainable Computing
por: Wu, Yanran, et al.
Publicado: (2025) -
GreenLLM: Disaggregating Large Language Model Serving on Heterogeneous GPUs for Lower Carbon Emissions
por: Shi, Tianyao, et al.
Publicado: (2024) -
Cache Your Prompt When It's Green: Carbon-Aware Caching for Large Language Model Serving
por: Tian, Yuyang, et al.
Publicado: (2025) -
Systematic Characterization of LLM Quantization: A Performance, Energy, and Quality Perspective
por: Shi, Tianyao, et al.
Publicado: (2025) -
Revisiting Computational Storage for Data Integrity and Security
por: Shi, Chao, et al.
Publicado: (2025)