Saved in:
| Main Author: | Morabia, Nataraj Agaram Sundar Tejas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2606.01472 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compliance-Scored Best-of-N Guardrail Orchestration for Multimodal Document Generation in Payments Dispute Defense
by: Sundar, Nataraj Agaram, et al.
Published: (2026)
by: Sundar, Nataraj Agaram, et al.
Published: (2026)
Caliper-in-the-Loop: Black-Box Optimization for Hyperledger Fabric Performance Tuning
by: Madhwal, Yash, et al.
Published: (2026)
by: Madhwal, Yash, et al.
Published: (2026)
From Prompts to Performance: Evaluating LLMs for Task-based Parallel Code Generation
by: Bantel, Linus, et al.
Published: (2026)
by: Bantel, Linus, et al.
Published: (2026)
Hierarchical Autoscaling for Large Language Model Serving with Chiron
by: Patke, Archit, et al.
Published: (2025)
by: Patke, Archit, et al.
Published: (2025)
Harnessing Deep Learning and HPC Kernels via High-Level Loop and Tensor Abstractions on CPU Architectures
by: Georganas, Evangelos, et al.
Published: (2023)
by: Georganas, Evangelos, et al.
Published: (2023)
PipeSpec: Breaking Stage Dependencies in Hierarchical LLM Decoding
by: McDanel, Bradley, et al.
Published: (2025)
by: McDanel, Bradley, et al.
Published: (2025)
Optimizing Data Distribution and Kernel Performance for Efficient Training of Chemistry Foundation Models: A Case Study with MACE
by: Firoz, Jesun, et al.
Published: (2025)
by: Firoz, Jesun, et al.
Published: (2025)
Power- and Fragmentation-aware Online Scheduling for GPU Datacenters
by: Lettich, Francesco, et al.
Published: (2024)
by: Lettich, Francesco, et al.
Published: (2024)
DP2FL: Dual Prompt Personalized Federated Learning in Foundation Models
by: Chang, Ying, et al.
Published: (2025)
by: Chang, Ying, et al.
Published: (2025)
Dual-Segment Clustering Strategy for Hierarchical Federated Learning in Heterogeneous Wireless Environments
by: Sun, Pengcheng, et al.
Published: (2024)
by: Sun, Pengcheng, et al.
Published: (2024)
Hardware-Aware Reformulation of Convolutions for Efficient Execution on Specialized AI Hardware: A Case Study on NVIDIA Tensor Cores
by: Bikshandi, Ganesh
Published: (2026)
by: Bikshandi, Ganesh
Published: (2026)
FedRAV: Hierarchically Federated Region-Learning for Traffic Object Classification of Autonomous Vehicles
by: Zhai, Yijun, et al.
Published: (2024)
by: Zhai, Yijun, et al.
Published: (2024)
Venus: An Efficient Edge Memory-and-Retrieval System for VLM-based Online Video Understanding
by: Ye, Shengyuan, et al.
Published: (2025)
by: Ye, Shengyuan, et al.
Published: (2025)
NeurLZ: An Online Neural Learning-Based Method to Enhance Scientific Lossy Compression
by: Jia, Wenqi, et al.
Published: (2024)
by: Jia, Wenqi, et al.
Published: (2024)
Hierarchical Federated Learning for Crop Yield Prediction in Smart Agricultural Production Systems
by: Abouaomar, Anas, et al.
Published: (2025)
by: Abouaomar, Anas, et al.
Published: (2025)
The Case for Co-Designing Model Architectures with Hardware
by: Anthony, Quentin, et al.
Published: (2024)
by: Anthony, Quentin, et al.
Published: (2024)
AI Benchmarks and Datasets for LLM Evaluation
by: Ivanov, Todor, et al.
Published: (2024)
by: Ivanov, Todor, et al.
Published: (2024)
D$^{2}$MoE: Dual Routing and Dynamic Scheduling for Efficient On-Device MoE-based LLM Serving
by: Wang, Haodong, et al.
Published: (2025)
by: Wang, Haodong, et al.
Published: (2025)
Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks
by: Chandrasekar, Ashok, et al.
Published: (2026)
by: Chandrasekar, Ashok, et al.
Published: (2026)
An Evaluation of LLMs Inference on Popular Single-board Computers
by: Tung, et al.
Published: (2025)
by: Tung, et al.
Published: (2025)
Evaluating the Efficacy of LLM-Based Reasoning for Multiobjective HPC Job Scheduling
by: Jadhav, Prachi, et al.
Published: (2025)
by: Jadhav, Prachi, et al.
Published: (2025)
Benchmarking Federated Learning in Edge Computing Environments: A Systematic Review and Performance Evaluation
by: Aribe Jr., Sales, et al.
Published: (2026)
by: Aribe Jr., Sales, et al.
Published: (2026)
Distributed LLM Pretraining During Renewable Curtailment Windows: A Feasibility Study
by: Wiesner, Philipp, et al.
Published: (2026)
by: Wiesner, Philipp, et al.
Published: (2026)
CloudEval-YAML: A Practical Benchmark for Cloud Configuration Generation
by: Xu, Yifei, et al.
Published: (2023)
by: Xu, Yifei, et al.
Published: (2023)
StreamWise: Serving Multi-Modal Generation in Real-Time at Scale
by: Qiu, Haoran, et al.
Published: (2026)
by: Qiu, Haoran, et al.
Published: (2026)
Adaptive Graph Pruning with Sudden-Events Evaluation for Traffic Prediction using Online Semi-Decentralized ST-GNNs
by: Kralj, Ivan, et al.
Published: (2025)
by: Kralj, Ivan, et al.
Published: (2025)
DGRAG: Distributed Graph-based Retrieval-Augmented Generation in Edge-Cloud Systems
by: Zhou, Wenqing, et al.
Published: (2025)
by: Zhou, Wenqing, et al.
Published: (2025)
ParaGAN: A Scalable Distributed Training Framework for Generative Adversarial Networks
by: Shi, Ziji, et al.
Published: (2024)
by: Shi, Ziji, et al.
Published: (2024)
Not All Errors Are Equal: A Systematic Study of Error Propagation in Large Language Model Inference
by: Huang, Yafan, et al.
Published: (2026)
by: Huang, Yafan, et al.
Published: (2026)
Accelerating Long-Tail Generation in Synchronous RLHF Training via Adaptive Tensor Parallelism
by: Zhao, Long, et al.
Published: (2026)
by: Zhao, Long, et al.
Published: (2026)
VibeCodeHPC: An Agent-Based Iterative Prompting Auto-Tuner for HPC Code Generation Using LLMs
by: Hayashi, Shun-ichiro, et al.
Published: (2025)
by: Hayashi, Shun-ichiro, et al.
Published: (2025)
ParaCodex: A Profiling-Guided Autonomous Coding Agent for Reliable Parallel Code Generation and Translation
by: Kaplan, Erel, et al.
Published: (2026)
by: Kaplan, Erel, et al.
Published: (2026)
HPCTransCompile: An AI Compiler Generated Dataset for High-Performance CUDA Transpilation and LLM Preliminary Exploration
by: Lv, Jiaqi, et al.
Published: (2025)
by: Lv, Jiaqi, et al.
Published: (2025)
An Upload-Efficient Scheme for Transferring Knowledge From a Server-Side Pre-trained Generator to Clients in Heterogeneous Federated Learning
by: Zhang, Jianqing, et al.
Published: (2024)
by: Zhang, Jianqing, et al.
Published: (2024)
DUAL-BLADE: Dual-Path NVMe-Direct KV-Cache Offloading for Edge LLM Inference
by: Jeong, Bodon, et al.
Published: (2026)
by: Jeong, Bodon, et al.
Published: (2026)
Clustered Federated Learning with Hierarchical Knowledge Distillation
by: Ahmad, Sabtain, et al.
Published: (2025)
by: Ahmad, Sabtain, et al.
Published: (2025)
Performance Evaluation of LLMs in Automated RDF Knowledge Graph Generation
by: Martin, Ioana Ramona, et al.
Published: (2026)
by: Martin, Ioana Ramona, et al.
Published: (2026)
Loop Improvement: An Efficient Approach for Extracting Shared Features from Heterogeneous Data without Central Server
by: Li, Fei, et al.
Published: (2024)
by: Li, Fei, et al.
Published: (2024)
Lodestar: An Online-Learning LLM Inference Router
by: Lim, Gangmuk, et al.
Published: (2026)
by: Lim, Gangmuk, et al.
Published: (2026)
Federated Hierarchical Clustering with Automatic Selection of Optimal Cluster Numbers
by: Zhang, Yue, et al.
Published: (2026)
by: Zhang, Yue, et al.
Published: (2026)
Similar Items
-
Compliance-Scored Best-of-N Guardrail Orchestration for Multimodal Document Generation in Payments Dispute Defense
by: Sundar, Nataraj Agaram, et al.
Published: (2026) -
Caliper-in-the-Loop: Black-Box Optimization for Hyperledger Fabric Performance Tuning
by: Madhwal, Yash, et al.
Published: (2026) -
From Prompts to Performance: Evaluating LLMs for Task-based Parallel Code Generation
by: Bantel, Linus, et al.
Published: (2026) -
Hierarchical Autoscaling for Large Language Model Serving with Chiron
by: Patke, Archit, et al.
Published: (2025) -
Harnessing Deep Learning and HPC Kernels via High-Level Loop and Tensor Abstractions on CPU Architectures
by: Georganas, Evangelos, et al.
Published: (2023)