Elevating Semantic Exploration: A Novel Approach Utilizing Distributed Repositories
Fuente:
arXiv
Saved in:
| Main Author: | Bellandi, Valerio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
by: Ma, Qianli, et al.
Published: (2025)
by: Ma, Qianli, et al.
Published: (2025)
Placement Semantics for Distributed Deep Learning: A Systematic Framework for Analyzing Parallelism Strategies
by: Mehta, Deep Pankajbhai
Published: (2026)
by: Mehta, Deep Pankajbhai
Published: (2026)
SemanticForge: Repository-Level Code Generation through Semantic Knowledge Graphs and Constraint Satisfaction
by: Zhang, Wuyang, et al.
Published: (2025)
by: Zhang, Wuyang, et al.
Published: (2025)
Distributed Inference on Mobile Edge and Cloud: A Data-Cartography based Clustering Approach
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
by: Bajpai, Divya Jyoti, et al.
Published: (2024)
Glia: A Human-Inspired AI for Automated Systems Design and Optimization
by: Hamadanian, Pouya, et al.
Published: (2025)
by: Hamadanian, Pouya, et al.
Published: (2025)
FlowKV: A Disaggregated Inference Framework with Low-Latency KV Cache Transfer and Load-Aware Scheduling
by: Li, Weiqing, et al.
Published: (2025)
by: Li, Weiqing, et al.
Published: (2025)
FedSEA-LLaMA: A Secure, Efficient and Adaptive Federated Splitting Framework for Large Language Models
by: Zhang, Zishuai, et al.
Published: (2025)
by: Zhang, Zishuai, et al.
Published: (2025)
Assortment of Attention Heads: Accelerating Federated PEFT with Head Pruning and Strategic Client Selection
by: Venkatesha, Yeshwanth, et al.
Published: (2025)
by: Venkatesha, Yeshwanth, et al.
Published: (2025)
Training Foundation Models on a Full-Stack AMD Platform: Compute, Networking, and System Design
by: Anthony, Quentin, et al.
Published: (2025)
by: Anthony, Quentin, et al.
Published: (2025)
LeMix: Unified Scheduling for LLM Training and Inference on Multi-GPU Systems
by: Li, Yufei, et al.
Published: (2025)
by: Li, Yufei, et al.
Published: (2025)
Optimus: Accelerating Large-Scale Multi-Modal LLM Training by Bubble Exploitation
by: Feng, Weiqi, et al.
Published: (2024)
by: Feng, Weiqi, et al.
Published: (2024)
Regulating Branch Parallelism in LLM Serving
by: Gandhi, Swapnil, et al.
Published: (2026)
by: Gandhi, Swapnil, et al.
Published: (2026)
GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems
by: Wawdhane, Sourish, et al.
Published: (2026)
by: Wawdhane, Sourish, et al.
Published: (2026)
Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM
by: Kodavanti, Sravanth, et al.
Published: (2026)
by: Kodavanti, Sravanth, et al.
Published: (2026)
KV-Runahead: Scalable Causal LLM Inference by Parallel Key-Value Cache Generation
by: Cho, Minsik, et al.
Published: (2024)
by: Cho, Minsik, et al.
Published: (2024)
Hybrid-RACA: Hybrid Retrieval-Augmented Composition Assistance for Real-time Text Prediction
by: Xia, Menglin, et al.
Published: (2023)
by: Xia, Menglin, et al.
Published: (2023)
Hardware Utilization and Inference Performance of Edge Object Detection Under Fault Injection
by: Pasandideh, Faezeh, et al.
Published: (2026)
by: Pasandideh, Faezeh, et al.
Published: (2026)
Model Agnostic Hybrid Sharding For Heterogeneous Distributed Inference
by: Angione, Claudio, et al.
Published: (2024)
by: Angione, Claudio, et al.
Published: (2024)
HPCTransCompile: An AI Compiler Generated Dataset for High-Performance CUDA Transpilation and LLM Preliminary Exploration
by: Lv, Jiaqi, et al.
Published: (2025)
by: Lv, Jiaqi, et al.
Published: (2025)
Alto: Orchestrating Distributed Compound AI Systems with Nested Ancestry
by: Raghavan, Deepti, et al.
Published: (2024)
by: Raghavan, Deepti, et al.
Published: (2024)
A Sparsity Predicting Approach for Large Language Models via Activation Pattern Clustering
by: Dhar, Nobel, et al.
Published: (2025)
by: Dhar, Nobel, et al.
Published: (2025)
Data Driven Optimization of GPU efficiency for Distributed LLM Adapter Serving
by: Agullo, Ferran, et al.
Published: (2026)
by: Agullo, Ferran, et al.
Published: (2026)
Lightweight Trustworthy Distributed Clustering
by: Li, Hongyang, et al.
Published: (2025)
by: Li, Hongyang, et al.
Published: (2025)
Pact: A Choreographic Language for Agentic Ecosystems
by: Gopinathan, Kiran, et al.
Published: (2026)
by: Gopinathan, Kiran, et al.
Published: (2026)
DLoRA: Distributed Parameter-Efficient Fine-Tuning Solution for Large Language Model
by: Gao, Chao, et al.
Published: (2024)
by: Gao, Chao, et al.
Published: (2024)
Tiny but Mighty: A Software-Hardware Co-Design Approach for Efficient Multimodal Inference on Battery-Powered Small Devices
by: Li, Yilong, et al.
Published: (2025)
by: Li, Yilong, et al.
Published: (2025)
Distributed LLM Pretraining During Renewable Curtailment Windows: A Feasibility Study
by: Wiesner, Philipp, et al.
Published: (2026)
by: Wiesner, Philipp, et al.
Published: (2026)
HETHUB: A Distributed Training System with Heterogeneous Cluster for Large-Scale Models
by: Xu, Si, et al.
Published: (2024)
by: Xu, Si, et al.
Published: (2024)
ParaGAN: A Scalable Distributed Training Framework for Generative Adversarial Networks
by: Shi, Ziji, et al.
Published: (2024)
by: Shi, Ziji, et al.
Published: (2024)
Cloudless-Training: A Framework to Improve Efficiency of Geo-Distributed ML Training
by: Tan, Wenting, et al.
Published: (2023)
by: Tan, Wenting, et al.
Published: (2023)
WORKSWORLD: A Domain for Integrated Numeric Planning and Scheduling of Distributed Pipelined Workflows
by: Paul, Taylor, et al.
Published: (2026)
by: Paul, Taylor, et al.
Published: (2026)
Communication-Efficient Large-Scale Distributed Deep Learning: A Comprehensive Survey
by: Liang, Feng, et al.
Published: (2024)
by: Liang, Feng, et al.
Published: (2024)
Distributed Speculative Inference (DSI): Speculation Parallelism for Provably Faster Lossless Language Model Inference
by: Timor, Nadav, et al.
Published: (2024)
by: Timor, Nadav, et al.
Published: (2024)
Distributed Neural Representation for Reactive in situ Visualization
by: Wu, Qi, et al.
Published: (2023)
by: Wu, Qi, et al.
Published: (2023)
Demystifying the Communication Characteristics for Distributed Transformer Models
by: Anthony, Quentin, et al.
Published: (2024)
by: Anthony, Quentin, et al.
Published: (2024)
Mesh-Attention: A New Communication-Efficient Distributed Attention with Improved Data Locality
by: Chen, Sirui, et al.
Published: (2025)
by: Chen, Sirui, et al.
Published: (2025)
Resource Allocation and Workload Scheduling for Large-Scale Distributed Deep Learning: A Survey
by: Liang, Feng, et al.
Published: (2024)
by: Liang, Feng, et al.
Published: (2024)
EDiT: A Local-SGD-Based Efficient Distributed Training Method for Large Language Models
by: Cheng, Jialiang, et al.
Published: (2024)
by: Cheng, Jialiang, et al.
Published: (2024)
MoESys: A Distributed and Efficient Mixture-of-Experts Training and Inference System for Internet Services
by: Yu, Dianhai, et al.
Published: (2022)
by: Yu, Dianhai, et al.
Published: (2022)
Slicing Is All You Need: Towards A Universal One-Sided Algorithm for Distributed Matrix Multiplication
by: Brock, Benjamin, et al.
Published: (2025)
by: Brock, Benjamin, et al.
Published: (2025)
Similar Items
-
VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
by: Ma, Qianli, et al.
Published: (2025) -
Placement Semantics for Distributed Deep Learning: A Systematic Framework for Analyzing Parallelism Strategies
by: Mehta, Deep Pankajbhai
Published: (2026) -
SemanticForge: Repository-Level Code Generation through Semantic Knowledge Graphs and Constraint Satisfaction
by: Zhang, Wuyang, et al.
Published: (2025) -
Distributed Inference on Mobile Edge and Cloud: A Data-Cartography based Clustering Approach
by: Bajpai, Divya Jyoti, et al.
Published: (2024) -
Glia: A Human-Inspired AI for Automated Systems Design and Optimization
by: Hamadanian, Pouya, et al.
Published: (2025)