Towards Message Brokers for Generative AI: Survey, Challenges, and Opportunities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Saleh, Alaa, Morabito, Roberto, Tarkoma, Sasu, Pirttikangas, Susanna, Lovén, Lauri |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neural Router: Semantic Content Matching for Agentic AI
von: Lovén, Lauri, et al.
Veröffentlicht: (2026)
von: Lovén, Lauri, et al.
Veröffentlicht: (2026)
Autonomic Federated-Market Orchestration for the Edge-Cloud Continuum
von: Lovén, Lauri, et al.
Veröffentlicht: (2026)
von: Lovén, Lauri, et al.
Veröffentlicht: (2026)
Agentic TinyML for Intent-aware Handover in 6G Wireless Networks
von: Saleh, Alaa, et al.
Veröffentlicht: (2025)
von: Saleh, Alaa, et al.
Veröffentlicht: (2025)
Naeural AI OS -- Decentralized ubiquitous computing MLOps execution engine
von: Bleotiu, Cristian, et al.
Veröffentlicht: (2023)
von: Bleotiu, Cristian, et al.
Veröffentlicht: (2023)
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
von: Lan, Guangchen, et al.
Veröffentlicht: (2024)
von: Lan, Guangchen, et al.
Veröffentlicht: (2024)
A Framework for testing Federated Learning algorithms using an edge-like environment
von: Schwanck, Felipe Machado, et al.
Veröffentlicht: (2024)
von: Schwanck, Felipe Machado, et al.
Veröffentlicht: (2024)
Combinatorial Client-Master Multiagent Deep Reinforcement Learning for Task Offloading in Mobile Edge Computing
von: Gebrekidan, Tesfay Zemuy, et al.
Veröffentlicht: (2024)
von: Gebrekidan, Tesfay Zemuy, et al.
Veröffentlicht: (2024)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
DAGER: Exact Gradient Inversion for Large Language Models
von: Petrov, Ivo, et al.
Veröffentlicht: (2024)
von: Petrov, Ivo, et al.
Veröffentlicht: (2024)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce
von: Han, Wenchen, et al.
Veröffentlicht: (2026)
von: Han, Wenchen, et al.
Veröffentlicht: (2026)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
von: Polyakov, Igor, et al.
Veröffentlicht: (2025)
von: Polyakov, Igor, et al.
Veröffentlicht: (2025)
Tram-FL: Routing-based Model Training for Decentralized Federated Learning
von: Maejima, Kota, et al.
Veröffentlicht: (2023)
von: Maejima, Kota, et al.
Veröffentlicht: (2023)
EPARA: Parallelizing Categorized AI Inference in Edge Clouds
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
von: Li, Zonghang, et al.
Veröffentlicht: (2025)
von: Li, Zonghang, et al.
Veröffentlicht: (2025)
Training LLMs on HPC Systems: Best Practices from the OpenGPT-X Project
von: Penke, Carolin, et al.
Veröffentlicht: (2025)
von: Penke, Carolin, et al.
Veröffentlicht: (2025)
ABACUS: A FinOps Service for Cloud Cost Optimization
von: Deochake, Saurabh
Veröffentlicht: (2024)
von: Deochake, Saurabh
Veröffentlicht: (2024)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
von: Erben, Alexander, et al.
Veröffentlicht: (2023)
von: Erben, Alexander, et al.
Veröffentlicht: (2023)
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
Roadmap for Edge AI: A Dagstuhl Perspective
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)
De-DSI: Decentralised Differentiable Search Index
von: Neague, Petru, et al.
Veröffentlicht: (2024)
von: Neague, Petru, et al.
Veröffentlicht: (2024)
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2025)
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2025)
AIvailable: A Software-Defined Architecture for LLM-as-a-Service on Heterogeneous and Legacy GPUs
von: Antunes, Pedro, et al.
Veröffentlicht: (2025)
von: Antunes, Pedro, et al.
Veröffentlicht: (2025)
Collective Communication for 100k+ GPUs
von: Si, Min, et al.
Veröffentlicht: (2025)
von: Si, Min, et al.
Veröffentlicht: (2025)
UserCentrix: An Agentic Memory-augmented AI Framework for Smart Spaces
von: Saleh, Alaa, et al.
Veröffentlicht: (2025)
von: Saleh, Alaa, et al.
Veröffentlicht: (2025)
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
von: Rosendal, Daan, et al.
Veröffentlicht: (2026)
von: Rosendal, Daan, et al.
Veröffentlicht: (2026)
Credibility Trilemma in Polymatroidal Service Markets
von: Lovén, Lauri, et al.
Veröffentlicht: (2026)
von: Lovén, Lauri, et al.
Veröffentlicht: (2026)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
Federated Learning Model Aggregation in Heterogenous Aerial and Space Networks
von: Dong, Fan, et al.
Veröffentlicht: (2023)
von: Dong, Fan, et al.
Veröffentlicht: (2023)
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
ADF-LoRA: Alternating Low-Rank Aggregation for Decentralized Federated Fine-Tuning
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
von: Motta, Steven, et al.
Veröffentlicht: (2026)
von: Motta, Steven, et al.
Veröffentlicht: (2026)
Knowledge Graphs-Driven Intelligence for Distributed Decision Systems
von: Napoli, Rosario, et al.
Veröffentlicht: (2026)
von: Napoli, Rosario, et al.
Veröffentlicht: (2026)
Cognitive Infrastructure: A Unified DCIM Framework for AI Data Centers
von: Sunkara, Krishna Chaitanya
Veröffentlicht: (2026)
von: Sunkara, Krishna Chaitanya
Veröffentlicht: (2026)
GraphBit: A Graph-based Agentic Framework for Non-Linear Agent Orchestration
von: Sarker, Yeahia, et al.
Veröffentlicht: (2026)
von: Sarker, Yeahia, et al.
Veröffentlicht: (2026)
Parameter-Efficient and Personalized Federated Training of Generative Models at the Edge
von: Khan, Kabir, et al.
Veröffentlicht: (2025)
von: Khan, Kabir, et al.
Veröffentlicht: (2025)
Mobile Traffic Prediction at the Edge Through Distributed and Deep Transfer Learning
von: Petrella, Alfredo, et al.
Veröffentlicht: (2023)
von: Petrella, Alfredo, et al.
Veröffentlicht: (2023)
Aergia: Leveraging Heterogeneity in Federated Learning Systems
von: Cox, Bart, et al.
Veröffentlicht: (2022)
von: Cox, Bart, et al.
Veröffentlicht: (2022)
Towards Optimal Heterogeneous Client Sampling in Multi-Model Federated Learning
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neural Router: Semantic Content Matching for Agentic AI
von: Lovén, Lauri, et al.
Veröffentlicht: (2026) -
Autonomic Federated-Market Orchestration for the Edge-Cloud Continuum
von: Lovén, Lauri, et al.
Veröffentlicht: (2026) -
Agentic TinyML for Intent-aware Handover in 6G Wireless Networks
von: Saleh, Alaa, et al.
Veröffentlicht: (2025) -
Naeural AI OS -- Decentralized ubiquitous computing MLOps execution engine
von: Bleotiu, Cristian, et al.
Veröffentlicht: (2023) -
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
von: Lan, Guangchen, et al.
Veröffentlicht: (2024)