Fox-1: Open Small Language Model for Cloud and Edge
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Zijian, Zhang, Jipeng, Pan, Rui, Xu, Zhaozhuo, Han, Shanshan, Jin, Han, Shah, Alay Dilipbhai, Stripelis, Dimitris, Yao, Yuhang, Avestimehr, Salman, Zhang, Tong, He, Chaoyang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TensorOpera Router: A Multi-Model Router for Efficient LLM Inference
by: Stripelis, Dimitris, et al.
Published: (2024)
by: Stripelis, Dimitris, et al.
Published: (2024)
Alopex: A Computational Framework for Enabling On-Device Function Calls with LLMs
by: Ran, Yide, et al.
Published: (2024)
by: Ran, Yide, et al.
Published: (2024)
ScaleLLM: A Resource-Frugal LLM Serving Framework by Optimizing End-to-End Efficiency
by: Yao, Yuhang, et al.
Published: (2024)
by: Yao, Yuhang, et al.
Published: (2024)
TorchOpera: A Compound AI System for LLM Safety
by: Han, Shanshan, et al.
Published: (2024)
by: Han, Shanshan, et al.
Published: (2024)
Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences
by: Han, Shanshan, et al.
Published: (2025)
by: Han, Shanshan, et al.
Published: (2025)
Kick Bad Guys Out! Conditionally Activated Anomaly Detection in Federated Learning with Zero-Knowledge Proof Verification
by: Han, Shanshan, et al.
Published: (2023)
by: Han, Shanshan, et al.
Published: (2023)
FedML-HE: An Efficient Homomorphic-Encryption-Based Privacy-Preserving Federated Learning System
by: Jin, Weizhao, et al.
Published: (2023)
by: Jin, Weizhao, et al.
Published: (2023)
LLM Multi-Agent Systems: Challenges and Open Problems
by: Han, Shanshan, et al.
Published: (2024)
by: Han, Shanshan, et al.
Published: (2024)
Understanding Communication Backends in Cross-Silo Federated Learning
by: Ziashahabi, Amir, et al.
Published: (2026)
by: Ziashahabi, Amir, et al.
Published: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
by: Zhang, Tuo, et al.
Published: (2024)
by: Zhang, Tuo, et al.
Published: (2024)
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
by: Han, Shanshan, et al.
Published: (2023)
by: Han, Shanshan, et al.
Published: (2023)
Personalized Visual Instruction Tuning
by: Pi, Renjie, et al.
Published: (2024)
by: Pi, Renjie, et al.
Published: (2024)
Edge Private Graph Neural Networks with Singular Value Perturbation
by: Tang, Tingting, et al.
Published: (2024)
by: Tang, Tingting, et al.
Published: (2024)
Toward Super Agent System with Hybrid AI Routers
by: Yao, Yuhang, et al.
Published: (2025)
by: Yao, Yuhang, et al.
Published: (2025)
Differentially Private Federated Learning without Noise Addition: When is it Possible?
by: Zhang, Jiang, et al.
Published: (2024)
by: Zhang, Jiang, et al.
Published: (2024)
Adapt-Pruner: Adaptive Structural Pruning for Efficient Small Language Model Training
by: Pan, Rui, et al.
Published: (2025)
by: Pan, Rui, et al.
Published: (2025)
LISA: Layerwise Importance Sampling for Memory-Efficient Large Language Model Fine-Tuning
by: Pan, Rui, et al.
Published: (2024)
by: Pan, Rui, et al.
Published: (2024)
Strengthening Multimodal Large Language Model with Bootstrapped Preference Optimization
by: Pi, Renjie, et al.
Published: (2024)
by: Pi, Renjie, et al.
Published: (2024)
MobiZO: Enabling Efficient LLM Fine-Tuning at the Edge via Inference Engines
by: Gao, Lei, et al.
Published: (2024)
by: Gao, Lei, et al.
Published: (2024)
Chain-Oriented Objective Logic with Neural Network Feedback Control and Cascade Filtering for Dynamic Multi-DSL Regulation
by: Han, Jipeng
Published: (2024)
by: Han, Jipeng
Published: (2024)
Intelligence Inertia: Physical Isomorphism and Applications
by: Han, Jipeng
Published: (2026)
by: Han, Jipeng
Published: (2026)
FedGrAINS: Personalized SubGraph Federated Learning with Adaptive Neighbor Sampling
by: Ceyani, Emir, et al.
Published: (2025)
by: Ceyani, Emir, et al.
Published: (2025)
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
by: Feng, Tiantian, et al.
Published: (2024)
by: Feng, Tiantian, et al.
Published: (2024)
Leveraging Uncertainty Estimation for Efficient LLM Routing
by: Zhang, Tuo, et al.
Published: (2025)
by: Zhang, Tuo, et al.
Published: (2025)
ATP: Enabling Fast LLM Serving via Attention on Top Principal Keys
by: Niu, Yue, et al.
Published: (2024)
by: Niu, Yue, et al.
Published: (2024)
The Instinctive Bias: Spurious Images lead to Illusion in MLLMs
by: Han, Tianyang, et al.
Published: (2024)
by: Han, Tianyang, et al.
Published: (2024)
Privacy-Preserving Federated Heavy Hitter Analytics for Non-IID Data
by: Shao, Jiaqi, et al.
Published: (2023)
by: Shao, Jiaqi, et al.
Published: (2023)
Image Textualization: An Automatic Framework for Creating Accurate and Detailed Image Descriptions
by: Pi, Renjie, et al.
Published: (2024)
by: Pi, Renjie, et al.
Published: (2024)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
by: Zhang, Jipeng, et al.
Published: (2024)
by: Zhang, Jipeng, et al.
Published: (2024)
‘INIVIT MC-2019’ nuevo cultivar de malanga isleña (Colocasia esculenta (L.) Schott) para la agricultura cubana
by: Alay Jiménez-Medina
Published: (2023)
by: Alay Jiménez-Medina
Published: (2023)
An Edge-Cloud Collaboration Framework for Generative AI Service Provision with Synergetic Big Cloud Model and Small Edge Models
by: Tian, Yuqing, et al.
Published: (2024)
by: Tian, Yuqing, et al.
Published: (2024)
MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance
by: Pi, Renjie, et al.
Published: (2024)
by: Pi, Renjie, et al.
Published: (2024)
OSMGen: Highly Controllable Satellite Image Synthesis using OpenStreetMap Data
by: Ziashahabi, Amir, et al.
Published: (2025)
by: Ziashahabi, Amir, et al.
Published: (2025)
Embracing Federated Learning: Enabling Weak Client Participation via Partial Model Training
by: Lee, Sunwoo, et al.
Published: (2024)
by: Lee, Sunwoo, et al.
Published: (2024)
EM-Aware Physical Synthesis: Neural Inductor Modeling and Intelligent Placement & Routing for RF Circuits
by: Huang, Yilun, et al.
Published: (2026)
by: Huang, Yilun, et al.
Published: (2026)
GEM: A Scale-Aware and Distribution-Sensitive Sparse Fine-Tuning Framework for Effective Downstream Adaptation
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
Hawk: Accurate and Fast Privacy-Preserving Machine Learning Using Secure Lookup Table Computation
by: Saleem, Hamza, et al.
Published: (2024)
by: Saleem, Hamza, et al.
Published: (2024)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
by: Ghasemi, Narges, et al.
Published: (2025)
by: Ghasemi, Narges, et al.
Published: (2025)
On Polynomial Approximations for Privacy-Preserving and Verifiable ReLU Networks
by: Ali, Ramy E., et al.
Published: (2020)
by: Ali, Ramy E., et al.
Published: (2020)
Effective Bilevel Optimization via Minimax Reformulation
by: Wang, Xiaoyu, et al.
Published: (2023)
by: Wang, Xiaoyu, et al.
Published: (2023)
Similar Items
-
TensorOpera Router: A Multi-Model Router for Efficient LLM Inference
by: Stripelis, Dimitris, et al.
Published: (2024) -
Alopex: A Computational Framework for Enabling On-Device Function Calls with LLMs
by: Ran, Yide, et al.
Published: (2024) -
ScaleLLM: A Resource-Frugal LLM Serving Framework by Optimizing End-to-End Efficiency
by: Yao, Yuhang, et al.
Published: (2024) -
TorchOpera: A Compound AI System for LLM Safety
by: Han, Shanshan, et al.
Published: (2024) -
Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences
by: Han, Shanshan, et al.
Published: (2025)