xLLM
Model DeploymentMLOps & Experiment Tracking

xLLM

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.

Open Source

About

xLLM is an efficient LLM inference framework, specifically optimized for Chinese AI accelerators, enabling enterprise-grade deployment with enhanced efficiency and reduced cost.

Open Source Health

Not enough history
Stars
1,568
Forks
298
License
Apache-2.0
Last commit
29 days ago
C++

Related Categories