Language Models
GradientHQ/parallax
Parallax is a distributed model serving framework that lets you build your own AI cluster anywhere
PythonApache-2.0updated Jul 1, 2026
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
$ git clone https://github.com/xLLM-AI/xllm.gitParallax is a distributed model serving framework that lets you build your own AI cluster anywhere
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
Machine Learning Engineering Open Book
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
High-speed Large Language Model Serving for Local Deployment
Data from GitHub ยท snapshot Sep 24, 2026