Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythongpustack/gpustack

gpustack

A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.

84.3/100
5.6KForks: 636
View on GitHubHomepage →
Loading report...

Similar Projects

sglang

91

SGLang is a high-performance serving framework for large language models and multimodal models.

Python35.6K

vllm

93

A high-throughput and memory-efficient inference and serving engine for LLMs

Python91.2K

unsloth

93

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

Python75.8K

CowAgent

96

Open-source super AI assistant & Agent Harness. Plans tasks, runs tools and skills, self-evolves with memory and knowledge. Multi-model, multi-channel. Lightweight, extensible, one-line install. (formerly chatgpt-on-wechat)

Python46.8K
Back to List