SGLang Team /
SGLang
Performance-optimized LLM inference server, offering an OpenAI-compatible API and support for many model formats and quantization methods. Best used with large recent datacenter GPUs and may require more manual configuration and tuning than simpler inference servers.
LLM InferenceAdvanced
Securely deploy SGLang on Laboratory OS.