SGLang deployment
SGLang deployment Articles
Browse 3 articles about SGLang deployment.

How to Run Nex-N2.5 Mini Locally on RunPod (Dual H100 Setup)
A practical guide to deploying Nex-N2.5 Mini on RunPod using dual H100 GPUs, an SGLang Docker template, and correct VRAM sizing.
run Nex-N2.5 MiniRunPod H100 setupSGLang deployment

How to Run Nex-N2.5-mini Locally on 2x H100 GPUs
Deploy Nex-N2.5-mini with SGLang and Docker on 2x H100 GPUs, covering tensor parallelism, reasoning modes, and tool-calling setup.
Nex-N2.5-minirun Nex-N2.5 locallySGLang deployment

How to Run Ornith-1.5-9B Locally with vLLM or SGLang
A practical guide to serving Ornith-1.5-9B, a 9B reasoning model with 262K context, on a single GPU using vLLM or SGLang.
Ornith-1.5-9B localrun on vLLMSGLang deployment