LM/SGLang: Difference between revisions
< LM
Jump to navigation
Jump to search
(→Run) |
(→Run) |
||
| Line 31: | Line 31: | ||
<syntaxhighlight lang="bash"> | <syntaxhighlight lang="bash"> | ||
sglang serve \ | sglang serve \ | ||
--model-path | --model-path unsloth/Qwen3.8-27B-FP8 \ | ||
--device xpu \ | --device xpu \ | ||
--attention-backend intel_xpu | --attention-backend intel_xpu | ||
Revision as of 07:38, 17 September 2026
Rebuild docker image from source
cd sglang
git pull
cd docker
docker build -t sglang-xpu:latest -f xpu.Dockerfile .
Run
docker run \
-it \
--privileged \
--ipc=host \
--network=host \
--user root \
--group-add $(getent group video | cut -d: -f3) \
--device /dev/dri \
-v /dev/dri/by-path:/dev/dri/by-path \
-v /dev/shm:/dev/shm \
-v ~/.cache/huggingface:/root/.cache/huggingface \
-p 30000:30000 \
-e "HF_TOKEN=$HF_TOKEN" \
sglang-xpu:latest /bin/bash
In container ...
sglang serve \
--model-path unsloth/Qwen3.8-27B-FP8 \
--device xpu \
--attention-backend intel_xpu
Structure