oneapi
Here are 6 public repositories matching this topic...
Field-tested guide: multi-GPU vLLM tensor-parallel (TP=2/TP=4) on Intel Arc Pro B70 (Battlemage BMG-G31, Xe2) on Linux. Driver setup (xe force_probe=e223), bare-metal vLLM + oneAPI 2025.3, the compute-runtime multi-root USM + triton-xpu init_devices fixes, FP8/int4-AutoRound quant, root-cause error reports. AI-agent readable (AGENTS.md).
-
Updated
Jun 13, 2026 - Shell
Docker Compose stack for serving a local, OpenAI-compatible LLM (vLLM on Intel XPU) on an Intel Arc Pro B60 GPU — reproducible config with operator and developer docs.
-
Updated
Jul 11, 2026 - Shell
Improve this page
Add a description, image, and links to the oneapi topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the oneapi topic, visit your repo's landing page and select "manage topics."