Make use of Intel Arc Series GPU to Run Ollama, StableDiffusion, Whisper and Open WebUI, for image generation, speech recognition and interaction with Large Language Models (LLM).
-
Updated
Aug 7, 2026 - Dockerfile
Make use of Intel Arc Series GPU to Run Ollama, StableDiffusion, Whisper and Open WebUI, for image generation, speech recognition and interaction with Large Language Models (LLM).
A minimal, cross-compatible CPU/GPU telemetry monitor with accurate data directly from vendor APIs and beautiful ASCII visualization.
Open recipes, engine patches, and benchmark harnesses for LLM inference on Intel Arc Pro B60/B70 (Battlemage, Xe2). MoE 35B at 160 t/s decode / 7.5K t/s prefill single-stream, 27B at 50~ t/s decode / 1.7K t/s prefill single stream. vLLM XPU MTP unlocked. Muse Glimmer recipe added!!
Benchmark results and performance data for the Intel Arc Pro B70 GPU (Xe2/Battlemage) - LLM inference, video generation, dual-GPU scaling.
End-to-end subtitle translation workstation with cloud and local OpenVINO model support.
Arc Power
An inference engine built with SYCL + oneDNN
Makes Intel Arc Pro B70 GPUs actually fast on Ubuntu Server. 11 llama.cpp cherry-picks that fix the big B70 bugs (MoE slot-init SEGV, Q8_0 reorder crash, OOM reorder, missing BF16 GET_ROWS, wrong Xe2 warptile, slow K-quant DMMV, etc.) + Mesa 26 + runtime env workarounds + SYCL/Vulkan backend-selection rules. 2-7x speedup on 4x B70, bench-verified.
An alternative to Arc Control by Intel®.
Custom fan curves + GPU power/clock tuning for Intel Arc Pro B60/B70 (Battlemage) on Linux. Confirmed on Arc Pro B60 (8086:e211).
Portable version of ComfyUI & stable-diffusion-webui for Arc B580
PosterChan AI is a Nostr-powered personal cloud and a self-hosted AI powerhouse, running on one box you own.
Proxmox VE GPU passthrough — recipes, scripts, and failure-mode notes (CPUID, IOMMU, WDDM).
Linux (Arch) optimization guide for Samsung Galaxy Book6 Pro (NP960XJG) — Panther Lake, Arc B390, fingerprint, BIOS analysis, idle power 7.5W→1.9W
Plug-and-play llama.cpp runtime for Intel Arc GPUs. Auto-detects your card, picks safe SYCL defaults, and exposes an OpenAI-compatible API.
Unofficial community hub for Intel Arc Pro B70/B-series, Intel XPU, oneAPI, OpenVINO, PyTorch XPU, vLLM, llama.cpp, setup guides, benchmarks, and patches.
videogen-b70: ComfyUI + PyTorch XPU recipes for MiniMax H3 T2V on Intel Arc Pro B70. Auto-detects one or two cards (CLIP on GPU1, UNet on GPU0). Not tensor parallel.
Field-tested guide: multi-GPU vLLM tensor-parallel (TP=2/TP=4) on Intel Arc Pro B70 (Battlemage BMG-G31, Xe2) on Linux. Driver setup (xe force_probe=e223), bare-metal vLLM + oneAPI 2025.3, the compute-runtime multi-root USM + triton-xpu init_devices fixes, FP8/int4-AutoRound quant, root-cause error reports. AI-agent readable (AGENTS.md).
To associate your repository with the intel-arc topic, visit your repo's landing page and select "manage topics."