<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0"><channel><title>LocalXPU</title><link>https://www.localxpu.com/</link><description>Notes from running local LLMs on Intel Arc cards: two Arc Pro B60s in a desktop, with the setup, the numbers and what broke.</description><language>en-us</language>
<item><title>Running a 27B model on two Arc Pro B60s</title><link>https://www.localxpu.com/articles/serving-27b-two-b60s/</link><guid isPermaLink="true">https://www.localxpu.com/articles/serving-27b-two-b60s/</guid><pubDate>Thu, 01 Oct 2026 12:00:00 GMT</pubDate><description>Two B60s in a normal desktop now serve Qwen3.8-27B at about 53 tokens/s for one user and 250 across eight, after a graph-capture change, a P2P kernel module and a oneCCL memory leak.</description></item>
</channel></rss>
