Field notes

Guides & analysis

In-depth guides on on-premise LLM deployment: hardware requirements, cost economics, and the compliance landscape for regulated organizations.

9 min read

Qwen3.8-Flash-Next Hardware Requirements for Self-Hosting

Qwen3.8-Flash-Next self-hosting: 125B MoE, 6B active, FP8 at 173 GiB, 4-bit GGUF at 111 GB, N-gram table in system RAM, and the Qwen Community License.

Qwen3.8-Flash-Nexthardware requirementsself-hostingMoEQwen
6 min read

GLM-5.3-Flash Hardware Requirements: Frontier AI on One Node

GLM-5.3-Flash sizing guide: 320B MoE, 18B active, multimodal, 1M context, MIT license. Quant-by-quant memory table (93–642 GB), node layouts, what it replaces.

GLM-5.3-Flashhardware requirementsMIT licensemultimodaldeployment

Start with the sovereignty assessment.

Two weeks, fixed fee. We map your data obligations and workloads, size the hardware, and hand you a written architecture with a real cost model — whether or not you build with us.

Book a sovereignty assessment