Local AI · Recipe · 1× Strix Halo

Qwen3 30B-A3B UD-Q4_K_XL with llama.cpp on 1× Strix Halo

by deseven/lhl (strixhalo.wiki)

85.1tok/sPeak

1 request, synthetic prompt

Source: strixhalo.wiki
Intelligence (original model) · Artificial Analysis8with thinkingno thinking 7

Engine

Engine
llama.cpp Build 2025 (nicht angegeben)
Quantization
GGUF UD-Q4_K_XL
Model family
Qwen3
Context
—
Parameters
30B MoE / 3B active
Creator
deseven/lhl (strixhalo.wiki)

Measurements

Every sourced value of this recipe, each with its condition and source. Bars relative to the largest value in the group.

No sourced measurements for this recipe.

What you need

Hardware
1 × AMD Strix Halo
Weights
Engine
llama.cpp Build 2025 (nicht angegeben)

Related recipes

← Back to overview

Local AI in your company?

In a workshop we work out which models and which hardware fit your tasks, and build the first agent on your infrastructure.