---
type: "Article"
title: "Qwen3.8-27B mit MLX auf 1× Mac"
description: "Qwen3.8-27B mit MLX auf 1× Mac: Rezept von Weschera mit Tempo, Bedingung und Quelle je Zahl, Gewichten und Lizenz."
resource: "https://www.contextstudios.ai/de/lokale-ki/weschera--qwen38-27b-omlx-ane-mtp3-mac-studio-m4-max"
language: "de"
generated:
  by: "process:contextstudios-md/1"
  at: "2026-10-02T23:10:12.030Z"
status: "stable"
---

# Qwen3.8-27B mit MLX auf 1× Mac

Qwen3.8-27B mit MLX auf 1× Mac: 53,3 tok/s laut github.com (Datensatz-Stand: 29. Sept. 2026).

- Kreator: [Weschera](https://www.contextstudios.ai/de/lokale-ki/kreatoren/weschera.md)
- Engine: oMLX 0.6.3rc2
- Quantisierung: oQ4e (4-bit affine g64, 166 sens. Tensoren 5-bit; Stock-Qwen3.8-27B-Konversion)
- Modellfamilie: Qwen3.8-27B
- Hardware: mac x1
- Kontext: 262144

## Alle Zahlen

- 53,3 tok/s (Alltag; Alltag, 1 Anfrage, Prosa-Prompt, mit MTP (nativ) (laut Rezept)) — [source](https://github.com/Weschera/Qwen3.8-27B-oMLX-MTP-Mac/blob/main/README.md)
- decode_tps: 53.3 (1 Anfrage) — [source](https://github.com/Weschera/Qwen3.8-27B-oMLX-MTP-Mac/blob/main/README.md)
- decode_tps: 72.1 (1 Anfrage) — [source](https://github.com/Weschera/Qwen3.8-27B-oMLX-MTP-Mac/blob/main/README.md)
- prefill_tps: 273.7 (1 Anfrage) — [source](https://github.com/Weschera/Qwen3.8-27B-oMLX-MTP-Mac/blob/main/README.md)

## Gewichte

- [Jundot/Qwen3.8-27B-oQ4e-mtp](https://huggingface.co/Jundot/Qwen3.8-27B-oQ4e-mtp) (main)

[https://github.com/Weschera/Qwen3.8-27B-oMLX-MTP-Mac](https://github.com/Weschera/Qwen3.8-27B-oMLX-MTP-Mac)

## Related

- [Lokale KI auf Ihrer Hardware.](https://www.contextstudios.ai/de/lokale-ki.md)
