1 request, prose prompt, with DFlash2 (per recipe) Source
- Custom kernel required
- Non-commercial
Il Suo obiettivo
Essere visibili dove risponde l'IA
Automatizzare i processi
Costruire un prodotto
Mettere al lavoro gli agenti IA
Collegare e modernizzare i sistemi
Sapere a che punto siamo
Casi d'Uso
CRMRafforzare relazioni clientiPopolareE-CommerceAumentare vendite onlineSistema prenotazioniPrenotazioni 24/7Gestione progettiCoordinare i teamFatturazionePagamenti più velociAnalyticsDecisioni basate sui datiLocal AI · Recipe · 4× DGX Spark
GLM-5.3-Flash 320B NVFP4 with vLLM on 4× DGX Spark: 45 tok/s according to github.com (dataset as of Sep 29, 2026).
by MiaAI-Lab
Every sourced value of this recipe, each with its condition and source. Bars relative to the largest value in the group.
What matters before you rebuild it.
1 request, prose prompt, with DFlash2 (per recipe) Source
1 request, prose prompt, with DFlash2 (per recipe) Source
1 request, prose prompt, with DFlash2 (per recipe) Source
HumanEval 97.6% / GSM8K 98.0% with FP8 KV and 4-bit dense Source
1 request, prose prompt, with DFlash2 (per recipe) Source
1 request, prose prompt, with MTP4 Source
In a workshop we work out which models and which hardware fit your tasks, and build the first agent on your infrastructure.