1 request, best value reported by source Source
Local AI · Recipe · 1× DGX Spark
Ornith-1.5 35B NVFP4 with SGLang on 1× DGX Spark
by vcruz305
Engine
- Engine
- SGLang / vLLM SGLang 0.5.18.dev760 / vLLM g75231eff2
- Quantization
- NVFP4 (ModelOpt, 21.4 GB)
- Model family
- Ornith-1.5
- Context
- 262,144
- Parameters
- ~35B MoE / ~3B active
- Creator
- vcruz305
- GitHub stars
- 3
- Repo updated
- Aug 27, 2026
Measurements
Every sourced value of this recipe, each with its condition and source. Bars relative to the largest value in the group.
No sourced measurements for this recipe.
What you need
- Hardware
- 1 × NVIDIA DGX Spark (GB10)
- Weights
- Main weightsornith-ai/Ornith-1.5-35B-A3B-NVFP4 MIT
- othervcruz305/Ornith-1.5-35B-A3B-GGUF License unclear
- Engine
- SGLang / vLLM SGLang 0.5.18.dev760 / vLLM g75231eff2
- Context
- 262,144 tokens
Related recipes
1 request, prose prompt, with MTP (per recipe) Source
- Custom kernel required
- License unclear
1 request, prose prompt, with MTP (per recipe) Source
- Access on request
- License unclear
1 request, prose prompt, prompt 50 tokens, with DSpark (per recipe) Source
- License unclear
1 request, prose prompt, with Drafting aus dem Prompt (Edit-Setup, lt. README 'drafts its guesses from your prompt') (per recipe) Source
- License unclear
1 request, story prompt, prompt 18 tokens, with MTP (per recipe) Source
- ≤ 3 bit: quantization may cost quality
Local AI in your company?
In a workshop we work out which models and which hardware fit your tasks, and build the first agent on your infrastructure.