1 request, prose prompt, context 19, with DSpark (per recipe) Source
- Custom kernel required
Your goal
Be visible where AI answers
Automate processes
Build a product
Put AI agents to work
Connect and modernize systems
Know where we stand
Use Cases
CRMStrengthen customer relationshipsPopularE-CommerceBoost online revenueBooking System24/7 appointment bookingProject ManagementCoordinate teamsInvoicingGet paid fasterAnalyticsData-driven decisionsLocal AI · Recipe · 2× DGX Spark
DeepSeek-V4-Flash FP8 with vLLM on 2× DGX Spark (1024k context): 83.8 tok/s according to github.com (dataset as of Sep 29, 2026).
by Weschera
Every sourced value of this recipe, each with its condition and source. Bars relative to the largest value in the group.
1 request, prose prompt, context 19, with DSpark (per recipe) Source
1 request, prose prompt, with DSpark Source
1 request, prose prompt, with DSpark (per recipe) Source
1 request, prose prompt, with DSpark (per recipe) Source
1 request, prose prompt, with DSpark (per recipe) Source
1 request, story prompt, prompt 18 tokens, with MTP (per recipe) Source
In a workshop we work out which models and which hardware fit your tasks, and build the first agent on your infrastructure.