Back openDesk Edu for a sovereign, open-source education β every vote counts.
Vote nowProduction infrastructure stacks, DGX Spark blueprints, and AI/Graph toolkits.
How we serve DeepSeek-V4-Flash β a 671B-parameter MLA-MoE model β across two NVIDIA DGX Spark (GB10) desksides at production quality, with every tuning decision, patch, and failure-mode documented. Powered by the Dual DGX Spark Cluster Blueprint.
Run Qwen3.8-27B β a 27B-parameter reasoning model β on a single DGX Spark (GB10) at high throughput with 128K+ context, tool calling, and reasoning. No second node, no RoCE fabric, no speculative-decoding patches. The accessible on-prem frontier.
The complete knowledge graph arsenal β GraphRAG pipelines, Neo4j deployment, and monitoring. 330 pages of production-tested guides plus infrastructure.
The complete blueprint for building a production-grade AI homelab. Connect Dify, n8n, Ollama, Qdrant, monitoring, and backup into one cohesive system.
Production Ansible deployment for two DGX Spark (GB10) nodes with vLLM TP2 serving DeepSeek-V4-Flash (671B), fronted by LiteLLM, monitored by Prometheus/Grafana, and wired over 200Gbps RoCE. The reference deployment for our Dual DGX Spark case study.
Production-ready Dify with PostgreSQL, Redis, Weaviate, Nginx, and SSRF protection β not a toy compose file.
Showing 6 of 9 products