back to home

tonyd2wild / DeepSeek-V4.1-Flash-vLLM-DGX-Spark

DeepSeek-V4.1-Flash (552B MoE, MXFP4 experts, 1M ctx) on four NVIDIA DGX Sparks with vLLM TP4: Engram-on-disk patch, sm121 kernel build, launchers, measured numbers

View on GitHub
62 stars
8 forks
6 issues
Python

AI Architecture Analysis

This repository is indexed by RepoMind. By analyzing tonyd2wild/DeepSeek-V4.1-Flash-vLLM-DGX-Spark in our AI interface, you can instantly generate complete architecture diagrams, visualize control flows, and perform automated security audits across the entire codebase.

Our Agentic Context Augmented Generation (Agentic CAG) engine loads full source files into context on-demand, avoiding the fragmentation of traditional RAG systems. Ask questions about the architecture, dependencies, or specific features to see it in action.

Source files are only loaded when you start an analysis to optimize performance.

Embed this Badge

Showcase RepoMind's analysis directly in your repository's README.

[![Analyzed by RepoMind](https://img.shields.io/badge/Analyzed%20by-RepoMind-4F46E5?style=for-the-badge)](https://repomind.in/repo/tonyd2wild/DeepSeek-V4.1-Flash-vLLM-DGX-Spark)
Preview:Analyzed by RepoMind