DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context

(β˜… 103)

Deploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache

  • .env.example
  • .gitignore
  • docker-compose.yml
  • README.md
  • start-deepseek-v4-flash.sh
  • stop-deepseek-v4-flash.sh

# Installation Guide

1. Get the code
git clone https://github.com/MiaAI-Lab/DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context

Downloads the entire project code from GitHub to your computer.

cd DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context

Moves into the project folder you just downloaded.

2. Docker

Easy Recommended
Prerequisites
  • Git Needed to download the project code from GitHub.
  • Docker Desktop Needed to build and run containers. Install it and keep it running in the background.
docker compose logs

Runs the command against the services defined in the compose file.

docker logs deepseek-v4-flash-vllm-1

Type this command into your terminal and run it.

βœ… Run docker compose ps to check the containers are Up. If the README mentions a port, open http://localhost:PORT in your browser.

Pulled directly from this repo's README.

// repository documentation