LLM model selection & deployment analysis tool for DeepSeek Harness: deployability, VRAM, TTFT, latency, throughput and power for 38 models × 20 GPUs/NPUs
dsh plugin --profile web add github:lhwwxy/dsh-model-deploy
View the source on GitHub