“Our engine is based on an early fork of vLLM from over a year ago.”

DeepSeek’s post about open-sourcing its inference engine announces that it will not. The code is a year-old vLLM fork tied to internal infrastructure. It will contribute pieces upstream instead. The title says open source and the body says maintenance burden.