r/LocalLLaMA
· Communities
DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark
https://github.com/yhfgyyf/vllm-deepseek-v4-sm89 I couldn't believe that someone actually got vLLM working with this particular set of GPUs, but here it is. The video is from right after I got it working with 64k context, but it is now running with 256k. submitted by /u/dangerous_inference [link] [comments]