Skip to content
Read the original: vLLM· vllm_project·Published · YesterdayAI score46/100

1/ Three weeks from day 0, DeepSeek-V4.1-Flash on vLLM runs 1.9× faster at low concurrency and delivers 5.3× the throughput at 150 TPS per user on @SemiAnalysis_ AgentX.

Original title1/ Three weeks from day 0, DeepSeek-V4.1-Flash on vLLM runs 1.9× faster at low concurrency and delivers 5.3× the throughput at 150 TPS pe...

AISummary

Here is how, with interactive figures you can step through 🧵

Read the original x.com

Source: vLLM · x.com