Skip to content

Add vllm benchmarking#9

Merged
manojdev-ai merged 3 commits into
masterfrom
add-vllm-benchmarking
Jun 18, 2026
Merged

Add vllm benchmarking#9
manojdev-ai merged 3 commits into
masterfrom
add-vllm-benchmarking

Conversation

@rsingh-alt

Copy link
Copy Markdown
Contributor

vLLM Benchmarking can now be done using the cli tool.

@rsingh-alt
rsingh-alt requested a review from manojdev-ai May 29, 2026 10:31
@manojdev-ai

manojdev-ai commented Jun 1, 2026

Copy link
Copy Markdown
Contributor

simplify benchmarking stack setup

give a sample command (default) showing how to run the benchmark produced on our blog: https://www.corespan.ai/resources/blog/the-smartest-inference-node-you-can-buy-right-now

4× RTX 5090 result: ~5,345 tok/s for Qwen2.5-32B-Instruct on a 4× RTX 5090 configuration using vLLM with pipeline parallelism (PP=4, TP=1), native Blackwell FP8, chunked prefill

@manojdev-ai
manojdev-ai merged commit 9e780db into master Jun 18, 2026
1 check passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants