Content by Mark Gitau and Azin Heidarshenas (1)

Mark Gitau and Azin Heidarshenas summarize Azure’s MLPerf Inference v6.1 submissions for DeepSeek-R1 on NVIDIA GB300/GB200 NVL72, highlighting throughput and latency-focused results across Interactive, Server, and Offline scenarios at 1-rack and 4-rack scale.
Community

End of content

Rejoining the server...

Rejoin failed... trying again in seconds.

Failed to rejoin.
Please retry or reload the page.

The session has been paused by the server.

Failed to resume the session.
Please reload the page.