Content by Sherry Xu, Prashant Ranjan, Torsten Hoefler (1)

Sherry Xu, Prashant Ranjan, and Torsten Hoefler explain how Azure Maia 200 targets efficient, predictable large-scale AI inference by making data movement explicit (SDLA) and extending that approach across an all-Ethernet scale-up network, with performance discussion across realistic matmul and collective-communication workloads.
Community

End of content

Rejoining the server...

Rejoin failed... trying again in seconds.

Failed to rejoin.
Please retry or reload the page.

The session has been paused by the server.

Failed to resume the session.
Please reload the page.