
18:48
Homa: The End of TCP for AI Clusters — John Ousterhout, Stanford
John Ousterhout
AI inference networking · GPU synchronization and tail latency · Incast
AI Engineer topic
1 AI Engineer conference talks about Byte streams and head-of-line blocking.