Engineering low-latency Cloud for next-gen compute

  • Watch on-demand now
  • The Cloud & Hybrid Channel
Speakers

You can now ask questions directly within our broadcasts using our new AskAI feature.

Real-time AI and HPC workloads are pushing Cloud infrastructure to new limits, requiring low-latency and rapid processing to keep models running efficiently. But achieving this is challenging due to the complexity of distributed compute and network bottlenecks. How can engineers best deliver the performance and responsiveness required for next-generation compute? In this session we address how to:

  • Design Cloud infrastructure to achieve low-latency performance
  • Scale infrastructure without compromising responsiveness
  • Overcome network bottlenecks for complex applications
  • Employ distributed compute for real-time workloads

Related Episodes