Uber Engineering
Read post

Scaling AI/ML Infrastructure at Uber

Uber has made significant progress in scaling their AI/ML infrastructure, transitioning from on-prem to cloud infrastructure and optimizing existing infrastructure. They have implemented a unified federation layer for batch workloads, upgraded network bandwidth for training efficiency, and upgraded memory to improve GPU allocation rates. They are also building new infrastructure by evaluating price-performance ratios of cloud SKUs and improving LLM training efficiency through memory offload.

    #cloud#machine-learning#gpu#uber
Apr 22, 2024•8m read time•From uber.com
Post cover image
Table of contents
Goal and Key MetricsOptimizing Existing On-prem InfrastructureBuilding New InfrastructureAcknowledgments
6 Impressions
Uber Engineering's image
Uber Engineering

The Uber Engineering Blog offers insights, technical deep dives, and updates on the engineering chal...

178 Followers

•

268 Upvotes

Would you recommend this post?

Copy link
WhatsApp
Facebook
X
New Squad
  • © 2026 Daily Dev Ltd.
  • Guidelines
  • Explore
  • Tags
  • Sources
  • Squads
  • Leaderboard