PyTorch
Read post

Driving the Future of Open Source AI: An Update from PyTorch Foundation Projects – PyTorch

The PyTorch Foundation, now a multi-project organization hosting PyTorch, vLLM, DeepSpeed, Ray, Helion, and Safetensors, shares Q2 2025 progress across all six projects. PyTorch 2.13 ships FlexAttention on Apple Silicon (~12x faster than SDPA), a memory-saving LinearCrossEntropyLoss, and FSDP2 communication overlap. vLLM completes Model Runner V2, publishes its Q3 2026 roadmap focused on agentic workloads, and announces its first conference. DeepSpeed integrates Ulysses parallelism into Hugging Face libraries and earns a best-paper honorable mention at ASPLOS 2026. Ray adds GB200/GB300 hardware support and a new high-performance Ray Data 2.57 engine. Helion delivers cross-hardware attention kernels outperforming FlashAttention-4 and introduces LLM-guided autotuning with 10x efficiency gains. Safetensors adds GIL-free serialization, Python 3.14 support, and MPS fast-load paths.

    #deep-learning#pytorch#ai-inference#vllm
Jul 22•7m read time•From pytorch.org
Post cover image
Table of contents
PyTorch Updates: 2.13 Release, ExecuTorch, and Apple Silicon OptimizationvLLM Updates: Model Runner V2, Q3 2026 Roadmap, and vLLM ConferenceDeepSpeed Updates: Ulysses Parallelism, Torch.xpu Integration, and ASPLOS 2026Ray Updates: GB200/GB300 Hardware Support, Frontier Model Scaling, and Ray Data 2.57Helion Updates: Cross-Hardware Attention Kernels and LLM-Guided AutotuningSafetensors Updates: GIL-Free Serialization, Python 3.14 Support, and MPS Fast-LoadingJoin the Community
138 Impressions
PyTorch's image
PyTorch

PyTorch offers insights into deep learning, neural network modeling, and machine learning research, ...

142 Followers

•

231 Upvotes

Would you recommend this post?

Copy link
WhatsApp
Facebook
X
New Squad
  • © 2026 Daily Dev Ltd.
  • Guidelines
  • Explore
  • Tags
  • Sources
  • Squads
  • Leaderboard