The PyTorch Foundation, now a multi-project organization hosting PyTorch, vLLM, DeepSpeed, Ray, Helion, and Safetensors, shares Q2 2025 progress across all six projects. PyTorch 2.13 ships FlexAttention on Apple Silicon (~12x faster than SDPA), a memory-saving LinearCrossEntropyLoss, and FSDP2 communication overlap. vLLM completes Model Runner V2, publishes its Q3 2026 roadmap focused on agentic workloads, and announces its first conference. DeepSpeed integrates Ulysses parallelism into Hugging Face libraries and earns a best-paper honorable mention at ASPLOS 2026. Ray adds GB200/GB300 hardware support and a new high-performance Ray Data 2.57 engine. Helion delivers cross-hardware attention kernels outperforming FlashAttention-4 and introduces LLM-guided autotuning with 10x efficiency gains. Safetensors adds GIL-free serialization, Python 3.14 support, and MPS fast-load paths.