Microsoft Build session covering how to build local AI-powered Windows apps using the Foundry on Windows stack, which includes Windows AI APIs, Foundry Local, and Windows ML. Key announcements include Windows AI APIs expanding to CPU and GPU support (previously NPU-only), Foundry Local reaching general availability with new Qwen vision-language and speech models, the new Windows ML CLI (in preview) for model conversion/optimization/benchmarking, Windows ML 2.0 with ONNX runtime improvements, and WebNN support for bringing hardware-accelerated AI to web apps in Chromium browsers. Live demos cover speech recognition, structured JSON extraction with Phi Silica on GPU, video super resolution in ClipChamp, inventory management with Qwen 3.5 VLM, and real-time voice transformation with VoiceMod — all running locally across NPU, GPU, and CPU without cloud dependencies.

42m watch time
7 Impressions