---
title: "Incident Review: Pageserver outage in us-east-1"
url: https://daily.dev/posts/incident-review-pageserver-outage-in-us-east-1-qojx2cozv
source_url: https://neon.com/blog/incident-review-pageserver-outage-in-us-east-1
type: article
source: "Neon"
published: 2026-07-09T06:57:47.855Z
updated: 2026-07-09T07:05:30.849Z
tags: ["postgresql", "distributed-systems", "aws-ec2"]
reading_time: 7
upvotes: 0
comments: 0
language: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# Incident Review: Pageserver outage in us-east-1

**[Neon](https://daily.dev/sources/neontech)** · 7 min read · 0 upvotes · 0 comments

## Summary

On August 14, 2024, about 0.4% of Neon customer projects in us-east-1 experienced up to 2 hours of downtime after an EC2 instance hosting a pageserver failed without warning. The incident review details the timeline: alerts fired 10 minutes after failure, migration of 17,037 projects began 30 minutes in, and full recovery took ~2 hours. Key delays included human-in-the-loop alerting, a semi-manual migration script, and stuck projects requiring manual intervention. Neon's new Storage Controller service — built with a reconciliation-loop model and autonomous heartbeat-based failure detection — can respond to node failures in seconds without human intervention. The fix is already in production for large projects (>64GiB) and Neon is accelerating rollout to all paying customers.

## Full article

daily.dev links to this article rather than hosting it. Read it at the original source: <https://neon.com/blog/incident-review-pageserver-outage-in-us-east-1>

## Similar posts on daily.dev

- [Post-incident review for 20th October 2025](https://daily.dev/posts/post-incident-review-for-20th-october-2025-5ec8gudkw) · Buildkite · 0 upvotes · 0 comments
- [How Neon's lakebase architecture stays resilient to cloud failures](https://daily.dev/posts/how-neon-s-lakebase-architecture-stays-resilient-to-cloud-failures-j7quhswdh) · Neon · 6 upvotes · 0 comments
- [Preparing for the worst: Our core database failover test](https://daily.dev/posts/preparing-for-the-worst-our-core-database-failover-test-16psrtfsf) · Vercel · 4 upvotes · 0 comments

---

Tags: [#postgresql](https://daily.dev/tags/postgresql), [#distributed-systems](https://daily.dev/tags/distributed-systems), [#aws-ec2](https://daily.dev/tags/aws-ec2)

[View this post on daily.dev](https://daily.dev/posts/incident-review-pageserver-outage-in-us-east-1-qojx2cozv)
