Home / Companies / Netlify / Blog / Post Details
Content Deep Dive

Learning Review for our 22 November API and Origin outage

Blog post from Netlify

Post Details
Company
Date Published
Author
Chris McCraw
Word Count
894
Company Posts That Month
7
Language
English
Hacker News Points
-
Post removed?
No
Summary

The downtime of a service's database, which occurred on November 22nd, was caused by the database filling up its disk space due to rapid growth in file deployment. The team had noticed an upward trend in database size days earlier and began preparing for potential issues, but ultimately failed to migrate data to a larger partition in time. Despite this, the CDN edge nodes continued to serve content, minimizing the impact of the outage. The team has since analyzed the situation, identified causes, and implemented measures to prevent similar outages, including revamping their monitoring system, deploying a new status page with incident history, and modifying their replication method to reduce space usage. They have also redesigned their database handling practices to ensure a live master and slave during potentially impactful operations.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Serverless 6 90 16 9 -12%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.