Postmortem of database outage of January 31
Blog post from GitLab
On January 31, 2017, GitLab.com faced a significant outage due to an accidental deletion of data from its primary database server, which led to several hours of downtime and the irreversible loss of production data, including around 5,000 projects, 5,000 comments, and 700 user accounts. The issue was compounded by a flawed backup process, caused by mismatched PostgreSQL versions and failed notifications, and inadequate disaster recovery procedures, which prolonged the restoration process to over 18 hours. GitLab's CEO issued an apology, and the company outlined efforts to improve its operations and recovery strategies, which included enhancing backup procedures, implementing better monitoring systems, and increasing database redundancy. The incident and recovery efforts were publicly documented and live-streamed, demonstrating GitLab's commitment to transparency and accountability.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Real-time | 2 | 261 | 68 | 31 | +55% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.