Not long ago we watched a production MySQL database melt down for sixteen minutes.The errors started as a trickle, a handful per minute, then fed on themselves:minute errors/min queries/s
0 5 15,000 <- a burst of work arrives
10 1,400 1,500 <- error peak = throughput trough
16 0 2,500 <- locks released, backlog drained
18 0 8,500 <- full recovery, throughput jumps 5x









