Who needs Latency ?
In simple terms, latency is the amount of time it takes to complete a given task. The most common example everyone is aware of is about application response time i.e. when you are using an application and request information to be processed or provided, latency is the term used for the amount of time it takes for the application to request the data and give you a response. Several major hyper-scale vendors have publicly stated that Latency matters. Amazon has reported that every 100ms of latency cost them 1% in sales. No wonder service level agreements for application owners often explicit spell out latency requirements for acceptable end user performance. Storage plays a critical role because the information being processed and requested is typically persisted on some form of non-volatile storage – this is true for both hard drives and SSDs, and the amount of time can differ for a variety of reasons. The table below highlights typical latency times associated with particular device types. I’ve tried to represent these speeds/feeds in a more meaningful “Human” scale form.
| Device | Latency | Human Scale |
| Srv RAM | 100ns | 1 sec |
| Flash Dimm | 2.5 – 5 microsecs | 25-30 secs |
| PCI-e SSD | 20 -50 microsecs | 3 – 8 mins |
| SATA / SAS drives | 100 – 300 microsecs | 16 – 50 mins |
| All Flash Arrays | 500 – 1100 microsecs | 1.5 – 3 hrs |
| Hybrid Arrays | 2 millsecs | 6 hrs |
| 15k RPM Drive | 5 – 15 millsecs | 15hrs |
- Howard Marks of DeepStorage.net

