> During this quarter 4 (four) drive models, from 3 (three) manufacturers, had 0 (zero) drive failures. None of the Toshiba 4TB and Seagate 16TB drives failed in Q1, but both drives had less than 10,000 drive days during the quarter. As a consequence, the AFR can range widely from a small change in drive failures. For example, if just one Seagate 16TB drive had failed, the AFR would be 7.25% for the quarter. Similarly, the Toshiba 4TB drive AFR would be 4.05% with just one failure in the quarter.
Backblaze should consider adding interval estimates in addition to point estimates. It might help the reader understand the uncertainty of the point estimate.
This data is significant and appreciated. Thank-you so much, and please keep up the good work.
It certainly is useful and interesting.
I mean "doing it well for an influencer and shafting everyone else" is far cheaper and easier than just making a solid and reliable product consistently. That's such old-fashioned thinking these days. :)
Exactly, as we just saw with the Western Digital debacle.
The unfortunate reality is, only a fortuneteller can say what the best hard drive to buy right now is.
I have a lot of old Hitachis that still work, but every last one of my Seagates died years ago. Yes, they've gotten better since then, but they're still reliably trounced by Hitachi.
Where are the failures coming from? What did they cheap out on? Has the C-suite decided it's not financially worth increasing reliability? What are the internal feelings on being the outlier, year in and year out, in these stats? Does anyone care?
From an interview last year, it sounds like they don't use SMR drives[1]. I would be interested if there is any good source for failure rates of those.
[1] https://www.backblaze.com/blog/how-backblaze-buys-hard-drive...
If it wasn't for Helium Sealed tech moving more platters into HDD, we would have zero capacity improvement in the past 4-5 years.
So, while the idea was ridiculed in the press, the advent of SSD's and the usecase for harddrives these days actually means it makes a lot of sense given the area (and therefor the capacity) increases with the square. Combined with the slower spin rates also increasing the density, and you get a device which is far more generally useful for bulk storage than SMR.
NAND prices fluctuate up and down and the current lows are not very sustainable.
It seems the days of computing getting cheaper every two or three years are gone.
What I really wish is that they would make an effort to go beyond just reporting their experiences and see what they are doing as also providing a service in the form of model reliability data. AKA toss a few pods of WD's/etc in there even if they are slightly more expensive/whatever. If the data were broken out by production location and manufacture date, its likely it would be something that they could sell on the open market for a small fee. I know I would have gotten the company I worked for to pay such a fee for a somewhat scientific look at the failure rates of certain models/etc.
AKA, pay a bit more for a broader set of drives, summarize the data for free, and then get people to pay for the detail data. If you work for a company buying a thousand or so drives a year, avoiding a 10% AFR is going to be worth a lot of money. AKA too small to have a good view of the state of things, but big enough that buying 1k bad drives and fighting daily RAID rebuilds for the next three years is real nightmare. Think of it as a bit of insurance, or at least validation of a problem when things start to go south.
> AKA toss a few pods of WD's/etc in there even if they are slightly more expensive/whatever.
If you see "low numbers" of one particular drive model in the stats, that is usually us trying out that drive model because at some point in the future the price might drop making it worth purchasing it in bulk worth it, or the price is ALREADY worth it but we're being careful in the rollout to make sure that drive model performs in our particular application - nobody else's application, just for us.
> make an effort to go beyond just reporting their experiences
We are VERY careful to report what we are seeing in our datacenter for our particular application, and no more. This isn't a scientific study, we aren't Consumer Reports, we're just publishing data we would collect whether or not we released it. We have a core business we super happy focusing on, somebody else is WELCOME to sell drive testing and drive predictions and we won't compete with them or get in their way. Heck, we would totally subscribe to that service!
I'm wondering why they keep maintaining two product lines that are sufficiently different to result in one basically being consistently the best, the other being consistently the worst, in terms of reliability.
Over the same features I find b2 to be way simpler and cheaper.
Thanks
Edit -> what are the feature's your missing most? We're collecting feedback at: b2feedback@backblaze.com if you want to send some notes over!
But perhaps the boot drives are just leftovers that are too small to be used any more for storage?
I tried to connect with s3cmd but it wasn't working. Unfortunately their support was not very helpful neither...
[0] https://www.backblaze.com/blog/announcing-our-first-european...
Seagate's drive longevity has improved a lot. Compare today's report to Backblaze's 2015Q1 report:
https://www.backblaze.com/blog/hard-drive-reliability-q1-201...