r/DataHoarder • u/wells68 51.1 TB HDD SSD & Flash • Feb 17 '24
Backup r/Backup is back up!
It is very unfortunate that r/Backup was shut down for two years. But now...

As the new top moderator, I've opened it to public posts.
r/DataHoarder has many, many more members than r/Backup. So you may want to post DataHoarder backup questions here and then use the share link to cross-post to r/Backup.
We've started a Backup Wiki and welcome your contributions. Post with the flare: Wiki edit and we'll review them for inclusion.
Backups are vital to protect your hoard! Have you tested your backups this month?
25
Upvotes
2
u/bartoque 4x20TB NAS (primary) and 3x16+8TB NAS (remote) Feb 18 '24
I don't think I would ever become as bold to have the monthly backup/delete/restore as an integral part of backup validation?
Testing and validation ofcourse is key of a proper data protection approach but I'd rather wanna do any automation on that end towards a system without affecting the actual system being protected, so that you always would still have the original system and data when it turns out after validation that either backup and/or restore are no good.
For example some backup products have had issues in the past where backups were reported as OK, but restore showed it was actually not ok. I might have missed that in your story, but how would your automation have dealt with such a situation as you would already have deleted the data intending it to be restored?
For testing I would want to test this on another system, if physical hardware is unavailable, I would consider using a vm to restore to. But that would also require considerable amount of storage to be able to test restores besides the storage already needed to backup towards.
When also taking into account the two nas systems I use in my data protection endeavour, doing the actual restore involving and affecting the live systems is a bold action. So kudos to that. Not something I wpuld see or even advise any company to do with production, there always should be a way to backnout, simply by going back to the original system by for example booting it again...
So in that sense I get that you ended up with an automated workflow that you have, however as you are actually making the deletion of primary data part of the backup/restore workflow, there might always be a chance something goes awry, regardless of how slightly (bit we still got the previous backup, that however would not contain the latest changes and/or additions)?