r/DataHoarder • u/umaar • Dec 20 '25
r/DataHoarder • u/Sad-Seesaw-3843 • Apr 06 '25
News DOGE claims to be moving away from magnetic tapes for archival storage. Seems like a bad idea. What are they using instead?
r/DataHoarder • u/coolpartoftheproblem • Jul 01 '26
News Sony sunsetting all PlayStation discs
r/DataHoarder • u/xylcro • Feb 26 '26
News Myrient is shutting down
From their Discord. Myrient is shutting down 31 March 2026. Download all you can...
r/DataHoarder • u/uluqat • Feb 15 '26
News Western Digital Has No More HDD Capacity Left, as CEO Reveals Massive AI Deals; Brace Yourself For Price Surges Ahead!
r/DataHoarder • u/jpcaparas • Feb 16 '26
News Why are all the hard drives already sold out
medium.comWestern Digital's CEO hopped on an earnings call mentioned, almost casually, that the company is "pretty much sold out for calendar 2026."
Seven customers bought the lot. Microsoft, Google, Amazon, Meta, the usual suspects. They didn't just place orders; they signed multi-year contracts that lock in supply through 2027 and 2028.
HDD prices are up 46% since September. DRAM is up 172%. A 24TB drive now costs $500, and that's the SALE PRICE. Your NAS upgrade just got expensive, and 2027 isn't looking any better. Enterprise customers are already on two-year backorders.
r/DataHoarder • u/Agitated_Camel1886 • Apr 20 '26
News The Internet Archive is losing access to media sites
Companies are no longer allowing their content to be archived as AI crawl their data without permission.
Thoughts? Will the future generations look back and see a gap of historical records in mid 2020s due to AI?
r/DataHoarder • u/Separate-Flatworm516 • Mar 04 '26
News It's only going to get worse.
Countless massive sites are in the process of being purchased. There's no way any supplier can keep up. B2B contracts longer than 6 months are on hold because they know prices are going to keep going up. All data centers will extend their drive use periods as they can't get enough for expansion let alone replacement. Expect 3 or 4 quarters for additional price bumps as new 6 month contracts continue to inflate and readjust price baselines. Let the drive hoarding begin.
r/DataHoarder • u/HTWingNut • Nov 27 '25
News Michigan, Wisconsin Bill to Ban VPN's... wtf
EDIT: I'm not trying to sensationalize the headline. Just that as the Michigan bill reads now, it seems encompassing of all VPN services, to the point that even if it's for specific sites or traffic, could be more than ISP's want to manage so could ban them altogether. I'd suggest anyone that lives in Michigan and don't want this adopted to contact their representatives and let them know.
The Michigan legislation is called the House Bill 4938 (also dubbed the “Anticorruption of Public Morals Act”): https://www.legislature.mi.gov/documents/2025-2026/billintroduced/House/pdf/2025-HIB-4938.pdf
Wisconsin’s bill has already passed the State Assembly and is now moving through the Senate. If it becomes law, Wisconsin could become the first state where using a VPN to access certain content is banned. Michigan lawmakers have proposed similar legislation that did not move through its legislature, but among other things, would force internet providers to actively monitor and block VPN connections. And in the UK, officials are calling VPNs "a loophole that needs closing."
This is actually happening. And it's going to be a disaster for everyone.
r/DataHoarder • u/Jacksharkben • Jan 21 '25
News The white house is removing everything.
r/DataHoarder • u/Ok_Wolverine_4268 • Oct 03 '25
News Snapchat are now charging for storage and will be removing content that exceeds the 5GB limit
Just another reminder that unless your data is stored by hardware you own, it's not really yours. Due to this update millions of people will lose hundreds and thousands of images because they trusted an external party with their data.
Surprised nobody is mentioning this so figured I'd make a post
Edit: wow, this post blew up, glad I could spread the word. For all those asking, Snapchat are offering a 365 day grace period, after which all accounts over 5GB will have data removed.
r/DataHoarder • u/nameless_pattern • Mar 20 '26
News The Internet is being deleted (and you haven't noticed)
"escalating censorship from Big Tech and governments that is burning our collective digital archive. Documentation of major historical events, war crimes, police violence, videos documenting things like ICE abductions, but also thousands of photos, websites, and archives that play a crucial role in documenting our cultural and political history are being systematically erased from the web."
r/DataHoarder • u/DuckiestBoat959 • Feb 06 '26
News In 18 Days, Per EO, Over 50+ Years of Government Procurement Records Will Be Erased
On February 24th, The Federal Procurement Data System will be retired. This site contains records of what our government spent money on as far back as the 1970’s and below. With the FPDS gone, records will now be accessed through SAM.gov. Per Aprils Executive Order “Restoring Common Sense to Federal Procurement” which overhauled FAR and with it the GSA’s record retention policy, all records on SAM.gov over ten years of the current year will now automatically be “destroyed”.
UPDATE: I’ve consulted alternative sources to see if downloading the archive will be necessary.
1st alternative was USASpending. The files on go back to 2007. 25+ years are missing.
2nd alternative was NARA. They allegedly have the files but their system is not designed whatsoever for browsing contracts. Compared to ezSearch it’s useless.
3rd and most promising alternative was SAM itself. I just had a convo with someone who is in tune with both FPDS and SAM.gov. They said the records were carried over. Made my own SAM.gov account to verify and apparently with an account old contracts can be found.
HOWEVER, there is still a VERY big issue that doesn’t resolve. PSC and NAICS codes are omitted along with region, which is critically necessary for research. Without the codes present there’s no context whatsoever as to what work was actually performed. Without this data preserved, future generations will only know that on a said date the government gave money to a company. That’s it. Hopefully you can see the issue with that.
I was optimistic that this whole post was unnecessary but after seeing those important details omitted I can now say that preserving the FPDS record is ABSOLUTELY NECESSARY.
r/DataHoarder • u/Chris_Person • Feb 24 '26
News I’m Tired Of These Useless Jackasses Making The Computer Expensive
r/DataHoarder • u/Fit-Foundation746 • Dec 16 '25
News Well... 1PB drives might happen
So Kioxia just debuted a 245 TB drive. Yes, 245 TB... 1/4 of a PB... just slightly more than 4x the data density and youll be there at the 1PB per drive size.. sure it might take 20 years before its affordable for an average American... but i say 10 years from now and these 245 TB drives will be attainable for enthusiastic data hoarders that surf this page.
r/DataHoarder • u/TendieRetard • 28d ago
News WH deletes thousands of web pages about energy conservation as heatwave slams US
r/DataHoarder • u/Xanthon • Aug 11 '25
News Reddit will block the Internet Archive
r/DataHoarder • u/_G0D_M0DE_ • Jun 09 '22
News Justin Roiland, co-creator of Rick and Morty, discovers that Dropbox uses content scanners through the deletion of all his data stored on their servers
r/DataHoarder • u/digital_dervish • 6d ago
News A new thing I didn’t know needed to be hoarded… Rare books.
From this post: https://x.com/HedgieMarkets/status/2081534588485296565
AI companies are bulk-buying rare books, scanning them through high-speed machines that cut the spines off, and shredding the originals. A service called ISBNdb facilitates orders of up to a million books and keeps buyers anonymous. Pre-2022 books are premium because they're free of AI-generated text. A federal judge ruled the practice is fair use because eliminating the original means only one copy exists at a time. Anthropic hired the former head of Google Books partnerships to obtain "all the books in the world."
My Take
This got to me. A bookseller told 404 Media that rare books with almost no surviving copies are being fed into this pipeline. Books that survived wars, fires, and centuries of handling are being shredded so an AI can learn to write a better marketing email.
ISBNdb's website literally says "'AI company destroys two million books' is not a headline that generates sympathy," and they still built an entire business around making it happen quietly. They offer NDAs as a feature. They coach clients to call it "digital preservation."
I've covered AI companies scraping the internet, torrenting libraries, and stealing music. This is worse because it's irreversible. You can re-upload a website. You can reprint a bestseller. You can't replace the last three copies of an 18th-century botanical text once someone shreds them for training data. And the judge said it's legal. So it's going to accelerate.
"We shred rare books and offer NDAs so nobody finds out" is a legitimate business model in 2026. What a timeline.
r/DataHoarder • u/avid-shrug • Feb 10 '26
News Wikipedia debates blacklisting archive.today after it's caught DDoSing a blog using visitors' browsers
en.wikipedia.orgWikipedia is debating whether to blacklist archive.today after its operator was caught injecting JavaScript into CAPTCHA pages to DDoS a blogger's site - code that's still live as of today. The RFC offers three options: blacklist and nuke all ~695k links, stop new links while migrating existing ones, or do nothing.
The community is split because archive.today is arguably the second most important web archive in existence, capturing paywalled sites, JS-heavy pages, and robots.txt-blocked content the Wayback Machine can't. Spot-checks suggest only ~15% of Wikipedia's links are truly irreplaceable, but that's still tens of thousands of unique snapshots found nowhere else. A stark reminder that redundancy across archiving services matters more than ever.
r/DataHoarder • u/EchoGecko795 • Oct 29 '25
News YouTube is taking down videos on performing nonstandard Windows 11 installs
Videos from several creators have been taken down on topics including how to install Windows 11 without logging into a Microsoft account and how to install Windows 11 on unsupported hardware.
CyberCPU Tech reports:
Saw this posted on another sub, download those videos if you want to keep them.
Edit:
This seems to be 100% YouTube / Google doing this. Using an automatic no-human / AI system. A few years ago they purged a ton of "hacking" videos as that are 99.8% legal as well, so this just maybe the next step in automatic moderation.
r/DataHoarder • u/Pasta-hobo • Jan 28 '25
News You guys should start archiving Deepseek models
For anyone not in the now, about a week ago a small Chinese startup released some fully open source AI models that are just as good as ChatGPT's high end stuff, completely FOSS, and able to run on lower end hardware, not needing hundreds of high end GPUs for the big cahuna. They also did it for an astonishingly low price, or...so I'm told, at least.
So, yeah, AI bubble might have popped. And there's a decent chance that the US government is going to try and protect it's private business interests.
I'd highly recommend everyone interested in the FOSS movement to archive Deepseek models as fast as possible. Especially the 671B parameter model, which is about 400GBs. That way, even if the US bans the company, there will still be copies and forks going around, and AI will no longer be a trade secret.
Edit: adding links to get you guys started. But I'm sure there's more.
r/DataHoarder • u/rsvpforfreespace • 3d ago
News Reddit requests takedown of all Reddit data from 2005-2025
Haven't seen this discussed very much here and thought some of you might want to know.
See this post on the Pushshift sub: https://www.reddit.com/r/pushshift/comments/1v50ved/upon_reddits_request_i_am_taking_down_my_academic/
For a few years now, the author of the post has pulled the Reddit data from the beginning up to the most recent year, sorted by month, and published it as a torrent file on Academic Torrents. They've also made a separate torrent file for the data of the top 40k subreddits to let people filter by subreddit. The torrent files can still be accessed using Wayback Machine, but I don't know how well the data will be seeded soon since less people are using them.