FBI Tries to Unmask Owner of Infamous Archive.is Site

silence7@slrpnk.net · 4 months ago

FBI Tries to Unmask Owner of Infamous Archive.is Site

Knock_Knock_Lemmy_In@lemmy.world · 4 months ago

The archive runs Apache Hadoop and Apache Accumulo. All data is stored on HDFS, textual content is duplicated 3 times among servers in 2 datacenters and images are duplicated 2 times. Both datacenters are in Europe, with OVH hosting at least one of them.

To avoid detection, archive.today runs via a botnet that cycles through countless IP addresses, making it quite difficult for grumpy webmasters to stop their sites getting scraped. Access to paywalled sites is through logins secured via unclear means, which need to be replenished constantly: here’s the creator asking for Instagram credentials. Finally, the serving of the website is also subject to a perpetual game of cat and mouse: “I can only predict that there will be approximately one trouble with domains per year and each fifth trouble will result in domain loss.” As of today, archive.today still works, but users are redirected to archive.md.

Phoenixz@lemmy.ca · 4 months ago

here’s the creator asking…

Where?

Knock_Knock_Lemmy_In@lemmy.world · 4 months ago

https://archive.ph/M8wEW