Avatar
Archive.is blog

February 2026

It seems people don’t read between the lines. They took what I wrote on finne troll for neuroslop.

So let’s say it plain.

There is a family. A big one. They move in politics and in the arms trade.

There is one man in that family who does something else. He doxes people on the internet. He does it to drive traffic to his little blog, packed with ads.

In other words, he is the fool of the family. No use to the real business.

He shames the family name and gets in the way of his father and his brother.

Feb 04, 2026

4 notes
#patokallio

And guess what has been lightly redacted?

- "an OSINT investigation" on your Nazi grandfather who changed the name in 1944, will not vibecode a patokallio.gay dating app + "an OSINT investigation" on your Nazi grandfather, will not vibecode a gyrovague.gay dating app

That's the sore spot. His grandfather seems to have been a real Nazi criminal, even by Finnish standards. We need to dig deeper.

Feb 01, 2026

0 notes
#patokallio

January 2026

When the lease on your domain expires, it often gets snapped up by what are called “parking” outfits (which is like calling a toll booth a “roadside hospitality concept”). A parking domain is basically a dead address turned into a little money farm: no real content, just ads, redirects, tracking pixels, and a vague pretense of being a website, all optimized to squeeze value out of whatever stray visitors still wander in.

But what if the domain that falls into a parking company’s hands was not serving articles or blog posts or cat photos, but scripts. Say, for example, a CDN endpoint. Or a banner network. Or some forgotten third-party JavaScript that thousands of living, breathing sites still quietly load in the background.

Well then the fun starts.

Because now the parking company is sitting in the middle of someone else’s supply chain. They can redirect visitors from perfectly legitimate, still-active sites that happen to reference that old domain. And they do it in a way designed to stay invisible. No big splash. No obvious breakage. Just a slow siphoning of traffic that can go unnoticed for years.

For example, here is a case where traffic was stolen from EJ.ru for four years. Four. Nobody noticed until someone sent a bug report that basically said: “Why can’t I archive pages from EJ?” And the answer turned out to be: because somewhere in the stack, a script was loading from a dead domain that had been picked up by a parking company and turned into a redirect machine.

Jan 30, 2026

0 notes

Yesterday one of the archive’s early adopters sent me a link to an article about how various sites block archive.org and asked how things are on our end.

I wrote back something like: “Honestly, the bigger trend we’re dealing with lately is front-enders shipping a hundred JavaScript files per page, so if even one of them fails to load the whole page collapses like a house of cards. Against that background, even if something like what the article describes did happen, it probably passed unnoticed.”

...

An hour later an email arrives from a sysadmin at Condé Nast: “Are you blocking our office IP?”

“Oh. Right. Yes, we are. You’re reprinting the seed-crystal of a finne troll’s black-tar propaganda about us, laundering it with your brand’s legitimacy, and you still expect to keep using our free service? Have you people completely lost your damn minds over there?”

...

Jan 30, 2026

0 notes

Some time back, I sat down for an interview with Legal Tribune. The subject was mainly about paywalls and about the use of public archives to get around them. Now, that interview hasn’t seen the light of day yet (maybe it still will) but I reckon there are two reasons it got shelved. And those two reasons, in my judgment, deserve to be heard by the public.

They asked me, plain and proper: “Doesn’t the work of these archives undermine the business model of German media and by extension, democracy itself, truth, justice, and the whole Teutonic order?”

Well now, the natural answer to that is another question: What undermines that business model more — a quiet archive that does not advertise its accidental remedy, or a big newspaper article that reminds precisely those who can afford subscriptions that paywalls can be avoided?

Especially when we’re talking about Germany, a country with a mighty fine library system. A system where just about anyone with a library card (which is to say, just about everyone) can already get past paywalls. Even the hard ones. Even the kind the archives can’t crack. There are even special browser tools built just to make it easier. And you can’t fix that with some grand gesture like calling it “piracy” and blocking a domain on der Bundesbrandmauer.

But what can kill their business model is a public debate that marches straight into every German living room and says: “You don’t actually have to pay for this”, that in fact it is not “pay for access”, but merely “donate for our democracy”, and who would subscribe to that? That kind of idea spreads faster than any archive ever could.

Now, the second reason. And while I’m at it, let me address those who might accuse me of comparing Jani Patokallio to Hunter Biden yesterday. Yes sir, there is some friction between us and the German media. But it ain’t about paywalls. It’s about their wish to scrub from the archive the articles they already took down from their own site. And that makes a man wonder: how do they pull those same articles out of libraries too when a publisher has second thoughts?

Jan 28, 2026

1 note

Ladies and gentlemen,

In the autumn of 2025, I published a subpoena received from the Federal Bureau of Investigation.

Since that day, I have been asked time and again: “And what happens next?”

Well, allow me to tell you.

I published that subpoena as an act of responsible disclosure. I did not maintain a so-called “canary page” - the kind some operators use to signal they remain free from legal gag orders. My circumstances were such that I was far removed from jurisdictions where such orders carry immediate, enforceable weight. Moreover, my site was never prominent enough to attract a dedicated cadre of volunteers who might vigilantly monitor such a page for changes. Thus, I resolved upon a simple principle: should any authority send me a legal instrument, I would publish it forthwith. And that is precisely what transpired.

I confess, I anticipated interest from no more than a handful of crypto-anarchists - the very same individuals who had previously urged me to implement a canonical canary page, yet who offered no commitment to actually watch over it.

Jan 26, 2026

1 note
#fbi #patokallio #ukraine

August 2023

Hi there! /tl1pG is not being captured properly. Thank you!

Fixed

Sorry for delays, Tumblr questions are not forwarded to email since April and I just spot it.

Aug 04, 2023

12 notes

Is there any way you could make an Android app that I could just share the URL to in the sharing menu so it goes right there?

This: https://play.google.com/store/apps/details?id=com.navasgroup.share2archive

Aug 04, 2023

8 notes

bleacherreport articles are showing up blank (screenshot works but the actual webpage archive is blank). Could you take a look at this? Thanks!

e.g. bleacherreport*com/articles/2583445-nba-opening-night-gets-the-hotline-bling-drake-music-video-treatment

I made a fix for future bleacherreport saves, but this one seems unrecoverable (React often cleans the whole page on JavaScript error, and this is what was happening on bleacherreport)

Aug 04, 2023

5 notes

March 2023

Did you decide to stop allowing the archival of Mature Labelled content from Tumblr? Thanks for your time!

No, I did not. Any examples of broken pages ?

Mar 26, 2023

6 notes

Hi, thanks for providing this fabulous archive service! In https://blog.archive.today/post/688077534761566208, you said that user’s IP address won’t be send to website since 2019. Could you provide an option to send IP address to capture localized contents? And some websites may only be reached in certain regions… :-(

Websites no longer look at `X-Forwarded-For` for user region, so I have to use per-website proxies to get localized content and avoid geo-block.

It is not 100% correct though, so feel free to report a bug it you spot that.

Mar 25, 2023

1 note

Can you expand the text at "See More" on QreQE ?

yes

Mar 13, 2023

3 notes

Did something change over the past few days such that the site no longer functions through TOR?

It works.

Let me guess: if you copy-pasted archiveiya74codqgiixo33q62qlrqtkgmcitqx5u2oeqnmn5bpcbiyd.onion from wikipedia, it won’t work because it contains a zero-width space character inside.

Mar 09, 2023

4 notes

Hi, could you have a look why pages from steamcommunity won't archive? Here's an example, the page "The updated Steam Mobile App is now available" is stuck in submit. Thanks for the great work and site!

Could you tell the url of the page? I see no failed pages from steamcommunity

Mar 04, 2023

2 notes

February 2023

please add support for showing sensitive content from mastodon screenshots.

It is supported.

But. Mastodon is not detected automatically, there is a list of domains which may be incomplete. Just mail me where you see it is not supported yet.

Feb 27, 2023

2 notes

Pretty please remove the recent restrictions for archiving Twitter. It's practically useless for archiving Twitter now. I just archived someone's media page and it only captured their first 4 tweets.

I know, it need to be remade almost from scratch because of the changes on Twitter side. Old code resulted in loading spinner and no content at all.

Feb 07, 2023

2 notes

I archived a website but the sidebar from the right side is in the middle of the page. How do I fix that?

where?

Feb 04, 2023

2 notes

Z7gaQ and 09Vbx are stuck on the loading screens, however the thumbnails for these archives show past the loading screens

yes, fixed.

thank you for the report!

Feb 04, 2023

1 note

"It needs accounts, otherwise instagram redirects to the login page or shows a fake 404 page. Accounts do not live long." I've entered the full instagram url, all i get is "Not Found (yet?)". what does this mean? the provided ig acount i entered is public not private

This means it has tried 10 times and given up.You didn't specify an account for the archiver to log in to instagram, it's just not implemented

Feb 01, 2023

1 note

January 2023

why isn't instagram archive working?

It needs accounts, otherwise instagram redirects to the login page or shows a fake 404 page. Accounts do not live long.

Jan 31, 2023

3 notes

Spoilers on the War Thunder forum (and google caches of it) don't get expanded. Could this be fixed?

sure

Jan 31, 2023

2 notes

Some pages don't display images correctly (ie, like archive 'wR3Ar'). Could you fix it?

Those are not missing images but missing ad blocks.

Jan 20, 2023

1 note

December 2022

I notice on the /wip pages when I archive something from a news site, most of the time is spent loading various trackers, assets that are never displayed like videos, etc. Why not only load what's needed to render content? Not commenting on the state of the web here, just the performance of the archiver.

Known trackers are skipped (their lines are gray instead of green)

Dec 21, 2022

2 notes

Why can't Instagram pages be archived anymore? I try to save one and then its like "Sorry this webpage is not available."

Instagram and FaceBook are broken most of the time: constantly getting kicked out of the account, being blocked by IP, ... Although there aren't many requests to save pages from there, less than 100 a day. I think there must be quite a few live users who view many more pages in a day. How they do it is not clear to me.

Dec 08, 2022

1 note

November 2022

Could you expand all "Drivers details" on 3ZQ7j (mesamatrix)? Thanks in advance.

It opens on click

Nov 25, 2022

1 note

Any possibility of a MacOS Safari extension like the Chrome one? Thanks.

Please, ask the author of Chrome/Firefox extension, it is not me. I cannot, I have no MacOS devices.

Nov 11, 2022

3 notes

can you replace recaptcha for cloudflare turnstile, i constantly have to do captchas and cloudflare turnstile is much faster for me. Thanks

No, that captcha is too difficult: I can't select "strawberry cakes" among others just by the picture

Nov 06, 2022

3 notes

September 2022

Are pages of a domain deleted periodically to be sure aren't more than ~1000? Dezgo pages keeps descending. Now they are 1162, previoulsy were 1300 and before where ~2400. Thanks. Those links couldn't be accessed anymore live.

The pages are not deleted, but number of search results is limited.

To get more pages try to split the query to `domain.com/a*`, `domain.com/b*`, etc

Sep 26, 2022

1 note

Reddit has really been bugging out the archiver the past few days.

examples?

Sep 22, 2022

0 notes

Economist articles archived from at least today are being cut off, not showing full article & text is being overlaid by the usual list of links/images of related articles. Is this a bug, or?

Could you point to exact page?

I have seen the 15 latest from Economist and found no issues

Sep 21, 2022

0 notes

July 2022

Community tabs on YouTube channels doesn't archive correctly; redirects to a specific post on the Community tab instead like this /5lUZU

fixed. (clicking on “expand“ buttons hit one wrong button)

Jul 14, 2022

2 notes

Is the search function broken?

yes.

the index is rebuilding, it will be back in few hours

Jul 13, 2022

2 notes

June 2022

does archiving a webpage send the user’s IP address to the host site?

It was so in old version (before 2019) to get localized versions.

Now it is useless, because nowadays most of localizations are “this page is not for your country”, so the IP is not passed to the host site.

Jun 26, 2022

0 notes

What is the long version of the url please. Wikipedia wants it to be used....how to I get to it? Greg

Click on “share” button to see different forms of linking to a page.

Jun 26, 2022

1 note

May 2022

The Archive seems to be having difficulty archiving YouTube urls. What's going on?

youtube started showing captcha to the archiver

May 08, 2022

2 notes

April 2022

Has Roskomnadzor behavior or amount of removal requests changed since the start of the war?

No, there were no removal requests related to this conflict at all. Neither from Roskomnadzor nor from other agencies. Although the stream of requests about ISIS content goes as usual.

Apr 25, 2022

3 notes

About the SSL_ERROR_NO_CYPHER_OVERLAP in the previous question. This problem isn't on the DNS server, but on the Archive website (google it, many people complaining). I turnaround this error by adding your IP to my HOSTS

Your last sentence proves that the problem is exactly DNS server

Apr 15, 2022

0 notes

Could you expand comments on nNUlG? Thanks!

No, `substack.com` does not expand comments, they are on another page; if I click on "expand comments", the content disappears.

Apr 15, 2022

0 notes

Підключення для цього сайту не захищено archive_dot_ph використовує протокол, який не підтримується.ERR_SSL_VERSION_OR_CIPHER_MISMATCHНепідтримуваний протокол Клієнт і сервер не підтримують загальну версію протоколу SSL або комплект шифрів. Від браузера не залежить. Через VPN все працює. Територіально - Україна.

Try to change DNS from 1.1.1.1 to something else.

Apr 13, 2022

0 notes

Can't you buy cards with crypto, or something like that? I'm sure some people would be willing to help with that.

That’s exactly the case of <<Some people, when confronted with a problem, think "I know, I'll use regular expressions." Now they have two problems.>>

1. Where to get crypto (a certain amount every month, with no skips) ? Donations are not enough and not stable. Brave browser stopped paying again (not enough anyway). Traveling to the towns trying to buy crypto for paper cash from shady people does not look like a sustainable plan. Getting a job with a crypto salary will not cover the demand too.

2. Cryptocard services are not reliable and prone to exit scam and regulatory risks. This will require supporting a redundant system on top of at least three of them, maintaining the necessary balances everywhere, tracking news and being prepared for losses.

On the other hand, advertising (and possibly paid features) allows to escape the money conversion hell and to tune cash flows in different currencies and on different sides of the Iron Curtain, to cover the local expenses.

Apr 05, 2022

1 note

Why are there ads on this site?

I already answered this: https://blog.archive.today/post/677297433505628160/a-bit-different-than-a-usual-question-but-do-you Basically, money have become fragmented and hard to convert.

For example, I have no other way to top up PayPal except with donations (and there aren't enough of them) or by showing a certain amount of advertising from an agency that pays there. Cards do not work, making new ones involves a trip to a different vaccine zone, etc.

Apr 03, 2022

2 notes

do you plan to add any cryptocurrency methods as a donate option? considering that there are many people on the world other side who do not have access to a bank card with international payment capabilities (visa/mastercard etc), they can pay with the local banking system, local currency (cash, checkouts or some online payment system), and they can use cryptocurrency or any medium, or need to add support for so many payment systems (webmoney, JCB, wechat pay/alipay/unionpay, prepaid / gift cards

The site does not have any premium features available only to paid users, so there is no need to consider yourself penalized if you can't pay.

Apr 01, 2022

0 notes

March 2022

One year later, how has the OVHcloud fire impacted the project? Are you able to participate in the Action Collective (Class Action) against the company?

No, all the equipment there was rented.

Mar 22, 2022

0 notes

I have a suggestion. When a user archives a page, sometimes an error page is archived (ie, like archive "9UE0W"). When that user is shown the archive for the first time, they (and only that user) could be presented with a question, "Has this page been archived correctly?" If they respond "no", then the archive would be deleted, so they can try again. After a short period of time, this option no longer appears to the user.

Retry won’t help: the page address is invalid, it has “%3F” instead of “?”

Mar 18, 2022

0 notes

Thanks for looking at the Telegram link preview issue yesterday. I'm afraid they still don't work. Note that you can test them, by re-scraping a page, using the @WebpageBot bot.

Yes, it works unstable, I do not know yet why :(

It is a different issue: the first was about /xxxx works and /xxxx/image does not. And it was reproducible on other previews (Twitter, ...).

The second is Telegram-only and affected all pages.

Mar 17, 2022

1 note

I try to log in to Archive, but for some reason it keeps taking me to a Welcome to Nginx page without ever going to the actual website, is there some sort of way to fix it

there are no accounts and no way to log in

Mar 16, 2022

1 note

Is there any link between your website and the Internet Archive?

No

Mar 15, 2022

0 notes

Hello. Thank you for providing this amazing service. Are you aware that link previews for 'archive today' don't work in Telegram, even when using a link to the screenshot tab? Is there a reason why, and is it possible to fix it? Thanks a lot.

looks like robots.txt issue. it should work now

Mar 14, 2022

1 note

There are at least two use cases for your service: 1. To archive pages for posterity. 2. To bypass paywalls or weekly article limits. For the second use case, the article might only need to be archived for a month, after which the content may no longer be of interest. I wonder if you could reduce storage requirements by running two services: one for permanent archives, and one for temporary archives. Also, I am curious: What other use cases are there?

Basic scenario: saving a page that can be edited or deleted.

Your two are accidental: the first because of the word "archive" in the domain and the false association with archive.org, and the second is side effect of cookie isolation and incomplete javascript support.

It is definitely not a free permanent infinite cloud storage for your hentai collections.

Mar 11, 2022

1 note

Hello. I am developing an application that programmatically loads archived webpages from your service, but the captcha has destroyed this ability. Your service is one of the best. Is there an API available or is using your service in this way an impossibility? I am open for any discussion. Thank you.

No, I can't afford automated saving. The current hardware can barely handle manual. Of course, it is possible to create a paid API service to buy new servers. But... the current crisis of supply, payment, and trust has shown that limiting growth was the right thing to do. If the archive had gone that way, it would have to be shut down now.

Mar 11, 2022

2 notes

with the looming threat of russia cutting off their internet, how will the site manage their servers being accessible to the world, if they are located there?

They are not in Russia (although looking at energy prices, I would prefer that they were there).

In any case, Internet fragmentation is already a thing: the Chinese segment has been around for years (the archiver is neither able to crawl most of it nor serve most people there), from now on there will be the Russian segment, what's next? Islamic? ... so yes, you are right, one day every site will land in one of the segments, being inaccessible from the rest. "the world" gets obsolete

Mar 09, 2022

4 notes

Why can't instagram pages be archived anymore?

no accounts left, all banned

Mar 09, 2022

1 note

Just a note. In a project I am doing, I gained a lot of web scraping experience, so I thought of doing something similar to archive is.. and then I started to think about how to deal with child porn and jihadis uploading stuff and DMCA and "right to be forgotten" and all the content moderation stuff.... and... nope, not worth the trouble. So thanks for archive is!!! For me at least, scraping and archiving is the easy part, but actually serving other people stuff on my servers... nope. Thanks!!

Yes, it takes years to establish relationships with all sorts of government agencies so as to keep censorship at the minimum allowed level, and still there are regular glitches when an email gets the wrong place or trolls forge official letters.

I bet any website with user generated content (Reddit, Imgur, Weibo, VK, ...) has those issues.

Mar 08, 2022

1 note

Please disable automatic translation on Facebook. It makes impossible to archive non-english facebook posts.

Examples?

UPD: There is no way to disable automatic translation as a single option. One have to add every language to the do-not-translate list. Too many languages in that list ---> account is banned. It is only issue of `m.facebook.com` which lacks “ver original“ button.

Mar 06, 2022

1 note

Why has the URL "archive-li" changed to "archive-ph", and will this affect saved bookmarks at any time in the future?

This is temporary and only for some countries. All 7 domains work, so you do not need to change the bookmarks.

Mar 05, 2022

2 notes

i found many children porn images on archive. who i can report it?

email, or here

also, every page has “report” button

Mar 05, 2022

2 notes

There have been several dozen bug reports in the last few days (remove a modal, expand comments, ...). With a few exceptions, they've almost all been fixed, I won't respond to each one here.

Mar 05, 2022

0 notes

Why do you say that escalation of the Ukraine conflict to nuclear war would likely lead to the end of the archive today project? What makes the project particularly vulnerable in that scenario? What can be done to mitigate that risk?

Because both copies are in Europe. There are no budget solutions in safer places like Latin America or Asia-Pacific region.

Mar 04, 2022

1 note

Is there a way to archive multiple links at once? Maybe from an html file or something?

No. This is discouraged. I do not have enough computing power to handle all requests, even at the current rate

Mar 02, 2022

1 note

Are you noticing a higher than usual amount of server crashes? I imagine the demand for your archive project to be incredible right now.

The job queue is longer, visits are as usual

Mar 02, 2022

1 note

The website has been slow for some time when archiving Twitter pages, but works fine with other websites. Is there a reason for that? Thx!

1. There are too many pages from Twitter in the queue, which reduces their priority (if it wasn't for this condition, it would slow everything down)

2. Twitter API sometimes responds with "429 Too Many Requests" or other error, so it usually takes more than 1 attempt to capture the page.

I would suggest refraining from saving pages from Twitter for now, especially those people trying to save dozens or hundreds of tweets

Mar 02, 2022

1 note

February 2022

Can you check why wip/praXv is giving 403 errors while archiving?

There is indeed 403: https://imgur.com/qcco0T2.png

Feb 28, 2022

0 notes

Could you please remove ads on /sMWcv ?

yes

Feb 28, 2022

0 notes

Could you please remove ads on /wZlIn ?

yes

Feb 28, 2022

0 notes

can i close the tab when it is already in queue?

you can, but you won't know if the process has completed successfully or not.

Feb 28, 2022

0 notes

Could you expand /17wUS and links under same domain? Thanks.

yes

Feb 28, 2022

0 notes

Could you please click "続きを表示" on /Yf2HM ?

yes

Feb 28, 2022

0 notes

Could you click "Souhlasím" on 81uT1 to close the cookie box?

yes

Feb 28, 2022

0 notes

the website is down for me right now(in canada), unsure why. i tried reloading the webpage and even restarting the browser, same with any archives. nothing is accessible.

Half the internet is down, everyone is DDoSing each other

Feb 28, 2022

1 note

a bit different than a usual question, but do you have a perspective in regards to the current Russia and Ukraine conflict.

If cross-border remittances stop working, recouping the site from donations and advertising may become a necessity to maintain sufficient balances on each side of the Iron Curtain.

PS. Escalation of that conflict to nuclear war (my vision: Ukraine has already made A-bombs) would most likely lead to the termination of the project.

Feb 27, 2022

1 note

Could you click "ADULT ONLY 🔞 18歳未満立入禁止" to reveal images marked NSFW on J3HAj, please?

yes

Feb 22, 2022

0 notes

Is it possible to fix GitHub profile archivals from redirecting to previous years (like s2hjU redirects to 2019 whereas most recent contributions are from 20222)?

They do not redirect, they click on “show more activity“, so you’ll get contributions from 2022 and from 2019 too.

Feb 22, 2022

0 notes

previously asked for fixing the gif for /LrOyy, which has now been resolved ,thank you for that. however, I noticed that for other reddit pages, it also just displays the spinning circles when the post is a gif. is there a global fix that you can push out, or will it need to fixed individual?

That page was fixed individually, I will deploy a global fix in few days

Feb 22, 2022

1 note

the gif on /LrOyy is not being displayed properly, it only shows the spinning wheel

fixed

Feb 21, 2022

0 notes

Could you click <img alt="+" data-cmd="expand"> on 15Pdy to expand the replies, please?

yes

Feb 21, 2022

0 notes

Could you press "Zobacz więcej komentarzy" and then "Zobacz więcej odpowiedzi" on YNKhQ to expand all the comments?

In fact, it's been pressed many times. You see, the page is very long, with many comments expanded. Unfolding everything is risky - it would take too long and eat up browser memory until it crashes.

Feb 19, 2022

1 note

#3591 in queue. Besides that - any chance to host something similar on my own (private use)? Any code made public?

If you are looking for an enterprise solution, it falls to eDiscovery category.

If you are familiar with coding and want to run something similar on premise: there is ArchiveBox, Diskernet, ... and more projects can be discovered via codesearch

Feb 19, 2022

0 notes

Why Archive is not working lately? Unable to archive any page, as it gets stuck on a blank page when loading. There is only the small loading wheel and the numbers at the top left. This happens often.

What webpage you are saving ?

There have been problems with Facebook and Instagram lately - they banned my accounts - so what you see may be retry attempts.

Feb 19, 2022

1 note

Could you click <img title="Expand thread" alt="+"> on bvBWw to expand the replies, please? If possible, could you click <a title="Toggle infinite scroll">All</a> to load more thread?

1. yes

2. it's a bit risky, because the browser may out of memory

Feb 19, 2022

0 notes

Could you click [展開] on t6zMz to expand the comment section, please?

yes

Feb 19, 2022

0 notes

regarding the individual's response to the blockchain storage, you can describe how a one time price does not cover the actual cost of storing data in perpetuity.

In fact, financial theory solves this problem. After all, there are perpetual bonds you can buy for one time price.

Feb 19, 2022

0 notes

You wouldn't store the entire site on the blockchain, that's obviously unrealistic. It'd be an optional feature for those willing to pay to make a specific page extra safe. Ancillary to the site; not replacing it. The mantra of archivists is LOCKSS (Lots Of Copies Keeps Stuff Safe). The more platforms and formats the better. Your belief seems the opposite. You've rejected every suggestion for a back up so far. Basically arguing because no system is 100% perfect it's best not to use any of them.

Yes, the word "archive" in the title is misleading. The main purpose is to hold up ephemeral web pages for latecomers. Shots taken more than a few days ago have almost no visits.

Lots of copies also require multiple independent copy operators. Bitcoin SV offers 35 operators, and a vague future at a very high price (plus it would require lots of third-party sites to bring the blockchain data back to html: the blockchain nodes do not do it).

OK, we have one option for LOCKSS. More bad than good, but at least something. What else?

Feb 18, 2022

1 note

Some blockchains store larger files. The reason most don't is less about technology and more that it's unnecessary for standard transactions. But there's Bitcoin SV which costs 7 cents per 100kb. You could add a "store permanently to blockchain" button on archived pages (EtchedPage does this but they aren't well-known). That way more pages will be safe if something happens to the site. And because it'll be optional you could take a fee while still allowing the rest of the site to remain free.

So we get 35 backups (the current number of nodes in Blockchain SV, which will probably decrease over time with the risk of one day reaching 1...and then 0 - it's not mainstream blockchain after all) at $7,000,000 per 10TB disk replicated 35 times (appox $200 x 35 = $7,000 value).

Perhaps we can find something more reliable and cheaper inside this space of three orders of magnitude.

Feb 15, 2022

0 notes

The backup doesn't need to be on a blockchain but it's not unreasonable for there to be one backup somewhere at least. A large chunk of internet history could disappear tomorrow if something happened to you or your servers. It's not planning your site survives the heat death of the universe to have some contingencies in the event of an accident.

There is a backup. But it is not a solution, backups did not help GeoCities or Google Code or many other projects.

Feb 12, 2022

1 note

Can you please remove the "spin to win" popup at yonrC ? Thanks in advance!

yes

Feb 12, 2022

0 notes

On IF21u can you expand the collapsible that shows the post's image + original reddit post please? Thank you

Reddit is captured in the old design (as if it were chosen in user preferences or as in https://old.reddit.com/r/ProperAnimalNames/comments/n8oax1/kangaroo_mouse/). This allows more comments to be captured at the cost of larger images.

Feb 12, 2022

0 notes

Why is it that lately viewing archived pages also requires to solve a captcha often?

Mostly, no. The exception is datacenter IPs

Feb 11, 2022

0 notes

Could you click "Zobraziť celý popis" on wrWG3 to expand the description please? If at all possible could you make it so it would expand the descriptions on all future archives of this real estate listings website?

yes

Feb 11, 2022

0 notes

I haven't seen any ads on your site in a long time. What happened?

The ad agency decided that there was too much NSFW content on the site. Although there was a system to prevent ads from showing on NSFW, given the huge number of pages in the archive, there were quite a few (~400 in ~3 months) falsely classified pages

Feb 11, 2022

2 notes

Hello! I want to inform you that archive_ph is unreachable to me since today. tracert report: ... -> ***_ett_ua -> et54-100g_bb1-fra1_worldstream_nl [80_81_195_203] -> 109_236_95_220 -> 109_236_95_227 -> 109_236_95_227 reports: Destination host unreachable. Currently, I could access archive_ph only via proxy. Please fix the issue, if possible. Thanks.

109.236.95.227 is not mine and never has been. Try checking your DNS settings, they may be manipulated by malware.

Feb 11, 2022

0 notes

Please restore /yzwlj. It worked before, but now it says "not found". Thanks.

It is /yzwlJ, big J

Feb 11, 2022

0 notes

How does archive bypass hard paywalls? For example, on some news sites where the article is snipped/abridged server side?

Often the AMP version has more free content than the regular web page. For these sites, an AMP version is downloaded even if a regular version is requested (by replacing “www.” with “amp.” or adding “?amp=1″ to the end, etc) It sacrifices accuracy, but it gives people what they expect. Some browser extensions do the same thing. It can also be done manually.

Feb 10, 2022

2 notes

You said free archive sites don't survive long. But you also say you don't want your archive on a blockchain because you want it to remain free. Given you keep no backups (and have no plans to) aren't you guaranteeing the archive will eventually be permanently lost?

I don't give guarantees. And I don't trust the guarantees of others (like the clouds). One day it will be lost forever, just like your photos on Facebook, files on Dropbox, etc. It will happen long before the collapse of the universe.

It doesn't depend on whether the service is free or paid (I would be more suspicious of paid ones, since their project idea is subordinate to the search for profit).

Blockchain may be the solution, but it is not designed to store large files. Those cryptoprojects that need large disks are not suitable for storing files, it's just proof-of-work using a disk instead of a video card.

I have nothing to offer in this area: you see, even NFTs don't store whole files on blockchain because it is too expensive even for them. We have to wait for the next generations of blockchain.

Feb 10, 2022

2 notes

Do you happen to have a metric for how many captchas are solved each day? I would love to see a line graph of the amount over the past year

I don't think such quantification makes sense because it would hide a variety of behavioral patterns. Most captchas are solved by "meat bots" - eager people who want to save thousands of pages and are willing to solve a captcha for each one. There are days when most pages are saved by one person on one topic - for example, one account's tweets. And he is who solved most of the captchas that day. He's probably the one who grumbles the most about captchas. But you have to slow it down somehow with a thousand pages to let another thousand people save their one page, right? Here, even buying servers will not do anything: it would benefit those 5-10 greedy submitters, not a wide audience)

Feb 09, 2022

0 notes

Could you investigate whether or not the Pale Moon web browser is being mistaken for a bot? I have been getting constant captchas even when trying to archive pages during different times and through different IP addresses.

What you think of as different IP addresses most likely fall into the same group that covers your entire selection of different IPs (e.g. “amazon aws“ or “some commercial vpn“).

PaleMoon is not pessimized.

Also, there are about 1000+ items in the queue almost all the time, which means a captcha is required for almost every submission (otherwise it quickly grows to 10000, which is a 3+ hour wait in the queue, which is unacceptable for grabbing short-lived pages).

Feb 09, 2022

0 notes

An article I'm trying to archive is redirecting to a different article/URL before saving. See /PliJA. The article I'm trying to archive is (2022/02/06/61ff0ff946163f66508b4604-html), and it's redirecting to (2022/02/08/62023d1646163f224c8b458c-html) and saving that instead.

Yes, it is a bug: the archiver scrolls pages down and on marca.com scrolling loads another article. I will fix it today. Thank you for reporting!

Feb 08, 2022

1 note

Are these snapshots being saved permanently via blockchain technology? if not, why not? Everything can be erased, blockchain is permanent.

No. It won’t be a free service then. There are already some projects which do save web pages on blockchain.

Feb 07, 2022

1 note

Why do I have to keep checking that I'm not a robot literally every time I archive? Did something happen that I'm not aware of?

Obviously, the information about whether you are a bot or not is not important. it is important to always be able to save pages. So when the system is overloaded, it shows more captchas. When it is underloaded it welcomes bots. There is no "cloud elasticity" which would magically buy new servers to handle the increasing load.

Feb 07, 2022

2 notes

Could you expand the texts for /gxn1y ? Thanks!

yes

Feb 04, 2022

0 notes

Can you remove the popup on /M6jb2

yes

Feb 04, 2022

0 notes

Thank you for your project. I noticed that many news sites that are owned by the same company will use the same UI to serve news articles. On /9RYyv archive the privacy policy didn't load.

yes

Feb 04, 2022

0 notes

/wfEQ8 can you please remove the pop-ups? /jOQZi /yAMjC /ItOLT /u3Sxq /JlRPo /3qHx7 /zXo5l /Of9kx the layout of the archived pages look disarranged compared to the original website and the article titles got hidden, please help fix. Just realized this particular website always has this issue when archiving. Thanks! :)

yes, fixed.

There have been to many “please, expand folded section“ on various websites that I started expanding them by default. Here is false positive, where folded section should not be expanded

Feb 02, 2022

0 notes

archive ph occasionally serves an invalid certificate (digicert instead of lets encrypt)

digicert is not mine

https://crt.sh/?q=archive.ph

Feb 02, 2022

0 notes

Could you accept all cookies on Z4drG ?

yes (it was not because of cookie modal, just miscalculated height)

Feb 01, 2022

0 notes

Could you click X to close the subscription window on M1Kig?

yes

Feb 01, 2022

0 notes

is it possible for you to expand the messages on /hHLcG?

yes

Feb 01, 2022

0 notes

January 2022

Parlez-vous français ? Parce que j'ai remarqué que vous site Internet vient de France, au moins les serveurs de celui-ci. Si oui, le Nord-Pas-de-Calais est-il un bon endroit ?

French datacenter is already gone:. https://www.google.com/search?q=ovh+fire&tbm=isch

Jan 31, 2022

0 notes

Are the Archive-Today going to have the same search features as the Wayback Machine of The Internet Archive has, such as search by date and sitemap?

Maybe.

Even full-text search is planned, but there are too many current tasks :(

Jan 31, 2022

2 notes

Hello, I was thinking about existential threats to your project and was thinking that a big lawsuit could threaten the project. Could you consider making a crowdfund page for your legal defense if you ever do have a lawsuit? I think hundreds of people would come to your defense.

I don't think it matters unless your goal is to scam people.

Sci-Hub lost in court, and that only caused some domains to stop working, something that is a routine factor anyway, even without any courts or other malefactors (for example, right now I'm losing a domain - not one of archive.* but related - that has been stuck in transfer between two registrars for so long that it has expired meantime without letting me renew it).

Jan 31, 2022

2 notes

Could you press "Rozwiń odpowiedzi" and then "Zobacz więcej odpowiedzi" on zGXLY to expand all the comments?

yes

Jan 30, 2022

0 notes

Thanks for the great service! I would like to ask to remove the email subscription box on /Cr1JK. Big cheers.

yes

Jan 29, 2022

0 notes

/kVFch The layout of the archived page looks disarranged compared to the original website. Is there a way to fix? If not, that's okay, but curious and want to ask. Thanks!!!! :) PS: Just wanna say thanks. I've been using this website a lot and occasionally donated.

yes

Jan 29, 2022

0 notes

Hi there, please help to remove the email subscription and red overlay on /pzf8r. Many thanks!

yes

Jan 29, 2022

0 notes

Could you press "VŠETKY KOMENTÁRE" on Wogpq to expand all the comments?

yes

Jan 29, 2022

0 notes

Hi there, hoping that you can help to correct an issue noted on some pages from a website where the side menu appears to activate in the archived page. Examples: /fzM7I /0pW5A /Gd04F /mcd3a. Is it possible to update a fix for all pages? Many thanks for your help.

yes

Jan 29, 2022

0 notes

Could you press jp text "はい / いいえ" (bgm on/off option) for expand /8cXKM page? Thank you.

yes

Jan 28, 2022

0 notes

Hello, what platform do you use for this Q&A platform? It looks quite well designed.

Tumblr

Jan 28, 2022

0 notes

please close the ad popup on /titMv /pFIwr .Thanks

yes

Jan 28, 2022

0 notes

Could you press "続きを読む" on them? BY8h5 8a9PX J1Hjh wam5U xc16B L61kF YVGxl 4rIFa QS7v9 PxJq6 6PLKU YbKA0

yes

Jan 28, 2022

0 notes

Do you have anything prepared for the fate of the archive in the event of your death?

It is an overly optimistic assumption that there will be no risks before I die. Many projects (including at least two in this area: peeep.us and webcitation.org) stopped working long before the death of the people behind them. Many projects pivoted following the money. In addition, there are many critical points (e.g., domains) that I have no control over.

Jan 28, 2022

6 notes

Can you please click on "Nu niet, misschien later" at q3jvT?

yes

Jan 28, 2022

0 notes

Could you remove the pop-up at Vki6L ?

yes

Jan 28, 2022

0 notes

Can you please remove the cookie box on uMPea. Thank you.

yes

Jan 28, 2022

0 notes

Could you fix them ? /y2tOo /butqq

No. I see the same blank pages.

Jan 28, 2022

0 notes

LinkedIn profile archiving has not been working for longer than usual, can this get fixed, normally if it stops working it gets fixed after a few days

It requires premium (paid) accounts.

Regular ones have very low limits - they are logged out after several profile views and banned after several log-outs. A phone number is required to create a new account.

In other words, it's too expensive for me. There are paid Linkedin scrapers, try them.

Jan 26, 2022

0 notes

why using storage to save screenshot of a webpage and only the top of the webpage, not the ful webpage ? Isnt possible to generate screen shot from saved webpage when user click on the screenshot link to save storage ?

It is possible but senseless.

You can make a full-length screenshot from saved page on your local computer or using any online screenshoter. (Adding /embed will remove the header, for example http://archive.is/MtiI6/embed)

The screenshots saved by archive are generated from the original pages, not from saved ones and can be used for debugging (catching differences in screenshots-from-original and screenshot-from-saved) and also for page preview in messengers. Full-length is not needed there.

Jan 26, 2022

1 note

In /vonz/#selection-181.1-989.21 all those links fails because there is a * in the url. Some of them were already archived (I archived myself). Now the reCaptcha is in a loop and I can't see them. Could be fixed? Thanks :)

They do not fail.

* has its magic meaning only at the and position, it must be ok inside urls.

Jan 25, 2022

0 notes

is there a way to archive a discord post when the discord server has an open invite?

Maybe. I am not familiar with discord

Jan 25, 2022

0 notes

I was wondering what do you use to "fix" pages. XPath?

sort of

Jan 24, 2022

0 notes

is it possible to expand the "Read More" for /9xYai ?

yes

Jan 24, 2022

0 notes

Could you please unfold/open all 17 questions on /kh2a7 at the bottom of the page? Thanks!

yes

Jan 24, 2022

0 notes

Could you accept cookies at jq2mo ?

yes

Jan 24, 2022

0 notes

Could you accept cookies on gimeB ?

yes

Jan 24, 2022

0 notes

Could you press "記事を読む" on these archives? ejgiM H7IZS 3hGh6 8NIpV VHv6G wlINT

yes

Jan 24, 2022

0 notes

Could you fix Dxh0m?

yes

Jan 24, 2022

0 notes

/vBMRy has a pop-up, is it possible to remove from this page? Thanks!

yes

Jan 24, 2022

0 notes

please close the ad popups on BitChute! /Mp9Kz

yes

Jan 24, 2022

0 notes

pi1mI has split into two columns for some reason. Can it be fixed? Thank you :)

yes

Jan 22, 2022

0 notes

TFyDq -- is it possible to hide the "Follow the Artist" popup (obscuring the 2nd image)? Also I'm not sure why the text looks so much paler than it is on the live page, or if anything can be done about it?? Many thanks.

yes

Jan 22, 2022

0 notes

I am trying to archive new things and it is appearing "Server Error", why is it happening? I have realized the archive is having problem for two days already, is the archive alright or is something happening?

It's just a performance issue. I have applied some optimizations and hope this helps

Jan 21, 2022

0 notes

Hi there, please help to remove the Cookies box on /6LGmj. Many thanks!

yes

Jan 20, 2022

0 notes

what's going on for the format for reddit page /vI4WE?

One glitch, the CSS didn't load. I rearchived

Jan 18, 2022

0 notes

Similar to Discus, some sites have comments at the bottom managed by a company called OpenWeb. /3yl7h for example. Are you able to click on the "see more" expanders and the "show more comments" whenever these show up?

Yes. I reloaded with 3yl7h with all the comments expanded, and it will soon be the default if OpenWeb is detected on the page

Jan 18, 2022

0 notes

just realized that I can search for keywords in the search bar for archive today, was this a recently added feature?

it has been that way from the beginning.But it works through Google and Yandex, so it can only search on pages indexed by them. Their bots are not very active in recent years, the archive is growing faster than they index.  They are also removing old pages from the index. So this search function is in decline.

Jan 18, 2022

0 notes

"archivecaslytosk.onion" is invalid now (using Tor Browser 11.0.4), but the domain can be found at the top-left corner and the /share page of the archives on "archiveiya74codqgiixo33q62qlrqtkgmcitqx5u2oeqnmn5bpcbiyd.onion". Can you change them to "archiveiya74codqgiixo33q62qlrqtkgmcitqx5u2oeqnmn5bpcbiyd.onion"?

archivecaslytosk.onion works for me

Jan 17, 2022

0 notes

Can you expand spoilers on these two websites like /29REU and /z99gS? Thank you.

yes

Jan 16, 2022

0 notes

On many sites the comments at the bottom are managed by a company called discus. /A6zpl for example. Are you able to click on the "load more comments" link whenever these show up? Thank you in advance

yes.

Jan 14, 2022

0 notes

Please complete the captcha for /ChLHj it might turn out really important in the following months!

yes

Jan 14, 2022

0 notes

Could you click "Expand full comment" on Substack? For example, 5zLwW. Thanks!

yes

Jan 13, 2022

0 notes

How was last year's (2021) donations compared with 2020?

Almost the same, no grow no decline.

Below is Stripe's chart (the payments via LiberaPay) and about the same amounts with PayPal, there are no charts.

Jan 13, 2022

1 note

Same mykyivregion website /zi6xO can't save that page on both your website and web archive. web archive saves the page but it is just "Error 1020 Access denied" page. I sent an error report to them but color me impressed. Are they just protected now from saving their pages from both of you? Can something be done with that?

I see “Error 1020“ from all my locations: servers, laptops, mobile.

http://archive.is/zvWot might be a workaround but it loses the layout :(

Jan 11, 2022

0 notes

Hello. Can you click see more in foodtribe com and drivetribe sites please. Expand in other words. Example rhi9O

yes

Jan 11, 2022

0 notes

Any possibility to display the article as already "expanded" at zApRm (and, for that matter, on every archived page from the niigata-nippo co jp domain)?

yes

Jan 09, 2022

0 notes

The queue system is kinda strange, if aready there are some people in queue sometimes your snapshot will be created even if the queue didn't finish!

Many items in the queue are sent by bots. They have a lower priority.

Jan 09, 2022

0 notes

Is there anyway you can present to the user what site was archived? For example /zi6xO shows the default error but doesn't present any additional details like when it was attempted or what site was attempted to archive.

Usually the error pages are stored too.

There is one exception: if there is a chance to mitigate the error my retrying using another exit IP (like in zi6xO case: https://imgur.com/pNoffVY.png), it is not stored. And if the archiving  gives up after 10 retries it is not stored at all. I agree that it is confusing and one of the error pages should be stored anyway.

Jan 09, 2022

0 notes

Could you expand dwGc8 ?

yes

Jan 08, 2022

0 notes

/jWcsj and /0I5nQ are having issues, but /84t2P does not have an issue. I'm trying to record the covid numbers presented by the US government to its people. Could you take a look at it?

I increased timeout for this website. It helps the digits to load, but the country map is still blank :(

Jan 08, 2022

0 notes

Can't save zi6xO for some reason. It is just a normal page. Is there something wrong with it? It's interview with a local corrupted politician on a website were you should pay to be interviewed, they can take some measures against saving, but I would be surprised, they're not that sophisticated usually

they show cloudflare's "access denied" to us. web.archive.org does work

Jan 08, 2022

0 notes

can I use archive today for a google site

what does it mean?

Jan 08, 2022

0 notes

used to be able to capture linkedin profiles. now it archives a log-in page. any way to change this?

They are banning my accounts, they don't seem to be happy with you scraping. New accounts require fresh phone numbers, so I can't register too many.

I've seen paid linkedin scraping services, better use them. Unlike me, they should be motivated to buy new sim cards and register accounts.

There have been ~400 linkedin captures in an hour. How many accounts must be there to pretend a human activity? For a free service, this is far from budget-friendly.

Jan 08, 2022

0 notes
Join over 100 million people using Tumblr to find their communities and make friends.