Register and share your invite link to earn from video plays and referrals.

Search results for InternetArchive
InternetArchive community
One keyword maps to one global community path.
Create community
People
Not Found
Tweets including InternetArchive
Zcash Ecosystem Digest | July 19th 🔸Introducing @ZakuraZcash Full Node 🔸@ZcashFoundation Zebra 6.1.0 Release 🔸zcashd is Sunset 🌆 🔸Zcash at @internetarchive DWeb Camp 🔸ZecHub Hackathon Projects In-Review! Full Digest + Network Stats:
Show more
You can query the entire internet, 100+ billion (!) rows, with @duckdb in under a minute. Yes, it's crazy, but you can basically select * the internet. You might be familiar with the Wayback Machine from @internetarchive. The @CommonCrawl is a similar project that also has all its data available in a S3 bucket. @dumkydewilde was interested to see how the 'vibe-coded' web has grown over the last few years. While @Lovable , @Vercel and @Cloudflare Pages have taken off tremendously, they still pale compared to the traditional Wordpress-Blogspot hegemony. See for yourself in the interactive Dive visualization, or read the full blog to do it yourself 👇. - Dive: - Blog:
Show more
Internet Archive Launches New Foundation in Switzerland, based in St. Gallen. This new nonprofit group operates independently under Swiss law while joining similar groups in Canada and Europe. • Initial priorities: > Endangered Archives initiative: Rescue and preserve vulnerable cultural heritage and historical records threatened by conflict, disaster, institutional collapse, or suppression. >Gen AI Archive project: In partnership with the University of St. Gallen’s School of Computer Science, aims to collect and preserve generative AI models and related systems
Show more
0
24
1.8K
178
Forward to community
We've been getting a lot of questions lately about how the Internet Archive digitizes books. The short answer: page by page, by hand. You may remember our viral 2021 video of Eliza Zhang scanning a book. That's still how we do it. Meet Eliza, and learn how we scan books:
Show more
0
1.1K
126.1K
17.1K
Forward to community
I just made sure this January, 2026 TIME Magazine article on the blatant sexual abuse by Effective Altruists to young women is saved at the Internet Archive. Why? It will get deleted soon. Speculate why.
Show more
I've been enjoying Victoria Whitworth's new work, The Book of Kells: Unlocking the Enigma. I've actually never seen the Book of Kells in person, somewhat to my embarrassment. I've been doing some reading about the origins of Christianity this year, however, and I figured I should know something about the most famous Irish manuscript. (Perhaps the most famous manuscript, full stop.) Reading the book, I was struck by how much the contents have suffered over the past ~1200 years (enduring everything from water damage to reckless malfeasance in attempted nineteenth century restoration), and I wondered whether AI could help give a sense for how the work might originally have appeared. I downloaded the Internet Archive's PDF and asked my friendly neighborhood agent to use gpt-image-2 to render each page the way it imagines it might have originally appeared. Remarkably, this all worked with a single prompt, with the agent spinning up 48 workers, since each page took a minute or two. (I'm sure that someone wiser than me could prompt the model better, ensuring somewhat more historical accuracy in color restoration and so forth. There is no gold leaf in the Book of Kells!) This part of the project went from conception to completion before I'd finished my morning coffee. I then wanted some easy way to view the results online, so I asked Stripe Projects ( to host the result on Vercel. That also worked in basically a single prompt: I also figured that people might want an easy way to download the full PDF of updated images, but it's a large (~200MB) file, so I decided that I should charge $0.10 to cover bandwidth costs using @MPP. I asked my agent to set this up, and it basically worked smoothly, though I had to tell it what MPP is (I guess it's not yet in the pretrain) and also manually set up the Cloudflare account that actually hosts the PDF and configure the API key. (Vercel seemingly has a 100MB limit.) The purchases now show up in my Stripe account alongside all other activity. The site now has a ready-made agent prompt for anyone who wants to download the whole thing. I'm guessing that we'll see a lot more UIs like this in the future. I remain pretty intrigued by the intersection of agents, micropayments, and stablecoins. I don't know much about managing crypto wallets from the CLI, but now AI can do that for me, while Stripe seamlessly handles turning it all back into fiat. So what is the moral of the story? • Whitworth's book is very good, and you should buy it. • The Internet Archive continues to be wonderful and a civilizational treasure. • While there are rough edges, setting up third-party services via the CLI now basically works. I'm pretty sure I wouldn't have bothered with any of this if I couldn't have outsourced almost all of the work to AI. • The image models have gotten very good. • There will probably continue to be all kinds of interesting applications of AI to history. (The Vesuvius Challenge of course being a shining pioneer.) • These days, I often find myself building single-use sites for things I'm learning or for books I'm reading. I think this is a cool new category of software.
Show more