HTTP Archive 💾

@httparchive.org

Public dataset that tracks how the web is built. Maintained by @patmeenan.com, @paulcalvano.bsky.social, @tunetheweb.com, @maxostapenko.com, and Nurullah Demir

Very pleased to have helped out as an editor on the Cookies chapter of the 2025 Web Almanac, which had 70 contributors this year. Congratulations to everyone who worked on it, and a massive thanks to @tunetheweb.com who does so much behind the scenes to get it published.

HTTP Archive 💾@httparchive.org · 7mo ago

The 2025 Web Almanac by HTTP Archive has been officially released! 🚀 We would like to thank all of our contributors from around the globe who made this extensive report possible! Check out the full report here: almanac.httparchive.org

The new Web Almanac is out. 🎉 If you don't know the Web Almanac, it's pretty much a summary and analysis of the state of the web based on real data from the HTTP Archive. As a yearly tradition, I'll go over it and highlight/comment on the things that stand out. Let's go! 🧵

HTTP Archive 💾@httparchive.org · 7mo ago

The 2025 Web Almanac by HTTP Archive has been officially released! 🚀 We would like to thank all of our contributors from around the globe who made this extensive report possible! Check out the full report here: almanac.httparchive.org

My contribution to this years Web Performance Calendar is all about Third Parties and Single Points of Failure (SPOF). Based on @httparchive.org data I found that 67% of websites have at least 1 render blocking third party - and quite a few of them are SPOF risks!

Stoyan Stefanov @stoyan.me · 7mo ago

Web Performance Calendar day 29 with @paulcalvano.bsky.social's research on 3rd parties, SPOF, and the need to test and monitor these. 67% of sites out there have at least 1 render-blocking external dependency. calendar.perfplanet.com/2025/third-p...

New @httparchive.org analysis about sites adding AI Bots and Crawlers to their robots.txt files. While robots.txt doesn't "block" bots by itself, it's a clear demonstration of the preferences of site owners and the sentiment towards AI crawlers on today's web. paulcalvano.com/2025-08-21-a...

AI Bots and Robots.txt

There’s been a lot of discussion lately around AI crawlers and bots, which are used to train LLMs and/or fetch content on behalf of their users. In the past few weeks I’ve seen blog posts about the am...

paulcalvano.com

We're please to welcome Nurullah Demir as the newest maintainer of the HTTP Archive project! Nurullah has helped lead the Web Almanac over the last two years and we look forward to seeing what this year's edition comes up with! Welcome to the team Nurullah!

It's that time of year where the web almanac is seeking contributors for the next edition. It takes a fair amount of work,but it's ridiculously rewarding, & guaranteed warm fuzzy feelings when you see the chapter you contributed towards published. I encourage you to give it a go see:

Contribute to the 2025 Web Almanac · HTTPArchive almanac.httparchive.org · Discussion #4062

Dear all, We are excited to announce the Call for Contributions for the 2025 Web Almanac (6th Edition)! The Web Almanac is an annual report that provides an overview of the state of the web, based ...

github.com

What do you think? Shall we do another Web Almanac this year? Diving deep into all the data we collect to see what’s changing in web trends? Check out this post if interested in getting involved: github.com/HTTPArchive/... Or nominate your favorite experts that you’d love to see author a chapter.

Contribute to the 2025 Web Almanac · HTTPArchive almanac.httparchive.org · Discussion #4062

Dear all, We are excited to announce the Call for Contributions for the 2025 Web Almanac (6th Edition)! The Web Almanac is an annual report that provides an overview of the state of the web, based ...

github.com