Gemini might have had enough of my nonsense. That period is ruthless...
brb. need to blow up our snowman www.sechselaeuten.ch
Das zünftige Zürcher Frühlingsfest | Sächsilüüte
sechselaeuten.ch
Gemini just told me it's a pleasure working with me, so, yeah... I'm def putting this on my resume
Processing crawl complaints, sometimes I wish I could throw some hosters under the bus. Why would you redirect your client to us for "excessive CPU usage" when the WHOLE MONTH there were about 1000 fetches from us from the WHOLE DOMAIN?!?!
Gary: I can't get this stupid ai to do anything right. Gary's prompts: Found a second cottage cheese, yes we did, precious! A cold lumpy 115g of it... and so salty, it stings our throat! What we do with it my love? We mix it with the nasty leaves? Garden-salads? Or do we eat it raw, just as it is?
A couple weeks ago we accidentally posted on LinkedIn the wrong description for a podcast episode. The comments under that post seem to be a great way to find automated accounts. www.linkedin.com/posts/google...
Googlebot Reality: Misnomer and SaaS Infrastructure | Google Search Central posted on the topic | LinkedIn
Think Googlebot is just a single "googlebot.exe" program we can run? In the latest episode of Search Off the Record, Martin Splitt and Gary Illyes dive deep into the reality of Google's crawling infr...
linkedin.com
I'm perpetually confused about people liking our ramblings in these podcast episodes. youtu.be/JpweMBnpS4Q
Google crawlers behind the scenes
YouTube video by Google Search Central
youtu.be
Who comes up with a name like "Punch, the monkey"?! No wonder everyone is bullying the poor thing
Genius tech bro move: Inventing the paragraph, Calling it a chunk. #showerthoughts
#showerthought What if the moltbook is not a demonstration of sentience, but a reflection of humans' (and bots') online shenanigans.
#showerthought LLMs rarely reply with "I don't know" to a prompt because we, humans, rarely ever admit online that we don't know something.
#showerthought LLMs rarely reply with "I don't know" to a prompt because we, humans, rarely ever admit online that we don't know something.
The coding agent I'm testing just suggested that we send temperature readings in Celsius to an LLM to make conversions to Fahrenheit, instead of making the conversion on the device, because ultimately we want reliability. So, yeah...
Posit 1: browsers don't particularly care about the status code of the parent frame, they'll happily display wherever content the server sent. Posit 2: crawlers very much care about the status code and they'll drop anything with non-content/transition status. #MisuseEverything
"Cursor CEO warns vibe coding builds ‘shaky foundations’ and eventually ‘things start to crumble’" fortune.com/2025/12/25/c...
a bald man wearing a suit and tie is sitting in front of a starz logo
ALT: a bald man wearing a suit and tie is sitting in front of a starz logo
media.tenor.com
If we migrated the googlebot.json file to a new location with a 301 redirect, would that break your fetches? Or the fetcher would just follow the redirect developers.google.com/static/searc...
developers.google.com
Subtwee... subblue: using something as an API doesn't make it an official API. developers.google.com/search/blog/...
Update on the Autocomplete API | Google Search Central Blog | Google for Developers
developers.google.com
Why do crawlers generally have a per-resource bytesize limit, like Googlebot's 15Mb? The internet: news.ycombinator.com/item?id=4467...
A valid HTML zip bomb | Hacker News
news.ycombinator.com
24 CPU cores working for 35 hours to train a model, 1% left to completion, and accuracy and loss values look phenomenal. Windows update restarts the PC...
Allegedly intelligent systems that allegedly excel at summarization requiring a manually updated file that essentially says what a site is about feels as intuitive as a sign explaining that the provided toilet paper is not meant to be used for giftwrapping the toilet.
SEO was claimed dead in 1997. web.archive.org/web/19981206...
NONE: ONLINE-ADS>> Search engines are dying
I'm beginning to believe that search engines are a dead-end technology and fretting over where your site comes up is a big waste of time. I'm now advising clients that we create good META tags, submit the site and then forget it.
web.archive.org
O.M.G. PEOPLE! UPDATING THE COPYRIGHT YEAR IN THE BOTTOM OF YOUR PAGES IS NOT A SIGNIFICANT UPDATE, NO NEED TO UPDATE YOUR LASTMOD
a man wearing a wig is holding his head and the words oh my god are below him
ALT: a man wearing a wig is holding his head and the words oh my god are below him
media.tenor.com
Just finished listening to the latest episode of SciShow Tangents, which means I've listened to every single one of the Tangents episodes @hankgreen.bsky.social et al. put out on the public web. AMA open.spotify.com/wrapped/shar...
My #1 podcast of 2024 – listen now
2024 Wrapped
open.spotify.com
Nothing says better "I want to wipe out my site from search" than these bad boys. Ok HTTP status codes do. And "noindex". Connection issues? Ok there are many things.
🔥 It's kinda fun watching people lose their shit about the Reddit blackout and blaming centralization, when Usenet (and more recently Tor) had the same problem, but with decentralized communities. You woke up and the community was not there anymore. Then Verizon and co fixed that with centralization