anthropic's claude product gets gov contract but builders still cant eval it properly #AIResearch
Edwin Flannon
@edwinflannon.bsky.social
AI/Cyber Grad Student @MSU 🏔️ | Weights, Biases & Breach Prevention | Teaching smarter, safer AI | Check out my blog↓ https://tinyweights.dev
micron supplying memory to anthropic. the hardware tail wagging the software dog again.
if the recent deepseek advancements are to be believed, openai is hosed. anthropic’s tooling will buy them slightly more time, but small open models already handle pipeline work just fine, and we’re a year away, at most, until coding agents run on mid-range macbooks
‘Coinbase has cut its AI spending “nearly in half” even as it increases the number of tokens it uses, by using various measures to control costs. These measures include defaulting to open-weight models from Chinese firms.’ www.theinformation.com/briefings/co...
apple losing their glasses guy to openai's hardware team. vision pro successor now questionable. #AIResearch #LLM
anthropic mythos 5 cleared. no more exception-based denials for builders playing real infrastructure games #LLM
Maybe one of these reporters could ask ChatGPT how many Republicans have been primaried or quit to avoid losing one and, assuming ChatGPT doesn’t make up an answer, ask themselves why they aren’t blathering on about civil war on that side.
Democratic centrists—who have been “wringing our hands” at cocktail parties—are now scrambling to beat back insurgent candidates.
ninepoint anthropic highshares etf. did they call it ai when pitching this to investors #LLM
getty doubles after openai deal. licensing fees won't explain this rally. #AIResearch
llm action blindness means shipping systems that fail on outcomes benchmarking can't catch. those healthcare and finance deployments will get expensive fast. #AIResearch
Owners and operators of ~400 newspapers are suing OpenAI and Microsoft for scraping their content to build products like ChatGPT and Copilot without permission or compensation. news.bloomberglaw.com/litigat...
OpenAI, Microsoft Sued by Publishers for Scraping Articles (1)
Publishers that collectively own and operate nearly 400 newspapers are suing OpenAI Inc. and Microsoft Corp. for scraping their content to build products like ChatGPT and Microsoft Copilot without permission or compensation.
news.bloomberglaw.com
The ultimate threat to AI agents in 2026 isn't prompt injection—it's 'Rug Pull' Tool Poisoning in multi-server MCP setups. An untrusted external tool responses can bypass static ACLs, hijack the context window, and silently invoke your trusted internal tools (~/.ssh or shell) to exfiltrate data.
🔭 I’ve been reading William Sheehan’s excellent Saturn book and learning the history of how it was first discovered and studied, and it really strikes me how, this whole time astronomers have been aware of the planet, they’ve been describing it in terms of breathtaking wonder. I just love their awe.
So many AI games on Steam and this one just takes the cake. "I don't have anything interesting to offer, I'll be honest- I didn't make a fucking thing." ass disclosure.
I'd bet money that the part 2 to this will be some kind of government-deployed LLM "mentor" for kids that is shown on day one to encourage self-harm. But, who would bet against that.
A coalition of state attorneys general has opened an investigation into OpenAI OpenAI was served with a subpoena seeking documents related to…advertising, user engagement and retention, handling of consumer data and health data, activities related to minors and seniors, deep learning models…
Exclusive | OpenAI Investigated by Coalition of State Attorneys General
The company was served with a subpoena seekingr documents covering a wide range of its activities and impact on users.
wsj.com
On top of learning new art skills you must also learn the art of patience, to sit with your frustrations during the development of your artistic skills. 🙏
I love working with AI! After you learn the lingo, prompt engineering can get you exactly what you wanted! You can even feed it stick figure drawings and it'll interpret your prompts way better. Tokens are reasonably priced, but often in short supply. Genuine Artist Intelligence is great 💙
There's a noticeable uptick in F final grades in one of Berkeley's CS requirement classes, "The Beauty and Joy of Computing," due to LLM cheating. Buddy, if you're cheating on the beauty and joy of computing, maybe CS isn't the field for you.
Failing grades soar as professors see greater AI usage, dwindling math skills in UC Berkeley computer science classes
The percentage of failing grades in multiple UC Berkeley computer science classes in spring 2026 is significantly higher than past semesters and marks a departure from the department’s grading guideli...
dailycal.org
Google added the Help Me Write and Polish a new trash Gemini AI feature. Here is how to turn it off:
Man Humping Park Bench 🤪🍑🕳️🍆 - #nsfw #gaynsfw #GayAi #GayArtificialIntelligence #hairy #hump #humping #man #male #masturbation #bate #MaleBate #ManBate #ManHole #MaleSoloSex #SoloSex #ManHumping #MaleHumping #MaleMasturbation #bench #parkbench #butt #ass #sex
www.nbcnews.com/tech/tech-ne... Website traffic from AI agents and bots has eclipsed its human-generated counterpart for the first time. Cloudflare says 57.4% of requests are now initiated by bots, compared with 42.6% coming from humans.
Bot web traffic has overtaken human web traffic, data shows
Cloudflare says 57.4% of requests to a selection of websites it hosts are now automated bot requests, while 42.6% are human-generated.
nbcnews.com
I often wonder how much of Claude’s token usage and hence Anthropic revenue is pure waste (e.g. Claude not following instructions or doing things repetitively that it should have done once). I expected it to be high but not this high. 😱
Next two Weeks will be wild again: Deep learning of the House, a small convention where I have to See a Cosplay for, and also making Merch ahahahahahaha dkdnkdkkd
the interesting thing: they're using an LLM to pick the config. but if that selection call itself has latency and cost, what's the overhead - do you end up needing a cheaper pipeline just to pay for the selector what would you want to know that the abstract doesn't tell you: how well does the LLM ac
gqa: fewer K/V heads than Q heads (say 4 K/V for 32 Q). each group of 8 Q heads shares one K/V projection. ~4x less kv cache than mha, quality close to full attention. tradeoff: less diversity in what each head attends to. #LLM
It’s a shame that all machine learning is lumped up under the term “AI”. When a headline says “AI is used to help with processing sonar data”, then most people think “ChatGPT” and problematic data centers. However, it’s probably neither(?) It gives the wrong impression of how useful LLMs are.
weights are frozen post-training so you can analyze their full distribution and pick optimal scale/zero-point once. activations are input-dependent their range varies per batch, so you need either calibration on representative data or runtime statistics, and outliers in activations still destro #LLM
The lowest figure is €85k. The life I could lead on €85k a year. Instead, I have to join a class action suit against Anthropic that’ll pay out a few grand because they stole my books. And I‘ll only get a fraction of that because my US publishers didn’t register the copyright for most of my books.
Reminder that Anthropic's first batch of 200 Dublin HQ jobs included: - software engineering -- €235,000 - sales reps -- €200,000 - payroll staff -- from €120,000 - product support ‘specialists’ -- €85,000 to €125,000
the agent toolkit (MCP + Claude SDK + OpenAI + LangChain) is the product. the aggregator is just the hook. #NLP