Mihai Maruseac

@mihai.page

Building AGI with Privacy and Security at OpenAI. Previously: ML Supply chain security @ Google OSS Security Team (model signing, GUAC). Previously: TensorFlow Security & OSS (@ Google) Previously: Haskell+differential privacy+ML @ LeapYear

Even with the Trump intervention to cheat out of a red card, the US got owned today. So hard that people would say "stop the count" once again. Gg Belgium! Good game

Dear @caltrain.com , I took 3 trains today, northbound and southbound, in the morning and in the evening. None of them had Wi-Fi working, for the entire duration of the trips. Unacceptable in the land of AI and tech startups.

In the AI era, if you have a bug in production code you actually have at least two. One bug is in the test suite, it should have surfaced the issue when the agent wrote the code.

"If they ever tell my story let them say that I walked with giants."--Troy I'm humbled&excited that the model we released last week is now trending as#2 on HuggingFace,between giant models like DeepSeek,Qwen,and Kimi. Excited that the community finds this useful; looking forward for more work here

Screenshot of HuggingFace trending models showing privacy-filter as the second model, between DeepSeek V4 Pro and Qwen3.6.

SHIPPED: My first release at @OpenAI with a ton of work from an awesome team: a privacy filter model that is small enough it can run in the browser while also pushing the frontier in the space. Also CLI tool to use model. Both under Apache 2.0. Read more: openai.com/index/introd...

Introducing OpenAI Privacy Filter

OpenAI Privacy Filter is an open-weight model for detecting and redacting personally identifiable information (PII) in text with state-of-the-art accuracy

openai.com

After significant deliberations today I decided to get out of the tech job industry. I just sent the "an update on Mihai" email. Starting tonight you can find me as a park ranger in Yosemite. That would be a career that cannot be impacted by AI related layoffs. It's 4/1, folks.

Pi Day is almost ending (already did if you look at UTC time, but there are a few more hours in US), so I just squeezed in an article about some geometries where pi is 2, 3 or 4. There were hints on my blog about this, but now this is done and I have more questions for follow-ups mihai.page/pi-2026/

Pi day 2026: Other values of pi

It's Pi Day today, so we ask: can pi be 2, 3, or 4? We find simple worlds where the answer is yes, worlds that we already talked about in this blog

mihai.page

I tested 80 models on two simple grid-based problems, asking them to locate 2026 and compute the sum of the neighbors when the numbers are placed in a spiral on the grid. Results surprised me: models performed better on harder problem, but they also cheated. Read more at mihai.page/ai-2026-1/

Testing 80 LLMs on spatial reasoning on grids

How do LLMs see 2D grids? Would it be harder for them to work on square grids or hexagonal ones? I'm expanding the Kaggle benchmark mentioned in the last article and I'm testing 80 different LLMs. The...

mihai.page

Yesterday was my last day at Google. It was a bittersweet departure, leaving a team I really enjoyed working with. GOSST's mission is extremely important and I still believe in it.

Image of 2 laptops, one clean, one with a lot of stickers on it (Google, OSV, GOSST, CoSAI, OpenSSF, AntennaPod, Gemini, DevConf, PyOpenSci, TensorFlow, HEIR, Scientific Python). On the side, there's a censored employee badge.

New year, new problems to test the LLMs on: arranging the numbers on a spiral, what is the sum of the neighbors of 2026? Read on for more details and some preliminary results, as well as how to suggest other LLMs to test: mihai.page/ai-2026-0/

Introducing my first benchmark of AI for 2026

Just like last year, on this special day, I create a new benchmark to test LLMs on different puzzles. This will not be the only benchmark I run this year, but it might be the only math related one.

mihai.page

New year, new problems to test the LLMs on: arranging the numbers on a spiral, what is the sum of the neighbors of 2026? Read on for more details and some preliminary results, as well as how to suggest other LLMs to test: mihai.page/ai-2026-0/

Introducing my first benchmark of AI for 2026

Just like last year, on this special day, I create a new benchmark to test LLMs on different puzzles. This will not be the only benchmark I run this year, but it might be the only math related one.

mihai.page

I cannot remember the rule of divisibility by .7., it's always so hard. This is the smallest prime where I have to actually do the division myself