Nish Tahir

@nishtahir.com

Principal Engineer (AI Research). Anti-hype. My opinions are my own. I try not to be, but I can and will be wrong sometimes. Blog: https://nishtahir.com Mastodon: social.nishtahir.com/@nish

Looks like Kimi K3 went in the direction Llama did with their license. > $20M in revenue for "Model as a Service" usecases requires a commercial license. Also if you have more than 100M MAU you have to prominently display "Kimi K3" 😂.

Bild

Why has software gotten worse? I think it's because with AI dev's don't have to refine ideas as much anymore. Constraints meant that they had to make choices really count. But now the cost of those choices are basically free. So uses get fed every bad idea they think of. ptrchm.com/posts/nothin...

Nothing Works and Everyone Is Euphoric

As I’m writing this, we’re in the middle of an AI-induced mass psychosis. People are literally token-maxxing themselves into hospital beds, scrambling to capture some of that market value before every...

ptrchm.com

I've seen a few posts showing Fable unable to count, after testing them myself I'm inclined to call them fake news. Strawberry (adjacent) mispelling tests on low and high

BildBild

I'm gradually switching up my local stack to lemonade - here are a few notes. I use Open WebUI as my frontend the docs provide good guidance on setting up lemonade server. However if you use/want task models, you need to set it up to load multiple models at once. max_loaded_models in config.json

I do not understand how despite ever improving tools and technology, (IVR) Interactive Voice Responses just consistently gets worse. Trying to get a human agent is like navigating a labyrinth where 1 wrong response sends you back to the beginning.

I've moved my local usage to qwen 3.6 35b and it is fantastic. The primary issue right now is inference is slow but it is very usable. My usecase today was working on an electron app for myself and it has been working fantastically even with moderately vague prompts

Bild

The "but you don't look at compiler outputs" argument needs to go away. They are not the same and saying they are the same is like comparing weighted dice to a calculator. Compiler outputs are computationally provable. At least for now, LLMs are not. skiplabs.io/blog/codegen...

Treat Agent Output Like Compiler Output | Skip

Why our discomfort with AI-generated code reveals exactly what we haven't built yet, and what the compiler analogy teaches us about trusting coding agents.

skiplabs.io

Trying Claude design for the first time and unfortunately this has been my entire experience. Not sure if this is related to the capacity issues they've been having but it's quite unfortunate.

5x Claude, Writing, [unknown] missing EndStreamResponse errors

My local LLM setup designed for low maintenance. GMKtec EVO-X2 (AMD AI Max 395+) 128GB, running Docker containers with Ollama and OpenWebUI. Web UI acts as an OpenAI compatible proxy so other machines can reach models over the network.

I was under the impression that linkedin is effectively a dead social network. Yet I was at an ML conference this weekend and heard more "Add me on LinkedIn" than any other method of communication.