@narsilou.bsky.social

Me: This function is too slow. Find a faster algorithm. Cursor: Hold my beer. Me: *Slacking off with colleagues* Cursor: Ping. Me: 🤯

Bild

Hot take: Rust is really good for vibe coding, much better than Python or JS. Why ? The compiler will not let crap pass. Yes the LLM can still get it wrong, and fail. The elegant error messages will nudge the LLM, so I don't have to do it constantly.

Want to run Deepseek R1 ? Text-generation-inference v3.1.0 is out and supports it out of the box. Both on AMD and Nvidia !

Text-generation-inference v3.0.2 is out. Basically we can run transformers models (that support flash) at roughly the same speed as native TGI ones. What this means is broader model support. Today it unlocks Cohere2, Olmo, Olmo2 and Helium Congrats Cyril Vallez github.com/huggingface/...

Release v3.0.2 · huggingface/text-generation-inference

Tl;dr New transformers backend supporting flashattention at roughly same performance as pure TGI for all non officially supported models directly in TGI. Congrats @Cyrilvallez New models unlocked: ...

github.com

I'm disheartened by how toxic and violent some responses were here. There was a mistake, a quick follow up to mitigate and an apology. I worked with Daniel for years and is one of the persons most preoccupied with ethical implications of AI. Some replies are Reddit-toxic level. We need empathy.

Daniel van Strien@danielvanstrien.bsky.social · 2y ago

I've removed the Bluesky data from the repo. While I wanted to support tool development for the platform, I recognize this approach violated principles of transparency and consent in data collection. I apologize for this mistake.

It's pretty sad to see the negative sentiment towards Hugging Face on this platform due to a dataset put by one of the employees. I want to write a small piece. 🧵 Hugging Face empowers everyone to use AI to create value and is against monopolization of AI it's a hosting platform above all.