Mike Trizna

@miketrizna.bsky.social

Data Scientist focused on AI/Data Literacy, responsible applications of AI for Libraries, Archives, and Museums

We're looking for a CV/ML Engineer to help us improve the machine learning systems that power iNaturalist's species identification and geographic range modeling. If you're excited to help build tools that help millions of people engage with nature, we'd love to hear from you! Apply: buff.ly/YZqaW6c

Image of flowers with text overlaid saying: "iNaturalist: We're hiring! Computer vision/machine learning engineer. Full time, remote in the United States."

AI agents finally have a proper CLI for Jupyter notebooks. nb-cli lets agents read, write, execute, and search notebooks without a running server, built in Rust, optimized for LLM context windows. Read the blog: blog.jupyter.org/nb-cli-a-com...

nb-cli: A Command-Line Interface for AI Agents and Notebook Automation

The rise of AI coding agents has transformed how we think about developer tools. Large language models like Claude, GPT, and others are…

blog.jupyter.org

Thanks so much for putting "The Case for Boring AI" right up front! I've been carrying that banner for a long time -- but just in conversations and meetings -- so now I have this chapter and amazing book-in-progress to link to.

Daniel van Strien@danielvanstrien.bsky.social · 3mo ago

What is "AI for libraries" beyond a catalogue chatbot? IMO: design patterns (OCR, extraction, classification, search), and agents that both run them and develop the small models behind them. As part of work with @natlibscot.bsky.social started a book on this: danielvanstrien.xyz/ai-patterns-...

IBM just released the R2 generation of their Granite multilingual embedding models for retrieval, and the jump over R1 is very notable. Two models, both Apache 2.0: - granite-embedding-97m-multilingual-r2 (384-dim) - granite-embedding-311m-multilingual-r2 (768-dim) 🧵

Bild

Good news for anyone working on their proposals for the next Fantastic Futures conference - the deadline is extended to April 16! ai4lam.org/submission-i... #FF2026 is in the US, but will be very hybrid so you don't need to travel there to present or attend many sessions #AI4LAM #MuseTech

Submission Instructions - Ai4lam

FF2026 will be both in‑person and hybrid! Submit your proposal via: Fantastic Futures 2026 – ConfTool Pro – Login The information below is for planning purposes and may change or expand. The Program…

ai4lam.org

EVoC is a library designed specifically for fast clustering of high dimensional embedding vectors. It can produce high quality clusters extremely efficiently, and requires little to no hyperparameter tuning. Better clustering than UMAP + HDBSCAN; faster clustering than KMeans.

The best path forward in AI requires technologists to be reflective/self-critical about how their work impacts society. Transparency helps this. Appreciate Bsky for flagging AI ethics &my colleague’s response. Let’s make informed consent a real thing. More later; Recommend: bsky.app/profile/cfie...

Daniel van Strien@danielvanstrien.bsky.social · 2y ago

I've removed the Bluesky data from the repo. While I wanted to support tool development for the platform, I recognize this approach violated principles of transparency and consent in data collection. I apologize for this mistake.

Super excited to announce our best open-source language models yet. OLMo 2. These instruct models are hot off the press -- finished training with our new RL method this morning and vibes are very good.

Bild

Bluesky uses AI internally to assist in content moderation, which helps us triage posts and shield human moderators from harmful content. We also use AI in the Discover algorithmic feed to serve you posts that we think you’d like. None of these are Gen AI systems trained on user content.