Sign in

Mike Trizna

@miketrizna.bsky.social
402 followers 932 following 9 posts

Data Scientist focused on AI/Data Literacy, responsible applications of AI for Libraries, Archives, and Museums

PostsRepliesMedia
Reposted by Mike Trizna
iNaturalist @inaturalist.bsky.social · 06/07/2026
We're looking for a CV/ML Engineer to help us improve the machine learning systems that power iNaturalist's species identification and geographic range modeling. If you're excited to help build tools that help millions of people engage with nature, we'd love to hear from you! Apply: buff.ly/YZqaW6c
Image of flowers with text overlaid saying: "iNaturalist: We're hiring! Computer vision/machine learning engineer. Full time, remote in the United States."
03921
Reposted by Mike Trizna
Daniel van Strien @danielvanstrien.bsky.social · 27/05/2026
Derived datasets are bigger on Hugging Face Hub than people realise. ~73% of analysed datasets on the Hub are derivatives of something else, i.e. cleaned, translated, extended, etc. Built an explorer that infers the missing lineage from content: huggingface.co/spaces/davan...
Screenshot of the overview of the lineage appScreenshot showing the lineage of one dataset with two steps of children.
1203
Reposted by Mike Trizna
Project Jupyter @jupyter.org · 19/05/2026
AI agents finally have a proper CLI for Jupyter notebooks. nb-cli lets agents read, write, execute, and search notebooks without a running server, built in Rust, optimized for LLM context windows. Read the blog: blog.jupyter.org/nb-cli-a-com...
blog.jupyter.org
nb-cli: A Command-Line Interface for AI Agents and Notebook Automation
The rise of AI coding agents has transformed how we think about developer tools. Large language models like Claude, GPT, and others are…
0184
Mike Trizna @miketrizna.bsky.social · 19/05/2026
Thanks so much for putting "The Case for Boring AI" right up front! I've been carrying that banner for a long time -- but just in conversations and meetings -- so now I have this chapter and amazing book-in-progress to link to.
152
Reposted by Mike Trizna
Rod Page @rdmpage.bsky.social · 09/05/2026
Exploring 218,567 pages of @bhl-au.bsky.social content using a "gilbert" curve. Just some of the content added by @nicolekearney.bsky.social and her team, HT to @cajunjoel.bsky.social for putting @biodivlibrary.bsky.social images on AWS which made this visualisation possible.
2136
Reposted by Mike Trizna
Tom Aarsen @tomaarsen.com · 01/05/2026
IBM just released the R2 generation of their Granite multilingual embedding models for retrieval, and the jump over R1 is very notable. Two models, both Apache 2.0: - granite-embedding-97m-multilingual-r2 (384-dim) - granite-embedding-311m-multilingual-r2 (768-dim) 🧵
182
Reposted by Mike Trizna
Nathan Lambert @natolambert.bsky.social · 08/04/2026
New report is out with the latest open model adoption data we have gathered for Interconnects & The ATOM Project. At the surface level, we can see Chinese models continuing to accelerate in adoption. The report details much more. atomproject.ai/report
2174
Reposted by Mike Trizna
mia ridge @miaout.bsky.social · 02/04/2026
Good news for anyone working on their proposals for the next Fantastic Futures conference - the deadline is extended to April 16! ai4lam.org/submission-i... #FF2026 is in the US, but will be very hybrid so you don't need to travel there to present or attend many sessions #AI4LAM #MuseTech
ai4lam.org
Submission Instructions - Ai4lam
FF2026 will be both in‑person and hybrid! Submit your proposal via: Fantastic Futures 2026 – ConfTool Pro – Login The information below is for planning purposes and may change or expand. The Program…
056
Reposted by Mike Trizna
Leland McInnes @lelandmcinnes.bsky.social · 31/03/2026
EVoC is a library designed specifically for fast clustering of high dimensional embedding vectors. It can produce high quality clusters extremely efficiently, and requires little to no hyperparameter tuning. Better clustering than UMAP + HDBSCAN; faster clustering than KMeans.
2133
Reposted by Mike Trizna
Daniel van Strien @danielvanstrien.bsky.social · 30/03/2026
You never know what data will be used for! I uploaded a @britishlibrary.bsky.social dataset to Hugging Face in 2022. IIRC one of my first PR to a HF repo! 4 years later, someone trains a Victorian chatbot on it More libraries should be sharing their public domain collections for AI to build on!
7837
Reposted by Mike Trizna
Eryk Salvaggio @eryk.bsky.social · 28/03/2026
Now that AI Literacy Day is over: mail.cyberneticforests.com/human-litera...
mail.cyberneticforests.com
Human Literacy
Something I Can Tell Students Now That I Am Not Teaching You and I probably both keep hearing that students should be working toward AI literacy. That you should know what to type into prompt windo...
1146
Reposted by Mike Trizna
Dominik Moritz @domoritz.de · 01/08/2025
🚀 We've just open-sourced Embedding Atlas – a tool for exploring large embedding spaces through rich, interactive visualizations 📊.
Screenshot of embedding atlas showing the embedding view on the left, a table at the bottom and charts on the right.
412133
Reposted by Mike Trizna
Margaret Mitchell @mmitchell.bsky.social · 27/11/2024
The best path forward in AI requires technologists to be reflective/self-critical about how their work impacts society. Transparency helps this. Appreciate Bsky for flagging AI ethics &my colleague’s response. Let’s make informed consent a real thing. More later; Recommend: bsky.app/profile/cfie...
611723
Reposted by Mike Trizna
Nathan Lambert @natolambert.bsky.social · 26/11/2024
Super excited to announce our best open-source language models yet. OLMo 2. These instruct models are hot off the press -- finished training with our new RL method this morning and vibes are very good.
59312
Reposted by Mike Trizna
Simon Willison @simonwillison.net · 25/11/2024
I like this new analogy for working with LLMs by @emollick.bsky.social "treat AI like an infinitely patient new coworker who forgets everything you tell them each new conversation, one that comes highly recommended but whose actual abilities are not that clear" www.oneusefulthing.org/p/getting-st...
oneusefulthing.org
Getting started with AI: Good enough prompting
Don't make this hard
618223
Reposted by Mike Trizna
Chris Holdgraf @choldgraf.com · 19/11/2024
Big milestone for Project Jupyter 🚨: Jupyter has finalized the creation of the Jupyter Foundation, hosted by @linuxfoundation.org. www.linuxfoundation.org/press/linux-...
linuxfoundation.org
Linux Foundation Announces Formation of the Jupyter Foundation
Linux Foundation Announces Formation of the Jupyter Foundation
16222
Reposted by Mike Trizna
Bluesky @bsky.app · 15/11/2024
Bluesky uses AI internally to assist in content moderation, which helps us triage posts and shield human moderators from harmful content. We also use AI in the Discover algorithmic feed to serve you posts that we think you’d like. None of these are Gen AI systems trained on user content.
350349603120