Reposted by Mike TriznaiNaturalist @inaturalist.bsky.social · 06/07/2026We're looking for a CV/ML Engineer to help us improve the machine learning systems that power iNaturalist's species identification and geographic range modeling. If you're excited to help build tools that help millions of people engage with nature, we'd love to hear from you! Apply: buff.ly/YZqaW6c 03921
Reposted by Mike TriznaDaniel van Strien @danielvanstrien.bsky.social · 27/05/2026Derived datasets are bigger on Hugging Face Hub than people realise. ~73% of analysed datasets on the Hub are derivatives of something else, i.e. cleaned, translated, extended, etc. Built an explorer that infers the missing lineage from content: huggingface.co/spaces/davan... 1203
Reposted by Mike TriznaProject Jupyter @jupyter.org · 19/05/2026AI agents finally have a proper CLI for Jupyter notebooks. nb-cli lets agents read, write, execute, and search notebooks without a running server, built in Rust, optimized for LLM context windows. Read the blog: blog.jupyter.org/nb-cli-a-com...blog.jupyter.orgnb-cli: A Command-Line Interface for AI Agents and Notebook AutomationThe rise of AI coding agents has transformed how we think about developer tools. Large language models like Claude, GPT, and others are… 0184
Mike Trizna @miketrizna.bsky.social · 19/05/2026Thanks so much for putting "The Case for Boring AI" right up front! I've been carrying that banner for a long time -- but just in conversations and meetings -- so now I have this chapter and amazing book-in-progress to link to. 152
Reposted by Mike TriznaRod Page @rdmpage.bsky.social · 09/05/2026Exploring 218,567 pages of @bhl-au.bsky.social content using a "gilbert" curve. Just some of the content added by @nicolekearney.bsky.social and her team, HT to @cajunjoel.bsky.social for putting @biodivlibrary.bsky.social images on AWS which made this visualisation possible. 2136
Reposted by Mike TriznaTom Aarsen @tomaarsen.com · 01/05/2026IBM just released the R2 generation of their Granite multilingual embedding models for retrieval, and the jump over R1 is very notable. Two models, both Apache 2.0: - granite-embedding-97m-multilingual-r2 (384-dim) - granite-embedding-311m-multilingual-r2 (768-dim) 🧵 182
Reposted by Mike TriznaNathan Lambert @natolambert.bsky.social · 08/04/2026New report is out with the latest open model adoption data we have gathered for Interconnects & The ATOM Project. At the surface level, we can see Chinese models continuing to accelerate in adoption. The report details much more. atomproject.ai/report 2174
Reposted by Mike Triznamia ridge @miaout.bsky.social · 02/04/2026Good news for anyone working on their proposals for the next Fantastic Futures conference - the deadline is extended to April 16! ai4lam.org/submission-i... #FF2026 is in the US, but will be very hybrid so you don't need to travel there to present or attend many sessions #AI4LAM #MuseTechai4lam.orgSubmission Instructions - Ai4lamFF2026 will be both in‑person and hybrid! Submit your proposal via: Fantastic Futures 2026 – ConfTool Pro – Login The information below is for planning purposes and may change or expand. The Program… 056
Reposted by Mike TriznaLeland McInnes @lelandmcinnes.bsky.social · 31/03/2026EVoC is a library designed specifically for fast clustering of high dimensional embedding vectors. It can produce high quality clusters extremely efficiently, and requires little to no hyperparameter tuning. Better clustering than UMAP + HDBSCAN; faster clustering than KMeans. 2133
Reposted by Mike TriznaDaniel van Strien @danielvanstrien.bsky.social · 30/03/2026You never know what data will be used for! I uploaded a @britishlibrary.bsky.social dataset to Hugging Face in 2022. IIRC one of my first PR to a HF repo! 4 years later, someone trains a Victorian chatbot on it More libraries should be sharing their public domain collections for AI to build on! 7837
Reposted by Mike TriznaEryk Salvaggio @eryk.bsky.social · 28/03/2026Now that AI Literacy Day is over: mail.cyberneticforests.com/human-litera...mail.cyberneticforests.comHuman LiteracySomething I Can Tell Students Now That I Am Not Teaching You and I probably both keep hearing that students should be working toward AI literacy. That you should know what to type into prompt windo... 1146
Reposted by Mike TriznaDominik Moritz @domoritz.de · 01/08/2025🚀 We've just open-sourced Embedding Atlas – a tool for exploring large embedding spaces through rich, interactive visualizations 📊. 412133
Reposted by Mike TriznaMargaret Mitchell @mmitchell.bsky.social · 27/11/2024The best path forward in AI requires technologists to be reflective/self-critical about how their work impacts society. Transparency helps this. Appreciate Bsky for flagging AI ethics &my colleague’s response. Let’s make informed consent a real thing. More later; Recommend: bsky.app/profile/cfie... 611723
Reposted by Mike TriznaNathan Lambert @natolambert.bsky.social · 26/11/2024Super excited to announce our best open-source language models yet. OLMo 2. These instruct models are hot off the press -- finished training with our new RL method this morning and vibes are very good. 59312
Reposted by Mike TriznaSimon Willison @simonwillison.net · 25/11/2024I like this new analogy for working with LLMs by @emollick.bsky.social "treat AI like an infinitely patient new coworker who forgets everything you tell them each new conversation, one that comes highly recommended but whose actual abilities are not that clear" www.oneusefulthing.org/p/getting-st...oneusefulthing.orgGetting started with AI: Good enough promptingDon't make this hard 618223
Reposted by Mike TriznaChris Holdgraf @choldgraf.com · 19/11/2024Big milestone for Project Jupyter 🚨: Jupyter has finalized the creation of the Jupyter Foundation, hosted by @linuxfoundation.org. www.linuxfoundation.org/press/linux-...linuxfoundation.orgLinux Foundation Announces Formation of the Jupyter FoundationLinux Foundation Announces Formation of the Jupyter Foundation 16222
Reposted by Mike TriznaBluesky @bsky.app · 15/11/2024Bluesky uses AI internally to assist in content moderation, which helps us triage posts and shield human moderators from harmful content. We also use AI in the Discover algorithmic feed to serve you posts that we think you’d like. None of these are Gen AI systems trained on user content. 350349603120