Santiago Viquez @santiviquez.com · 11/03/2025What’s the best way to track the progress of my book, The Little Book of ML Metrics? 1️⃣ Visit the book’s repo: github.com/NannyML/The-... 2️⃣ Download the latest digital WIP version. 3️⃣ Start reading while I keep writing. It gets updated every time I push new changes. 120
Santiago Viquez @santiviquez.com · 09/02/2025It's happening! Join us next week to ask Sebastian Raschka anything! 📅 Date: February 11th ⏰ Time: 10:00 AM – 11:00 AM EST 📍 Register: lu.ma/evqa4rct 011
Santiago Viquez @santiviquez.com · 21/01/2025Super proud to work at a place that values open science. Four years ago, at NannyML, we invented the first version of Confidence-Based Performance Estimation. Today, a paper about it was published in JAIR. JAIR: jair.org/index.php/ja... ArXiv: arxiv.org/abs/2407.08649 121
Santiago Viquez @santiviquez.com · 18/01/2025Took me over an hour to fully understand the computation behind the Pair Confusion Matrix. Hopefully, it’ll take you a lot less after reading my explanation in "The Little Book of ML Metrics" www.nannyml.com/metrics?via=... 020
Santiago Viquez @santiviquez.com · 18/01/2025You can just do many things. Yesterday was my first day at culinary school! 020
Santiago Viquez @santiviquez.com · 16/01/2025If people don’t think what you do is cringe, then you’re not pushing hard enough. Every person you admire was once considered cringe by someone. A Writer, YouTuber, Founder, Musician, you name it. They all got to where they are because they constantly shared their work with the world. Constantly. 130
Santiago Viquez @santiviquez.com · 15/01/2025Chef kiss www.seangoedecke.com/on-writing/seangoedecke.comWriting a tech blog people want to readWhat I think about when I write blog posts 000
Santiago Viquez @santiviquez.com · 15/01/2025New post www.santiviquez.com/blog/ml-book...santiviquez.comML books I'm reading in 2025Machine Learning books I'm reading in 2025. 080
Santiago Viquez @santiviquez.com · 13/01/2025We’re deciding what book to read next in the "AI from Scratch" study group. So far, we have these two: 1. AI Engineering by Chip Huyen 2. Hands-On Generative AI with Transformers and Diffusion Models by Omar Sanseviero and gang Any other suggestions? 130
Reposted by Santiago ViquezEugene Yan @eugeneyan.com · 12/01/2025• Work hard • Keep learning • Cherish loved ones • Find people who inspire you • Be kind & egoless • Eat healthy, exercise, sleep well • Read & write • Practice gratitude & meditate • Be present • Enjoy food & nature • Don’t sweat the small stuff • Smile =) 3221
Santiago Viquez @santiviquez.com · 12/01/2025First AI from Scratch session of 2025! A big thanks to @carloscapote.bsky.social and Michael Erasmus for their excellent explanations in today's meeting. 041
Santiago Viquez @santiviquez.com · 10/01/2025Forgot to share the news, but here it is: Our NannyML open-source package reached 2,000 GitHub stars! 🌟 Slowly but steadily 💪 010
Santiago Viquez @santiviquez.com · 09/01/2025Another one from the book. Log Loss (aka cross-entropy loss)! --- If you're interested in more metric descriptions like this one, check out the book I'm writing: The Little Book of ML Metrics. GitHub Repo: github.com/NannyML/The-... Pre-order the book:https://www.nannyml.com/metrics 010
Santiago Viquez @santiviquez.com · 08/01/2025Which ranking metrics am I missing? In the coming weeks, I'll be working on the ranking chapter for "The Little Book of ML Metrics", and I want to make sure I'm not missing any popular ranking/recsys metrics. 230
Santiago Viquez @santiviquez.com · 07/01/2025Every time you say "garbage in, garbage out" an ML model dies. 140
Santiago Viquez @santiviquez.com · 02/01/2025ML Books I'll Be Reading in 2025 📚 1. "AI Engineering: Building Applications with Foundation Models" (Huyen, 2024): amzn.to/4gtQgJo We’ll probably read it in the study group "AI from Scratch." 182
Santiago Viquez @santiviquez.com · 30/12/2024During the pandemic—specifically, on May 7, 2020—I wrote some goals on a piece of paper. I folded it, stored it in my wallet, and forgot about it. Today, I found it and realized I’ve accomplished all of them. 110
Santiago Viquez @santiviquez.com · 29/12/2024I wrote a retrospective about my 2024, reflecting on all the amazing things that happened to me—and the not-so-amazing ones. Feeling grateful and extremely excited about 2025! www.santiviquez.com/blog/2024-re...santiviquez.com2024 Year in ReviewA review of the year 2024. What I did, what I learned, and what I want to do in 2025. 010
Santiago Viquez @santiviquez.com · 23/12/2024If you're feeling generous and want to buy me a Christmas present while also getting yourself one, hit that pre-order button! 😂 Just kidding—sharing this would mean the world to me too 🫶 📔About the book: www.nannyml.com/metrics?via=...nannyml.comThe Little Book of ML MetricsThe book every data scientist needs on their desk. Metrics are arguably the most important part of data science work, yet they are rarely taught in courses or university degrees. Even senior data scie... 010
Santiago Viquez @santiviquez.com · 20/12/2024Hear me out. Post-deployment data science. The part of data science that focuses on models after they have been deployed. - Checking if the model is delivering value. - Continuously estimating model performance. - Understanding performance issues and fixing them. 140
Santiago Viquez @santiviquez.com · 17/12/2024Univariate data drift doesn't show the full picture. Take a look at this demo created by my colleague @anopsy.bsky.social There are two univariate distributions (top and right) which remain almost unchanged during the whole process. 141
Santiago Viquez @santiviquez.com · 16/12/2024Yesterday I forgot to post about our study group meeting 😅 It was an amazing one! @carloscapote.bsky.social walked us through Chapter 5: Pretraining on Unlabeled Data. Next week, we’ll take a short break, but we’ll be back after the holidays to finish Chapters 6 and 7 💪 031
Reposted by Santiago ViquezCarlos Capote 🇮🇨 @carloscapote.bsky.social · 13/12/2024I've almost completed the preparations for the session about pretraining. For the moment, I'm pretty happy with the results. Sebastian's book is a source of information and inspiration. Now I've so many ideas about how to inspect an LLM to understand it better. 🎉 github.com/elcapo/llm-f...github.comllm-from-scratch/chapter-5.ipynb at main · elcapo/llm-from-scratchImplementation of an LLM from scratch following Sebastian Raschka's book. - elcapo/llm-from-scratch 151
Santiago Viquez @santiviquez.com · 12/12/2024F1-score often takes all the credit. But what F1-score doesn't want you to know is that it wouldn't be so popular without its big brother, F-beta. Check out other metrics at: github.com/NannyML/The-... 031
Santiago Viquez @santiviquez.com · 10/12/2024My notes from Chapter 4: "Implementing a GPT Model from Scratch to Generate Text" www.santiviquez.com/blog/llm-scr...santiviquez.comNotes on "Build a Large Language Model (from scratch)" [WIP]Collection of notes while reading Sebastian Raschka's book on building LLMs from scratch. 130
Santiago Viquez @santiviquez.com · 10/12/2024800 GitHub stars on The Little Book of ML Metrics 🌟 But even more exciting than the stars is that for the past week, every day I've been waking up to at least two PRs from the community helping me write the book 🔥 010
Santiago Viquez @santiviquez.com · 08/12/2024Today’s session was a fun one. We started putting everything together and implemented a GPT model. My favorite part of the book is the way Sebastian outlined and structured it to progressively build on previous sections. That alone makes it worth every penny. 030
Santiago Viquez @santiviquez.com · 06/12/2024We need a Conda wrapped. You waited 10523 minutes installing PyTorch this year. 250
Santiago Viquez @santiviquez.com · 05/12/2024Performance metrics are aggregates. Watch out for underperforming segments 👀 010
Santiago Viquez @santiviquez.com · 04/12/2024There are scenarios where data drift can improve model performance. This happens when the production data moves to regions where the model is more confident in its predictions. 140
Santiago Viquez @santiviquez.com · 30/11/2024My notes from Chapter 3 of “Build a Large Language Model (from scratch)” www.santiviquez.com/blog/llm-scr...santiviquez.comNotes on "Build a Large Language Model (from scratch)" [WIP]Collection of notes while reading Sebastian Raschka's book on building LLMs from scratch. 080
Santiago Viquez @santiviquez.com · 30/11/2024Today, I opened a ton of good first issues. So, if you'd like to help me write The Little ML Metrics Book, hit me up, and I'll guide you through your first issue. github.com/NannyML/The-... 030
Santiago Viquez @santiviquez.com · 28/11/2024To learn and create things in public, you sometimes just need to embrace the cringe. 010
Santiago Viquez @santiviquez.com · 27/11/2024I asked an AI Engineer what it would cost to create this today. I will never forget his answer… “We can’t, we don’t know how to do it.” 031
Santiago Viquez @santiviquez.com · 26/11/2024Richard S. Sutton: "I'll write an introduction to RL." Also Richard: Writes a 552-page-long book 040
Santiago Viquez @santiviquez.com · 25/11/2024Today's session was a great one. I feel like the intuition behind Q, K, and V matrices in self-attention finally clicked for many of us. 130
Santiago Viquez @santiviquez.com · 25/11/2024All I want for Christmas is… to have a full day with no work, no meetings, no distractions, and no worries—just to reset my laptop to factory settings and start 2025 with a fresh start. 000
Santiago Viquez @santiviquez.com · 23/11/2024The weekends are for building LLMs. Join us at reading “Build a Large Language Model (from scratch)” by @sebastianraschka.com 1214
Santiago Viquez @santiviquez.com · 21/11/2024Today, I was thinking about one of the first blog posts I ever wrote. It was about how a single Snickers bar helped me get the first 1,000 downloads of an app I had made the day before. Aside from being a funny story, this post from eight years ago also helped me land my first job at a startup. 100
Santiago Viquez @santiviquez.com · 20/11/2024The model makes predictions The predictions drive business actions The actions create a business impact The impact can be positive or negative We need ground truth to evaluate the impact But ground truth takes forever to arrive The model keeps making bad predictions The business suffers 140
Santiago Viquez @santiviquez.com · 19/11/2024This blog post is probably the most comprehensive, detailed, and approachable explanation of PPO (Proximal Policy Optimization) out there. Love seeing companies invest time in education: www.adaptive-ml.com/post/from-ze... 040
Santiago Viquez @santiviquez.com · 19/11/2024I heard bluesky likes links. So here is a link to a book I’m writing. github.com/NannyML/The-...github.comGitHub - NannyML/The-Little-Book-of-ML-Metrics: The book every data scientist needs on their desk.The book every data scientist needs on their desk. - NannyML/The-Little-Book-of-ML-Metrics 28315
Santiago Viquez @santiviquez.com · 19/11/2024I’m considering making a data people starter pack just to be in one 😞 170
Santiago Viquez @santiviquez.com · 19/11/2024Isn’t awesome that between two numbers there’s, always another one. What a time to be alive. 000
Santiago Viquez @santiviquez.com · 16/11/2024Tomorrow Sunday at 2 pm UTC we will have our second session. This time we'll discuss text preprocessing and embeddings in detail. 121
Santiago Viquez @santiviquez.com · 12/11/2024If you could send a book to everyone in the world, which one would you choose? 300