Eugene Yan @eugeneyan.com · 21/05/2025Some thoughts on leadership: eugeneyan.com/writing/lead... • What makes an exceptional leader? • What do exceptional leaders do? • Leadership styles: Commando, soldier, police 1100
Eugene Yan @eugeneyan.com · 19/05/2025converted all images to webp and hopefully made the site faster. something i wouldn't have bothered in the past 030
Eugene Yan @eugeneyan.com · 18/05/2025Had a fun couple of hours this weekend with Codex & Windsurf • Migrated off deprecated jekyll-algolia to official sdk (better indexing) • Added recommendations + relevance scores to each post • Improved site responsiveness; fixed dark mode flicker • Marie Kondo-ed unused files & dead code 151
Eugene Yan @eugeneyan.com · 07/05/2025Here's a three-minute demo of news-agents in action. It's pretty cool at the 30-second mark how the sub-agents get spawned! We then see the main agent assigning tasks and polling for progress, and finally shutting the sub-agents down when they're done with their assigned tasks. 130
Eugene Yan @eugeneyan.com · 30/04/2025@hamel.bsky.social & @sh-reya.bsky.social are two of the world's best on evals. They've built evals for 35+ AI apps & helped teams ship confidently. Now they'll teach everything they know on building evals that work. Enrollment closes in 4 days. Secret 35% discount code: maven.com/parlance-lab... 042
Eugene Yan @eugeneyan.com · 28/04/2025The Art of Doing Science and Engineering: Learning to Learn by Richard Hamming only $1.99 for the Kindle version today: amazon.com/dp/B088TMLQDC 080
Eugene Yan @eugeneyan.com · 15/04/2025Great example of generate -> validate loop + error analysis > "the most effective route to improve outcomes was brute force: retry steps until they passed or reached a limit. We give the validation errors ... to the LLM and built a loop runner" 1101
Eugene Yan @eugeneyan.com · 12/04/2025Stumbled on the first(?) RAG in NarrativeQA from 2017. Because books & movies were too large for LSTMs to do Q&A on, they embedded 200-word chunks and retrieved similar snippets to answer questions. "Chunking and cosine similarity retrieval is so 2017." arxiv.org/abs/1712.07040 0171
Eugene Yan @eugeneyan.com · 09/04/2025If you were building a Q&A feature (or chatbot) based on very long documents (like books), what evals would you focus on? 2180
Eugene Yan @eugeneyan.com · 08/04/2025Can't wait for when I can vibe code a production recommender system. Until then, here's some system designs: • Retrieval vs. Ranking: eugeneyan.com/writing/syst... • Real-time retrieval: eugeneyan.com/writing/real... • Personalization: eugeneyan.com/writing/patt... 1484
Eugene Yan @eugeneyan.com · 02/03/2025Been querying gpt-4.5 and it's better in ways we can't quantify yet: creativity, humor, world knowledge, wisdom, nuance, based, etc. Excited about how we'll discover new ways to evaluate gpt-4.5 on these aspects which will also transfer to product / application related evals 3191
Eugene Yan @eugeneyan.com · 28/01/2025♥️ it's tricky to separate what i do on the job (at the bookstore i work at) and what i hack on in my personal time. out of abundance of caution, to not discuss possible proprietary info, i won't be sharing more about the backend of aireadingclub.com 😔 1100
Eugene Yan @eugeneyan.com · 22/01/2025Thanks to the hundreds of readers who've tried aireadingclub.com and interacted with Dewey. If you've tried aireadingclub and have feedback, feature ideas, or thoughts on how AI can help you get more out of reading, please comment or dm me 🙏 060
Eugene Yan @eugeneyan.com · 17/01/2025> Nobody tells you the variables you should be regressing. What's the target? What's the source? Do you notice when results are rubbish? ... That's why I think you need smart people who appear to do something technically easy but actually not so easy. news.ycombinator.com/item?id=1906... 1211
Eugene Yan @eugeneyan.com · 15/01/2025Finally, if we need help with a term or character that was previously mentioned, Dewey can help with a summary of the term so we don’t have to look it up ourselves. 110
Eugene Yan @eugeneyan.com · 15/01/2025If you've stopped reading a book for a while, it can be challenging to pick it up again and remember what you've read. To help with this, it can help with summarizing the book up to the current page and refresh our memory, highlighting major themes, characters, and concepts. 110
Eugene Yan @eugeneyan.com · 15/01/2025It can also help with creating quizzes / flashcards. The goal here is to test our knowledge and improve retention. 110
Eugene Yan @eugeneyan.com · 15/01/2025With the context, it can answer simple queries via "Explain" and "Discuss". The goal is to keep us in flow while reading, instead of having to reread other sections of the book or open a web browser for our queries. 110
Eugene Yan @eugeneyan.com · 15/01/2025At the heart of AI Reading Club is Dewey, your AI reading companion. It understands context via selected text or the page we're on. This explicit context is displayed during discussions. At the same time, behind the scenes, it can retrieve and consider the rest of the book as implicit context. 130
Eugene Yan @eugeneyan.com · 19/12/2024phone usage has halved since starting december detox and deleting all social media apps off my phone 2210
Eugene Yan @eugeneyan.com · 11/12/2024to the latter point, the anti-ai comment were really strong in several comments, even those that didn't have anything to do with ai (these images are part of the appendix of the writeup) 010
Eugene Yan @eugeneyan.com · 11/12/2024also, while _some_ accounts that didn't like their data being scrapped had anti-AI explicitly posted on their profiles, not all of did. i hope i expressed this nuance sufficiently, and not a sweeping "people objecting to their data being stolen without permission as anti-AI" 100
Eugene Yan @eugeneyan.com · 11/12/2024ah i see your point now, thank you for clarifying! my point was that there was no stealing of data from bluesky's database, and no scraping of html. instead, the data was simply downloaded via the api. i deliberately avoiding comment on license or legality. 000
Eugene Yan @eugeneyan.com · 11/12/2024oh, perhaps it's just me that isn't used to such reactions to the release of a dataset, and the comments against the training of AI. haven't come across such negative reactions elsewhere tbh 120
Eugene Yan @eugeneyan.com · 07/12/2024Repeat after me: I will build evals for my tasks. I will build evals for my tasks. I will build evals for my tasks. 4649
Eugene Yan @eugeneyan.com · 07/12/2024Learning about quantization suffixes while `ollama pull llama3.3` download completes (fyi, quantization for the default 70b is q4_K_M) • make-ggml .py: github.com/ggerganov/ll... • pull request: github.com/ggerganov/ll... 3234
Eugene Yan @eugeneyan.com · 03/12/2024Here's my attempt at something similar—machine learning systems and applications in industry—a couple years ago. Isn't as fancy as the one above though lol applyingml.com/papers/ 2130
Eugene Yan @eugeneyan.com · 03/12/2024Wow, this is such a useful resource of industry LLM applications! And filtering via search/tags is so responsive. I was thinking of compiling something like this over the holidays (ala applied-ml) but thanks to @strickvl.bsky.social I can spend the time reading instead ♥️ zenml.io/llmops-datab... 3504
Eugene Yan @eugeneyan.com · 01/12/2024@hamel.bsky.social is 💯: hamel.dev/blog/posts/a... Write for yourself. Assume no one reads it. Write about topics to learn, clarify your thoughts, put out a bat signal. This makes it sustainable. Writing is primarily a single player game; multiplayer benefits (e.g., audience building) are bonuses. 4595
Eugene Yan @eugeneyan.com · 30/11/2024Great explainer on sinusoidal positional encoding and rotary positional embedding (RoPE). fleetwood.dev/posts/you-co... 2616
Eugene Yan @eugeneyan.com · 28/11/2024This Thanksgiving, I'm grateful for peace, health, and happiness for my family and me—what are you thankful for? 11413
Eugene Yan @eugeneyan.com · 27/11/2024And if you're looking for more learning over the long thanksgiving weekend, this could be a good place to start: eugeneyan.com/start-here/ 1145
Eugene Yan @eugeneyan.com · 27/11/2024Feels good to be mentioned on HN for engineers learning AI 🥰 Helping others is a big reason I write. Here's a list on ML/AI: ## Building AI systems • Patterns for Building LLM-based Systems: eugeneyan.com/writing/llm-... • What We’ve Learned From A Year of Building with LLMs: applied-llms.org 3709
Eugene Yan @eugeneyan.com · 23/11/2024It's helpful to distinguish between crisp vs. fuzzy tasks: • Crisp: Answers are verifiable, like math or code • Fuzzy: Answers are subjective, like reasoning, summarization, translation, multi-turn dialogue The latter is far harder to evaluate reliably aligned.substack.com/p/crisp-and-... 3470
Eugene Yan @eugeneyan.com · 20/11/2024yea some standard commands / packages need workarounds for fish, such as rbenv 110
Eugene Yan @eugeneyan.com · 17/11/2024probably data, evals, and the flywheel that mixes the cake batter 010
Eugene Yan @eugeneyan.com · 17/11/2024Eight years later, Yann LeCun’s cake 🍰 analogy was spot on: self-supervised > supervised > RL > “If intelligence is a cake, the bulk of the cake is unsupervised learning, the icing on the cake is supervised learning, and the cherry on the cake is reinforcement learning (RL).” 109413