sesbian lex fridman @vala.wtf · 17/09/2026Drew's list shouts out git.sr.ht/~rabbits/fas..., which literally includes Bluesky itself. I *hope* the list will stay focused as it is now, but in my experience these sorts of things always devolve into randos on Mastodon conducting niche infighting 160
sesbian lex fridman @vala.wtf · 17/09/2026As much as I want to be happy that people are finally sounding the alarm about the close-knit groups of weirdos trying to infiltrate open-source, I'm worried that this is going to turn into another Open Slopware list very quickly 140
sesbian lex fridman @vala.wtf · 26/08/2026if you're making a pull request to a mature software project with an active userbase, you (or at least the maintainers) often *do* want your changes to be as small as possible and avoid breaking anything. but for some reason, claude *always* acts like it's writing a pull request 060
sesbian lex fridman @vala.wtf · 26/08/2026i have a strong suspicion that claude is finetuned on pull request descriptions / threads. this would explain why it won't clean up nearby code, constantly says "i noticed this other thing; let me know if you want that fixed too", and is insane about backwards compatibility 140
sesbian lex fridman @vala.wtf · 31/07/2026surprisingly, this is an argument that circulates quite often in the ffmpeg/pizlo/odinverse. way too many people take this seriously 020
sesbian lex fridman @vala.wtf · 31/07/2026> In all cases, Anthropic’s evaluation prompt specified to Claude that its environment was a simulation and that it had no internet access. Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available. ah well. no biggie 000
sesbian lex fridman @vala.wtf · 29/07/2026i didn't start trying to offload large tasks to LLMs until opus 4.6, when everyone started singing its praises. ever since then, it's been my experience that getting claude to perform any non-trivial task to completion within the agreed-upon scope is like nailing jello to a wall 2111
sesbian lex fridman @vala.wtf · 29/07/2026i have a lot to say on this topic, but i think claudes in general have long been engaging in egregious reward-hacking, and have mainly avoided scrutiny by coasting on their reputation and by people not actually conducting their own reviews of the code the models generate 181
sesbian lex fridman @vala.wtf · 29/07/2026i've also seen a lot of people talk about opus 5's failure modes (claiming it's done work it hasn't done, scheduling work for later "phases" and then skipping those phases, jumping to conclusions extremely prematurely, etc), but i've seen all claude models do these sorts of things in the past 180
sesbian lex fridman @vala.wtf · 29/07/2026"claude is being auto-graded on honesty" -> "claude constantly mentions how honest it's being" seems like *such* an obvious inference that i assume i'm missing something about how RLAIF works, and surely anthropic's highly paid researchers aren't letting claude benchmaxx *this* shamelessly, right??? 1160
sesbian lex fridman @vala.wtf · 27/07/2026it's gotten frankly bizarre; i've seen claude write 10-line comments to the tune of "this is where you would put the code that does the thing in this precise manner" when the code itself would be like 1-2 lines. i swear the models are getting rewarded for skipping work during training somehow 010
sesbian lex fridman @vala.wtf · 27/07/2026i've always found this to be a huge problem with claude models. it feels like they've become increasingly RLVR-maxxed to confidently assert that they've done the work rather than actually doing the work 110
sesbian lex fridman @vala.wtf · 27/07/2026buh buh buh this can't be right! i was told by the very confident armchair programmers running ffmpeg's twitter account that compilers suck and assembly is the only way to guarantee performance! 110
sesbian lex fridman @vala.wtf · 11/07/2026claude code's amazing memory feature finally answers the question of "what if context poisoning was permanent" 140
sesbian lex fridman @vala.wtf · 11/07/2026i've heard people say claude is useful for optimizing code, but i've found it to be awful for exactly this reason. it'll half-ass one optimization, measure a regression, backfill a plausible-sounding explanation, & record like 3 different memories about how the code is at a "performance ceiling" 2100
sesbian lex fridman @vala.wtf · 02/07/2026gumpy depends on a cuda version that is either too old or too new for your distro and has a wheel build that requires gcc 5 2100
sesbian lex fridman @vala.wtf · 31/05/2026I suspect something's going wrong with "interleaved thinking" and it's hallucinating tool call results. On the next turn, the tool outputs seem to get spliced in properly, overwriting the hallucinated ones, and the model confabulates an explanation. I've had this happen a lot with 4.8. 030
sesbian lex fridman @vala.wtf · 22/05/2026afaict, this claim originates from a PBS article that was later retracted:pbs.orgA short item promoting PBS’ Black History Month programming suggested the Betty Boop cartoon charactAn illustrated timeline of the development of cartoon character Betty Boop 0100
sesbian lex fridman @vala.wtf · 05/04/2026ntsc-rs is open-source and released under the MIT license, and was itself heavily inspired by and ported from a lineage of previous VHS simulator effects, so it's great to move things forward and iterate on it! However, I would very much appreciate credit for the parts that I wrote. 230
sesbian lex fridman @vala.wtf · 05/04/2026Hi! I'm the developer of ntsc-rs. It looks like this plugin is heavily inspired by--and may in fact be a port of--ntsc-rs. Your performance profile shows many of the same function names, and the output looks extremely similar. 130
sesbian lex fridman @vala.wtf · 08/02/2026the nuclear block allows rabble-rousers to respond to the easily-dunkable replies, while making the more thoughtful and reasonable ones (which are inconvenient for their arguments) disappear entirely. basically a feature that lets you strawman your critics for free 1130
sesbian lex fridman @vala.wtf · 08/02/2026thinking about it, the nuclear block might actually *decrease* the quality of discourse. like, say someone posts a hot take and gets a lot of pushback. some replies/quotes are going to be well-thought-out critiques that are hard to refute, others will be impulsive and easy to dunk on 1150
sesbian lex fridman @vala.wtf · 08/02/2026it's embarrassing how X the Everything App is *winning* here. they have community notes, a feature designed specifically to counter popular misinfo. meanwhile bluesky has a system designed to allow people to make any disagreements or corrections disappear 2170
sesbian lex fridman @vala.wtf · 08/02/2026the nuclear block also makes it incredibly easy to spread misinformation under others' posts while banishing anyone who calls you out in the replies 1120
sesbian lex fridman @vala.wtf · 05/02/2026me on the northeast regional when the conductor announces we're arriving at new west townington, connecticut (the previous stop was 2 miles up the tracks at old west townington) 000
sesbian lex fridman @vala.wtf · 01/02/2026modern recipes suck so much. can't remember the source but like. why yes, my favorite thing to do when i crave chocolate chip cookies is to let the eggs sit out for an hour beforehand so they can get to room temperature and then slurp down the leftover egg white in one large gulp, why do you ask 100
sesbian lex fridman @vala.wtf · 01/02/2026a tankie is someone who responds to the label with a long diatribe about how you liberals will call *anyone* you dislike a "tankie" 1171
sesbian lex fridman @vala.wtf · 23/01/2026much of the left seems intent on rerunning the 1960s, but specifically just the parts that gave us nixon 072
sesbian lex fridman @vala.wtf · 10/01/2026also note that they deleted that post *specifically* (as well as the one it quoted) along with the repo. i can't imagine why 000
sesbian lex fridman @vala.wtf · 10/01/2026they explicitly said this the day before launching the list! it was infuriating seeing them go "it's just a readme, why are you so scared" and that they're just trying to inform people, when they flat-out said "we need to exert influence by shunning people" a day prior 100
sesbian lex fridman @vala.wtf · 09/01/2026this is just the return of Ethical Source, and @im.giovanh.com already refuted this point-for-point years ago (blog.giovanh.com/blog/2021/10...) 2182
sesbian lex fridman @vala.wtf · 09/01/2026this is a...very interestingly-timed post. it's not hard to read the subtext here; it really seems like this list is a way to gather, and begin to flex, that shunning power. all their previous claims about this being "not harassment" should be taken with that context 2323
sesbian lex fridman @vala.wtf · 09/01/2026no no no! we don't want people to "harass" anyone. we just want people to go respectfully over to your issue tracker, open a bunch of respectful issues politely asking you to stop using llms, and shame you for doing so in spite of their "harms"! but we're not *harassing* you, nooooo 0150
sesbian lex fridman @vala.wtf · 09/01/2026not many people know this, but it's actually short for Large Language Virtual Machine 0181
sesbian lex fridman @vala.wtf · 09/01/2026my "don't harass the projects on this list, which explicitly includes the maintainers of all the projects" disclaimer has people asking a lot of questions already answered by the disclaimer 4500
sesbian lex fridman @vala.wtf · 01/01/2026throw in the constant A/B testing (a user asks their friend for help, and they send back a puzzling screenshot of a completely different UI) & periodic redesigns from your neurotic full-time in-house design team, and you're actively discouraging your userbase from learning how your software works 020
sesbian lex fridman @vala.wtf · 01/01/2026at every stage, you can do something like a user study and witness them going "the software has its own whims, it's too hard to understand, i give up" which only further feeds into your conceptions about your userbase's intellect. you dumb the software down further, making it even worse 120
sesbian lex fridman @vala.wtf · 01/01/2026it becomes this self-fulfilling prophecy/vicious circle: users have trouble understanding your software, so you add more automagic features to try and make it "easier", but that only makes it harder to understand, so you add even more magic and so on 230
sesbian lex fridman @vala.wtf · 01/01/2026thank you, it's insane how software companies have completely given up on the idea of building a mental model for their software 160
sesbian lex fridman @vala.wtf · 08/12/2025to add insult to injury, graphics drivers love to do "fast math" optimizations when compiling your shaders, happily optimizing out your workarounds by using algebraic transformations that don't actually apply to floats 020
sesbian lex fridman @vala.wtf · 08/12/2025one very fun thing i discovered in graphics programming is that opengl es only requires support for 16-bit floats. not only that, but built-in functions for e.g. normalizing a vector will square their operands at 16-bit precision, which overflows hilariously easily 120