Shadab Choudhury @namer.bsky.social · 03/10/2026Sauce: arxiv.org/abs/2609.33150arxiv.orgGeneralization Dynamics of LM Pre-trainingPeople typically assume that LMs stably mature from pattern-matching parrots to generalizable intelligence during pre-training. We build a toy eval suite and show this mental model is wrong: throughou... 000
Shadab Choudhury @namer.bsky.social · 25/09/2026I think Jev is incredible. It's a lightning rod for people who jump on bandwagons and pump out slop, so you can just go on arXiv or on ICLR's openreview page and look for Jev papers and just blacklist every author you see on there. 000
Shadab Choudhury @namer.bsky.social · 25/09/2026It's actually incredible how this is the single most facebook thing they could've possibly done and then they actually went and did it. 021
Shadab Choudhury @namer.bsky.social · 29/08/2026Also the second order effect of people just doing data annotation and sourcing for him for free. 001
Shadab Choudhury @namer.bsky.social · 29/08/2026AI overviews displacing first-party publishers was pretty well known and has been the source of like half of the lawsuits, so it's nice to have research helping them. Though, re; the SE analogy, Coding agents became popular from late 2024 onwards (Devin was March '24). SE fell off long before that. 060
Shadab Choudhury @namer.bsky.social · 23/07/2026One of the main comorbidities I've noticed is the sheer number of very simple questions/basic help posts about EMNLP. It's fine to ask questions. I've been answering a bunch. But it shows just how many people submitted with minimal prior ACL/ARR experience on either their or their advisors' parts. 110
Shadab Choudhury @namer.bsky.social · 19/07/2026They call me 007 0 shots on goal 0 shots on target 7 yellow cards 010
Shadab Choudhury @namer.bsky.social · 07/07/2026crazy how Egypt is still ahead despite playing 11 vs 13 genuinely filthy review from probably the coolest goal in this world cup so far 000
Shadab Choudhury @namer.bsky.social · 28/06/2026Someone in management decided it takes a thief to catch a thief lol But yeah, human annotation/eval is a messss, since it's not like anyone's evaluating the evaluators. 120
Shadab Choudhury @namer.bsky.social · 22/06/2026Yeah, and the amount of whining that relatively low bar has been creating on the other site, LinkedIn and Reddit has been eye-opening. And in 90% of cases, most of those 'papers' are just little experiments worth a blog post, stretched to a paper by an LLM. ngl, I appreciate the gate more. 000
Shadab Choudhury @namer.bsky.social · 22/06/2026Any idea if CHI planning to go ahead with the rubic-based desk reject that was part of the restructuring proposal earlier this year? Would be nice to see more experiments beyond "let's shuffle around how many Chairs and levels we have and ask people nicely to volunteer more." 120
Shadab Choudhury @namer.bsky.social · 07/06/2026As opposed to an LLM, which will definitely "have time" to read them~! Maybe the website's presuming much of your abilities- 000
Shadab Choudhury @namer.bsky.social · 29/05/2026To be fair, Gabriel Knight: Sins of the Fathers itself already sounds like a steamy M4M romance ebook lmao 000
Shadab Choudhury @namer.bsky.social · 29/04/2026I'm open to reviewing it! Your DMs aren't open, but you can DM or email me instead if you want to send me the details and paper. 000
Shadab Choudhury @namer.bsky.social · 21/04/2026It's almost twice as old now as the AlexNet paper was when it initially came out, to put things into perspective. 000
Shadab Choudhury @namer.bsky.social · 19/04/2026surely a superintelligence can make itself equally bad at both bad cyber and 'being-bad-at-good-cyber' 000
Shadab Choudhury @namer.bsky.social · 19/04/2026I think a good way to go about it is to show how Claude *will* give them a response no matter how good the current code is. It doesn't have an 'endpoint'. If you keep telling it to improve/review some code, it will potentially just endlessly offer new suggestions and feedback. 110
Shadab Choudhury @namer.bsky.social · 18/04/2026Honestly, if they don't shutter the game entirely four weeks from now they'll be ahead of the curve. 000
Shadab Choudhury @namer.bsky.social · 04/03/2026My position paper "The Perceptual Gap: Why We Need Accessible XAI for Assistive Technologies" has been conditionally accepted as a Poster in CHI '26! (arXiv: arxiv.org/abs/2603.024...) tl;dr: Folks with sensory disabilities need XAI, but XAI for the models used in assistive tech aren't accessible. 000
Shadab Choudhury @namer.bsky.social · 04/03/2026(images from www.smithsonianmag.com/history/what...)smithsonianmag.comWhat the Luddites Really Fought AgainstThe label now has many meanings, but when the group protested 200 years ago, technology wasn't really the enemy 000
Shadab Choudhury @namer.bsky.social · 04/03/2026Nah. Misusing the term "luddite" deserves to be dunked on as much as using "stochastic parrots". They're two sides of the same derogatory coin that drags down any discussions about the real concerns. 200
Shadab Choudhury @namer.bsky.social · 10/02/2026That's not remotely the issue here though. LLMs can already generate stories/code/designs. World models aren't the same thing as persistent memory. It can't generate *good* stories for reasons of verifiability an un-RL-ability, as said above, and world models won't change that at all. 000
Shadab Choudhury @namer.bsky.social · 06/02/2026The neatest thing is just how much it looks like mold growing on the fruits. A nicely picked example. 000
Shadab Choudhury @namer.bsky.social · 22/01/2026Like, if you don't know how to I'm happy to show you. It's not hard or something and it just speeds up your workflow. Copying and pasting each of your citations into ChatGPT is just silly. (img source: the other site) 000
Shadab Choudhury @namer.bsky.social · 22/01/2026Why on earth would you even do this in the first place? Pasting the details into ChatGPT, then asking it to generate the citation is a way bigger hassle than simply having Scholar + Zotero (or another bib manager) extensions set up correctly to grab metadata and generate citations. 130
Shadab Choudhury @namer.bsky.social · 08/01/2026Lowkey, I think this is intentional, and it seems like the kind of thing I would do if I was making Claude more attractive to students or other people who don't want their code to be easily recognized as LLM-generated. 000
Shadab Choudhury @namer.bsky.social · 04/01/2026It was great listening to Dr's Ishtiaque, Ferdous, and Sultana talk about their experiences! There's a *lot* of HCI work to be done in the scope of Bangladesh, and IMO not enough folks working on them, so HCCS should be a great addition. Looking forward to the work soon to come out of the lab. 010
Shadab Choudhury @namer.bsky.social · 02/01/2026lowkey I appreciate folks aren't posting linkedinisms like "In 2025 I achieved X, Y, Z" this time around. It's fine to celebrate your wins, but I imagine for most people, 2025 was... A Year. ...and I think that's all that needs to be said. 010
Shadab Choudhury @namer.bsky.social · 01/01/2026Nahhh my high school friends would've also found that name funny as fuck 10 years ago 060
Shadab Choudhury @namer.bsky.social · 25/12/2025Seen on LinkedIn. How does someone raise $5 million and then write job ads like this? smh... 010
Shadab Choudhury @namer.bsky.social · 21/12/2025I don't think "money" is the simple answer, since every frontier lab is a black hole of money rn, and gene editing could've also been ludicrously profitable. Was it the political climate? The everyday accessibility of GenAI? Lower levels of scruples in the community (not to accuse anyone directly)? 000
Shadab Choudhury @namer.bsky.social · 21/12/2025In the mid 2010s biology folks figured out how to do human gene editing. The community took one look, realized the consequences would be so dire for humanity, and put a hard stop to it. People like Jiankui He were excoriated for illegally editing embryos. Why wasn't this the case with GenAI? 132
Shadab Choudhury @namer.bsky.social · 17/12/2025And it's not ML, it's "GenAI" that invokes certain concerns. People don't mind when ML's used to detect cancer or study whale speech. Those models aren't trained on human inputs and used to take human jobs. GenAI specifically, however, is trained on human inputs and used to take human jobs. 000
Shadab Choudhury @namer.bsky.social · 17/12/2025Nah. That article's 8mo old and says they're "not there yet". Vincke's recent comment, however, explicitly mentions flesh out PowerPoint presentations, develop concept art" and these tasks are explicitly creative jobs lost. 100
Shadab Choudhury @namer.bsky.social · 16/12/2025by perchance is it specifically like these 5 companies? 110
Shadab Choudhury @namer.bsky.social · 04/12/2025I hope the Peer Reviewer Recognition Policy actually puts the <$30~ per paper reviewed (based on 3 reviewers per $100 paper minus waived papers) to good use. 000
Shadab Choudhury @namer.bsky.social · 04/12/2025How was no one talking about this?* IJCAI-ECAI 2026 @ijcai.org levying a $100 fee per submission unless every author on the paper is only on that one submitted paper. * rhetorical question. I assume the ICLR drama drowned it 110
Shadab Choudhury @namer.bsky.social · 29/11/2025Basically, if your Altmetric score is higher than your Accesses, you just got ratioed. 000
Shadab Choudhury @namer.bsky.social · 27/11/2025Took me about 5 minutes to dig out the identity of the 40 questions reviewer — after someone posted it on the other site; I dunno how to search Xiaohongshu directly. Honestly, I don't think western academics are going to feel a fraction of the shitstorm that the Chinese ML community's probably in. 010
Shadab Choudhury @namer.bsky.social · 20/11/2025This is the site I unfortunately have to be professional on 😭 110
Shadab Choudhury @namer.bsky.social · 16/11/2025I understand the sentiment behind this, but I'm just extremely bearish on rankings that do fancy mathematical tricks or use black-box algorithms because the more of that you do, the more you can bias it towards a specific outcome. The closer the metric is to the raw data instead, the better. 010
Shadab Choudhury @namer.bsky.social · 16/11/2025Two issues: first, this is not transparent. There's NO way to tell *which* papers were counted, nor how 'most important papers to this paper' is computed. The papers contributing to the ranking should be listed. Second, CSRankings is CC BY-NC-ND 4.0 so I'm pretty sure this is copyright infringement 220
Shadab Choudhury @namer.bsky.social · 15/11/2025I can't believe I spent the evening skimming through every Visual Reasoning paper at ICLR instead of finishing my SoP. 000
Shadab Choudhury @namer.bsky.social · 15/11/2025It may be in bad taste to call out a reviewer like this, but I don't believe anyone who gives weaknesses like this, as if it isn't empirical common sense for anyone working with MLLMs that larger models give better outcomes when inference speed isn't relevant, is acting in good faith. 120