Sign in

Shadab Choudhury

@namer.bsky.social
415 followers 464 following 315 posts

(He/Him). PhD at LARA Lab @UMBC, previously at Mohsin Lab @BRACU | Accessibility, Explainability and Multimodal DL. My opinions are mine. www.shadabchy.com

PostsRepliesMedia
Shadab Choudhury @namer.bsky.social · 03/10/2026
Sauce: arxiv.org/abs/2609.33150
arxiv.org
Generalization Dynamics of LM Pre-training
People typically assume that LMs stably mature from pattern-matching parrots to generalizable intelligence during pre-training. We build a toy eval suite and show this mental model is wrong: throughou...
000
Shadab Choudhury @namer.bsky.social · 03/10/2026
All papers should start like this.
A screenshot of a paper showing just the intro and the first line: "Building general AI without generalization is doable but meh."
100
Shadab Choudhury @namer.bsky.social · 25/09/2026
I think Jev is incredible. It's a lightning rod for people who jump on bandwagons and pump out slop, so you can just go on arXiv or on ICLR's openreview page and look for Jev papers and just blacklist every author you see on there.
000
Shadab Choudhury @namer.bsky.social · 25/09/2026
It's actually incredible how this is the single most facebook thing they could've possibly done and then they actually went and did it.
021
Shadab Choudhury @namer.bsky.social · 29/08/2026
Also the second order effect of people just doing data annotation and sourcing for him for free.
001
Shadab Choudhury @namer.bsky.social · 29/08/2026
mog them back by tokenizing 'tokens'
000
Shadab Choudhury @namer.bsky.social · 29/08/2026
AI overviews displacing first-party publishers was pretty well known and has been the source of like half of the lawsuits, so it's nice to have research helping them. Though, re; the SE analogy, Coding agents became popular from late 2024 onwards (Devin was March '24). SE fell off long before that.
060
Shadab Choudhury @namer.bsky.social · 23/07/2026
One of the main comorbidities I've noticed is the sheer number of very simple questions/basic help posts about EMNLP. It's fine to ask questions. I've been answering a bunch. But it shows just how many people submitted with minimal prior ACL/ARR experience on either their or their advisors' parts.
A list of 18 reddit posts posts under the result of searching for "EMNLP" in the r/LanguageTechnology subreddit.
110
Shadab Choudhury @namer.bsky.social · 19/07/2026
They call me 007 0 shots on goal 0 shots on target 7 yellow cards
010
Shadab Choudhury @namer.bsky.social · 07/07/2026
crazy how Egypt is still ahead despite playing 11 vs 13 genuinely filthy review from probably the coolest goal in this world cup so far
000
Shadab Choudhury @namer.bsky.social · 28/06/2026
Someone in management decided it takes a thief to catch a thief lol But yeah, human annotation/eval is a messss, since it's not like anyone's evaluating the evaluators.
120
Shadab Choudhury @namer.bsky.social · 22/06/2026
Yeah, and the amount of whining that relatively low bar has been creating on the other site, LinkedIn and Reddit has been eye-opening. And in 90% of cases, most of those 'papers' are just little experiments worth a blog post, stretched to a paper by an LLM. ngl, I appreciate the gate more.
000
Shadab Choudhury @namer.bsky.social · 22/06/2026
Any idea if CHI planning to go ahead with the rubic-based desk reject that was part of the restructuring proposal earlier this year? Would be nice to see more experiments beyond "let's shuffle around how many Chairs and levels we have and ask people nicely to volunteer more."
120
Shadab Choudhury @namer.bsky.social · 07/06/2026
As opposed to an LLM, which will definitely "have time" to read them~! Maybe the website's presuming much of your abilities-
000
Shadab Choudhury @namer.bsky.social · 01/06/2026
Must've been on the wrong tokenizer.
000
Shadab Choudhury @namer.bsky.social · 29/05/2026
To be fair, Gabriel Knight: Sins of the Fathers itself already sounds like a steamy M4M romance ebook lmao
000
Shadab Choudhury @namer.bsky.social · 19/05/2026
lamp heads are load bearing here
000
Shadab Choudhury @namer.bsky.social · 29/04/2026
I'm open to reviewing it! Your DMs aren't open, but you can DM or email me instead if you want to send me the details and paper.
000
Shadab Choudhury @namer.bsky.social · 21/04/2026
It's almost twice as old now as the AlexNet paper was when it initially came out, to put things into perspective.
000
Shadab Choudhury @namer.bsky.social · 19/04/2026
surely a superintelligence can make itself equally bad at both bad cyber and 'being-bad-at-good-cyber'
000
Shadab Choudhury @namer.bsky.social · 19/04/2026
I think a good way to go about it is to show how Claude *will* give them a response no matter how good the current code is. It doesn't have an 'endpoint'. If you keep telling it to improve/review some code, it will potentially just endlessly offer new suggestions and feedback.
110
Shadab Choudhury @namer.bsky.social · 18/04/2026
Honestly, if they don't shutter the game entirely four weeks from now they'll be ahead of the curve.
000
Shadab Choudhury @namer.bsky.social · 04/03/2026
My position paper "The Perceptual Gap: Why We Need Accessible XAI for Assistive Technologies" has been conditionally accepted as a Poster in CHI '26! (arXiv: arxiv.org/abs/2603.024...) tl;dr: Folks with sensory disabilities need XAI, but XAI for the models used in assistive tech aren't accessible.
000
Shadab Choudhury @namer.bsky.social · 04/03/2026
(images from www.smithsonianmag.com/history/what...)
smithsonianmag.com
What the Luddites Really Fought Against
The label now has many meanings, but when the group protested 200 years ago, technology wasn't really the enemy
000
Shadab Choudhury @namer.bsky.social · 04/03/2026
Nah. Misusing the term "luddite" deserves to be dunked on as much as using "stochastic parrots". They're two sides of the same derogatory coin that drags down any discussions about the real concerns.
200
Shadab Choudhury @namer.bsky.social · 10/02/2026
That's not remotely the issue here though. LLMs can already generate stories/code/designs. World models aren't the same thing as persistent memory. It can't generate *good* stories for reasons of verifiability an un-RL-ability, as said above, and world models won't change that at all.
000
Shadab Choudhury @namer.bsky.social · 06/02/2026
The neatest thing is just how much it looks like mold growing on the fruits. A nicely picked example.
000
Shadab Choudhury @namer.bsky.social · 22/01/2026
Like, if you don't know how to I'm happy to show you. It's not hard or something and it just speeds up your workflow. Copying and pasting each of your citations into ChatGPT is just silly. (img source: the other site)
A twitter post:

Kyunghyun Cho
@kchonyc
i was made aware of miscitations thanks to the GPTZero team (cc 
@alexcdot
). ji won and i quickly checked them ourselves and have posted what happened on openreview: https://openreview.net/forum?id=IiEtQPGVyV&noteId=W66rrM5XPk. we have already notified NeurIPS'25 PC's about this issue. 

i truly thank the GPTZero team for bringing this to our attention as well as raising the awareness of this serious issue (https://gptzero.me/news/neurips/), and at the same time i sincerely apologize to all for our error.

There is an image inside the post:

We identify the causes of miscitations in this work and provide fixes for them. These miscitations were identified and reported by the GPTZero team. The full report can be found at

Nazar Shmatko, Alex Adam and Paul Esau. GPTZero finds 100 new hallucinations in NeurIPS 2025 accepted papers. January 2026. https://gptzero.me/news/neurips/

We (Park & Cho) used a large-scale language model (LLM; specifically ChatGPT) to generate the citations after giving it author-year in-text citations, titles, or their paraphrases. Most, if not all, of these miscitations, either hallucinated, typoed or misattributed, are incorrect or non-existent bibtex entries fetched by the LLM. We categorize and analyze these citations below.

We are submitting an updated version to arXiv (https://arxiv.org/abs/2510.21310) and have already uploaded the updated version at https://tinyurl.com/ymb99d7s. We will shortly reach out to the program chairs of NeurIPS'25 as well.

We sincerely apologize to the whole community, especially the authors affected, for this grave mistake and thank the GPTZero team for bringing this issue to our attention.

(1) Prior work directly relevant to the paper
Below are miscitations for papers. Methods in these papers were implemented, or considered for implementation, in our paper as baseline uncertainty quantification methods, baseline sampling methods, or as the base LMs.
000
Shadab Choudhury @namer.bsky.social · 22/01/2026
Why on earth would you even do this in the first place? Pasting the details into ChatGPT, then asking it to generate the citation is a way bigger hassle than simply having Scholar + Zotero (or another bib manager) extensions set up correctly to grab metadata and generate citations.
130
Shadab Choudhury @namer.bsky.social · 08/01/2026
Lowkey, I think this is intentional, and it seems like the kind of thing I would do if I was making Claude more attractive to students or other people who don't want their code to be easily recognized as LLM-generated.
000
Shadab Choudhury @namer.bsky.social · 04/01/2026
It was great listening to Dr's Ishtiaque, Ferdous, and Sultana talk about their experiences! There's a *lot* of HCI work to be done in the scope of Bangladesh, and IMO not enough folks working on them, so HCCS should be a great addition. Looking forward to the work soon to come out of the lab.
Poster titled "Launching HCCS: Inauguration and research dialogue with the HCI Pioneers of Bangladesh". On the bottom left, there are 3 portraits, named Dr Syed Ishtiaque Ahmed, Dr. Hasan Shahid Ferdous and Dr. Sharifa SultanaA stage with 4 people sitting on chairs in the middle. A poster in the background shows the previous image's poster, as well as a QR code to the panel discussion question submission.
010
Shadab Choudhury @namer.bsky.social · 02/01/2026
lowkey I appreciate folks aren't posting linkedinisms like "In 2025 I achieved X, Y, Z" this time around. It's fine to celebrate your wins, but I imagine for most people, 2025 was... A Year. ...and I think that's all that needs to be said.
010
Shadab Choudhury @namer.bsky.social · 01/01/2026
Nahhh my high school friends would've also found that name funny as fuck 10 years ago
060
Shadab Choudhury @namer.bsky.social · 25/12/2025
Seen on LinkedIn. How does someone raise $5 million and then write job ads like this? smh...
010
Shadab Choudhury @namer.bsky.social · 21/12/2025
I don't think "money" is the simple answer, since every frontier lab is a black hole of money rn, and gene editing could've also been ludicrously profitable. Was it the political climate? The everyday accessibility of GenAI? Lower levels of scruples in the community (not to accuse anyone directly)?
000
Shadab Choudhury @namer.bsky.social · 21/12/2025
In the mid 2010s biology folks figured out how to do human gene editing. The community took one look, realized the consequences would be so dire for humanity, and put a hard stop to it. People like Jiankui He were excoriated for illegally editing embryos. Why wasn't this the case with GenAI?
132
Shadab Choudhury @namer.bsky.social · 17/12/2025
And it's not ML, it's "GenAI" that invokes certain concerns. People don't mind when ML's used to detect cancer or study whale speech. Those models aren't trained on human inputs and used to take human jobs. GenAI specifically, however, is trained on human inputs and used to take human jobs.
000
Shadab Choudhury @namer.bsky.social · 17/12/2025
Nah. That article's 8mo old and says they're "not there yet". Vincke's recent comment, however, explicitly mentions flesh out PowerPoint presentations, develop concept art" and these tasks are explicitly creative jobs lost.
100
Shadab Choudhury @namer.bsky.social · 16/12/2025
by perchance is it specifically like these 5 companies?
110
Shadab Choudhury @namer.bsky.social · 04/12/2025
I hope the Peer Reviewer Recognition Policy actually puts the <$30~ per paper reviewed (based on 3 reviewers per $100 paper minus waived papers) to good use.
000
Shadab Choudhury @namer.bsky.social · 04/12/2025
How was no one talking about this?* IJCAI-ECAI 2026 @ijcai.org levying a $100 fee per submission unless every author on the paper is only on that one submitted paper. * rhetorical question. I assume the ICLR drama drowned it
Primary Paper Initiative: IJCAI-ECAI 2026 is launching the Primary Paper Initiative in response to the international AI research community’s call to address challenges and to revitalize the peer review process, while strengthening the reviewers and authors in the process. Under the IJCAI-ECAI 2026 Primary Paper Initiative, every submission is subject to a fee of USD 100. That paper submission fee is waived for primary papers, i.e., papers for which none of the authors appear as an author on any other submission to IJCAI-ECAI 2026. The initiative applies to the main track, Survey Track, and all special tracks, excluding the Journal Track, the Sister Conferences Track, Early Career Highlights, Competitions, Demos, and the Doctoral Consortium. All proceeds generated from the Primary Paper Initiative will be exclusively directed toward the support of the reviewing community of IJCAI-ECAI 2026. To recognize the reviewers’ contributions, the initiative introduces Peer Reviewer Recognition Policy with clearly defined standards (which will be published on the conference web site). The initiative aims to enhance review quality, strengthen accountability, and uphold the scientific excellence of the conference. Details and the FAQ will be published on the IJCAI-ECAI 2026 website.
110
Shadab Choudhury @namer.bsky.social · 29/11/2025
Basically, if your Altmetric score is higher than your Accesses, you just got ratioed.
000
Shadab Choudhury @namer.bsky.social · 27/11/2025
Case in point. Good grief.
A twitter post by Lorenzo Xiao with the text: "this shit is real...". An image is attached, and the following text is from the image, translated from Chinese to English: "got an anonymous death threat, need to cancel my trip to NeurIPS. Looking for someone who will take my hotel reservation."
010
Shadab Choudhury @namer.bsky.social · 27/11/2025
Took me about 5 minutes to dig out the identity of the 40 questions reviewer — after someone posted it on the other site; I dunno how to search Xiaohongshu directly. Honestly, I don't think western academics are going to feel a fraction of the shitstorm that the Chinese ML community's probably in.
010
Shadab Choudhury @namer.bsky.social · 20/11/2025
This is the site I unfortunately have to be professional on 😭
110
Shadab Choudhury @namer.bsky.social · 20/11/2025
many such cases Simpsons-tier prescience
001
Shadab Choudhury @namer.bsky.social · 16/11/2025
I understand the sentiment behind this, but I'm just extremely bearish on rankings that do fancy mathematical tricks or use black-box algorithms because the more of that you do, the more you can bias it towards a specific outcome. The closer the metric is to the raw data instead, the better.
010
Shadab Choudhury @namer.bsky.social · 16/11/2025
Two issues: first, this is not transparent. There's NO way to tell *which* papers were counted, nor how 'most important papers to this paper' is computed. The papers contributing to the ranking should be listed. Second, CSRankings is CC BY-NC-ND 4.0 so I'm pretty sure this is copyright infringement
220
Shadab Choudhury @namer.bsky.social · 15/11/2025
I can't believe I spent the evening skimming through every Visual Reasoning paper at ICLR instead of finishing my SoP.
000
Shadab Choudhury @namer.bsky.social · 15/11/2025
It may be in bad taste to call out a reviewer like this, but I don't believe anyone who gives weaknesses like this, as if it isn't empirical common sense for anyone working with MLLMs that larger models give better outcomes when inference speed isn't relevant, is acting in good faith.
120