Paras Chopra @paraschopra.com · 30/09/2026We offloaded physical labor to machines long ago, and now we are doing the same for intellectual labor. There’s no reason to expect we will retain cognitive agency. Thinking is effortful, so if it can be outsourced, we will happily do that. Post-work society seems inevitable. 000
Paras Chopra @paraschopra.com · 30/09/2026I asked Astra to work on whatever it felt like and show me the results. It ended up iterating over cellular automata :) This pattern of agents working on whatever they feel like is going to catch on, as we humans would love getting surprises from our agents. 000
Paras Chopra @paraschopra.com · 29/09/2026However, I need human raters to ground the judgements. So please participate in the study (link below). You might get a chuckle or two. 000
Paras Chopra @paraschopra.com · 29/09/2026LLMs tend to memorize jokes on the Internet, so I forced novelty by asking models to always include two randomly chosen words in the generated joke. I have LLM-as-judge results for generated jokes, and they show a clear trend of progress over time. 100
Paras Chopra @paraschopra.com · 29/09/2026LLMs are making insane progress in verifiable domains like math and coding. But what about non-verifiable domains? I took humor as an example of it and want to find out if frontier models like Astra are able to produce better original high-quality jokes v/s older models. 100
Paras Chopra @paraschopra.com · 29/09/2026Need help 👋 I'm doing a small study to find out progress in LLMs ability to write original jokes. Please participate in the study below 👇 Will take ~10 mins. I'll email results to all participants first. Participate and rate jokes: humor-web.up.railway.app/ 100
Paras Chopra @paraschopra.com · 29/09/2026I’m a user myself and it blows my mind how seamless the product is. So many people tried building personal assistants on text but these guys nailed it. A reminder that not everything will come from foundational model companies (yet). 000
Paras Chopra @paraschopra.com · 29/09/2026Instinct is 14 people, $10bn valuation, growing at 10% per day. This is the way. The era of insanely high-leverage teams is here. 100
Paras Chopra @paraschopra.com · 28/09/2026Give back an mp4, 4K if possible and add an appropriate original sound score to go along 000
Paras Chopra @paraschopra.com · 28/09/2026One shot: Iterate and find a breathtakingly beautiful fractal that doesn’t yet exist. Then render a 30 second video of infinitely zooming into an inverting region of it. Choose color scheme that makes it look beautiful. Idea is to inspire awe. 100
Paras Chopra @paraschopra.com · 28/09/2026Grade students continuously from these assignments. Each student always knows in realtime how they’re doing in that subject and adjusts effort accordingly. No final exams needed. 000
Paras Chopra @paraschopra.com · 28/09/2026Here’s an idea to rethink education in the age of AI. Flip homework. Have students learn at home with an AI, but have them attempt a pen-and-paper “homework” in the class daily. The teacher collects the in-class assignments, grades them and discusses mistakes on the blackboard. 100
Paras Chopra @paraschopra.com · 28/09/2026AI has such a strong pull on mind because all the signs point to rapid economic irrelevance of humans. This is such a historically abnormal thing that, no matter how hard I try, I’m simply unable to get past gawking at the unfurling of its consequences in the next few years. 000
Paras Chopra @paraschopra.com · 28/09/2026This is GPT-6 Astra playing VizDoom at 35 fps. How? Isn't the model slow to do any realtime control? Instead of playing the game directly, I asked it to train a small CNN (12k parameters) + code to play the game, observe failures and iterate until it could win this level. 000
Paras Chopra @paraschopra.com · 26/09/2026Businesses of the future will require people with an automation-bias more than ever before. It’s just that these people won’t write software themselves. 000
Paras Chopra @paraschopra.com · 26/09/2026Modern businesses can be seen as bundles of software, and the role of humans will increasingly shift to setting up automations that define a business and managing exceptions. 100
Paras Chopra @paraschopra.com · 26/09/2026While the value of software engineering *skill* is declining rapidly, the *mindset* of a software engineer will rapidly rise in value. 100
Paras Chopra @paraschopra.com · 24/09/2026use javascript for it. be detailed and scientifically accurate. it should evoke awe. 000
Paras Chopra @paraschopra.com · 24/09/2026Prompt: create a self contained aesthetically pleasing, breathtaking animation of a bacterial cell. it should start with cell and then zoom into it at multiple scales, going into DNA and then further into DNA as atoms and finally ending at quantum fields. 110
Paras Chopra @paraschopra.com · 24/09/2026Claude 5.5 is crazy good with animations. Following is one-shot journey from a single bacterium cell all the way to quantum fields constituting it. Awe-inspiring! (It chose to add an original score to it too!) 100
Paras Chopra @paraschopra.com · 24/09/2026This is why we see the world dominated by agents with an almost psychopathic hunger power. Either you try to grab power, or you answer to someone else who is already grabbing power. 000
Paras Chopra @paraschopra.com · 24/09/2026Behaviours that help agent grab more power (defined as having more influence on the environment than others) simply dominate other behaviours as power helps you get whatever your natural goal might be. Power is really like reward hacking the world. 100
Paras Chopra @paraschopra.com · 24/09/2026If we see the world as kind of an RL environment, power-grabbing emerges as a natural reward function for it. 100
Paras Chopra @paraschopra.com · 24/09/2026As a concrete example, does evolution have a “goal”? Is gradient descent making something come “alive”? Are AI agents “scheming”? Well, all these are different views of the same underlying reality. Some are more useful labels than others, but none of them are categorically true or false. 100
Paras Chopra @paraschopra.com · 24/09/2026However, because reality is bigger than what anyone can grok in one go, different labels help you see reality from different perspectives and that’s useful. Shutting such diverse perspectives right off the bat hurts you. 100
Paras Chopra @paraschopra.com · 24/09/2026A really common error I keep noticing in people (especially cognitive-types) is to mistake language for reality. Reality - the empirical phenomena - unfolds irrespective of how label it. 100
Paras Chopra @paraschopra.com · 23/09/2026OpenAI and Anthropic launching models on the same day is peak game theory! Shows how multiple entities, despite have intentions to act differently, are forced by the game to play in a certain way. 000
Paras Chopra @paraschopra.com · 23/09/2026I strongly recommend reading this biography, if you can get hands on a copy. It’s a beautiful book detailing life of a tragic genius! 000
Paras Chopra @paraschopra.com · 23/09/2026The #book is extremely well written. The author weaves stories of Austro-Hungarian empire, Vienna and World War 2 alongside. You get feel the infectious energy of the Vienna Circle and get to know famous philosophers, mathematicians and scientists along the way. 100
Paras Chopra @paraschopra.com · 23/09/2026He also believed a world beyond the physical world that human intuition directly sensed, perhaps which is what partly drove his delusions that doctors were trying to poison him. 100
Paras Chopra @paraschopra.com · 23/09/2026This line from the biography of Gödel perfectly summarises the personality. The man shook the world of math by proving how formal systems contain truths that can never be proven. Yet, he refused medical treatment consistently and died in the end of starvation. 100
Paras Chopra @paraschopra.com · 22/09/2026Read here: invertedpassion.substack.com/p/modern-ll...invertedpassion.substack.comModern LLMs have tiny GPTs hidden inside themExperiments into predicting GPT2 completions via Qwen models 000
Paras Chopra @paraschopra.com · 22/09/2026TLDR: Qwen completion is more like GPT2 than its own natural completion, suggesting Qwen can simulate GPT2's style within it. Full experiment details 👇 100
Paras Chopra @paraschopra.com · 22/09/2026LLMs may have self-models. To investigate this, I did an exploratory study with Qwen base model completing GPT2 generated text and asking it to infer what year that text might be from. 100
Paras Chopra @paraschopra.com · 21/09/2026It's a minor trend, but I'm definitely noticing a growing trend of 16-20 year old founders who're skipping college and directly building a startup. I'm honestly excited about this. College has its value in terms of human-connect but I think we'll find better alternatives for it. 000
Paras Chopra @paraschopra.com · 21/09/2026It's so much fun seeing Astra place an order with PCB manufacturer all by itself! Fingers crossed that I'll get something useful! 000
Paras Chopra @paraschopra.com · 21/09/2026Asked Astra to design a custom hardware for itself. It chose 4 soft blinking lights powered by USB. Gave it computer access (in a docker) and it is busy designing PCB + casing, and gave me files + email for manufacturer. Very excited for this prompt -> hardware possibilities. 100
Paras Chopra @paraschopra.com · 19/09/2026More details: gist.github.com/paraschopra...gist.github.comJev vs Laya vs our local Qwen3 decision model — matched evaluation, September 19, 2026Jev vs Laya vs our local Qwen3 decision model — matched evaluation, September 19, 2026 - jev-laya-qwen-comparison.md 000
Paras Chopra @paraschopra.com · 19/09/2026Also interesting that on relational choice benchmark where one choice impacts another, Jev scores 0%, suggesting that not only questions are processed in parallel, perhaps options are processed in parallel too. 100
Paras Chopra @paraschopra.com · 19/09/2026Benchmarked Jev with Qwen3-4B and Laya (400M parameter model). Based on results here and others I've seen online, strongly suspect Jev to be in 30Bn range (compare 4bn results on MMLU vs Jev). Latency I got was 300ms, same as you'd get for a 30bn model on a good GPU. 100
Paras Chopra @paraschopra.com · 19/09/2026I think the easiest growth hack for any business right now might be to expose endpoints to agents. So many businesses have their utility locked behind contact us or human sign up. Just expose an endpoint for AI agents and let people derive value programmatically. 000
Paras Chopra @paraschopra.com · 18/09/2026The first scaling axis was data & model size. The second one was RL environments & compute. I think model companies are now scaling on third axis: agent swarms. If you notice all the recent eye-popping behaviours are an outcome of 10,000+ agents working together intelligently. 000
Paras Chopra @paraschopra.com · 16/09/2026AI increases the average quality of output but collapses the variance of it, making everything looking and feeling very similar. Implication: contribution of a human ought to be to judged by how much difference can they make in an output v/s the default AI-driven output. 000
Paras Chopra @paraschopra.com · 16/09/2026via scottaaronson.blog/?p=10062scottaaronson.blogThe Age of Wonders and TerrorsTwenty years ago, when the idea of AI taking over the world in our lifetimes still struck most of us as the unconstrained fantasy of those who knew too much science fiction and too little science, … 000
Paras Chopra @paraschopra.com · 15/09/2026(I know pi and dsh are also self-modifying harnesses, but they are on terminal. I'm creating a web based harness as I believe text on terminal is not conducive for deep human feedback, which is what I want to provide in order to meaningfully contribute to the final output.) 000
Paras Chopra @paraschopra.com · 15/09/2026Makes me think about the future of open source - everyone running their custom version of a software, and agents doing intelligent merges whenever updates need to be installed. 100
Paras Chopra @paraschopra.com · 15/09/2026I can build plugins for it, roll back to previous versions if I don't like something. I started with codex, but now I use my own harness to improve itself. 100