Janu Verma @januverma.bsky.social · 29/01/2026Been thinking about the trends at the intersection of AI and RecSys. Where are we heading. Based on my own work, extensive research, and lots of analysis/thinking, I have put together my thoughts in a detailed article on substack. Link: open.substack.com/pub/januverm... 000
Janu Verma @januverma.bsky.social · 05/01/2026There is a specific kind of friction that hits on January 1st. I call it "Cold Shyness." It’s that reluctance to be out there—whether physically in the cold or metaphorically in a new skill—exposed and uncomfortable. 110
Janu Verma @januverma.bsky.social · 04/11/2025New work: Multi-turn tool use using RL Link: open.substack.com/pub/januverm...open.substack.comMulti-Turn Tool Use with RLThink → Code → Check → Answer 000
Janu Verma @januverma.bsky.social · 11/07/2025Add. Related. Context. Often, the most significant performance gains come from enriching models with related, contextual info. Models get better by being exposed to auxiliary signals that deepen their understanding of the task. 100
Janu Verma @januverma.bsky.social · 09/07/2025My latest blog post dives into the protein folding problem - a fundamental question in molecular biology that puzzled scientists for decades, until deep learning models like AlphaFold changed the game. I walk through the biological and computational roots of the problem. 120
Janu Verma @januverma.bsky.social · 04/02/2025As a personal research project, I’m exploring the efficacy of LLMs for Recommendation System tasks. Check out my experments at januverma.substack.comjanuverma.substack.comIncomplete Distillation | Janu Verma | Substackpersonal research journal containing articles based on my explorations with cutting edge AI. Click to read Incomplete Distillation, by Janu Verma, a Substack publication. Launched 11 days ago. 100
Janu Verma @januverma.bsky.social · 24/01/2025Recently, I’ve been exploring the potential of LLMs for recommendation tasks. Sharing the first report of my project where I experiment with the ability of Llama 1B model to understand user preferences from their past behavior. open.substack.com/pub/januverm...open.substack.comLarge Language Models for Recommender SystemsCan LLMs reason over user behaviour data to decipher preferences? 000
Janu Verma @januverma.bsky.social · 16/01/2025Have we swapped “reasoning” for “agentic” as the new shibboleth 010
Janu Verma @januverma.bsky.social · 10/01/2025Just came back after a month in India, no-laptop family time. Any tips on how to motivate myself to do any work are highly appreciated 🙏 110
Reposted by Janu VermaJohanna Franklin @johannamath.bsky.social · 08/12/2024The queen of examples and counterexamples! 1449
Reposted by Janu VermaThomas Wolf @thomwolf.bsky.social · 08/12/2024The FineWeb team is happy to finally release "FineWeb2" 🥂🥳 FineWeb 2 extends the data driven approach to pre-training dataset design that was introduced in FineWeb 1 to now covers 1893 languages/scripts Details: huggingface.co/datasets/Hug... A detailed open-science tech report is coming soon 310613
Janu Verma @januverma.bsky.social · 04/12/2024Nothing like waking up to see your models training in a nice way. #neuralnets 000
Reposted by Janu VermaVicki @vickiboykis.com · 03/12/2024This seems like… what we started with, no? arxiv.org/abs/2410.02724arxiv.orgLarge Language Models as Markov ChainsLarge language models (LLMs) have proven to be remarkably efficient, both across a wide range of natural language processing tasks and well beyond them. However, a comprehensive theoretical analysis o... 161659
Reposted by Janu VermaYoav Artzi @yoavartzi.com · 02/12/2024I am seriously behind uploading Learning Machines videos, but I did want to get @jonathanberant.bsky.social's out sooner than later. It's not only a great talk, it also gives a remarkably broad overview and contextualization, so it's an excellent way to ramp up on post-training youtu.be/2AthqCX3h8Uyoutu.beJonathan Berant (Tel Aviv University / Google) / Towards Robust Language Model Post-trainingYouTube video by Yoav Artzi 15312
Reposted by Janu VermaAlexander Doria @dorialexander.bsky.social · 29/11/2024Won’t help with my reputation but since I worked on social network analysis/regulation: if Bluesky ever is a success, they are extremely likely to retrain AI models (not necessarily LLM) on user data. 66019
Janu Verma @januverma.bsky.social · 30/11/2024If you are trying to fine-tune (instruction sft) a LLM for a specific task, how much should you work on refining the prompt? The general alpaca format seem suboptimal and far from how we use these models. 110