Sign in

Lj Miranda

@ljvmiranda921.itch.io
502 followers 98 following 27 posts

PhD student at the University of Cambridge ljvmiranda921.github.io

PostsRepliesMedia
Lj Miranda @ljvmiranda921.itch.io · 27/09/2026
Rediscovering my love for electronics, first step is trying to deploy a model on an RPi 5: ljvmiranda921.github.io/notebook/202...
ljvmiranda921.github.io
Edge LM Devlog: Running a language model on a Raspberry Pi 5
A collection of notes, projects, and essays.
111
Lj Miranda @ljvmiranda921.itch.io · 03/09/2026
it's cool that 'teacher' and 'student' have the same number of letters so they align in my code
000
Reposted by Lj Miranda
Tom McCoy @rtommccoy.bsky.social · 20/08/2026
Since many are starting grad school soon, let me re-share my One Big Tip™️ for research! Research involves many skills - collaborating, writing, presenting, etc. But many of these skills can be unified under a single overarching ability: theory of mind Blog post link in reply
Illustration of the blog post's main argument, summarized as: "Theory of Mind as a Central Skill for Researchers: Research involves many skills.If each skill is viewed separately, each one takes a long time to learn. These skills can instead be connected via theory of mind – the ability to reason about the mental states of others. This allows you to transfer your abilities across areas, making it easier to gain new skills."
28019
Lj Miranda @ljvmiranda921.itch.io · 03/08/2026
And a quick blog post on my experience developing this game with Claude from the perspective of a hobbyist game dev: ljvmiranda921.github.io/projects/202...
ljvmiranda921.github.io
Postscript: Idle Frontier and on developing games with AI
Idle Frontier is a clicker game that challenges you to build a frontier language model in 1000 days! In this blog post, I'll talk about some reflections on d...
000
Lj Miranda @ljvmiranda921.itch.io · 02/08/2026
Currently writing a blog post talking about my experience in using AI to augment my game development workflow...in the meantime, check out my other games on itch: ljvmiranda921.itch.io
ljvmiranda921.itch.io
Lj V. Miranda
020
Lj Miranda @ljvmiranda921.itch.io · 02/08/2026
Building this game has been quite helpful to brush up my Godot and Aseprite skills. AI Disclosure: Claude has been super helpful in migrating some Godot 3 patterns I learned 2 yrs ago to Godot 4. I also learned new design patterns to make better games!
100
Lj Miranda @ljvmiranda921.itch.io · 02/08/2026
Summer fun project! ljvmiranda921.itch.io/idle-frontier Idle Frontier is a one-click idle incremental game about the economics of training models. Annotate data, run experiments, chase grants and conference deadlines, and spend everything you earn on the next training run!
110
Lj Miranda @ljvmiranda921.itch.io · 06/07/2026
Yes, I still do ol' reliable NLP. Just released a maintenance version of calamanCy, a spaCy-based open source toolkit for Tagalog. Useful for dep. parsing, NER, tagging, etc. You can now install via `uv`! github.com/ljvmiranda92...
121
Reposted by Lj Miranda
Multilingual Representation Workshop @ EMNLP 2026 @mrl-workshop.bsky.social · 27/05/2026
📢 Call for Papers: 6th Multilingual Representation Learning Workshop at EMNLP in Budapest, Hungary! Join us and submit your works relating to multilingual NLP Speakers to be announced, so stay tuned! 👀 More info in the CFP: 🔗 sigtyp.github.io/ws2026-mrl.html
153
Lj Miranda @ljvmiranda921.itch.io · 20/08/2025
Finally, I want to thank the folks from HuggingFace for helping draft the official blog post (special shoutout to @clefourrier , @vanstriendaniel, @nathanhabib1011) and @Cohere_Labs for the research credits. :)
000
Lj Miranda @ljvmiranda921.itch.io · 20/08/2025
Evals are often the first step, we hope FilBench paves the way for language-specific adaptation especially for Philippine languages! I've written some of my thoughts here: ljvmiranda921.github.io/projects/20...
100
Lj Miranda @ljvmiranda921.itch.io · 20/08/2025
Here's the link to the paper and leaderboard: 📜 Paper: arxiv.org/abs/2508.03523 📊 Leaderboard: ud-filipino-filbench-leaderboard.hf.space/
100
Lj Miranda @ljvmiranda921.itch.io · 20/08/2025
This collaboration is exciting, it felt like assembling the Avengers of Filipino NLP. @acocodes and Conner are great collaborators, and I was happy to team-up with @jcblaisecruz and @josephimperial_, who are working on Filipino NLP for longer than I did!
100
Lj Miranda @ljvmiranda921.itch.io · 20/08/2025
🇵🇭 One of my research interests is improving the state of Filipino NLP Happy to share that we're taking a major step towards this by introducing FilBench, an LLM benchmark for Filipino! Also accepted at EMNLP Main! 🎉 Learn more: huggingface.co/blog/filbench
huggingface.co
🇵🇭 FilBench - Can LLMs Understand and Generate Filipino?
140
Lj Miranda @ljvmiranda921.itch.io · 02/08/2025
ljvmiranda921.github.io/notebook/202...
ljvmiranda921.github.io
Field Report: ACL 2025
A collection of notes, projects, and essays.
020
Reposted by Lj Miranda
Ai2 @ai2.bsky.social · 28/07/2025
Ai2 is excited to be at #ACL2025 in Vienna, Austria this week. Come say hello, meet the team, and chat about the future of NLP. See you there! 🤝📚
093
Lj Miranda @ljvmiranda921.itch.io · 24/07/2025
I was also part of a large-scale @seacrowd.bsky.social collaboration on building a vision-language dataset tailored for Southeast Asian Languages :) Also at ACL Main - aclanthology.org/2025.acl-lo... July 29 Hall 4/5 10:30-12:00 #ACL2025 #ACL2025NLP
aclanthology.org
Crowdsource, Crawl, or Generate? Creating SEA-VL, a Multicultural Vision-Language Dataset for Southeast Asia
Samuel Cahyawijaya, Holy Lovenia, Joel Ruben Antony Moniz, Tack Hwa Wong, Mohammad Rifqi Farhansyah, Thant Thiri Maung, Frederikus Hudi, David Anugraha, Muhammad Ravi Shulthan Habibi, Muhammad Reza Qorib, Amit Agarwal, Joseph Marvin Imperial, Hitesh Laxmichand Patel, Vicky Feliren, Bahrul Ilmi Nasution, Manuel Antonio Rufino, Genta Indra Winata, Rian Adam Rajagede, Carlos Rafael Catalan, Mohamed Fazli Mohamed Imam, Priyaranjan Pattnayak, Salsabila Zahirah Pranida, Kevin Pratama, Yeshil Bangera, Adisai Na-Thalang, Patricia Nicole Monderin, Yueqi Song, Christian Simon, Lynnette Hui Xian Ng, Richardy Lobo Sapan, Taki Hasan Rafi, Bin Wang, Supryadi, Kanyakorn Veerakanjana, Piyalitt Ittichaiwong, Matthew Theodore Roque, Karissa Vincentio, Takdanai Kreangphet, Phakphum Artkaew, Kadek Hendrawan Palgunadi, Yanzhi Yu, Rochana Prih Hastuti, William Nixon, Mithil Bangera, Adrian Xuan Wei Lim, Aye Hninn Khine, Hanif Muhammad Zhafran, Teddy Ferdinan, Audra Aurora Izzani, Ayushman Singh, Evan Evan,
020
Lj Miranda @ljvmiranda921.itch.io · 24/07/2025
3️⃣ The UD-NewsCrawl Treebank: Reflections and Challenges from a Large-scale Tagalog Syntactic Annotation Project (Main) - aclanthology.org/2025.acl-lo... July 29 Hall 4/5 10:30-12:00 Collab with folks from UP Diliman #ACL2025 #ACL2025NLP
100
Lj Miranda @ljvmiranda921.itch.io · 24/07/2025
2️⃣ M-RewardBench: Evaluating Reward Models in Multilingual Settings (Main) - aclanthology.org/2025.acl-lo... July 28 Hall 4/5 11:00-12:30 Collab with folks from @cohereforai.bsky.social #ACL2025 #ACL2025NLP
100
Lj Miranda @ljvmiranda921.itch.io · 24/07/2025
1️⃣ Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback (Main) - aclanthology.org/2025.acl-lo... 7/29 Hall 4/5 10:30-12:00 My project here at @ai2.bsky.social! #ACL2025NLP
100
Lj Miranda @ljvmiranda921.itch.io · 24/07/2025
I'll be at @aclmeeting.bsky.social‬ in Vienna! I'm going to present the ff first/co-first author works:
130
Lj Miranda @ljvmiranda921.itch.io · 20/07/2025
fun learning stuff (+ phew i haven't blogged in a long time!): ljvmiranda921.github.io/notebook/202...
ljvmiranda921.github.io
‘Draw me a swordsman’: Can tool-calling LLMs draw pixel art?
Just a fun weekend experiment on model-context protocol (MCP): I asked several tool-calling LLMs to draw a 4-frame spritesheet of a swordsman performing a sl...
010
Reposted by Lj Miranda
SEACrowd @seacrowd.bsky.social · 16/05/2025
We’re thrilled that SEA-VL has been accepted to the ACL 2025 (Main)! Thank you to everyone who contributed to this project 🥳 Paper: arxiv.org/abs/2503.07920 Project: seacrowd.github.io/seavl-launch/ #ACL2025NLP #SEACrowd #ForSEABySEA
024
Reposted by Lj Miranda
Benjamin Minixhofer @bminixhofer.bsky.social · 02/04/2025
We created Approximate Likelihood Matching, a principled (and very effective) method for *cross-tokenizer distillation*! With ALM, you can create ensembles of models from different families, convert existing subword-level models to byte-level and a bunch more🧵
Image illustrating that ALM can enable Ensembling, Transfer to Bytes, and general Cross-Tokenizer Distillation.
12514
Reposted by Lj Miranda
Arduin Findeis @arduin.io · 17/03/2025
🕵🏻💬 Introducing Feedback Forensics: a new tool to investigate pairwise preference data. Feedback data is notoriously difficult to interpret and has many known issues – our app aims to help! Try it at app.feedbackforensics.com Three example use-cases 👇🧵
172
Reposted by Lj Miranda
Daniel van Strien @danielvanstrien.bsky.social · 13/03/2025
OLMo 2 0325 32B Preference Mixture: Solves AI alignment challenges through diverse preferences - Combines 7 datasets - Filters for instruction-following capability - Balances on-policy and off-policy prompts - Enabled successful DPO of OLMo-2-0325-32B model huggingface.co/datasets/all...
huggingface.co
allenai/olmo-2-0325-32b-preference-mix · Datasets at Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
162
Reposted by Lj Miranda
Ai2 @ai2.bsky.social · 30/01/2025
Here is Tülu 3 405B 🐫 our open-source post-training model that surpasses the performance of DeepSeek-V3! It demonstrates that our recipe, which includes RVLR scales to 405B - with performance on par with GPT-4o, & surpassing prior open-weight post-trained models of the same size including Llama 3.1.
The logo for Tülu 405B.
29221
Reposted by Lj Miranda
Kyle Lo @kylelo.bsky.social · 03/01/2025
kicking off 2025 with our OLMo 2 tech report while payin homage to the sequelest of sequels 🫡 🚗 2 OLMo 2 Furious 🔥 is everythin we learned since OLMo 1, with deep dives into: 🚖 stable pretrain recipe 🚔 lr anneal 🤝 data curricula 🤝 soups 🚘 tulu post-train recipe 🚜 compute infra setup 👇🧵
26917
Reposted by Lj Miranda
Tom Aarsen @tomaarsen.com · 19/12/2024
BERT is BACK! I joined a collaboration with AnswerAI and LightOn to bring you the next iteration of BERT. Introducing ModernBERT: 16x larger sequence length, better downstream performance (classification, retrieval), the fastest & most memory efficient encoder on the market. 🧵
1487
Reposted by Lj Miranda
Melissa Heikkilä @melissahei.bsky.social · 18/12/2024
New research reveals a worrying trend: AI's data practices risk concentrating power overwhelmingly in the hands of dominant technology companies. With analysis from @shaynelongpre.bsky.social @sarahooker.bsky.social @smw.bsky.social @giadapistilli.com www.technologyreview.com/2024/12/18/1...
technologyreview.com
This is where the data to build AI comes from
New findings show how the sources of data are concentrating power in the hands of the most powerful tech companies.
28530
Reposted by Lj Miranda
Jennifer Hu @jennhu.bsky.social · 07/12/2024
Stop by our #NeurIPS tutorial on Experimental Design & Analysis for AI Researchers! 📊 neurips.cc/virtual/2024/tutorial/99528 Are you an AI researcher interested in comparing models/methods? Then your conclusions rely on well-designed experiments. We'll cover best practices + case studies. 👇
neurips.cc
NeurIPS Tutorial Experimental Design and Analysis for AI ResearchersNeurIPS 2024
68614
Reposted by Lj Miranda
David Mimno @dmimno.bsky.social · 10/12/2024
We just updated the AI for Humanists guide to model selection to include Llama 3.3, and a recommended best cost/capability tradeoff, llama 3.1 8B. What have you tried, and what would you suggest? aiforhumanists.com/guides/models/
aiforhumanists.com
Models
The AI for Humanists project is developing resources to enable DH scholars to explore how large language models and AI technologies can be used in their research and teaching. Find an annotated biblio...
25419
Reposted by Lj Miranda
Kyle Lo @kylelo.bsky.social · 10/12/2024
the science of LMs should be fully open✨ today @akshitab.bsky.social @natolambert.bsky.social and I are giving our #neurips2024 tutorial on language model development. everything from data, training, adaptation. published or not, no secrets 🫡 tues, 12/10, 9:30am PT ☕️ neurips.cc/virtual/2024...
neurips.cc
NeurIPS Tutorial Opening the Language Model Pipeline: A Tutorial on Data Preparation, Model Training, and AdaptationNeurIPS 2024
514517
Reposted by Lj Miranda
Ian Magnusson @ianmagnusson.bsky.social · 10/12/2024
Come chat with me at #NeurIPS2024 and learn about how to use Paloma to evaluate perplexity over hundreds of domains! ✨We have stickers too✨
1214
Reposted by Lj Miranda
Grassroots Science @grassroots-science.bsky.social · 09/12/2024
⭐️ We're going to launch Grassroots Science, a year-long ambitious, massive-scale, fully open-source initiative aimed at developing multilingual LLMs aligned to diverse and inclusive human preferences in Feb 2025. 🌐 Check our website: grassroots.science. #NLProc #GrassrootsScience
grassroots.science
Grassroots Science
A global initiative focused on developing state-of-the-art multilingual language models through grassroots efforts.
175
Lj Miranda @ljvmiranda921.itch.io · 04/12/2024
Thank you @oxykodit.bsky.social !
010
Lj Miranda @ljvmiranda921.itch.io · 04/12/2024
Happy to share this and excited to bring this to the public! Nice collab with folks from the University of the Philippines (UP), @angelaquino_ph and Elsie Or, for this impactful work :) Hoping to have the official UD release next year as well.
010
Lj Miranda @ljvmiranda921.itch.io · 04/12/2024
We're releasing the largest Universal Dependencies (UD) treebank for Tagalog, UD-NewsCrawl! This dataset has been a long time coming, but glad to see this through: 15k+ sentences versus the previous ~150 sents from older Tagalog treebanks. 🤗 : huggingface.co/datasets/UD-... 📝 : Paper soon!
huggingface.co
UD-Filipino/UD_Tagalog-NewsCrawl · Datasets at Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
1142
Reposted by Lj Miranda
Yoav Artzi @yoavartzi.com · 02/12/2024
I am seriously behind uploading Learning Machines videos, but I did want to get @jonathanberant.bsky.social's out sooner than later. It's not only a great talk, it also gives a remarkably broad overview and contextualization, so it's an excellent way to ramp up on post-training youtu.be/2AthqCX3h8U
youtu.be
Jonathan Berant (Tel Aviv University / Google) / Towards Robust Language Model Post-training
YouTube video by Yoav Artzi
15312
Lj Miranda @ljvmiranda921.itch.io · 26/11/2024
My favorite part about this release is that we were able to replicate our findings from the Tülu 3 post-training recipe here (e.g., on-policy preferences, RLVR) and found significant performance gains in our -DPO and -Instruct models! Find all artifacts here: huggingface.co/collections/...
huggingface.co
OLMo 2 - a allenai Collection
Artifacts for the second set of OLMo models.
030
Lj Miranda @ljvmiranda921.itch.io · 21/11/2024
So many exciting things in the paper! To learn more about Tülu 3, visit the website: allenai.org/tulu Or better yet, check out the paper! allenai.org/papers/tulu-...
000
Lj Miranda @ljvmiranda921.itch.io · 21/11/2024
The synthetic preference pipeline was a nice research (and engineering!) challenge: we ablated Ultrafeedback to figure out parts that made it awesome, and made changes such as the inclusion of on-policy data, source of prompts, and many other things.
100
Lj Miranda @ljvmiranda921.itch.io · 21/11/2024
Happy to be part of Tülu 3! Great effort to make the post-training stage open-source 😄 I worked on scaling our synthetic preference data (around 300k preference pairs for 70B) that led to performance gains when trained on using DPO.
170
Reposted by Lj Miranda
Ai2 @ai2.bsky.social · 21/11/2024
Meet Tülu 3, a set of state-of-the-art instruct models with fully open data, eval code, and training algorithms. We invented new methods for fine-tuning language models with RL and built upon best practices to scale synthetic instruction and preference data. Demo, GitHub, paper, and models 👇
211131
Reposted by Lj Miranda
Nathan Lambert @natolambert.bsky.social · 21/11/2024
I've spent the last two years scouring all available resources on RLHF specifically and post training broadly. Today, with the help of a totally cracked team, we bring you the fruits of that labor — Tülu 3, an entirely open frontier model post training recipe. We beat Llama 3.1 Instruct. Thread.
821343
Lj Miranda @ljvmiranda921.itch.io · 24/10/2024
Oh this means a lot! Thanks for the warm welcome here Vincent! Happy to see a lot of folks (and people I look up to) here as well~ 🍻
010