Sign in

Mohammed Hamdy

@mmhamdy.bsky.social
185 followers 346 following 40 posts

A curious explorer of human and machine learning 🧐 🤝🤖

PostsRepliesMedia
Mohammed Hamdy @mmhamdy.bsky.social · 27/07/2026
Finally!
000
Mohammed Hamdy @mmhamdy.bsky.social · 24/07/2026
The universe works in mysterious ways!
000
Mohammed Hamdy @mmhamdy.bsky.social · 12/01/2026
Original essay: papers.ssrn.com/sol3/papers....
papers.ssrn.com
On the Slow Death of Scaling
For the last decade, it has been hard to stray off the beaten path of accepted wisdom for what drives innovation. We have been held hostage to a painfully simpl
000
Mohammed Hamdy @mmhamdy.bsky.social · 12/01/2026
Read it here: open.substack.com/pub/surfingm...
open.substack.com
Scaling: A Never-ending Saga
My thoughts on the slow death of scaling
100
Mohammed Hamdy @mmhamdy.bsky.social · 12/01/2026
New article! My thoughts on the slow death of scaling essay by Sara Hooker
100
Mohammed Hamdy @mmhamdy.bsky.social · 08/06/2025
Ok, I'll confess! I too like Roland Emmerich's Godzilla. I even like the creature design in this film!
010
Mohammed Hamdy @mmhamdy.bsky.social · 02/06/2025
The one Frankenstein film to rule them all! Thank you, @realgdt.bsky.social 🙏
010
Mohammed Hamdy @mmhamdy.bsky.social · 31/05/2025
The Consciousness API: What if consciousness isn't contained within us, but rather we are temporary antennas, tuning into a vast, universal broadcast of awareness?
000
Mohammed Hamdy @mmhamdy.bsky.social · 30/03/2025
huggingface.co/blog/mmhamdy...
huggingface.co
Pandemonium: The Transformers Story
A Blog post by Mohammed Hamdy on Hugging Face
000
Mohammed Hamdy @mmhamdy.bsky.social · 30/03/2025
In this article, I explore the story behind some of the ideas introduced in the Transformer paper. Exploring things from the fundamental attention mechanism that lies at its heart to the surprisingly simple explanation for its name. You may find it interesting! 🙂 👇link below
110
Reposted by Mohammed Hamdy
Cohere Labs @cohereforai.bsky.social · 05/03/2025
We're particularly proud to release Aya Vision 8B - it's compact 🐭 and efficient 🐎, outperforming models up to 11x its size 📈. Releasing open weights helps to make breakthroughs in VLMs accessible to the research community.
1144
Mohammed Hamdy @mmhamdy.bsky.social · 27/01/2025
📅 Event on Mozilla AI discord: discord.gg/QTCRfefF?eve... 📄 ProGen paper: www.biorxiv.org/content/10.1...
discord.gg
Join the Mozilla AI Discord Server!
A global space for sharing and advancing open-source AI. | 3757 members
000
Mohammed Hamdy @mmhamdy.bsky.social · 27/01/2025
🧬 Join us this Wednesday on @mozilla.ai discord server in our second session of the Biological Representation Learning series where we discuss landmark papers in the field! We will be presenting the ProGen protein language model paper from Salesforce. See you there! 😃
100
Reposted by Mohammed Hamdy
mozilla.ai @mozilla.ai · 20/01/2025
📢 Join us on Discord for our first Blueprints Hub event 📢 Discover Blueprints and learn how to transform text into podcast-style conversations using entirely open source tools. 🗓️ Wednesday, Jan. 22nd ⏰ 1:30-2:00 PM EST 🔗 Event: discord.gg/BaYFBaeh?eve... #OpenSource #AI #Blueprints #MozillaAI
discord.gg
Join the Mozilla AI Discord Server!
A global space for sharing and advancing open-source AI. | 3695 members
051
Reposted by Mohammed Hamdy
Sara Hooker @sarahooker.bsky.social · 18/01/2025
As the @cohereforai.bsky.social joins the Bluesky family — we will be sharing paper gems from when we first started as a lab. This paper is part of a larger research agenda where we have focused on how to better represent the long tail = making AI work for almost all real world distributions.
0253
Reposted by Mohammed Hamdy
Kyutai @kyutai-labs.bsky.social · 13/01/2025
Meet Helium-1 preview, our 2B multi-lingual LLM, targeting edge and mobile devices, released under a CC-BY license. Start building with it today! huggingface.co/kyutai/heliu...
huggingface.co
kyutai/helium-1-preview-2b · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
1165
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
And lastly, big thanks to you for making it this far 🤗, don’t forget to read the paper! www.dataprovenance.org/Multimodal_D... 11/n
dataprovenance.org
000
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
Big thanks to Melissa Heikkilä for featuring our work in MIT Tech Review. www.technologyreview.com/2024/12/18/1...
technologyreview.com
This is where the data to build AI comes from
New findings show how the sources of data are concentrating power in the hands of the most powerful tech companies.
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
Xuhui Zhou, Caiming Xiong, Luis Villa, @stellaathena.bsky.social, Alex Pentland, @sarahooker.bsky.social, Jad Kabbara 9/n
110
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
An Dinh, Shrestha Mohanty, Deividas Mataciunas, Tobin South, Jianguo Zhang, @arielnlee.bsky.social , Campbell S. Lund, Christopher Klamm, Damien Sileo, Diganta Misra, Enrico Shippole, Kevin Klyman, Lester JV Miranda, Niklas Muennighoff, Seonghyeon Ye, Seungone Kim, Vipul Gupta, Vivek Sharma 8/n
110
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
🎉 big thanks to all the contributors to this huge and magnificent effort. I'm truly honored for the chance to work alongside all of you: Manan Dey, Nayan Saxena, Ahmad Mustafa Anis, Emad A. Alghamdi, Vu Minh Chien, Naana Obeng-Marnu, Da Yin, Kun Qian, Yizhi Li, Minnie Liang 7/n
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
This work was supported by the Mozilla Foundation Data Futures Lab, and was lead by: @shaynelongpre.bsky.social, Nikhil Singh, Manuel Cherep, Kushagra Tiwary, Joanna Materzynska, William Brannon, and Robert Mahari 6/n
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
4️⃣ Linguistic representation has not improved by most measures: Gini Coefficients for text and speech datasets show significant concentration, indicating limited progress in diversifying data sources. 5/n
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
3️⃣ Geographical representation has not improved for a decade: Datasets from African and South American organizations account for < 0.2% of all modality content, while North American or European organizations span 93% of text tokens and 60%+ hours of speech and video. 4/n
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
2️⃣ Inconsistent dataset licenses: While ~30% of datasets have permissive licenses, 78%+ of their sources carry hidden anti-crawling or licensing restrictions, making compliance a minefield. 3/n
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
📌 Key Findings 1️⃣ The web is still the primary source: The internet, social media platforms, and synthetically generated data are increasingly becoming the predominant sources for multimodal data, compared to curated sources. 2/n
100
Mohammed Hamdy @mmhamdy.bsky.social · 19/12/2024
✨ Excited to share our latest work from The Data Provenance Initiative ☸️ This is the most comprehensive audit of multimodal training data, auditing ~4000 datasets between 1990 and 2024, and covering more than 400 unique tasks in 608 languages! 🧵 1/n
131
Mohammed Hamdy @mmhamdy.bsky.social · 11/12/2024
EPIC! 🤗
010
Reposted by Mohammed Hamdy
Johannes @johko.bsky.social · 01/12/2024
🌟 500! 🌟 Our Community Computer Vision Course Repo just reached 500 stars on GitHub: github.com/johko/comput... 🤩 I'm really proud of all the amazing content people from the community have contributed here and that they still keep on adding very cool and helpful material 💪
github.com
GitHub - johko/computer-vision-course: This repo is the homebase of a community driven course on Computer Vision with Neural Networks. Feel free to join us on the Hugging Face discord: hf.co/join/disc...
This repo is the homebase of a community driven course on Computer Vision with Neural Networks. Feel free to join us on the Hugging Face discord: hf.co/join/discord - johko/computer-vision-course
0112
Mohammed Hamdy @mmhamdy.bsky.social · 29/11/2024
The Hudsucker Proxy is the most underrated Coen Brothers film!
230
Mohammed Hamdy @mmhamdy.bsky.social · 27/11/2024
It can go even further... arxiv.org/abs/1611.04499
arxiv.org
Post Training in Deep Learning with Last Kernel
One of the main challenges of deep learning methods is the choice of an appropriate training strategy. In particular, additional steps, such as unsupervised pre-training, have been shown to greatly im...
020
Mohammed Hamdy @mmhamdy.bsky.social · 27/11/2024
Funny thought: if "post-training" refers mostly to supervised instruction-tuning and alignment of a "pre-trained" model, then where does the actual "training" happen! 😀
020
Reposted by Mohammed Hamdy
Nathan Lambert @natolambert.bsky.social · 26/11/2024
Super excited to announce our best open-source language models yet. OLMo 2. These instruct models are hot off the press -- finished training with our new RL method this morning and vibes are very good.
59312
Reposted by Mohammed Hamdy
Andi @andimara.bsky.social · 26/11/2024
Let's go! We are releasing SmolVLM, a smol 2B VLM built for on-device inference that outperforms all models at similar GPU RAM usage and tokens throughputs. SmolVLM can be fine-tuned on a Google collab and be run on a laptop! Or process millions of documents with a consumer GPU!
410422
Mohammed Hamdy @mmhamdy.bsky.social · 27/11/2024
Nice ones! 😀 They probably were created after this post. Someone has to create a new super mega starter pack!😅
020
Reposted by Mohammed Hamdy
Eric Topol @erictopol.bsky.social · 24/11/2024
A vertical takeoff of life science with #AI LLLMs. Publication of 10 new foundation models of Proteins, DNA, RNA, methylation, cells, and interactions, evolution, and design in the past couple of weeks! Unprecedented progress, reviewed in the new Ground Truths erictopol.substack.com/p/learning-t...
Covers of Science and Nature journals in the past 2 weeks denoting remarkable progress of life science with A.I. tools
1534486
Mohammed Hamdy @mmhamdy.bsky.social · 25/11/2024
A Data-centric AI starter pack (Needless to say, this is by no means exhaustive). Please feel free to mention anyone I've missed in this list and I will update it. go.bsky.app/SNj7M2Y
010
Reposted by Mohammed Hamdy
Thomas Wolf @thomwolf.bsky.social · 24/11/2024
It's Sunday morning so taking a minute for a nerdy thread (on math, tokenizers and LLMs) of the work of our intern Garreth By adding a few lines of code to the base Llama 3 tokenizer, he got a free boost in arithmetic performance 😮 [thread]
527134
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
10. ML OSS & Open Source / Science go.bsky.app/8MFcfXd
150
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
9. Grumpy Machine Learners go.bsky.app/6ddpivr
130
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
8. Women in ML/CV/NLP go.bsky.app/9Khutjf
130
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
7. Machine Learners who were once children go.bsky.app/F6mM37U
120
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
6. Image/Video Gen related researchers go.bsky.app/SP1uWoE
130
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
5. Google DeepMind starter Pack go.bsky.app/GZ4hZzu
150
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
4. RLHF / Alignment go.bsky.app/MqRGAf2
130
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
3. Cracked ML Engineers & Researchers go.bsky.app/H9nj9nJ
130
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
2. Hugging Face Folks go.bsky.app/2VGyqGt
1102
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
🤖 ML/AI Mega Starter Pack 1. Open-source LLMS go.bsky.app/FELkyDr 🧵
3249
Mohammed Hamdy @mmhamdy.bsky.social · 22/11/2024
NLP? RL? Why not both?! arxiv.org/abs/2411.14251
arxiv.org
Natural Language Reinforcement Learning
Reinforcement Learning (RL) mathematically formulates decision-making with Markov Decision Process (MDP). With MDPs, researchers have achieved remarkable breakthroughs across various domains, includin...
000