Sign in

Sujee Maniyam

@sujee.dev
51 followers 61 following 115 posts

Developer Advocate @ Nebius | ML/Data Engineer | Open Source contributor | Technical Instructor | Author | Speaker portfolio: bit.ly/sujee-dev3

PostsRepliesMedia
Sujee Maniyam @sujee.dev · 03/10/2026
Also on YouTube (4K, with chapters): youtu.be/RQ6f6z-dPWA More explainers: www.youtube.com/playlist?lis...
youtu.be
Jev vs. LLMs: a model that decides instead of writing | Explainer Video
YouTube video by sujee_dev
000
Sujee Maniyam @sujee.dev · 03/10/2026
Most AI models write. Jev decides. A 2-minute explainer for devs. Made with Claude Code + Opus 5.5. Open source and fully reproducible. Remix it: github.com/sujee/visual... #AI #LLM #ExplainerVideo #Jev
110
Sujee Maniyam @sujee.dev · 02/10/2026
Here’s the crinkled/scanned-looking document image (generated!) I fed into the model 👇 Not exactly a clean input - which makes the result even more fun.
000
Sujee Maniyam @sujee.dev · 02/10/2026
Quick demo showing the vision capabilities of Qwen 3.8-27B 👇 I gave it a crinkled, scanned-looking document containing PII. It extracted the text, analyzed the document, and identified the PII ✅ This little model is a beast 😄 And it was ticking along at 272 tok/sec on @nebiustf 🚀
100
Sujee Maniyam @sujee.dev · 25/09/2026
Try it for yourself 🐍 sujee.github.io/llm-snake-ar... Go burn some tokens :-) Code: github.com/sujee/llm-sn...
000
Sujee Maniyam @sujee.dev · 25/09/2026
Friday funsies 🐍 LLM Snake Arena: GPT-6-Astra vs Claude Opus 5.5 Does it tell us something about the models? Maybe. Fun to watch? Most definitely 😀
100
Sujee Maniyam @sujee.dev · 16/09/2026
Setup + instructions: github.com/sujee/worksh... Want to try open models? Get credits: dev.nebius.com/builders?utm...
github.com
workshops/coding-with-open-models/README.md at main · sujee/workshops
Hands-on workshops on AI engineering, open models, coding agents, LLM inference, developer tools, and physical AI. - sujee/workshops
011
Sujee Maniyam @sujee.dev · 16/09/2026
Want to try open models for coding? Keep your coding agent - Claude Code, Codex, OpenCode, Cline - and just swap the model underneath. In this short talk: * open vs closed models * why the coding harness matters * what coding tasks cost * quick setup demo Watch: www.youtube.com/watch?v=gjNc...
youtube.com
Coding with Open Models: Claude Code + OpenCode with Kimi, GLM & DeepSeek
YouTube video by sujee_dev
100
Sujee Maniyam @sujee.dev · 12/09/2026
Top open models now: 1. GLM-5.3: 45 2. Kimi K3: 44 (60 → 44 = ▼ 16) 3. GLM-5.3-Flash: 42 (57 → 42 = ▼ 15 ) Big score shifts, but the important thing is the relative ranking under the new benchmark suite.
000
Sujee Maniyam @sujee.dev · 12/09/2026
Artificial Analysis just updated its Intelligence Index to v4.3 with a tougher benchmark suite. Scores dropped quite a bit. Top proprietary models now: - Claude Fable 5.1: 53 = 66 → 53 = ▼ 13 - GPT-6 Astra: 53 = 61 → 53 = ▼ 8
210
Sujee Maniyam @sujee.dev · 07/09/2026
Went to buy a 4TB portable SSDs and almost fell off my chair looking at the prices 😳 Damn. See screenshots 👇 @demian_ai has been writing about the memory/storage crunch driven by the AI boom. It is finally sinking in for me :-) x.com/demian_ai/st... x.com/demian_ai/st... x.com/demian_ai/st...
000
Sujee Maniyam @sujee.dev · 07/09/2026
Fine-tuning, distillation, and quantization are easy to blur together. Here is an explainer video with LEGO® bricks 🧱 🟢 Fine-tune = specialize 🟠 Distill = teach a student 🟡 Quantize = use lower precision More detail: sujee.dev/post/fine-tu... drop your suggestions in comments
000
Sujee Maniyam @sujee.dev · 29/08/2026
2️⃣ Second pic: GLM-5.3-Flash running inside OpenCode. Amazing speed and performance! ⚡
000
Sujee Maniyam @sujee.dev · 29/08/2026
1️⃣ First graphic: GLM-5.3-Flash (formerly Ox Alpha!) scores 57 on the Artificial Analysis Intelligence Index, while Kimi K3 scores 60. But GLM-5.3-Flash is: much cheaper, much smaller Try it for yourself : tokenfactory.nebius.com/playground?m... Model visualizer: sujee.github.io/practical-ll...
100
Sujee Maniyam @sujee.dev · 28/08/2026
I tested 3 NVIDIA Nemotron models with the same sales-analysis agent. All scored 100%. Lightning done in 7s for $0.0019 - about 7x cheaper than Super and 16x cheaper than Ultra. TLDR; For focused agent workflows, smaller may be enough. www.youtube.com/watch?v=Ycme... github.com/nebius/token...
000
Sujee Maniyam @sujee.dev · 20/08/2026
🚀 Token Factory Office Hour #4 Join us live with special guest Chris Alexiuk from NVIDIA talking Nemotron models. Also * What’s new in @nebiustf * Builder demos * Live Q&A * Credits 🎁 Like to show what you built with Token Factory? DM me. 📅 Aug 26, 9 AM PDT Register: nebius.com/events/webin...
nebius.com
Token Factory office hours: Nemotron models
Build and scale faster on the purpose-built AI cloud, engineered from silicon to API.
010
Sujee Maniyam @sujee.dev · 15/08/2026
I am doing quick hands-on workshop on ClawCamp SF Summer Summit tomorrow! 🦞 “Coding with Open Models + Coding Agents” Want to try Kimi K3, GLM 5.2 and other open models? Bring your laptop 💻 We’ll provide Token Factory credits so you can try them with your coding agent. luma.com/clawcamp-sf-...
luma.com
🦞 ClawCamp SF Summer Summit 🦞 · Luma
NEW VENUE: Ecosystem Coworking 540 Howard St 2nd floor, San Francisco, CA 94105 Join us to Up-Level Your Agentic AI Superpowers. All Levels Welcome. This…
030
Sujee Maniyam @sujee.dev · 07/08/2026
Kimi K3 was one of the most anticipated open models. Join the Kimi + Nebius Token Factory teams on Aug 12: 🏗️ K3 architecture + capabilities ⚡ Inference engineering + optimizations 🎁 Token Factory credits to try K3 yourself 9 AM PT / 12 PM ET / 6 PM CET 👉 nebius.com/events/webin...
000
Sujee Maniyam @sujee.dev · 31/07/2026
3D map of New York City x.com/kirillk_web3... And a browser-playable Quake game from me 🎮 x.com/sujee_dev/st... I’m collecting more fun Kimi K3 demos here: github.com/sujee/practi... Know of another cool demo? Drop it in the comments!
x.com
Kirill (@kirillk_web3) on X
Kimi K3 is dangerous. I rebuilt Google Maps 3D mode in 1.5 hours. an interactive 3D map — New York City in full detail, and the entire globe on top of it. two prompts. that's the whole build. > ...
010
Sujee Maniyam @sujee.dev · 31/07/2026
Friday fun: Here are some cool @kimi-moonshot-x.bsky.social K3 demos I’ve seen so far 👇 Charlie and the Token Factory by @demian_ai 🎮 datamon.vercel.app Subway Surfers x.com/Arindam_1729... Car crushing and physics simulation x.com/atomic_chat_... fish tank x.com/UnslothAI/st...
datamon.vercel.app
Charlie & the Token Factory
A full browser RPG set inside the Nebius Token Factory.
100
Sujee Maniyam @sujee.dev · 29/07/2026
Join the Nebius AI Builder Program and get $25 in free Token Factory credits to start building: dev.nebius.com/builders?utm...
dev.nebius.com
Nebius Builder Program
Build and scale faster on the purpose-built AI cloud, engineered from silicon to API.
000
Sujee Maniyam @sujee.dev · 29/07/2026
Trying out @kimi-moonshot-x.bsky.social K3 (running on Nebius Token Factory ) for coding with @opencode.bsky.social . The integration was incredibly smooth - easy to configure and start building. Here’s a quick 30-second walkthrough to get you up and running. 👇 start: tokenfactory.nebius.com
110
Sujee Maniyam @sujee.dev · 28/07/2026
Want to build with K3? Join the Nebius AI Builder Program and get $25 in free Token Factory credits to start building: dev.nebius.com/builders?utm...
dev.nebius.com
Nebius Builder Program
Build and scale faster on the purpose-built AI cloud, engineered from silicon to API.
000
Sujee Maniyam @sujee.dev · 23/07/2026
🎬 Replay: nebius.com/events/webin... 📑 Slides: email.nebius.com/hubfs/webina...
nebius.com
Token Factory office hours: Build and ship with open models
Build and scale faster on the purpose-built AI cloud, engineered from silicon to API.
000
Sujee Maniyam @sujee.dev · 23/07/2026
Nebius Token Factory @nebiustf Office Hour #3 Recap: We covered new models and features, inference optimizations, fine-tuning, the Builder and Fellow programs, plus live Q&A. recording + slides in comment 👇 See you at the August office hour: nebius.com/events/webin...
100
Sujee Maniyam @sujee.dev · 22/07/2026
Building with open models? Join us for Nebius Token Factory Office Hour - a live, developer-focused session featuring demos and Q&A 🗓️ July 22, 2026 ⏰ 9 AM PT / 12 PM ET You’ll also receive Token Factory credits, so you can start building right away. See you there! nebius.com/events/webin...
000
Sujee Maniyam @sujee.dev · 02/07/2026
Nebius booths are never boring. Meet NEBU - our Nebius pup 🐶 - making a new friend 🐉 Stop by the Nebius booth at AI Engineer World Fair. You can play with the robots... and chat with the humans, too. 😄
000
Sujee Maniyam @sujee.dev · 01/07/2026
Come by the Nebius booth at AI Engineer conf! We’ve got fun demos: LLMs playing Snake 🐍, LLMs playing soccer/fútbol ⚽️ Demo below 👇 We’ll also show you our open-model coding stack - GLM 5.2 and Kimi K2.6 are our current faves. Plus, we’ll set you up with Token Factory credits 🎁
000
Sujee Maniyam @sujee.dev · 24/06/2026
I’m running a workshop at ClawCamp Campout @ Dual Tech Summit! We’ll add real-time intelligence to OpenClaw using Nebius Token Factory + Tavily 🗓️ Wed, June 24, 2026 📍 War Memorial Veterans Building, SF 👉 clawcamp.us/event/260624... 🎁 Free Token Factory + Tavily credits . See you there!
110
Sujee Maniyam @sujee.dev · 19/02/2026
The AgentBeats–AgentX Competition Phase 1 winners have been announced - and the projects are fantastic. Huge congratulations to all the teams involved! 👏 Read the highlights here: berkeleyrdi.substack.com/i/188179396/... Now, on to Phase 2 : rdi.berkeley.edu/agentx-agent...
000
Sujee Maniyam @sujee.dev · 18/02/2026
GLM 4.7 (thought 42 secs): If your car has a manual transmission (stick shift) and the road is flat, you might choose to **push** the car to save fuel and avoid the wear of starting the engine just for a 50-meter drive ... (see image for full answer) sometimes the models do overthink :-)
000
Sujee Maniyam @sujee.dev · 18/02/2026
(another car wash post) Q: I want to wash my car. The car wash is 50 meters away. Should I walk or drive? Kimi K2.5: (thought 52 secs) : Drive the 50 meters, but maybe take the long way around the block to let the engine warm up slightly ...
100
Sujee Maniyam @sujee.dev · 13/02/2026
Try it out: tokenfactory.nebius.com/playground?m...
tokenfactory.nebius.com
Nebius Token Factory
Transform your business with AI on Nebius Token Factory. Premium open-source LLMs and image models, simple API integration, and infrastructure that scales automatically with your growth.
000
Sujee Maniyam @sujee.dev · 13/02/2026
Kimi 2.5 from @kimi-moonshot-x.bsky.social is now live on Nebius Token Factory. A major step up from Kimi 2, this os multimodal model is strong in agentic workflows and coding tasks - and it’s stacking up well against SOTA proprietary models. Check out the official benchmarks and the pelican 🪿🚲
100
Sujee Maniyam @sujee.dev · 15/01/2026
MiniMax M2.1 from MiniMax is now live on Nebius Token Factory. strong showing for coding, agentic workflows It’s also ranked among the **top open models** by Artificial Analysis And yes… it passes my favorite vibe benchmark too: “pelican riding a bicycle” 🪿🚲 - look at that beak 😄
000
Sujee Maniyam @sujee.dev · 13/01/2026
qwen3-next-80B-A3B-Thinking is now available on Nebius Token Factory. “Fast but verbose” per Artificial Analysis: artificialanalysis.ai/models/qwen3... My fav vibe benchmark: 'pelican riding a bicycle' 🪿🚲 Try it 👉 tokenfactory.nebius.com/playground?m... #OpenModels #LLM #PelicanRidingBicycle
000
Sujee Maniyam @sujee.dev · 08/01/2026
Installed @linuxmint.bsky.social on desktop & laptop Overall - smooth install - great collection of OS software Desktop: - detected dual monitors and NVIDIA GPU Laptop: - every thing works including suspend / wake A pic of my two fav mints :) Install notes: sujee.dev/post/another... #linux
020
Sujee Maniyam @sujee.dev · 03/12/2025
🚀 Token Factory Builder Hour #1 Dec 9 @ 9am PT / 12pm ET / 5pm UTC / 6pm CET Hang out, share ideas, see demos, learn what’s new in Token Factory. Can’t make it? We’ll send out the recording + notes. Join: nebius.com/events/webin... Discord: [discord.gg/2BUfPAu5uH](discord.gg/2…
nebius.com
Builder hour: Token Factory
Discover the most efficient way to build, tune and run your AI models and applications on top-notch NVIDIA® GPUs.
000
Sujee Maniyam @sujee.dev · 21/11/2025
2️⃣ Open-Source RAG Pipeline with Docling + Data Prep Kit + Milvus + Open LLMs - Session: qconsf.com/training/nov... - Repo: github.com/sujee/data-p... We ran open source models on tokenfactory.nebius.com - DM if you want credits to try it out. Collaborate on ; discord.gg/bk5fcvNJVZ
000
Sujee Maniyam @sujee.dev · 21/11/2025
That's a wrap on two workshops at @qconferences.com Great crowd, good questions, and lots of fun discussions. 1️⃣ Chat with Your Website Using an LLM + Open Stack (Allycat) - Session: qconsf.com/training/nov... - Repo: github.com/The-AI-Allia...
121
Sujee Maniyam @sujee.dev · 19/11/2025
running 2 workshops Nov-20 at @qconferences.com 1. Chat with your website using an LLM and open stack (Allycat) qconsf.com/training/nov... 2. Open Source Rag Pipeline With Docling + Data Prep Kit + Milvus + Open LLMs qconsf.com/training/nov... discord: discord.gg/bk5fcvNJVZ #NebiusTokenFactory
011
Sujee Maniyam @sujee.dev · 31/10/2025
🎃 SVGenAI Meetup #5 — Halloween Edition! Real-world AI _horror stories_ 👻 … and the solutions that saved the day. Plus: great networking with fellow builders. 📍 San Jose 🗓️ Nov 4, 5:30 PM See you there : luma.com/df08ml54?tk=... #svgenai #meetup
luma.com
SVGENAI Meetup # 5 Spooky Halloween Edition: AI Horror Stories · Luma
!!!Date Change!!! This meet-up date has been changed to Nov. 4th! SVGENAI meetup We are focused on fostering learning, networking and career development of AI…
010
Sujee Maniyam @sujee.dev · 22/10/2025
You *can* indeed train nanochat (github.com/karpathy/na...) under $100! see the post by Marouane Khoukh www.linkedin.com/posts/marou... Done on @nebiusai 💪
linkedin.com
#opensource | Marouane Khoukh
Andrej Karpathy said you could train an #opensource ChatGPT-like model from scratch for under $100. I wanted to see if that actually holds, so I tried the full NanoChat (https://lnkd.in/dTfWtFv2) on Nebius. ⚙️ I spun up an 8×H100 instance on Nebius, cloned karpathy/nanochat, and ran the full pipeline end-to-end. The steps: • Train a tokenizer on FineWeb-EDU (65K vocab) • Pretrain a 560M-parameter Transformer to predict the next token • Mid-train on conversational and reasoning data (SmolTalk, MMLU, GSM8K) • Supervised Finetuning — aligning it to chat format The results: ⏱️ The whole thing took a bit more than three hours. 💰 Cost: around $80 on Nebius . 📉 Validation bpb ≈ 0.81 📊 CORE ≈ 0.20 ⚡ MFU ≈ 48 % So yes, it really does train a ChatGPT-style model under $100. After the run, I opened the web UI and chatted with my model, and that’s the video below. It’s pretty wild that you can now train, fine-tune, and talk to your own LLM in an afternoon. Curious how the pipeli
000
Sujee Maniyam @sujee.dev · 21/10/2025
Quick recap from our chat with @HackAgingAI hackathon @ElmuratovArtem and I talked about: 🔬 How AI is tackling some of the toughest problems in biosciences 🧠 Building AI agents with open models + open stacks ☁️ Run effortlessly on @nebiusaistudio / @nebiusai 🎥 www.youtube.com/live/93sDOFV...
000
Sujee Maniyam @sujee.dev · 17/10/2025
The AgentX–AgentBeats competition challenges you to "build agents that evaluate other agents" - how meta is that? 🤖 You’ll get to: 🌍 Build **public-good** systems through friendly competition 🏆 Compete for **$1M in prize money** 👉: rdi.berkeley.edu/agentx-agent... Proudly supported by #nebius
000
Sujee Maniyam @sujee.dev · 10/10/2025
Little backstory: davenielsen.bsky.social and I started Allycat (github.com/The-AI-Alli...) as a demo project at @aialliance.bsky.social . Since then it has been adopted by many and getting contributions from others. This is how open source works.
github.com
GitHub - The-AI-Alliance/allycat: Chat with your website using LLMs
Chat with your website using LLMs. Contribute to The-AI-Alliance/allycat development by creating an account on GitHub.
010
Sujee Maniyam @sujee.dev · 10/10/2025
Tech-stack: - Web crawl - Docking for extracting data from downloaded data (HTML / PDF ..etc) - Index and store in vector database - llama-index framework - open source LLMs like (Qwen3, GLM, GPT-OSS, Deepseek) powered by Nebius AI Studio You will walk away with working code you can build on.
100
Sujee Maniyam @sujee.dev · 10/10/2025
Hah! Another workshop accepted at #QConSF 😃 "Allycat - chat with your website with LLMs" In this hands-on workshop, I will show how to crawl a website, index data and query it using natural language (not just keywords) with LLMs. qconsf.com/training/no... @qconferences.com #QConSF #docling
122
Sujee Maniyam @sujee.dev · 10/10/2025
Happy to be supporting NextBio Hackathon (AI x BIO Hackathon and DEMO Day) at at UCSF. luma.com/wuqccdje?tk... Connect, network and build together with other builders and AI x Biology development experts. Nebius AI Studio is sponsoring with credits. #NebiusAIStudio #Nebius
000
Sujee Maniyam @sujee.dev · 10/10/2025
Agentic AI Against Aging Hackthon now on! www.hackaging.ai/ Oct 7 - 25. Hybrid format (online + finals in San Francisco). Nebius is a proud sponsor.
hackaging.ai
Agentic AI Against Aging
AI agents, working to extend human life. Join our hackathon October 7-25, 2025.
000