Sign in

ngrok

@ngrok.com
375 followers 14 following 95 posts

One gateway for all your traffic. Service status available at status.ngrok.com.

PostsRepliesMedia
Pinned
ngrok @ngrok.com · 09/09/2026
The best compressors are... LLMs? Compression and language modeling share a job: predicting what comes next. @anniesexton.com shows how those predictions become fewer bits, and why gzip still gets to keep its all-important job. ↓ www.youtube.com/watch?v=HkUy...
youtube.com
The best compressors are...LLMs?
YouTube video by ngrok
0121
Reposted by ngrok
fry69 @fry69.dev · 10/09/2026
Highly recommended blog and video by @anniesexton.com about the relation between compression algorithms and large language model token prediction. -> ngrok.com/blog/compres...
ngrok.com
Compression is prediction | ngrok blog
Compression and LLMs are trying to solve the exact same problem: predicting what comes next. Learn the fundamentals of compression and how better prediction leads to better shrinkage.
161
Reposted by ngrok
Annie Sexton @anniesexton.com · 09/09/2026
I made a video about the compression article, it has animations, go watch it and admire my sick new studio.
2245
ngrok @ngrok.com · 02/09/2026
GOPHRS is here. AI gateway migration is mandatory. The arch is final (as in "no more feature requests" final). Please watch the full architecture briefing before submitting questions. www.youtube.com/watch?v=UVe1...
youtube.com
GOPHRS is here. AI gateway migration is mandatory.
YouTube video by ngrok
131
Reposted by ngrok
Marcus Noble @averagemarcus.bsky.social · 25/08/2026
This post from Sam Rose is absolutely fantastic! 💙 A real deep dive into Kubernetes probe with amazing interactive demos, advice and even a bug discovered in Kubernetes! 🤯 ngrok.com/blog/probes
ngrok.com
How Kubernetes probes work | ngrok blog
Learn how probes work interactively on simulated Kubernetes clusters running in your browser.
0165
Reposted by ngrok
Sean Killeen @seankilleen.com · 20/08/2026
Phenomenal article from @samwho.dev at @ngrok.com on #kubernetes probes ngrok.com/blog/probes With interactive visualizations! Really nicely done. I learned a ton.
ngrok.com
How Kubernetes probes work | ngrok blog
Learn how probes work interactively on simulated Kubernetes clusters running in your browser.
021
Reposted by ngrok
Sam Rose @samwho.dev · 19/08/2026
Impossible to overstate how much of my soul is imbued into this post. I wanted to push the boundaries of the web as an education platform. Why _shouldn’t_ we run Kubernetes in the browser to demonstrate for-real what happens, and let you influence it? Hope you love it. 🫶
89218
ngrok @ngrok.com · 19/08/2026
What do a port of Kubernetes to TypeScript and a high-priority bug in the kubelet have in common? They both feature in @samwho.dev's guide to probing, out now. We're experimenting with what next-generation k8s education looks like, and we hope you love it.
1466
Reposted by ngrok
Pieter van der Westhuizen 💙 @mythicalmanmoth.com · 12/08/2026
One of the best explained and clearly written pieces I've read in a long time. Love the interactive examples. Very cool. Nice one @anniesexton.com !
011
ngrok @ngrok.com · 11/08/2026
Cosigned.
000
Reposted by ngrok
Annie Sexton @anniesexton.com · 11/08/2026
This is the coolest article I've ever written. Genuinely. All of you should read it.
1329545
ngrok @ngrok.com · 11/08/2026
Compression and LLMs are trying to solve the exact same problem: predicting what comes next. @anniesexton.com's interactive essay walks us through the basics of compression to reveal its surprising overlap with every language model you've ever used.
513225
ngrok @ngrok.com · 05/08/2026
You should be able to send inference: → to any public provider → privately to models you self-host → to any number of fallbacks → with full access control and observability built in ngrok.ai is live on Product Hunt today ↓ www.producthunt.com/products/ngr...
ngrok.ai
One gateway, every model — ngrok.ai
Route, secure, and manage traffic to any LLM—cloud or local—with one unified platform. Monitor usage, optimize costs, and keep your AI products online.
010
ngrok @ngrok.com · 30/07/2026
"You need four of this pod... how hard can that be?" — @samwho.dev Turns out it's several thousand lines of code hard. Part 4 of our video series on his Kubernetes-in-the-browser project tiptoes through that and other footguns ↓ youtu.be/y4XAOcCT9dY
youtu.be
Webernetes: How we got Kubernetes running in the browser
YouTube video by ngrok
131
Reposted by ngrok
ScyllaDB @scylladb.com · 28/07/2026
💡 Tuesday tech tip: At our free & virtual #P99CONF, @ngrok.com's Sam Rose will dive into the mathematical process applied to models to make them smaller and introduce our audience to tools that gauge how much a quantized model differs from its original. www.p99conf.io?latest_sfdc_... #ScyllaDB
041
ngrok @ngrok.com · 13/07/2026
Validating JWTs, restricting IPs, and running a WAF on your ngrok endpoints just got 10x cheaper. Traffic Policy is now $0.10 per 100k TPUs, down from $1. OAuth didn't get cheaper. It was already free. ↓ ngrok.com/blog/pricing...
ngrok.com
Traffic Policy is now 10x cheaper, and idle custom domains are free | ngrok blog
Traffic Policy pricing drops by 10x, custom domains move to active-only billing, and Endpoint Pooling now bills per active endpoint. Here's what changed and why.
000
ngrok @ngrok.com · 09/07/2026
The next interactive essay from @samwho.dev is about k8s probes, but even starting it took… 90k lines of TypeScript. Here's webernetes episode 3: the first demo of a cluster in a browser tab. Deployment controllers due in 3 days, when Sam's summer holiday starts ↓ www.youtube.com/watch?v=gxo3...
youtube.com
The demo that proves Kubernetes can run in the browser
YouTube video by ngrok
171
ngrok @ngrok.com · 01/07/2026
You can now call models running on your own hardware through a hosted AI gateway. Privately, with one command: `ngrok http 8000 --url https:⁠//vllm.internal` Plus all the public models you know and love. ↓ ngrok.ai
ngrok.ai
Route, Secure & Manage Any LLM with ngrok AI Gateway
Route, secure, and manage traffic to any LLM—cloud or local—with one unified platform. Monitor usage, optimize costs, and keep your AI products online.
110
Reposted by ngrok
Sam Rose @samwho.dev · 30/06/2026
I promise this is my last post this week sharing work stuff but we’ve condensed a bunch of new stuff @ngrok.com can do into a 6 and a half minute video! This is for the folks that know how to run `ngrok http 8080` but haven’t ventured further into the product than that. 🫶 youtu.be/trXqyNXJa1k?...
youtu.be
This is ngrok in 2026
YouTube video by ngrok
0103
ngrok @ngrok.com · 30/06/2026
Yes, ngrok gives localhost a public URL. It's also load-bearing production infrastructure. Reach into customer networks. Put every bit of ingress behind one front door. Run one gateway for devices, APIs, and LLMs. Here's what ngrok really is today ↓ www.youtube.com/watch?v=trXq...
youtube.com
This is ngrok in 2026
YouTube video by ngrok
000
Reposted by ngrok
Sam Rose @samwho.dev · 30/06/2026
Inviting controversy with this take, which is outside of my comfort zone, but I do stand by it. You can generate good quality code with LLMs, it “just” requires a tonne of human oversight and automated guard rails. Got a week off then I’ll be building my first proper post with webernetes. Stoked.
9794
ngrok @ngrok.com · 30/06/2026
Our own @samwho.dev ported Kubernetes to the browser. Like a real flippin' cluster with lifecycles, DNS, and a simulated network. It's ~100k lines of TypeScript, almost all written by LLMs, but with every line reviewed by hand to keep it slop-free ↓ ngrok.com/blog/i-porte...
ngrok.com
I ported Kubernetes to the browser | ngrok blog
Almost 100,000 lines of LLM-generated code in 2 months, and none of it is slop.
3274
ngrok @ngrok.com · 18/06/2026
Yup, @samwho.dev is a wild one for this. But we want to teach the invisible layers of modern software in ways grounded in reality. Ways that'll last for a long, long time. Sometimes that means 100k lines of TypeScript just to earn the right to start telling the story at all.
110
ngrok @ngrok.com · 03/06/2026
Our very own @samwho.dev decided the only honest way to explain how Kubernetes probes work, using his signature interactive essay style, was to build himself a cluster. In the browser. 👇
1260
ngrok @ngrok.com · 30/03/2026
We gave ngrok domains a `resolves_to` property and, what do you know, made the *routing* part of data residency a one-click. Pin your traffic to specific PoPs so European data stays in Europe, then rearrange the whole thing tomorrow if you want.
200
ngrok @ngrok.com · 26/03/2026
Expose local apps and APIs to the internet with your coding agent and our new `expose-localhost` skill. Then ask it to add in OAuth, OWASP protection, rate limits, and more if you're feeling plucky. $ npx skills add ngrok/agent-skills github.com/ngrok/agent-...
github.com
GitHub - ngrok/agent-skills: Official repository for ngrok agent skills.
Official repository for ngrok agent skills. Contribute to ngrok/agent-skills development by creating an account on GitHub.
1531
ngrok @ngrok.com · 25/03/2026
Quantization can make an LLM 4x smaller and 2x faster, with barely any quality loss. But what *is* it? @samwho.dev crafted a beautiful interactive essay explaining it from first principles, aimed at coders, not mathematicians. ngrok.com/blog/quantiz...
ngrok.com
Quantization from the ground up | ngrok blog
A complete guide to what quantization is, how it works, and how it's used to compress large language models
052
Reposted by ngrok
Sam Rose @samwho.dev · 25/03/2026
I spent 2 months learning about quantization and am extremely proud of the post I've written about it. I think these are some of the nicest visuals I've ever made, and I love how this compression technique invented in 1898 is being used on the bleeding edge in 2026. ngrok.com/blog/quantiz...
ngrok.com
Quantization from the ground up | ngrok blog
A complete guide to what quantization is, how it works, and how it's used to compress large language models
1425155
ngrok @ngrok.com · 02/03/2026
Route between OpenAI and Anthropic with one API key without an account for either. ngrok's AI Gateway now has its own API keys and prepaid credits. One key covers both, and you can just set your model to "ngrok/auto" if you don't even want to pick. ngrok.com/blog/ai-gate...
ngrok.com
AI Gateway: use Anthropic and OpenAI with one API key | ngrok blog
Create an AI Gateway API Key, add credits, and start making requests to OpenAI and Anthropic—ngrok handles the rest.
021
ngrok @ngrok.com · 19/02/2026
ngrok's AI gateway now supports the Anthropic SDK natively. Change your base_url, keep your prompt caching, extended thinking, all of it. The diff is one line & there's no step two. ngrok.com/blog/native-...
021
Reposted by ngrok
Sam Rose @samwho.dev · 19/02/2026
What do LLMs see? I wrote a lil' tool that extracts the attention matrices out of open models and creates this typing visual, with each token's opacity changing according to its average attention score as the prompt progresses. Dimmer words are considered less important to the model.
1726444
Reposted by ngrok
Overcommitted Podcast @overcommitted.dev · 10/02/2026
This week on Overcommitted, we got to sit down with Bluesky's favorite tech blogger @samwho.dev and it did not disappoint! Sam makes some of the coolest tech content on the internet, and if you haven't heard from him yet, you should! Full episode out now: overcommitted.dev/interactive-...
0215
ngrok @ngrok.com · 09/02/2026
We're hiring a full-time video creator. Looking for someone who always aims for quality above cadence, is eager to deep-dive on AI and networking, wants to be part of a small and taste-obsessed team. Details: job-boards.greenhouse.io/ngrokinc/job...
job-boards.greenhouse.io
Sr. Developer Educator, Video
United States
030
ngrok @ngrok.com · 29/01/2026
Did you know that SWE-bench only tests how good models are at Python? Or the big drama around the FrontierMath benchmark? We didn't either! But now you will, because @samwho.dev put together an extensive explanation of 14 popular AI benchmarks. ngrok.com/blog/ai-benc...
ngrok.com
What those AI benchmark numbers mean | ngrok blog
An explanation of 14 benchmarks you're likely to see when new models are released.
0132
ngrok @ngrok.com · 28/01/2026
Three ships coming soon we're excited about: ◆ region pinning for data residency ◆ similarly, dedicated IPs instead of our default range of multi-tenant IPs ◆ faster & more reliable certificate provisioning for ngrok/custom/wildcard domains ↓
111
ngrok @ngrok.com · 27/01/2026
A favorite beat on the new ngrok.com: We put ngrok on top of ngrok.com so you can ngrok while you learn about ngrok. If we're going to ask you to use us for anything, we better do it ourselves first, in prod, and also show our work.
011
ngrok @ngrok.com · 26/01/2026
ngrok is one gateway for all your traffic. So, time to introduce the new ngrok.com For years we've helped you secure, transform, and route to services running anywhere. localhost to prod, APIs to AI models, devices in the field to databases in customer networks.
ngrok.com
ngrok - All your traffic. One gateway. | API Gateway, Secure Tunnels, Traffic Management
ngrok is an all-in-one cloud networking platform that secures, transforms, and routes your traffic to services running anywhere.
182
ngrok @ngrok.com · 15/01/2026
We heard you like paging through dozens or hundreds of objects in API responses to find the one you want. Just kidding, nobody's ever said that, so thank goodness that server-side filtering of the ngrok API is now GA. ngrok.com/blog/api-fil...
020
Reposted by ngrok
Dave Rupert @davatron5000.bsky.social · 18/12/2025
An enjoyable look at how LLMs work under the hood from @samwho.dev. The whole chain of tokenizers (text chunks), embeddings (vectorized text chunks), transformers (the "T" in ChatGPT), caching (reused chunks), and all the math that goes into returning a response. ngrok.com/blog/prompt-...
ngrok.com
Prompt caching: 10x cheaper LLM tokens, but how? | ngrok blog
A far more detailed explanation of prompt caching than anyone asked for.
1428
Reposted by ngrok
Owen Lacey @owenlacey.dev · 18/12/2025
Learned more about LLM's under the hood from this post than my stupid $30 online course ngrok.com/blog/prompt-... @samwho.dev doing @samwho.dev things 🙌
ngrok.com
Prompt caching: 10x cheaper LLM tokens, but how? | ngrok blog
A far more detailed explanation of prompt caching than anyone asked for.
0303
Reposted by ngrok
Sam Rose @samwho.dev · 16/12/2025
New post! ✨ Prompt caching ✨ My first big project post for @ngrok.com. 5 weeks, 12217 lines of code, 195 commits. I poured a lot into this one, and learned a lot in the process. I really hope you enjoy it ❤️
79520
ngrok @ngrok.com · 16/12/2025
Yesterday we launched ngrok.ai into early access. Today we're bringing you a deep dive into LLM internals with beautiful visuals crafted by our very own @samwho.dev. Discover exactly what gets cached to offer you 10x cheaper input tokens. ngrok.com/blog/prompt-...
1243
ngrok @ngrok.com · 15/12/2025
One gateway for every AI model. Change your baseURL and we handle routing, failover, key rotation, and more. Early access is open—and we're building the roadmap with you. 🔗 ngrok.com/blog/ngrok-a...
011
ngrok @ngrok.com · 12/12/2025
At ngrok, we embedded the OWASP Core Rule Set and Coraza into ngrok's own Traffic Policy engine, and then dogfooded it across 300M+ requests to ngrok.com. 🛡️ In this post, Ben Chan walks you through how we built (and battle-tested) ngrok’s WAF! ⚔️ 🔗 ngrok.com/blog/ngrok-w...
010
ngrok @ngrok.com · 17/11/2025
Shape #1: The Database Gateway Gateways aren’t just for APIs. This one lives between your users and the database you need to (securely) put online, offering mTLS, rate limiting, query filtering, and secrets support, all at the edge. 🔗 ngrok.com/docs/univers... 1/5
010
ngrok @ngrok.com · 17/11/2025
To celebrate the many shapes of gateways (🔗 RE: ngrok.com/blog/api-gat...), this week we'll walk through a few of our favorite gateway patterns; how they work, what they solve, and where ngrok fits in each one!
ngrok Documentation: The many shapes of API Gateways and where ngrok fits in
031
ngrok @ngrok.com · 31/10/2025
If your secrets live on a platform like HashiCorp Vault, AWS Secrets Manager, or Google Secret Manager, you may want to sync them into ngrok Vaults via the External Secrets Operator so your Traffic Policies reference a single, rotated source of truth. 🔑 🔗 Check: ngrok.com/blog/kuberne...
ngrok Blog: Sync secrets from external sources to ngrok with Kubernetes External Secrets
011
ngrok @ngrok.com · 31/10/2025
Last week’s AWS outage took down half the internet. ngrok stayed up, and a bunch of people asked "HOW?" 🤔 Our Senior Customer Engineer, Peter Yoakum, covers just that in the latest on our blog. 🕸️ 🔗 Read here: ngrok.com/blog/dont-us...
ngrok Blog: Why didn't ngrok go down in last week's AWS outage?
000
ngrok @ngrok.com · 27/10/2025
“The AI gold rush has led to a very modern problem: too many shovels, too little gold.” 🪏 AI gateways are how teams keep their models, costs, and chaos under control. 💸 Here’s a guide on how they work and whether you actually need one... 🔗 ngrok.com/blog/ai-gate...
ngrok Blog: What are AI gateways, and do you even need them?
000
ngrok @ngrok.com · 27/10/2025
Start from scratch or skip straight to simple. 💯 In this blog post, @samwho.dev helps you build up the moving parts of self-hosting your app first, then shows how ngrok makes it all effortless. 🔗 Read here: ngrok.com/blog/self-ho...
ngrok Blog: Self-hosting with and without ngrok
020