Sign in

stanislavkozlovski.bsky.social

@stanislavkozlovski.bsky.social
106 followers 404 following 176 posts
PostsRepliesMedia
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 07/04/2026
@jauntywk.bsky.social you posted back in the day about SlateDB making a very fine streaming system: I wanted to notify you - it is already WIP - github.com/opendata-oss...
120
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 23/01/2026
I sat down with Denis Magda to talk about Postgres. He literally just published a book called “Just Use Postgres”. In this conversation, that turned out to be super fun, we talked about everything from what MySQL got wrong to the recent PG dev explosion Watch here👇 www.youtube.com/watch?v=LxQy...
youtube.com
just use postgres - the podcast (/w Denis Magda)
YouTube video by 2 Minute Streaming
020
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 01/01/2026
Have you heard about Tansu? Tomorrow, I will post the longest-ever podcast about it. It’s an open-source (Apache), stateless, leaderless, single-binary Kafka broker supporting pluggable storage backends (S3, PostgreSQL, SQLite) with built-in schema registry and support for open-table formats
110
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 18/11/2025
Why Kafka performs so well 👇 🧵
110
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 11/11/2025
github.com/tansu-io/tansu This is something to keep an eye on. Should be doable to add within a pg extension and bundle Kafka into Postgres?
000
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 08/11/2025
the data bible dropping soon
051
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 02/11/2025
Just use Postgres until it breaks 🧘‍♂️
010
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 26/10/2025
This was posted 15 years ago (part 1/2)
100
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 29/09/2025
**taps the sign**
000
Reposted by @stanislavkozlovski.bsky.social
austin @aparker.io · 28/09/2025
when I say “storage is cheaper now” this is what I mean topicpartition.io/definitions/...
topicpartition.io
Small Data
Small Data Small data appears to be a very exciting movement that is moving the overton window away from Big Data onto much simpler and cheaper solutions ...
23411
Reposted by @stanislavkozlovski.bsky.social
Chris @chris.blue · 21/04/2025
“KIP-1150 introduces Diskless Kafka topics that write directly to S3 instead of replicating between brokers.” “Even using the expensive S3 Express (which a week ago lowered its prices by more than 50%) still saves 73% compared to traditional Apache Kafka.” /ht @ananthdurai.bsky.social
topicpartition.io
KIP-1150 in Apache Kafka is a big deal (Diskless Topics)
TL;DR KIP-1150 introduces Diskless Kafka topics that write directly to S3 instead of replicating between brokers. It literally reduces costs by 97% (from $1.8M to $20K annually for a 1GiB/s cluster) a...
01910
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 14/07/2025
a new 2 minute streaming post is sitting patiently in your inbox... open it to learn when: • Kafka decides what messages are visible to Consumers • acks=all Producers receive responses
100
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 18/05/2025
Apache Kafka has been on a Diskless craze in the last two years: • 2023: WarpStream launched • 2024: Confluent bought them for $220M+ • 2025: Aiven published a KIP to the open source project to introduce the same type of leaderless, direct-to-S3 topics
100
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 11/04/2025
Yesterday, CloudFlare dropped a bomb that I believe may change the future of Lakehouse storage. R2 + Iceberg should become the de-facto choice for hybrid and multi-cloud data lakehouse architectures. Here's why it may break the cloud monopoly 🧵
110
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 08/04/2025
+1 to this future
000
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 21/02/2025
This is the most impactful Apache Kafka mentor you’ve never heard of: Chia-Ping Tsai. In just 18 months, he bootstrapped a large Taiwanese open source community boasting: • 5000 participants • 15,000 Slack messages/month • 10 meetings/week • 20+ Apache committers
151
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 18/02/2025
why doesn't Kafka use Protobuf or another popular serialization format? why is everything custom? afaict the decision to go custom was taken back when it was first created. I just assume we never questioned it again? ever since it's been a one way street - since upgrading clients will be a pain
000
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 14/02/2025
What if I told you that a 1 GiB/s Kafka topic streamed directly into your S3 data lakehouse as an Iceberg table could cost you... $10/hr? Bufstream does it. It's literally too good to be true. I spent 20 hours researching them. Here's their story (2 minute read) 🧵
164
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 03/02/2025
Confluent runs tens of thousands of Kafka clusters across 93 regions in 3 clouds - AWS, GCP, and Azure. How do they manage all of that? Here are 13 innovative modifications they did to run Kafka smoothly at that scale. 👇
000
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 31/01/2025
Nobody does data infrastructure like Uber does. Here are their numbers 👇
000
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 17/01/2025
a new edition of 2 minute streaming is waiting patiently in your inbox probably the simplest and fastest explanation of AWS networking costs out there on the internet (plus a gift announced at the end) 🎁 ✅ blog.2minutestreaming.com
000
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 16/01/2025
how do you read this? you pay $0.02 when you cross AZ. do you pay an extra $0.02 if it's going through a public IPv4? or does it imply you pay $0 when going cross-AZ though private ip? ipv6: do you pay an extra charge when going cross-VPC? or is it $0 cross-AZ same-VPC? 🤷‍♂️🤷‍♂️
210
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 12/01/2025
Genius is making complex ideas simple, not making simple ideas complex. Over 70% of Fortune 500 companies have used Apache Kafka. At its core, it’s just a distributed commit log. A log (a.k.a. {write-ahead, commit, transaction} log) is a simple but efficient data structure:
010
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 08/01/2025
The largest performance improvements often come the easiest. Here’s an Apache Kafka config tweak to increase your performance by 50% 🔥 (a 1-minute 🧵)
210
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 14/12/2024
man I can't get it to post here without errors, so you'll have to go on the bird app if you want to see it. But I broke down the WarpStream story in tweet format: x.com/BdKozlovski/... I also explained very explicitly what my position is regarding the situation. So before you judge - 👀
020
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 13/12/2024
The spiciest post I’ve published yet is out now. 🌶️ "The Brutal Truth About Kafka Cost Calculators" I also announced a new product that I’ve been heads down coding for the last two months. Interested? 👀 bigdata.2minutestreaming.com/p/the-bruta...
083
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 12/12/2024
I'm exposing something about the Kafka industry tomorrow. I usually don't hype these up but tomorrow's newsletter edition will be the most impactful one I've released yet. It will be spicy. It will change how you think. You don't wanna miss it. 🌶️
010
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 06/12/2024
Every data engineer is talking about S3’s new features from this reInvent. But can you remember all the others? Here is a small cheatsheet with AWS S3’s top 12 features to help you keep up👇
040
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 06/12/2024
Looking at the API, it wouldn't be that hard to extend Kafka's Tiered Storage plugin to write to S3 Tables in an Iceberg format directly. The only question would be - where do you get the topic schema from? Which makes me question... why doesn't Kafka have first-class schema support?
120
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 05/12/2024
More thoughts about the S3 Iceberg release in a new long-form piece, including: • whether we can call Iceberg the winner of the table format wars • how I calculate it 37% more expensive than regular S3 • where the lock-in may be • + lots of references! (link in reply)
110
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 04/12/2024
Yesterday, AWS shook the data lake world by releasing two new S3 features that will forever cement its place there. 👑 Every data engineer must become familiar with them. A short thread on these game-changers 🧵 (2 minute read)
120
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 29/11/2024
What people think Kafka compaction is: • unique keys ❌ What it actually is: • a retention strategy denoting how you preserve existing data 💡 A compacted topic is one where you have a more granular retention strategy. You're basically saying "retain at least one record per key".
100
stanislavkozlovski.bsky.social @stanislavkozlovski.bsky.social · 22/11/2024
Got a juicy Kafka letter ready to go out tomorrow Here's a leak:
100