Sign in

daniel

@daniel-j-h.bsky.social
25 followers 236 following 95 posts
PostsRepliesMedia
daniel @daniel-j-h.bsky.social · 14h
#jev is fun! you don't have to care about evaluating it on your domain's test set, understand how and when it fails, bias, or explainability. We all decided to yolo it this time.
100
Reposted by daniel
Will Oremus @willoremus.com · 02/10/2026
feels like there is a group of smart and well-intentioned people who long ago made up their minds that AI Is Bad, swore off using it, and now get all their info about AI from each other. not saying they're wrong that ai is bad! but i think they're gradually losing touch with what it is and does
233946104
Reposted by daniel
wariziles @riziles.bsky.social · 02/10/2026
Please stop anthropomorphizing computers. I cannot "turn on" a computer because I just don't have that dog in me.
412818
Reposted by daniel
hailey @hailey.at · 01/10/2026
fuck yea
huggingface.co
Cloudflare/clef · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
1013412
Reposted by daniel
Armin Ronacher @mitsuhiko.at · 01/10/2026
We released Pi 1.0! earendil.com/posts/pi-1-0/
earendil.com
Pi 1.0 | Earendil
Today we are shipping Pi 1.0, a hardened, minimal, extensible agent harness, alongside Pi Durable, a new experimental substrate for long-running agentic applications.
1634137
Reposted by daniel
antirez @antirez.bsky.social · 25/09/2026
I'm worried to see people that think the right move is to question AI companies and narratives around AI regardless of logic and the actual significance of events we are living. They act just as an inverted flock, but still a flock. Sometimes it's fear of friends judgement.
4192
daniel @daniel-j-h.bsky.social · 26/09/2026
LOVING SON & BROTHER CHRISTOPHER JOHN GOUINE NOV 10 1965 – MAR 27 2016 FOREVER IN OUR HEARTS AGENTWORLD V2.4 85.7% DEEPWORLD V3 87.2%
000
Reposted by daniel
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 22/09/2026
Now that LLMs are good at math, we can solve other long-standing verifiable problems like the law, institution building, community, and management
816112
Reposted by daniel
Simon Willison @simonwillison.net · 18/09/2026
Being a computer scientist who refuses to find anything about LLMs interesting right now is a bit like being a geneticist who refuses to find anything interesting about the recently opened Jurassic Park
4382182
Reposted by daniel
antirez @antirez.bsky.social · 16/09/2026
At this point DwarfStar contains many fast fused kernels for important model families: feel free to steal everything you want from there according to the MIT license, in order to improve your own implementation.
0654
Reposted by daniel
antirez @antirez.bsky.social · 14/09/2026
Qwen3.8 Flash Next is now supported in DwarfStar, covering 64GB Mac systems very well and with very fast inference of 50~70 t/s and > 1400 t/s prefill. For now this is Metal only. Thanks to @ivanfioravanti.bsky.social for all the cool work in the PR. N-grams on SSD like for DS4.1F.
3778
Reposted by daniel
Thorne 🌸 @ens0.me · 12/09/2026
DeepSeek-4.1-Flash and GLM-5.3-Flash feel way too cheap for how much of a punch they have I'm still just in shock at how efficient they are for how good they are
3765
Reposted by daniel
Adina Yakup @adinayakup.bsky.social · 10/09/2026
DeepSeek v4.1 Flash is just another level 🤯 huggingface.co/deepseek-ai/... - Asymmetric Causal-Encoder-Decoder: 550B MoE, input 8B / output 16B - Native vision merged into one endpoint - KV cache crushed: ~1/4 the HBM vs last one, 437× smaller than their first model
huggingface.co
deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
2754
Reposted by daniel
Astra ⎔ @astrra.space · 10/09/2026
but yeah they released V4.1 flash weights huggingface.co/deepseek-ai/... and a tech report huggingface.co/deepseek-ai/...
huggingface.co
deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
3754
Reposted by daniel
kira ☆ @kira.ws · 11/09/2026
the czosndog has a whole family behind it. grandmother: the zapiekanka. open baguette, mushrooms, cheese, ketchup, sold from windows on kraków's plac nowy since the 70s. communism's answer to pizza.
1172
Reposted by daniel
antirez @antirez.bsky.social · 10/09/2026
DwarfStar running DeepSeek v4.1 Flash on a 128GB M5 Max at steady 15 t/s. I didn't expect with SSD streaming it could be so fast. Recent SSD streaming changes to retain the right experts surely helped, but also maybe DS4.1 uses the same experts more. Will push online when ready QA > ASAP.
29413
Reposted by daniel
antirez @antirez.bsky.social · 06/09/2026
DwarfStar in the latest two weeks was improved in almost every aspect for Metal, DGX Spark and Strix Halo. It is simpler to say: update, you will hopefully see speed and correctness improvements in many areas. Also DSpark with DeepSeek v4 Flash now works much better overall.
2343
Reposted by daniel
codewright @codewright.bsky.social · 06/09/2026
PSA for everybody who has not fully arrived on AI-Bluesky: Follow these Starter Packs: bsky.app/starter-pack... bsky.app/starter-pack... Pin and use this feed (instead of Discover): bsky.app/profile/spac... Stay friendly and respectful, even while strongly disagreeing or criticizing ... ❤️
1385
daniel @daniel-j-h.bsky.social · 06/09/2026
I don't know what's worse - that we wake up the #AI and let it solve benchmarks - or that we humans wake up and there are simply no benchmarks
100
daniel @daniel-j-h.bsky.social · 06/09/2026
GPT-6 Astra so good it just predicted the PlayStation 6
000
daniel @daniel-j-h.bsky.social · 05/09/2026
from blog.google/innovation-a... can you spot #GLM 5.3 Flash and #DeepSeek V4 Flash 🕵️‍♂️
cost vs capability; gemini 3.8 on the pareto frontier; but also deepseek v4 flash and glm 5.3 flash on the scale being wildly cheaper than anything else
010
daniel @daniel-j-h.bsky.social · 23/08/2026
#Claude Code 🦧💡 #DeepSeek harness 🤓✨
100
daniel @daniel-j-h.bsky.social · 23/08/2026
how do #LLM #AI samplers from the last five years (e.g. top-n-sigma) behave on quantized weights / kv cache, of varying degrees? in the context of coding agents? let's read MaDiqXx99's post on /r/LocalLlama as the authorative source
020