Sign in

daniel

@daniel-j-h.bsky.social
29 followers 243 following 123 posts
PostsRepliesMedia
daniel @daniel-j-h.bsky.social · 2h
The OpenStreetMap ecosystem as a direct connection to everything around us
000
daniel @daniel-j-h.bsky.social · 14h
Is this the brain thinking about itself or something
000
daniel @daniel-j-h.bsky.social · 15h
hm something is off, I get faster decode speeds on a framework 13 amd with vulkan or rocm backend re. speculative drafter give this one a try huggingface.co/z-lab/Qwen3.... I have no experience with the nvidia backend (all running on amd here) unfortunately, just sounds very slow for nvidia hw
huggingface.co
z-lab/Qwen3.8-27B-DFlash2 · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
030
Reposted by daniel
ponder @ponder.ooo · 22h
the future of gaming is that your friend yaps at you about their favorite new game, the one their agent made just for them, tailored to their preferences, and you zone out because why would you give a shit, all you want to talk about is your favorite new game, the one your agent just made for you,
713410
daniel @daniel-j-h.bsky.social · 15h
all layers on gpu? speculative drafters set? flash attwntion on? compiled for nvidia backend? there are sooo many things that can go wrong, it takes quite a while to tune this; could try auto-research an optimal setup overnight
130
daniel @daniel-j-h.bsky.social · 16h
How does it do vs Clef Flash, as both are Qwen3.5-9B fine-tunes blog.cloudflare.com/clef-decisio... Anyone got numbers or vibes on this?
blog.cloudflare.com
Introducing Clef: our open-source decision models, and new RL fine-tuning platform
We are introducing Clef and Clef-flash, open-source decision models hosted on Workers AI for high-speed classification and agentic workflows. Also launching: a new reinforcement learning platform that...
000
daniel @daniel-j-h.bsky.social · 16h
awesome work, would the same work by fine-tuning small projection heads on top of the frozen encoders? what are the pros and cons vs this approach here?
000
daniel @daniel-j-h.bsky.social · 17h
TIL how to pronounce dinosaur
000
daniel @daniel-j-h.bsky.social · 17h
Have you seen the related work in ghostty itself gist.github.com/jake-stewart... hachyderm.io/@mitchellh/1...
110
daniel @daniel-j-h.bsky.social · 17h
the hardest part of AGI will be getting audio capture to work reliably on linux
010
daniel @daniel-j-h.bsky.social · 09/10/2026
y no vulnerability-fixing tho
010
daniel @daniel-j-h.bsky.social · 09/10/2026
let's grep and find through code bases like it's 1999
000
daniel @daniel-j-h.bsky.social · 08/10/2026
they just opened up how they did it github.com/goodreads/math
github.com
080
daniel @daniel-j-h.bsky.social · 08/10/2026
but can opus 5.5 call my parents saying that I love them
120
daniel @daniel-j-h.bsky.social · 08/10/2026
How much faster is it? How much better is it? How much money does it save? How should we accept this without any serious evaluation?
000
daniel @daniel-j-h.bsky.social · 07/10/2026
wonderful! what's latest in terms of approx. nearest neighbor searches on MaxSim / late-interaction re-ranker outputs? regular ann libs like faiss won't do, no?
100
daniel @daniel-j-h.bsky.social · 07/10/2026
have you tried the dflash speculative drafter huggingface.co/z-lab/Qwen3.... depending on your workload that should speed things up a bit more 😍
huggingface.co
z-lab/Qwen3.6-35B-A3B-DFlash · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
110
daniel @daniel-j-h.bsky.social · 07/10/2026
oof, in Germany I see it starting at ~4k EUR without any disks or extras how's the thermal regulation? On a bosgame m5 I have seen it getting limited to 80W after a few minutes as it's getting too hot that and the slowdown from large context makes ds v4 (not even 4.1) barely usable there
350
daniel @daniel-j-h.bsky.social · 07/10/2026
Also re-rankers and especially late-interaction re-rankers e.g. of the colbert kind. That's on the way from embeddings to llms but seems like not super well known?
010
daniel @daniel-j-h.bsky.social · 06/10/2026
Does that mean passing a watermarked text through another LLM and having it rephrase things a little gets rid of watermarking; while for images this will not work by design?
100
daniel @daniel-j-h.bsky.social · 05/10/2026
woah can it choice on morse code 🦧💡
110
daniel @daniel-j-h.bsky.social · 05/10/2026
$ klaus --weitermachen
160
daniel @daniel-j-h.bsky.social · 05/10/2026
Where's GLM 5.3, MiMo v2.6, DeepSeek v4.1 in those charts tomcompare. They're at roughly ~300b params. There's a table that shows some of them in comparison if you scroll all the way to the right. Why try to bend reality. This comes out sooner or later anyways.
030
daniel @daniel-j-h.bsky.social · 05/10/2026
~500b is also quite the chonker the competitors (deepseek v4 flash, glm 5.3 flash, mimo v2.6 flash) are all around ~300b moes they're amazing and wonderful for self-hosting on datacenter grade gpus (single b300 for example) or even unified memory devices
010
daniel @daniel-j-h.bsky.social · 05/10/2026
and for extra tinfoil hat points you can self-host both clef models, too! the flash one fits on a laptop or mini-pc; no dedicated gpu required
010
daniel @daniel-j-h.bsky.social · 05/10/2026
Same for vision machine learning. Everyone trained and benchmarked on ImageNet but no one asked about the licensing aspect of those images. Maaybe fine for research but at some point this stuff turned out to be actually useful but we still kept training on these images (and later also videos).
0110
daniel @daniel-j-h.bsky.social · 04/10/2026
16 GB is limiting; recent open weights models are really good but they're not that small try the gemma4 e2b, e4b, or 12b versions and see how things go the problem is the good ones like deepseek v4 flash, glm 5.3 flash, mimo v2.6 flash all require 128 GB or more while being slow to run
020
daniel @daniel-j-h.bsky.social · 04/10/2026
it makes me less scared starting a project connecting to IOT devices via BLE on Linux for sure
100
daniel @daniel-j-h.bsky.social · 04/10/2026
Fascinating to see it pulling OpenStreetMap data through Overpass. It's meant for simple queries. Interesting to see if the OpenStreetMap people see more load due to agents.
110
daniel @daniel-j-h.bsky.social · 03/10/2026
this, and also qwen3.5 is a prev. prev. generation model omg what are they doing over there
010
daniel @daniel-j-h.bsky.social · 03/10/2026
i'm just sad it's yet another polarizing topic as if we don't have enough of that already; we need more nuanced takes and less us vs them
030
daniel @daniel-j-h.bsky.social · 03/10/2026
what's the economic incentive to do any of that?
110
daniel @daniel-j-h.bsky.social · 03/10/2026
qwen3.5, gpt oss, .. guys we're in late 2026 what is this
140
daniel @daniel-j-h.bsky.social · 03/10/2026
what pareto frontier are they talking about
120
daniel @daniel-j-h.bsky.social · 03/10/2026
you're already ahead of most people out there using jev if you do this 😛
000
daniel @daniel-j-h.bsky.social · 03/10/2026
#jev is fun! you don't have to care about evaluating it on your domain's test set, understand how and when it fails, bias, or explainability. We all decided to yolo it this time.
110
daniel @daniel-j-h.bsky.social · 02/10/2026
calculators they're putting them in there
010
Reposted by daniel
Will Oremus @willoremus.com · 02/10/2026
feels like there is a group of smart and well-intentioned people who long ago made up their minds that AI Is Bad, swore off using it, and now get all their info about AI from each other. not saying they're wrong that ai is bad! but i think they're gradually losing touch with what it is and does
235966105
daniel @daniel-j-h.bsky.social · 02/10/2026
s/AI/engineers
020
daniel @daniel-j-h.bsky.social · 02/10/2026
anyone tried distilling Jev into a modernbert based classifier??
000
daniel @daniel-j-h.bsky.social · 02/10/2026
👸✨🧎🏻🧎🏻🧎🏻
010
daniel @daniel-j-h.bsky.social · 02/10/2026
what's their moat tho?
100
daniel @daniel-j-h.bsky.social · 02/10/2026
looks like clef flash is qwen35 9b from 6+ months earlier this year and I guess there's just hasn't been an upgrade in that size category unfortunately
000
daniel @daniel-j-h.bsky.social · 02/10/2026
told you they're putting calculators in there
010
Reposted by daniel
wariziles @riziles.bsky.social · 02/10/2026
Please stop anthropomorphizing computers. I cannot "turn on" a computer because I just don't have that dog in me.
413018
daniel @daniel-j-h.bsky.social · 02/10/2026
docs.sglang.io/docs/support... the speed at which things are moving is incredible, what a time to be alive
docs.sglang.io
Decision models - SGLang Documentation
Answer typed choice, score, and yes or no questions with a probability for every option from a chat model, without generating text.
000
Reposted by daniel
hailey @hailey.at · 01/10/2026
fuck yea
huggingface.co
Cloudflare/clef · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
1013412
daniel @daniel-j-h.bsky.social · 02/10/2026
poof goes the moat
100
daniel @daniel-j-h.bsky.social · 02/10/2026
pi is strategically important for open source and the open llm ecosystem let's see how things are progressing especially with commercial models getting rl-trained on their harnesses
030
Reposted by daniel
Armin Ronacher @mitsuhiko.at · 01/10/2026
We released Pi 1.0! earendil.com/posts/pi-1-0/
earendil.com
Pi 1.0 | Earendil
Today we are shipping Pi 1.0, a hardened, minimal, extensible agent harness, alongside Pi Durable, a new experimental substrate for long-running agentic applications.
1835038