Will Kurt @willkurt.bsky.social · 13/02/2026An experiment in “AI dreaming” Starting with an initial image, 5 sec segments are created prompted by another AI and then stitched together! Generated entirely with open models (including the code to automate the process) 010
Will Kurt @willkurt.bsky.social · 25/11/2024The current messiness around LLM evaluations is ultimately caught up in the limits of working under conditions of pure empericism. We’ll never dig ourselves entirely out of this hole until theory starts to catch up with practice. Paper after paper overreaches and attempts impossible general claims 150
Reposted by Will Kurtsaganite.bsky.social @saganite.bsky.social · 13/11/2024LLM observation of the day: I think that guided/constrained generation gets a bad rap. There was one paper making the rounds about how guided generation harms reasoning ability that everyone took as gospel. 252
Reposted by Will Kurt.txt @dottxtai.bsky.social · 21/11/2024A new paper, "Let Me Speak Freely" has been spreading rumors that structured generation hurts LLM evaluation performance. Well, we've taken a look and found serious issue in this paper, and shown, once again, that structured generation *improves* evaluation performance! 412323
Reposted by Will Kurt.txt @dottxtai.bsky.social · 21/11/2024Our new blog post is out! @willkurt.bsky.social provides a rebuttal for a reasonably well known paper which concluded that structured generation with LLMs always resulted in worse performance. We do not find the same thing. blog.dottxt.co/say-what-you... 1137
Will Kurt @willkurt.bsky.social · 12/11/2024First post! Created this account awhile ago, but things seem to be picking up and it has a very nice "old Twitter" feel to it here! 030