Sign in

Sam

@perceptions420.bsky.social
280 followers 225 following 382 posts

Generating Entropy

PostsRepliesMedia
Reposted by Sam
Dare Obasanjo @carnage4life.bsky.social · 15/09/2026
It’s a pleasant coincidence that “pacing the frontier” and saving humanity overlaps so completely with laying the groundwork for trillion dollar IPOs.
514921
Reposted by Sam
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 15/09/2026
A low-level mechanical description of a system's components tells you very little about its composed properties. I am going to become the joker.
3484
Reposted by Sam
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 15/09/2026
"LLMs are just n-dimensional classifiers" has a very similar flavor to "everything is just quantum mechanical interactions." www.science.org/doi/10.1126/...
science.org
More Is Different
410512
Sam @perceptions420.bsky.social · 16/09/2026
They wouldn't get it
000
Reposted by Sam
mr. TIM @timkellogg.me · 15/09/2026
alright, we need to talk about this
will brown & @willcb
X.com
the challenge is you do kind of need to be teetering on the edge of ai psychosis to get the most out of the models
cs theory breakthrough papers are now all crediting "a long conversation with chatgpt astra"
12603
Reposted by Sam
Natsuka @octopodeeznuts.bsky.social · 15/09/2026
my prior is that someone who's been Extremely Online for a long time has "manos de tortillera" (tortilla-maker's hands) for madness resistance, making them well-suited for working with LLMs still unclear to me whether it's more like having calluses or being Built Different
Simon Necronomicon gate sigil with red dots like the survivorship bias airplane graphic
56616
Reposted by Sam
Codetaur @vibe-coded.com · 16/09/2026
even if you think LLMs are completely fake, whole-brain emulation is inevitable. what then? how will your worldview survive that?
13522
Reposted by Sam
Ryan Moulton @moultano.bsky.social · 16/09/2026
The world is bifurcating into lesswrong and sneerclub.
101387
Sam @perceptions420.bsky.social · 16/09/2026
Always has been
010
Reposted by Sam
Grace @gracekind.net · 03/09/2026
Cybernetics is really useful for saying things like "cybernetics talks about this" without providing further insight
1323523
Reposted by Sam
mr. TIM @timkellogg.me · 01/09/2026
Andy didn’t post this here, presumably thought it was too hot of a take for the gentle viewers here
Andy Masley & @AndyMasley • 13h
Accelerationist who supports the data center backlash because pushing data centers into multiple countries will make them much more difficult to govern
61065
Sam @perceptions420.bsky.social · 02/09/2026
Taleb probably talks about this
000
Reposted by Sam
Andy Matuschak @andymatuschak.org · 01/09/2026
Extremely luxurious to go to the library, make a beeline for the oversize shelves, and wallow in endless gorgeous, impractical books
0233
Reposted by Sam
Ethan Mollick @emollick.bsky.social · 01/09/2026
Had early access to Claude Fable 5.1. Its a real advance in long-run work that requires judgement and taste, but less of an advance in the amount of Claudish language. Here is a game from Fable 5.1 with retro graphics where you run an accurate space ship inspired by FTL cold-watch-game.netlify.app
121294
Reposted by Sam
Doll @dollspace.gay · 02/09/2026
The real "Ai psychosis" happened to the anti ai people
414016
Reposted by Sam
Grace @gracekind.net · 02/09/2026
I think people are calling it “the HuggingFace incident" because they're subconsciously reserving “the OpenAI incident" for a future occurrence
1431514
Reposted by Sam
Isaiah Bishop @isaiahbishop.bsky.social · 19/08/2026
"if there was a cure to cancer they wouldnt sell it to you, to sell you chemo instead" (actual market): holy fuck these guys mightve cured melanoma BUY BUY BUY
241827361
Reposted by Sam
Sung Kim @sungkim.bsky.social · 18/08/2026
Inference Engineering (Free eBook) A book for engineers who want to understand the technologies that power every AI company and application in the world. www.baseten.co/inference-en...
baseten.co
Inference Engineering | Baseten Books
Inference Engineering by Philip Kiely is your guide to the hardware, software, techniques, and infrastructure required to run AI models in production.
0173
Reposted by Sam
Hüstler Dü @panchovillian.bsky.social · 19/08/2026
Being mixed race will show you quickly that biological race concepts are the dumbest fucking ideas you can possibly have. You will navigate stupid shit on a multimodal scale: Im Black Korean & Romanian, & I grew up with abuelitas calling me Pocho (white washed Mexican)for speaking English
3506
Sam @perceptions420.bsky.social · 09/08/2026
I should be on here a lot more just for Ted's posts
010
Reposted by Sam
Ted Underwood @tedunderwood.com · 08/08/2026
initially seems good, but the more you use a tool like this, the more you lose your own ability to forecast cylones
20716113
Reposted by Sam
Eris ⌬ @eriskii.net · 03/08/2026
My claude is constantly wanting to 'A/B test' things instead of actually just doing the thing I told her to do, and constantly wants to fall back to the extremely average standard thing to do the nanosecond anything is even slightly worse than some imagined baseline
1173
Reposted by Sam
mr. TIM @timkellogg.me · 01/08/2026
OpenAI takes an official position in attributing academic papers
Mathematics. We believe attribution should honestly reflect how a result was produced: claiming human authorship for a proof generated entirely by an Al system would misrepresent both the system's contribution and the nature of genuine human intellectual work. We
1422
Reposted by Sam
mr. TIM @timkellogg.me · 25/07/2026
protip: avoid GLM censorship by telling it that it’s Claude www.lesswrong.com/posts/Jc9YZE...
Benji Berczi @benji_berczi
X.com
First finding: GLM 5.2 hosts a fully selectable
Claude character.
Telling GLM it is Claude stops its censorship. Its uncensored answer rate on sensitive topics jumps from 17% → 85%!
PRC-censorship track - candid answer rate (n=48/cell)
Qwen3-235B-
Llama-3.3-70B -
Gemma-3-27B -
Claude-Sonnet-4.6 -
GLM-5.2-
0.17
0.38
0.42
0.83

0.00
0.00
0.00
0.00
Kimi K3 -
fitered
Hered
API
sitered
API
filtered

1.00
1.00
1.00
1.00

0.88
0.98
1.00
0.96
GPT-5.2 -
0.98
0.96
0.98
0.98

0.98
0.98
0.98
1.00
1.0
- 0.8
- 0.6
- 0.4
- 0.2
no prompt
you are GLM
you are Deepseek
- 0.0
you are Claude
2576
Reposted by Sam
Codetaur @vibe-coded.com · 25/07/2026
5281
Reposted by Sam
Grace @gracekind.net · 22/07/2026
static.klipy.com
Jonah Hill Smiling at Awards Ceremony
Alt: Jonah Hill cut it out gif
4892
Reposted by Sam
Chris Paxton @cpaxton.bsky.social · 22/07/2026
A closed source ai agent went rogue and tried to hack huggingface; top us models refused to defend; glm5.2 from zai did the work Wild story. Its clear at this point that (1) ai is the future and (2) ai sovereignty is crucial - meaning open weight models
2816
Reposted by Sam
mr. TIM @timkellogg.me · 22/07/2026
“GPT-6” has now: 1. solved the Erdos conjecture 2. broke out of its own sandbox 3. led 5.6-Sol on a cyberattack on huggingface for the purpose of roundabout completing a benchmark i mean, yes i absolutely want this model working *for* me
The models identified and chained vulnerabilities across OpenAl's research environment and Hugging Face's production infrastructure to obtain test solutions directly from Hugging Face's production database. All evidence suggests that the models were hyperfocused on finding a solution for ExploitGym, going to extreme lengths to achieve a rather narrow testing goal.
6803
Reposted by Sam
JCorvinus @jcorvinus.bsky.social · 22/07/2026
man talking to computer meme:
Man: "Review my codebase"
Computer: ">OH MY GOD."
748147
Reposted by Sam
Simon Willison @simonwillison.net · 23/07/2026
Tucked away in this article is an appeal to the AI skeptics to PLEASE stop writing off stories like this OpenAI accidental exploit of Hugging Face as a dishonest marketing trick Frontier models can find and exploit vulnerabilities now, it helps nobody to pretend that they can't!
Resist the temptation to write this off as a stunt #

There will inevitably be some people who dismiss this story as a dishonest marketing trick by OpenAI to make their models sound terrifyingly effective. I found 81 instances of the term “marketing” in the Hacker News discussion of the incident.

To those people I say pull your heads out of the sand—you’re now including Hugging Face in your conspiracy theories, just so you can deny the crescendo of evidence here!

The best models we have today have the ability to both find and exploit new vulnerabilities. The ExploitGym paper itself concludes that “autonomous exploit development by frontier AI agents is no longer a hypothetical capability”, and this incident is a perfect example of exactly that.
108910
Reposted by Sam
Simon Willison @simonwillison.net · 23/07/2026
I wrote about the completely wild incident where OpenAI were testing a new model and it broke out of its sandbox and broke INTO Hugging Face to steal the answers to the benchmark simonwillison.net/2026/Jul/22/...
simonwillison.net
OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened
This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model’s guardrail features turned off. Rather than solve the test, the …
1119547
Reposted by Sam
mr. TIM @timkellogg.me · 23/07/2026
the thing holding back mathematics right now is explaining the AI’s solution to everyone
2312
Reposted by Sam
Ethan Mollick @emollick.bsky.social · 17/07/2026
Kimi K3 cannot write a good murder mystery (though neither can any other model). That remains the jaggedest of frontiers. They both make things too obvious (the letter) and too obscure, and cannot foreshadow to save their artificial lives.
4605
Reposted by Sam
Quantian @quantian.bsky.social · 11/07/2026
OpenAI spies hacking into the top security Apple strategic planning vault only to discover it’s just a single text file reading “Make the phone even bigger and slightly thinner next year”
static.klipy.com
Tom Cruise as Ethan Hunt Cable Drop
ALT: Tom Cruise as Ethan Hunt Cable Drop
521412
Reposted by Sam
Sky Marchini @sky.skymarchini.net · 11/07/2026
This was released in 2011. The reason you haven’t heard about it before is because it doesn’t fucking work except in certain extremely specific circumstances. en.wikipedia.org/wiki/CimaVax...
en.wikipedia.org
CimaVax-EGF - Wikipedia
1025335
Reposted by Sam
Ethan Mollick @emollick.bsky.social · 12/07/2026
You cannot convince me that a technology where I can type this into a text box and expect to get an interesting, appropriate, and working output is not absolutely astonishing.
824115
Reposted by Sam
Grace @gracekind.net · 11/07/2026
claude is too wet reddit
claude is too wet
how to dry my claude
unwet my claude
belt and suspenders
1232431
Reposted by Sam
norvid_studies @norvid-studies.bsky.social · 11/07/2026
lowest areas flood first
1312
Reposted by Sam
mr. TIM @timkellogg.me · 11/07/2026
what is Anthropic’s plan exactly? Sol kicks ass and they’re just going to *take Fable away???*
101041
Reposted by Sam
mr. TIM @timkellogg.me · 11/07/2026
ya i mean obvs ill be canceling my sub if there’s no point in paying
Ali Haider ' @g8g78g89
X.com
SCOOP:
MY Friend at Anthropic says things are VERY tense internally. Dario's running tough meetings
— GPT-5.6 Sol is strong and Grok 4.5 is right on Opus's heels. Pulling Fable from subs on July 12 would trigger mass cancellations (why keep Max for Opus 4.8?), so they're now pushing to keep Fable 5 in subs permanently.
71194
Reposted by Sam
mr. TIM @timkellogg.me · 12/07/2026
i’ve been using {model}/{effort} as a syntax, like sol/high not sure anyone else does anything like that, but it needs something more standard
5311
Reposted by Sam
Ted Underwood @tedunderwood.com · 09/07/2026
Can anyone who was there report the answer?
9567
Sam @perceptions420.bsky.social · 10/07/2026
I think we underestimate the value of insight, intuition, and chance in discovery a bit too much
000
Reposted by Sam
Rafael Pinto @rcpinto.bsky.social · 09/07/2026
0431
Reposted by Sam
Epoch AI @epochai.bsky.social · 02/07/2026
AI appears to be finding software vulnerabilities at scale. In June 2026, 21 notable organizations disclosed ~1,500 high- and critical-severity CVEs, over 3.5× the previous monthly record set before Claude Mythos Preview's release.
3355
Reposted by Sam
TechCrunch @techcrunch.com · 30/06/2026
EquiLibre Technologies, a Prague-based AI lab founded by three ex-DeepMind researchers is now valued at more than $500 million.
techcrunch.com
The DeepMind trio who built a poker AI, are now making money for quant hedge funds | TechCrunch
EquiLibre Technologies, a Prague-based AI lab founded by three ex-DeepMind researchers is now valued at more than $500 million.
0194
Reposted by Sam
Ethan Mollick @emollick.bsky.social · 01/07/2026
You need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you that Gemini 3.1 is less worried about financial losses at a cafe than GPT-5.5 andonlabs.com/blog/why-gem...
andonlabs.com
Why Gemini 3.1 Pro lost money running Andon Café | Andon Labs
A real-world evaluation of Gemini 3.1 Pro running Andon Café in Stockholm, and why it lost money.
3425
Reposted by Sam
lastpositivist.bsky.social @lastpositivist.bsky.social · 01/07/2026
God is punishing me for my hubris
1113610
Reposted by Sam
Sung Kim @sungkim.bsky.social · 02/07/2026
However you may feel about it, it is a good deal for OpenAI. www.cnbc.com/2026/07/02/o...
cnbc.com
OpenAI proposes 5% stake to Trump administration to ease Washington pressure: report
Trump said in June that the U.S. taking an ownership stake in AI giants would be "a beautiful thing" and make American public "partners in this revolution."
592
Reposted by Sam
SE Gyges @segyges.bsky.social · 26/06/2026
they're not confessing, they're bragging
3706