Sign in

magnesit

@magnesit.bsky.social
102 followers 56 following 1K posts

llm go brrrr

PostsRepliesMedia
magnesit @magnesit.bsky.social · 20/08/2026
People who oppose or hinder mRNA research are literally traitors against humanity, and today showed just that again
000
magnesit @magnesit.bsky.social · 18/08/2026
The Pixel 10 is a surprisingly shitty smartphone. I bought it, coming from my Pixel 7a, and it's genuinely not an upgrade. Can absolutely not recommend. Battery life is, at best, mid for a newly bought phone with almost no apps installed, and 128 GB of storage is genuinely tight for the price.
100
magnesit @magnesit.bsky.social · 02/08/2026
Reinforcement Learning From Human Feedback
000
magnesit @magnesit.bsky.social · 14/07/2026
Why does it have that Starmer look on the eyes
000
magnesit @magnesit.bsky.social · 09/07/2026
Hmm. Reading the numbers, Grok 4.5 seems to be an absurdly good model. And relatively cheap. The high price you pay isn't in dollars, though... Credit where credit is due
100
magnesit @magnesit.bsky.social · 09/07/2026
What the fuck, EU. Losing respect at a record pace. You know you'll never be able to get this through in practice. What the fuck.
100
magnesit @magnesit.bsky.social · 24/06/2026
Listening to Ribs by Lorde when a cool breath of air streams through your window while laying on the bed in the heat is corny but peak as fuck and you ain't gonna change my mind
010
magnesit @magnesit.bsky.social · 17/06/2026
Can Veritasium talk about a different topic than Prime Numbers for more than 2 consecutive months ffs 😭
000
magnesit @magnesit.bsky.social · 16/06/2026
Mistral can do the funniest thing now... Pleeeease let them release Le Chaton Fat
000
magnesit @magnesit.bsky.social · 06/06/2026
Gotta love these GitHub accounts that fork every repository in existence, do nothing to it and let leave, refusing to elaborate Absolute cinema
000
magnesit @magnesit.bsky.social · 04/06/2026
lol it's happening, GitHub copilot is pulling the rug so hard
010
magnesit @magnesit.bsky.social · 16/05/2026
Sorry but can anyone confirm this? They released more models than ever last month???
010
magnesit @magnesit.bsky.social · 12/05/2026
How does this keep happening? Bluesky is a platform full of very loud anti-AI people, but still, it's the single platform that drowns in slop and and everybody keeps liking these images. Again, what do you want, man? How does this keep happening?
100
Reposted by magnesit
J.L. Worrad @worrad.bsky.social · 05/05/2026
252400303
magnesit @magnesit.bsky.social · 02/05/2026
Something's odd about the launch of Mistral Medium 3.5, and it's not their fault. Artificial Analysis paid close to $1k to benchmark the model but still defaults to showing Mistral Small 4, a much worse model, on their front page. Medium 3.5 scores way better (Deepseek V3.2 level).
100
magnesit @magnesit.bsky.social · 29/04/2026
The new pricing of Medium 3.5 suggests Mistral is in need of money, probably because of failed training runs or low demand. They were able to provide Devstral 2 (123B dense) at a much lower cost than Medium 3.5 (128B dense), even though it should effectively be almost identically heavy on hardware.
000
magnesit @magnesit.bsky.social · 29/04/2026
Mistral dropped the bomb: a 128B DENSE Medium 3.5. They indeed seemed to take a Devstral-like architecture (same hidden size, attn heads/dims, among other things), maybe even literally a Devstral checkpoint, and added a vision encoder on top. But not just that: They added reasoning 1/5
100
magnesit @magnesit.bsky.social · 29/04/2026
Gemma 4 E2B reaches a whopping 620 tok/s prefill speed on the dated Tensor G4 GPU! This is truly impressive. Shows that simply interleaving sliding window attention with full attention is still a worthy technique. Wouldn't have guessed that, to be honest.
000
magnesit @magnesit.bsky.social · 29/04/2026
White background; title text: "When your mom calls you by your full name". Qwen logo in the lower-middle left. Speech bubble coming from the right, saying "Tongyi Qianwen!" Qwen saying "Shit".
000
magnesit @magnesit.bsky.social · 28/04/2026
New open-weights, likely multimodal+dense Mistral Medium model (128B) inbound for tomorrow, hopes are high as always! It's still unclear whether the model has reasoning baked in, as you can't see that from the architecture alone, but if it has, it could become a real banger!
110
magnesit @magnesit.bsky.social · 25/04/2026
Deepseek V4 finally released today, if you haven't noticed yet. I've spent some time reading their technical report and benchmarks and my feelings are extremely mixed. I am full of thoughts, but I feel like it's too early to pinpoint most things for V4 yet.
100
magnesit @magnesit.bsky.social · 23/04/2026
Peak deutscher Humor ist so fucking cursed dass er für jeden, der die Sprache im Nachhinein lernt, der reinste Fiebertraum sein muss, ich liebs
000
magnesit @magnesit.bsky.social · 23/04/2026
I am, by no means, a Trump supporter. I was shattered when I saw Kamala lose to this guy. But we shouldn't give up clean research. That said, is there any credible source for the claim on the image? I haven't been able to find anything. Genuinely curious.
020
magnesit @magnesit.bsky.social · 23/04/2026
This is extremely ironic. Like, comically ironic.
020
magnesit @magnesit.bsky.social · 23/04/2026
I usually don't post these "hype"-ish posts, but I need to say it: Qwen 3.6 35B A3B is genuinely impressive. I'm able to run it quantized to Q4_K_M @ 50+ tok/s on my RTX 4070 Ti (12 GB VRAM) with FULL (!) 256k context length. It's smart and actually useful. Local models are evolving really quickly.
200
magnesit @magnesit.bsky.social · 20/04/2026
Reasoning LLMs are the biggest compromise in machine learning - converting all those rich, contextual embeddings into discrete token IDs and then feeding them back. Awful. I'm confident "they" will come up with something more efficient soon. Please.
110
magnesit @magnesit.bsky.social · 15/04/2026
I don't like the chains of thought of the newer open-weight LLMs on the market. They just don't try to be efficient anymore. I know, it's supposed to be more structured and stuff, but I think leaving all the distillation artifacts from bigger models like the one marked in the image is unacceptable.
Image depicting a chain-of-thought of Gemma 4 E2B generating "Thinking process: 1. **Analyze the request**: ..."
100
magnesit @magnesit.bsky.social · 10/04/2026
I came up with a tremendous, nearly undetectable method to cheat in exams; instead of trying to smuggle in an AI assistant akin to ChatGPT to help you out, simply *remember* (!) the weights of an open, frontier LLM like, say, GLM 5.1.
110
Reposted by magnesit
Ursula von der Leyen @vonderleyen.ec.europa.eu · 08/04/2026
I welcome the two-week ceasefire the US and Iran agreed last night. It brings much-needed de-escalation. I thank Pakistan for its mediation. Now it is crucial that negotiations for an enduring solution to this conflict continue. We will continue coordinating with our partners to this end.
7239064
magnesit @magnesit.bsky.social · 18/03/2026
This is such an astronomically great thing! 🥹
100
magnesit @magnesit.bsky.social · 11/03/2026
Why in the name of fuck is Gemini 3.1 Flash Lite priced at $1.50/M output tokens? lil bro it's not that good 🥀🥀
010
magnesit @magnesit.bsky.social · 06/03/2026
I'm really, really looking forward to Deepseek V4! Let's just hope it releases soon, because the competition is evolving a lot right now...
020
Reposted by magnesit
GrapheneOS @grapheneos.org · 02/03/2026
We're happy to announce a long-term partnership with Motorola. We're collaborating on future devices meeting our privacy and security standards with official GrapheneOS support. motorolanews.com/motorola-thr...
motorolanews.com
Motorola News | Motorola's new partnership with GrapheneOS
Motorola announces three new B2B solutions at MWC 2026, including GrapheneOS partnership, Moto Analytics and more.
691036240
magnesit @magnesit.bsky.social · 10/02/2026
Gemini 3 Flash (Fast mode) is literally just a reasoning model that pretends like it's not and any comparison between instruct models is inherently unfair. Even minimal reasoning is still reasoning and you can clearly feel the difference in the response quality my opinion.
100
magnesit @magnesit.bsky.social · 30/01/2026
Graphene Is All You Need.
000
magnesit @magnesit.bsky.social · 26/01/2026
Every VLM Implementation except for Qwen's and Gemini's feels botched.
010
magnesit @magnesit.bsky.social · 20/12/2025
I'm slightly disappointed in Mistral Large 3 being solely based on a recycled Deepseek architecture with minor changes. Mistral has a lot of potential, and while trying out custom architectures is risky for smaller ML startups, it's the only way to remain independent in the long term.
210
magnesit @magnesit.bsky.social · 20/12/2025
I can't help but think that Gemini 3 Flash has been even more benchmaxxed than other models... Besides that, they waited way too with publishing Gemini 3. The models are only barely SOTA a few weeks after their initial release. What was the point of all that?!
100
magnesit @magnesit.bsky.social · 01/12/2025
Brace, Mistral might be dropping bombshell LLMs very soon. 3B, 8B, and one proprietary.
100
magnesit @magnesit.bsky.social · 02/11/2025
I know the Artificial Analysis Index can be inaccurate as hell But boy, is that a hell of a beautiful sight.
Chart showing Magistral Medium 1.2 outperforming Gemini 2.5 Flash and Deepseek R1 2508.
010
magnesit @magnesit.bsky.social · 14/10/2025
YouTube is rage baiting everyone, yet again: They finally introduced a selector on MOBILE WEB for the "Audio track" which allows you to get rid of the terrible auto-translation... BUT you can only use this for Shorts. Not for full videos. They don't have that option there. How the fuck.
Screenshot of the YouTube mobile web UI's "Options" menu on a YouTube short showing the "Audio track" selector being listed.
010
magnesit @magnesit.bsky.social · 29/09/2025
No app is as good at keeping you logged out as Bluesky. My mailbox is bursting more and more with dozens of log in mails and I'm tired of pretending it's not.
000
magnesit @magnesit.bsky.social · 07/09/2025
Okay, so Android is "de jure open source" but de facto closed source right now?? Great job, Google 😮‍💨
110
magnesit @magnesit.bsky.social · 07/09/2025
Yesterday, I talked with a more or less influential Shiftphone employee on the IFA about their phone and future support for @grapheneos.org. They said the question isn't new - and that they'd happily work on supporting GrapheneOS. 1/5
110
magnesit @magnesit.bsky.social · 04/09/2025
Had my Pixel 7a get fixed yesterday, but now I'm totally out of the loop! Did Android 16 QPR1 finally get released?? 😳 Can I expect a One UI clone in a few days?
000
magnesit @magnesit.bsky.social · 01/09/2025
My smartphone broke (Pixel 7a, fell on concrete) and there's nothing better than seeing GrapheneOS Attestation work more flawless than ever! Give it a shot at attestation.app!
Email by alert@attestation.app saying:

This is an alert for the account '...'.

The following devices have failed to provide valid attestations before the expiry time:

* ...

Log in to https://attestation.app/ for more information.

If you do not want to receive these alerts and cannot log in to the account,
email contact@attestation.app from the address receiving the alerts.
150
Reposted by magnesit
magnesit @magnesit.bsky.social · 31/08/2025
> GPT-4 I thought you said GPT-5 was so great?
media.tenor.com
a close up of a man in a suit and tie with a question mark on his neck .
Alt: Confused Joe Biden GIF
001
magnesit @magnesit.bsky.social · 28/08/2025
Wie bitte????? "Unbehagen"? No shit sherlock #NATO #Finland
Screenshot eines Artikels der Zeitung "DER SPIEGEL".

Titel: Finnlands Militär will Hakenkreuzflaggen vollständig abschaffen

Beschreibung: Das Hakenkreuz hat in der finnischen Luftwaffe Tradition, noch immer verwenden einzelne Brigaden das Symbol. Das soll sich nun offenbar ändern, wegen »Unbehagen« bei Nato-Partnern.

28.08.2025, 17.30 Uhr
120
magnesit @magnesit.bsky.social · 27/08/2025
He can't keep getting away with it.
Screenshot of post by Junyang Lin (Handle @JustinLin610) writing "Qwen".
110
magnesit @magnesit.bsky.social · 24/08/2025
HEAR ME OUT GUYS AND GALS THEY DID FUCKING COOK. lmarena.ai/leaderboard/... EU may be back in business
Screenshot from the LMArena Text Overal leaderboard (style control disabled), showing Mistral Medium 3.1 (2508) at place 2, only outperformed by Gemini 2.5 Pro.
120