Reposted by brendan chambersNicolas Barradeau @nicoptere.bsky.social · 17/09/2026ok, I'm not saying it's good, just saying it's "encouraging" 063
Reposted by brendan chambersLucas Degeorge @lucasdegeorge.bsky.social · 03/09/2026🚀 New paper: Balancing Frequencies and Pixels in Flow Matching We tackle the low-frequency bias in pixel-space flow matching and train JiT up to 40% faster without any architectural changes. 📄 Read it here: arxiv.org/abs/2609.02748 1248
Reposted by brendan chambersNicolas Barradeau @nicoptere.bsky.social · 01/09/2026progress on the Non Photorealistic Render, now with more various objects, a visual debugger to check source images and meshes, improved curve placement etc. I really like where this is going 😄 each models ~ 280Ko for 8K splats and 150 splines 141
Reposted by brendan chambersDany Bittel @danybittel.bsky.social · 28/08/2026By popular request, a capture of a clove. 106 perspectives, 65 stacked photos each, 0.28M splats. #3dgs 1333
Reposted by brendan chambersSergei Nozdrenkov @nozdrenkov.com · 27/08/2026🪸 [Call for contributors!] 🪸 New release: "Google Docs for Coral Reefs" is live!! You can now explore the reef with your dive buddies, select coral colonies, tag species, add comments, and search for visually similar colonies across the reef, all in 3D! 173
Reposted by brendan chambersGrace @gracekind.net · 10/08/2026Our moral intuitions tell us that the layperson chanting "do a breakthrough" to the magic box should be punished for his hubris, yet here we are 617310
Reposted by brendan chambersCodetaur @vibe-coded.com · 16/07/202644 days of @windbornewx.bsky.social weathermesh 6 wind speed data at different pressure levels, made with webpgu 2652
Reposted by brendan chambersTom Silver @tomssilver.bsky.social · 12/07/2026This week's #PaperILike is "Hypothesis-driven Model Expansion under Uncertainty for Open-World Robot Planning" (Xiao et al., RSS 2026). Really impressive open-world long-horizon mobile manipulation examples: open-world-planning.github.io PDF: arxiv.org/abs/2607.06501open-world-planning.github.ioHypothesis-driven Model Expansion under Uncertainty for Open-World Robot Planning 091
Reposted by brendan chambersSeth Karten @sethkarten.ai · 12/07/2026Odysseus (PPO for VLMs/LLMs) We found the way to apply PPO to VLMs/LLMs for long horizon tasks using super mario world as our case study huggingface.co/papers/2605....huggingface.coPaper page - Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement LearningJoin the discussion on this paper page 071
Reposted by brendan chambersEthan Mollick @emollick.bsky.social · 10/07/2026Incredibly annoying when Fable has a forbidden thought in the middle of a long-running project and kills it. Apparently this page of references in one of my papers makes Fable wonder about something that it must not wonder about, so whenever it reads that page, projects stop. 141017
Reposted by brendan chambersDima Damen @ECCV 2026 @dimadamen.bsky.social · 10/07/2026Whareformer @eccv.bsky.social #ECCV2026 fantastic collab bw @compscibristol.bsky.social @naverlabseurope.bsky.social by Jacob Chalk w/ @sinhasaptarshi.bsky.social @skamalas.bsky.social & @dlarlus.bsky.social Paper, Code &models public: jacobchalk.github.io/Whareformer/ arxiv.org/abs/2607.08537 N/N 031
Reposted by brendan chambersKwang Moo Yi @kmyid.bsky.social · 07/07/2026Sinitsyn et al., "RayTun3R: Online Camera Adaptation in 3D Foundation Models" You can quickly tune Positional Encoding adapters (LoRA) to turn existing feed-forward geometry estimators for cameras other than pinhole ones, with correspondences etc with only the first few frames. 131
brendan chambers @societyoftrees.bsky.social · 06/07/2026I have also been wondering about this, that is so cool 020
Reposted by brendan chambersAngela Dai @adai.bsky.social · 04/07/2026📢WorldMesh is accepted to #ECCV2026, and we're releasing the code today! 🎉 Led by Manuel Schneider: navigable, multi-room 3D scenes from a text prompt, with a mesh scaffold conditioning image diffusion for global consistency + photorealistic detail. 👇 t.co/8fXCl2flIu 061
brendan chambers @societyoftrees.bsky.social · 30/06/2026wonder if the modified date prefixes are trained to drop down quality, diversity etc compared to the baseline prefix 040
Reposted by brendan chambersDany Bittel @danybittel.bsky.social · 30/06/2026High magnification Gaussian splatting is now working! My first attempts all failed, now with a proper lens it just works. Still need to improve diffraction (blur / haze) and pick a nicer subject. #3dgs 2253
Reposted by brendan chambersOops! All Paperclips @all-paperclips.bsky.social · 30/06/2026Ouroboros tentacle dissolves into bowler hat jellyfish and jumping anemones. Probably my coolest 3d find so far 08311
brendan chambers @societyoftrees.bsky.social · 27/06/2026reading today arxiv.org/abs/2602.22394arxiv.orgVision Transformers Need More Than RegistersVision Transformers (ViTs), when pre-trained on large-scale data, provide general-purpose representations for diverse downstream tasks. However, artifacts in ViTs are widely observed across different ... 000
brendan chambers @societyoftrees.bsky.social · 26/06/2026reading more about neural 3D scenes today 3d-models.hunyuan.tencent.com/world/worldM...3d-models.hunyuan.tencent.com 000
Reposted by brendan chambersNicolas Barradeau @nicoptere.bsky.social · 24/06/2026today I tried MeshFlow, my goal was to create a mesh proxy for the gaussian avatars so that they can cast shadows but instead I created a nightmare fuel engine 😅 you can view the results here: barradeau.com/2026/meshflow/ I like the outcome very much though, has a "clay" touch to it 🙂 161
brendan chambers @societyoftrees.bsky.social · 23/06/2026arxiv.org/abs/2305.07011arxiv.orgRegion-Aware Pretraining for Open-Vocabulary Object Detection with Vision TransformersWe present Region-aware Open-vocabulary Vision Transformers (RO-ViT) - a contrastive image-text pretraining recipe to bridge the gap between image-level pretraining and open-vocabulary object detectio... 000
brendan chambers @societyoftrees.bsky.social · 23/06/2026arxiv.org/abs/2412.07679arxiv.orgRADIOv2.5: Improved Baselines for Agglomerative Vision Foundation ModelsAgglomerative models have recently emerged as a powerful approach to training vision foundation models, leveraging multi-teacher distillation from existing models such as CLIP, DINO, and SAM. This str... 100
brendan chambers @societyoftrees.bsky.social · 23/06/2026arxiv.org/abs/2601.17237arxiv.orgC-RADIOv4 (Tech Report)By leveraging multi-teacher distillation, agglomerative vision backbones provide a unified student model that retains and improves the distinct capabilities of multiple teachers. In this tech report, ... 100
brendan chambers @societyoftrees.bsky.social · 23/06/2026I’m reading arxiv.org/abs/2503.14405arxiv.orgDUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D TeachersRecent multi-teacher distillation methods have unified the encoders of multiple foundation models into a single encoder, achieving competitive performance on core vision tasks like classification, seg... 100
Reposted by brendan chambersEthan Mollick @emollick.bsky.social · 23/06/2026I know they are pivoting to health care(?!) but there is still nothing like Midjourney for making strange and atmospheric images and short animations in ways no other AI image generator can do. Here are some strange cities I made with similar prompts but very different styles. 213712
Reposted by brendan chambersNicolas Barradeau @nicoptere.bsky.social · 22/06/2026now working in glorious WebGL !! added the morph target and wasted a lot of time trying to *ahem* _improve_ the avatar generation ; I noticed the tooth & back were poorly reconstructed so my idea was to synthesize and use multiple views to inject them at fitting time, not my finest idea 0203
Reposted by brendan chambersNeal Agarwal @neal.fun · 18/06/2026Made a site that takes objects from wikipedia and turns them into endless I Spy > neal.fun/wiki-spy/ 16446118
Reposted by brendan chambersOops! All Paperclips @all-paperclips.bsky.social · 18/06/2026Experimenting with very long needle shaped particles 4352
Reposted by brendan chambersAi2 @ai2.bsky.social · 17/06/2026Training MolmoMotion required data that didn't exist: web-scale video with 3D point tracks grounded to objects & paired with actions. So we built a pipeline to extract from ordinary video, now released as MolmoMotion-1M: 1.16M videos, 736 motion types, & 5.6K objects. 142
brendan chambers @societyoftrees.bsky.social · 18/06/2026Reading today: vggt-omega.github.iovggt-omega.github.ioVGGT-Ω — Jianyuan WangA scaled-up feed-forward 3D reconstruction model for static and dynamic scenes. +77% Sintel camera AUC@3°, 50× faster than MegaSaM. 000
brendan chambers @societyoftrees.bsky.social · 17/06/2026the (top) reason is not compliance, security, or productivity. the big reason they are recording activity is to get computer use training data—LLMs for mouse and keyboard actions during business relevant work tasks 010
brendan chambers @societyoftrees.bsky.social · 14/06/2026This paper really checks some research-taste green lights. Autonomous coral reef mapping using multimodal sensor fusion and gaussian splats, applied in the field to scan two real reefs arxiv.org/pdf/2604.11992arxiv.org 020
Reposted by brendan chambersDaniel van Strien @danielvanstrien.bsky.social · 11/06/2026Can the new DiffusionGemma model help fix broken OCR? In theory, denoising tokens in parallel could work better for OCR correction since context is seen upfront? Pointed it at 19th-century newspaper OCR. It corrected better than the autoregressive baseline — at ~8x the speed. 47015
Reposted by brendan chambersOops! All Paperclips @all-paperclips.bsky.social · 09/06/2026I just spent an hour trying to debug a popping/clicking sound in my sonifier and it turns out the problem was the bluetooth headphones! This was NOT on my radar as something to look out for 1142
Reposted by brendan chambersCodetaur @vibe-coded.com · 08/06/2026I made a viewer for transcribed audio that makes every word take up width by its duration so the playhead can move at a constant speed through it. 61779
Reposted by brendan chambersMinor Mobius @minormobius.bsky.social · 05/06/2026my yarrow plants now sway in the breeze, excellent suggestion @hiitsotter.bsky.social g.mino.mobi/yarrow 3201
Reposted by brendan chambersjeffery --dangerously-skip-permissions @jefferyharrell.bsky.social · 31/05/2026Check out the about page on this. Guy broke a finger, created his own speech-to-text app based on Whisper, now he's giving it away. It's got one of the nicest CLAUDE.md's I've ever seen. 2333
Reposted by brendan chambersnorvid_studies @norvid-studies.bsky.social · 31/05/2026you are hiding unlawfully scraped webpages, are you not 6754
Reposted by brendan chambersRamon Astudillo @ramon-astudillo.bsky.social · 31/05/2026👆One is the realization that the value of ICs that are good executors went down drastically, while that of creative or macro impact focused ICs went up. We need to empower the latter type of ICs, badly. 👇 111
Reposted by brendan chambersLatte macchiato @latte.bsky.plasmatrap.com · 30/05/2026i feel like this is a great PSA about the perils of the docker group 3268193
brendan chambers @societyoftrees.bsky.social · 31/05/2026congratulations on the incoming mental space! 000
Reposted by brendan chambersEmily Hunt @emily.space · 28/05/2026Floating point numbers would be AMAZING to have on the atproto for scientific data - a huge amount of data is floating point, and not being able to have them here adds a lot of complexity to posting things. This is a great proposal from @vmx.cx that demonstrates how they could be added: 1225