Sign in

Will Smith

@willsmithvision.bsky.social
1.2K followers 567 following 92 posts

Professor in Computer Vision at the University of York, vision/graphics/ML research, Boro @mfc.co.uk fan and climber 📍York, UK 🔗 www-users.york.ac.uk/~waps101

PostsRepliesMedia
Reposted by Will Smith
Alistair Foggin @alistairfoggin.bsky.social · 23/04/2026
Very excited to be in Rio for #ICLR2026 presenting my poster for CroCoDiLight! If anyone is also here, please pop by and say hi! I'd love to chat. I'll be at poster session 2 in pavilion 4 this afternoon
151
Will Smith @willsmithvision.bsky.social · 01/04/2026
Excited to announce our latest (submitted to) SIGBOVIK 2026 @harryqbovik.bsky.social paper: "SchmidhubAI: Accurate Historical Paper Attribution". We built an AI system that, given any modern AI paper, automatically determines which of its ideas were already published by Jürgen Schmidhuber.
34510
Reposted by Will Smith
Gabby DaRienzo @gabbydarienzo.com · 06/03/2026
wait no this is actually incredible youraislopbores.me
7262118
Will Smith @willsmithvision.bsky.social · 03/03/2026
I am delighted (that pun will make sense in a second) that @alistairfoggin.bsky.social's first paper, CroCoDiLight, has been accepted to ICLR. The idea came from a group discussion on the CroCo paper from @naverlabseurope.bsky.social and realising it might implicitly already understand relighting.
191
Will Smith @willsmithvision.bsky.social · 23/02/2026
The LLM obsession with em-dashes has created a weird sort of paradox. I see so much AI generated content that I now see how I should have been using em-dashes all along. But if I start using them, everyone will assume what I've written is AI generated.
130
Reposted by Will Smith
Paul @paulpw.bsky.social · 22/01/2026
Excited to share our Paper - VENI: Variational Encoder for Natural Illumination 🌐 🔗 Project page: paul-pw.github.io/veni/ 📄 Paper: arxiv.org/abs/2601.14079 👩‍💻 Code: github.com/paul-pw/veni 🧵 (1/5)
Image comparing the uniqueness of VENI to the Uniqueness of RENI++ by showing an interpolation between two optimized images and showing the "Reconstruction Cosistency": Image error vs Latent error of the two models.
132
Will Smith @willsmithvision.bsky.social · 03/12/2025
My students @fhudson.bsky.social and @jadgardner.bsky.social are presenting TAPVid-360 at NeurIPS this week. We introduce an interesting new problem, a benchmark dataset and a baseline adaptation of an existing TAP model for our task. More importantly, they've also created a genre-defining poster...
140
Will Smith @willsmithvision.bsky.social · 11/05/2025
Has anyone ever tried a very non-standard tone for a rebuttal? I'm thinking something like "Hey reviewers! Sit back, relax and let me convince you that you actually want to accept this paper..." or "You wouldn't let a little thing like that stop you accepting the paper would you? WOULD YOU?!!!"
220
Will Smith @willsmithvision.bsky.social · 01/04/2025
So, we wrote a neural net library entirely in LaTeX...
38415
Reposted by Will Smith
Jon Barron @jonbarron.bsky.social · 18/02/2025
I just pushed a new paper to arXiv. I realized that a lot of my previous work on robust losses and nerf-y things was dancing around something simpler: a slight tweak to the classic Box-Cox power transform that makes it much more useful and stable. It's this f(x, λ) here:
210924
Will Smith @willsmithvision.bsky.social · 14/01/2025
#CVPR2025 Area Chair update: depending on which time zone the review deadline is specified in, we are past or close to the review deadline. Of the 60 reviews needed for my batch, I currently have 52 and they have been coming in quite fast this morning. In general, review standard looks good.
030
Reposted by Will Smith
Dmytro Mishkin @ducha-aiki.bsky.social · 02/01/2025
Image matching and ChatGPT - new post in the wide baseline stereo blog. tl;dr: it is good, even feels like human, but not perfect. ducha-aiki.github.io/wide-baselin...
ducha-aiki.github.io
ChatGPT and Image Matching – Wide baseline stereo meets deep learning
Are we done yet?
2348
Reposted by Will Smith
Gabriele Berton @berton-gabri.bsky.social · 19/12/2024
This simple pytorch trick will cut in half your GPU memory use / double your batch size (for real). Instead of adding losses and then computing backward, it's better to compute the backward on each loss (which frees the computational graph). Results will be exactly identical
3547
Will Smith @willsmithvision.bsky.social · 18/12/2024
Me and my friend-since-before-school @ekd.bsky.social (a law academic) have written a blog post about the NotebookLM podcast generator in the style of, well, a corny podcast dialogue: slsablog.co.uk/blog/blog-po... 1/5
slsablog.co.uk
Turning scholarship into an "engaging" podcast using AI: interdisciplinary perspectives
by Dr Edward Kirton-Darling, Senior Lecturer at the University of Bristol Law School, and Professor Will Smith, Department for Computer Science, University of York, Ed & Will in 1987 (or is it?) ...
110
Reposted by Will Smith
Keenan Crane @keenancrane.bsky.social · 09/12/2024
Entropy is one of those formulas that many of us learn, swallow whole, and even use regularly without really understanding. (E.g., where does that “log” come from? Are there other possible formulas?) Yet there's an intuitive & almost inevitable way to arrive at this expression.
22543128
Reposted by Will Smith
Chris Offner @chrisoffner3d.bsky.social · 10/12/2024
"Sora is a data-driven physics engine." x.com/chrisoffner3...
1213716
Reposted by Will Smith
Kyoto University Computer Vision Lab @kyotovision.bsky.social · 10/12/2024
Multistable Shape from Shading Emerges from Patch Diffusion #NeurIPS2024 Spotlight X. Nicole Han, T. Zickler and K. Nishino (Harvard+Kyoto) Diffusion-based SFS lets you sample multistable shape perception! Nicole at poster on Th 12/12 11am East A-C 1308 vision.ist.i.kyoto-u.ac.jp/research/mss...
031
Reposted by Will Smith
Maurice Fallon @mauricefallon.bsky.social · 09/12/2024
To kick off using Bluesky: our new dataset called Oxford Spires. Synchronised, multi-color cameras and lidar in multiple Oxford colleges. Ground-Truth highly accurate 3D maps from tripod scanners. The ideal basis for NeRF/3DGS SLAM research. dynamic.robots.ox.ac.uk/datasets/oxf...
views of Oxford colleges from a Leica scanner
24713
Reposted by Will Smith
zhengqili.bsky.social @zhengqili.bsky.social · 06/12/2024
Introducing MegaSaM! Accurate, fast, & robust structure + camera estimation from casual monocular videos of dynamic scenes! MegaSaM outputs camera parameters and consistent video depth, scaling to long videos with unconstrained camera paths and complex scene dynamics!
16818
Reposted by Will Smith
Tim Behrens @behrenstimb.bsky.social · 01/12/2024
OK If we are moving to Bluesky I am rescuing my favourite ever twitter thread (Jan 2019). The renamed: Bluesky-sized history of neuroscience (biased by my interests)
14634205
Reposted by Will Smith
Philip Bontrager @pbontrager.bsky.social · 04/12/2024
Really cool new work out of Deep Mind for video game world generation using latent diffusion! Soon you'll be able to speed run a game just by tricking a model to morph you from one location to another. deepmind.google/discover/blo...
deepmind.google
Genie 2: A large-scale foundation world model
Generating unlimited diverse training environments for future general agents
13910
Reposted by Will Smith
Jia-Bin Huang @jbhuang0604.bsky.social · 01/12/2024
How to drive your research forward? “I tested the idea we discussed last time. Here are some results. It does not work. (… awkward silence)” Such conversations happen so many times when meetings with students. How do we move forward? You need …
19118
Reposted by Will Smith
Andrew Davison @ajdavison.bsky.social · 28/11/2024
For my first post on Bluesky, this recent talk I did at the recent BMVA one day meeting on World Models is a good summary of my work on Computer Vision, Robotics and SLAM, and my thoughts on a bigger picture of #SpatialAI. youtu.be/NLnPG95vNhQ?...
youtu.be
1 Andrew Davison, Imperial College London - BMVA Symposium: Robotics Foundation & World Models
YouTube video by BMVA: British Machine Vision Association
59024
Will Smith @willsmithvision.bsky.social · 28/11/2024
I am a first time Area Chair for #CVPR2025 so, in the interests of transparency, I'll post some updates here on the various stages of the process. There are 708 (!) ACs (not that long ago, CVPR could have coped with 708 *reviewers*!) We've been allocated 18.27 papers on average (I have 20).
1221
Reposted by Will Smith
International Conference on 3D Vision @3dvconf.bsky.social · 26/11/2024
Hello Bluesky! 🔵 We start our account by having our third guess for Ask Me Anything session #3DV2025AMA! Noah Snavely @snavely.bsky.social from Cornell & Google DeepMind! 🌟 🕒 You have now 24 HOURS to ask him anything — drop your questions in the comments below! Keep it engaging but respectful!
134514
Reposted by Will Smith
Jia-Bin Huang @jbhuang0604.bsky.social · 26/11/2024
Introducing Generative Omnimatte: A method for decomposing a video into complete layers, including objects and their associated effects (e.g., shadows, reflections). It enables a wide range of cool applications, such as video stylization, compositions, moment retiming, and object removal.
313420
Reposted by Will Smith
Lucas Beyer (bl16) @giffmana.ai · 23/11/2024
A real-time (or very fast) open-source txt2video model dropped: LTXV. HF: huggingface.co/Lightricks/L... Gradio: huggingface.co/spaces/Light... Github: github.com/Lightricks/L... Look at that prompt example though. Need to be a proper writer to get that quality.
68910
Reposted by Will Smith
Jon Barron @jonbarron.bsky.social · 22/11/2024
I used 📍🔗 emojis to maximize Twitter/Bluesky parity in my profile. This is definitely pointless, but it's fun.
5173
Reposted by Will Smith
NeurIPS Conference @neuripsconf.bsky.social · 22/11/2024
NeurIPS Conference is now Live on Bluesky! -NeurIPS2024 Communication Chairs
1127669
Reposted by Will Smith
Gene Chou @gene-chou.bsky.social · 21/11/2024
We've released our paper "Generating 3D-Consistent Videos from Unposed Internet Photos"! Video models like Luma generate pretty videos, but sometimes struggle with 3D consistency. We can do better by scaling them with 3D-aware objectives. 1/N page: genechou.com/kfcw
311017
Reposted by Will Smith
Michael J. Black @michael-j-black.bsky.social · 20/11/2024
For those who missed this post on the-network-that-is-not-to-be-named, I made public my "secrets" for writing a good CVPR paper (or any scientific paper). I've compiled these tips of many years. It's long but hopefully it helps people write better papers. perceiving-systems.blog/en/post/writ...
perceiving-systems.blog
Writing a good scientific paper
426065
Will Smith @willsmithvision.bsky.social · 20/11/2024
As we approach the one year anniversary of a T-PAMI submission still waiting for first reviews, I imagine "With Associate Editor" to mean they sit in a lotus position atop a Himalayan peak, our paper and the reviews in their hand as they meditate (indefinitely) on what recommendation to make.
3192
Reposted by Will Smith
David Picard @davidpicard.eurosky.social · 20/11/2024
🍏 New preprint alert! 🍏 PoM: Efficient Image and Video Generation with the Polynomial Mixer arxiv.org/abs/2411.12663 This is my latest "summer project" and it was so big I had to call in reinforcements (Thanks @nicolasdufour.bsky.social) TL;DR Transformers are for boomers, welcome to the future 🧵👇
arxiv.org
PoM: Efficient Image and Video Generation with the Polynomial Mixer
Diffusion models based on Multi-Head Attention (MHA) have become ubiquitous to generate high quality images and videos. However, encoding an image or a video as a sequence of patches results in costly...
19323
Reposted by Will Smith
David Picard @davidpicard.eurosky.social · 19/11/2024
I'm slowly putting my intro to ML course material on github, starting with the lab sessions: github.com/davidpicard/... These are self-contained notebooks in which you have to implement famous algorithms from the literature (k-NN, SVM, DT, etc), with a custom dataset that I (painstakingly) made!
49815
Will Smith @willsmithvision.bsky.social · 18/11/2024
I recently gave a tutorial on the DUSt3R paper (web: dust3r.europe.naverlabs.com, paper: tinyurl.com/5t2ks575, code: github.com/naver/dust3r) in a research group meeting. In case you missed it, didn’t understand it or would like to hear some perspectives on why it’s such a cool idea, read on… 1/23
dust3r.europe.naverlabs.com
DUSt3R: Geometric 3D Vision Made Easy
67817
Reposted by Will Smith
Serge Belongie @serge.belongie.com · 18/11/2024
The auspicious appearance of @jbhuang0604.bsky.social on Bluesky inspires me to look back at my notes on “how to get cool research ideas,” which I started jotting down way back in 2001, at the start of @belongielab.bsky.social at UC San Diego (1/18)
28819
Will Smith @willsmithvision.bsky.social · 17/11/2024
Life hack: if you have fluff stuck in a USB-C port, the tooth pick in a Swiss army knife is exactly the right size (and non-conducting!)
120
Will Smith @willsmithvision.bsky.social · 15/11/2024
Well, my first post here might as well be an ode the passing of another CVPR deadline. May you get some sleep, be pleased with your submissions and your #CVPR2025 papers be sent to positive reviewers!
050