Sign in

Anand Bhattad

@anandbhattad.bsky.social
229 followers 199 following 68 posts

Incoming Assistant Professor at Johns Hopkins University | RAP at Toyota Technological Institute at Chicago | web: anandbhattad.github.io | Knowledge in Generative Image Models, Intrinsic Images, Image-based Relighting, Inverse Graphics

PostsRepliesMedia
Anand Bhattad @anandbhattad.bsky.social · 08/09/2026
1/ What do video models know about high school physics? Less than you'd think. We built Principia: relational physics tests for video models. We evaluate on eight laws, over 500+ real scenes, and one simple idea. principiabench.github.io paper: arxiv.org/abs/2609.04200
100
Anand Bhattad @anandbhattad.bsky.social · 23/06/2026
📢 Super excited about our "Thinking in Boxes" Idea: fit 3D boxes around objects in a single "real" photo, then move the boxes. Objects translate & rotate; two of them can swap places; the camera moves; and the scene holds up. Enjoy the video below. More on our website: thinking-in-boxes.github.io
010
Anand Bhattad @anandbhattad.bsky.social · 02/06/2026
"Bitter Lessons" workshop tomorrow (Jun 3rd) starting from 08:45 am in Room 3A-3D at #CVPR2026. Website and schedule: sites.google.com/view/bitterl...
073
Reposted by Anand Bhattad
Vision and Graphics Trends @si-cv-graphics.bsky.social · 04/05/2026
𝗚𝗲𝗻𝗲𝗿𝗮𝗹𝗶𝘇𝗮𝗯𝗹𝗲 𝗦𝗽𝗮𝗿𝘀𝗲-𝗩𝗶𝗲𝘄 𝟯𝗗 𝗥𝗲𝗰𝗼𝗻𝘀𝘁𝗿𝘂𝗰𝘁𝗶𝗼𝗻 𝗳𝗿𝗼𝗺 𝗨𝗻𝗰𝗼𝗻𝘀𝘁𝗿𝗮𝗶𝗻𝗲𝗱 𝗜𝗺𝗮𝗴𝗲𝘀 Vinayak Gupta, Chih-Hao Lin, Shenlong Wang ... Jia-Bin Huang arxiv.org/abs/2604.28193 Trending on www.scholar-inbox.com
011
Anand Bhattad @anandbhattad.bsky.social · 25/04/2026
At #ICLR2026, Vaibhav Vavilala presented our paper, "Generative Blocks World"! This is one of my favorites from recent times. We spent two years hacking around with diffusion models to get them to do object editing. The solution turned out to be way simpler. A few blocks in 3D space is enough ...
140
Reposted by Anand Bhattad
Zhenjun Zhao @ericzzj.bsky.social · 14/04/2026
SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization Deming Li, Abhay Yadav, Cheng Peng, Rama Chellappa, @anandbhattad.bsky.social tl;dr: joint conditional over multiple views to enforce consistency during denoising arxiv.org/abs/2604.11797
041
Reposted by Anand Bhattad
Johns Hopkins Data Science and AI Institute @hopkinsdsai.bsky.social · 19/12/2025
Join us in advancing data science and AI research! The Johns Hopkins Data Science and AI Institute Postdoctoral Fellowship Program is now accepting applications for the 2026–2027 academic year. Apply now! Deadline: Jan 23, 2026. Details and apply: apply.interfolio.com/179059
0119
Reposted by Anand Bhattad
JHU Computer Science @jhucompsci.bsky.social · 08/10/2025
10 new CS professors! 🥳 @anandbhattad.bsky.social @uthsav.bsky.social @gligoric.bsky.social @murat-kocaoglu.bsky.social @tiziano.bsky.social
0106
Anand Bhattad @anandbhattad.bsky.social · 06/10/2025
I decided not to travel to #ICCV2025 because it coincides with Diwali (Oct 20). Diwali often falls near the #CVPR deadline window, but this year overlaps with ICCV. I understand it’s hard to avoid all global holidays, but I hope future conferences can keep this in mind when selecting dates.
020
Anand Bhattad @anandbhattad.bsky.social · 06/10/2025
I will be recruiting a few students for Fall 2026. In particular, I will strongly consider a PhD applicant with training in applied/computational mechanics and computer vision/machine learning. If you or someone you know has this background, please contact me.
041
Anand Bhattad @anandbhattad.bsky.social · 21/09/2025
So You Want to Be an Academic? A couple of years into your PhD, but wondering: "Am I doing this right?" Most of the advice is aimed at graduating students. But there's far less for junior folks who are still finding their academic path. My candid takes: anandbhattad.github.io/blogs/jr_gra...
anandbhattad.github.io
So You Want to Be an Academic? What I Wish I Knew Early in Graduate School
Blog for junior PhD students on work, visibility, community, and sanity—long before the faculty job market is on the horizon.
1204
Reposted by Anand Bhattad
ML for Science @ml4science.bsky.social · 30/06/2025
On our blog: Science is moving fast. How do we keep up? #ScholarInbox, developed by the Autonomous Vision Group led by @andreasgeiger.bsky.social, helps researchers stay ahead - by making the discovery of #openaccess papers smarter and more personal: www.machinelearningforscience.de/en/scholar-i...
machinelearningforscience.de
Scholar Inbox: Daily Research Recommendations just for You
Science is moving fast. How can we keep up? Scholar Inbox helps researchers stay ahead by making the discovery of open access papers more personal.
12714
Anand Bhattad @anandbhattad.bsky.social · 30/06/2025
All slides from the #cvpr2025 (@cvprconference.bsky.social ) workshop "How to Stand Out in the Crowd?" are now available on our website: sites.google.com/view/standou...
000
Anand Bhattad @anandbhattad.bsky.social · 02/06/2025
I’m thrilled to share that I will be joining Johns Hopkins University’s Department of Computer Science (@jhucompsci.bsky.social, @hopkinsdsai.bsky.social) as an Assistant Professor this fall.
182
Reposted by Anand Bhattad
Zhenjun Zhao @ericzzj.bsky.social · 08/05/2025
FastMap: Revisiting Dense and Scalable Structure from Motion Jiahao Li, Haochen Wang, @zubair-irshad.bsky.social, @ivasl.bsky.social, Matthew R. Walter, Vitor Campagnolo Guizilini, Greg Shakhnarovich tl;dr: replace BA with epipolar error+IRLS; fully PyTorch implementation arxiv.org/abs/2505.04612
0102
Reposted by Anand Bhattad
Noah Snavely @snavely.bsky.social · 30/03/2025
This is really cool work!
181
Anand Bhattad @anandbhattad.bsky.social · 29/03/2025
[1/10] Is scene understanding solved? Models today can label pixels and detect objects with high accuracy. But does that mean they truly understand scenes? Super excited to share our new paper and a new task in computer vision: Visual Jenga! 📄 arxiv.org/abs/2503.21770 🔗 visualjenga.github.io
75914
Anand Bhattad @anandbhattad.bsky.social · 27/03/2025
I can’t believe this! Mind-blowing! There are small errors (a flipped logo, rotated chairs), but still, this is incredible!! Xiaoyan, who’s been working with me on relighting, sent this over. It’s one of the hardest examples we’ve consistently used to stress-test LumiNet: luminet-relight.github.io
030
Reposted by Anand Bhattad
Jia-Bin Huang @jbhuang0604.bsky.social · 15/03/2025
Check out UrbanIR - Inverse rendering of unbounded scenes from a single video! It’s a super cool project led by the amazing Chih-Hao! @chih-hao.bsky.social is a rising star in 3DV! Follow him! Learn more here👇
0102
Anand Bhattad @anandbhattad.bsky.social · 16/03/2025
Can we create realistic renderings of urban scenes from a single video while enabling controllable editing: relighting, object compositing, and nighttime simulation? Check out our #3DV2025 UrbanIR paper, led by @chih-hao.bsky.social that does exactly this. 🔗: urbaninverserendering.github.io
021
Reposted by Anand Bhattad
Johan Edstedt @parskatt.bsky.social · 11/03/2025
Introducing DaD (arxiv.org/abs/2503.07347), a pretty cool keypoint detector. As this will get pretty long, this will be two threads. The first will go into the RL part, and the second on the emergence and distillation.
46211
Reposted by Anand Bhattad
Ben Waber @bwaber.bsky.social · 06/03/2025
Last was an intriguing talk by @anandbhattad.bsky.social on the degree to which physical properties are encoded in generative image models at the @grasplab.bsky.social www.youtube.com/watch?v=Idcb... (7/7)
youtube.com
Fall 2024 GRASP Seminar Anand Bhattad, Toyota Technological Institute at Chicago
YouTube video by GRASP Lab
011
Anand Bhattad @anandbhattad.bsky.social · 28/02/2025
ZeroComp will be presented as an Oral at #WACV2025! Train a diffusion model as a renderer that takes intrinsic images as input. Once trained, we can perform zero-shot object compositing & can easily extend this to other object editing tasks, such as material swapping. arxiv.org/abs/2410.08168
140
Anand Bhattad @anandbhattad.bsky.social · 27/02/2025
LumiNet is going to be at #CVPR2025! Excited about this work—fully data-driven, with zero reliance on computer graphics datasets. It builds upon our previous works—StyLitGAN provided the training data, while Latent Intrinsics abstracted lighting and albedo for diffusion models. And more to come! 🤓
040
Reposted by Anand Bhattad
Elliott / Shangzhe Wu @elliottwu.bsky.social · 12/02/2025
Really excited to put together this #CVPR2025 workshop on "4D Vision: Modeling the Dynamic World" -- one of the most fascinating areas in computer vision today! We've invited incredible researchers who are leading fantastic work at various related fields. 4dvisionworkshop.github.io
1223
Reposted by Anand Bhattad
Aaron Hertzmann @aaronhertzmann.com · 06/02/2025
New paper about pictures: I identify trends in geometric perspective in my own drawings and photos, and compare them to how the original scenes looked. I discuss what these trends might say about art history and vision science. Published in _Art & Perception_. #visionscience psyarxiv.com/pq8nb
psyarxiv.com
OSF
3196
Anand Bhattad @anandbhattad.bsky.social · 06/02/2025
Encouraging to see a civilized & thoughtful discussion on a few papers in my AC batch. We need more reviewers like this! Also chasing down a few reviewers to update their justification beyond 'my concerns were not addressed in the rebuttal, so I’ll maintain my original score.' 😅 #CVPR2025
030
Anand Bhattad @anandbhattad.bsky.social · 05/02/2025
Not a big fan of Borderline ratings in the review process. They offer an easy way for reviewers to be non-committal—or even lazy. I’ve never liked this, as it rarely adds value. Clear, decisive ratings with concrete justifications of a paper’s contributions are far more constructive. #CVPR2025
170
Anand Bhattad @anandbhattad.bsky.social · 01/02/2025
It's great to see renewed interest in Barrow and Tenenbaum's seminal 1978 work on "Intrinsic Images." This foundational paper has significantly influenced the field, even though it has historically received fewer citations than it deserves.
010
Anand Bhattad @anandbhattad.bsky.social · 17/01/2025
Scholar Inbox provides the best paper recommendations—I rarely miss papers relevant to my interests or related to my work. I wish this could be integrated with paper submission services to flag missing related works.
120
Anand Bhattad @anandbhattad.bsky.social · 13/01/2025
🧵 1/3 Many at #CVPR2024 & #ECCV2024 asked what would be next in our workshop series. We're excited to announce "How to Stand Out in the Crowd?" at #CVPR2025 Nashville - our 4th community-building workshop featuring this incredible speaker lineup! 🔗 sites.google.com/view/standou...
1418
Reposted by Anand Bhattad
James Tompkin @jamestompkin.bsky.social · 10/01/2025
Can GANs compete in 2025? In 'The GAN is dead; long live the GAN! A Modern GAN Baseline', we show that a minimalist GAN w/o any tricks can match the performance of EDM with half the size and one-step generation - github.com/brownvc/r3gan - work of Nick Huang, @skylion.bsky.social, Volodymyr Kuleshov
36914
Anand Bhattad @anandbhattad.bsky.social · 11/12/2024
[1/2] Latent Intrinsics Emerge from Training to Relight. We discovered an easy way to extract albedo without any albedo-like images. By questioning why intrinsic image representations must be 3-channel maps, we trained a simple relighting model where intrinsics & lighting are latent variables.
100
Anand Bhattad @anandbhattad.bsky.social · 11/12/2024
[1/3] From a single image to the 3D world—it’s possible with the right data, & we have it for you🚀 We’re thrilled to release our 360° video dataset. Training a simple conditional diffusion model with explicit camera control can synthesize novel 3D scenes—all from a single image input! #NeurIPS2024
120
Reposted by Anand Bhattad
Jon Barron @jonbarron.bsky.social · 11/12/2024
I've never seen a more polished paper teaser video than this one. The 3D Gaussians even cast shadows! fnzhan.com/Evolutive-Re...
16211
Anand Bhattad @anandbhattad.bsky.social · 11/12/2024
Spotlight Poster #NeurIPS2024 📍 East Exhibit Hall A-C, #1500 🗓️ Wed 11 Dec, 16:30 – 19:30 PT Latent Intrinsics: Intrinsic image representation as latent variables, resulting in emergent albedo-like maps for free. Code is now available: github.com/xiao7199/Lat... arXiv: arxiv.org/abs/2405.21074
000
Reposted by Anand Bhattad
Anand Bhattad @anandbhattad.bsky.social · 05/12/2024
I'm super stoked to share LumiNet! 🎉 💡LumiNet combines latent intrinsic representations with a powerful generative image model. It lets us relight complex indoor images by transferring lighting from one image to another. It is completely data-driven: no labels, mark-up, not even depth or normals.
161
Anand Bhattad @anandbhattad.bsky.social · 05/12/2024
Our latest work is on arXiv now and the thread is updated with the movies :) arXiv: arxiv.org/abs/2412.00177 project: luminet-relight.github.io
010
Reposted by Anand Bhattad
Roni Sengupta @ronisen.bsky.social · 03/12/2024
💡Introducing ScribbleLight, a generative model that can relight an image with simple scribbles - making it easier for virtual staging and interior design of homes. Led by Jun-Myeong, collabs with Anand (TTIC), Pieter (W&M), and Annie. @unccs.bsky.social 👉 chedgekorea.github.io/ScribbleLight/
194
Anand Bhattad @anandbhattad.bsky.social · 05/12/2024
I'm super stoked to share LumiNet! 🎉 💡LumiNet combines latent intrinsic representations with a powerful generative image model. It lets us relight complex indoor images by transferring lighting from one image to another. It is completely data-driven: no labels, mark-up, not even depth or normals.
161