Yash Bhalgat @ysbhalgat.bsky.social · 23/03/2025Excited to announce the 1st Workshop on 3D-LLM/VLA at #CVPR2025! 🚀 @cvprconference.bsky.social Topics: 3D-VLA models, LLM agents for 3D scene understanding, Robotic control with language. 📢 Call for papers: Deadline – April 20, 2025 🌐 Details: 3d-llm-vla.github.io #llm #3d #Robotics #ai 061
Reposted by Yash BhalgatAndreas Geiger @andreasgeiger.bsky.social · 22/02/2025Our beginner's oriented accessible introduction to modern deep RL is now published in Foundations and Trends in Optimization. It is a great entry to the field if you want to jumpstart into RL! @bernhard-jaeger.bsky.social www.nowpublishers.com/article/Deta... arxiv.org/abs/2312.08365 26214
Yash Bhalgat @ysbhalgat.bsky.social · 18/02/2025"LLaDA: Large Language Diffusion Models" Nie et al. Just read this fascinating paper. Scaled up Masked Diffusion Language Models to 8B params, and show that it can match #LLMs (including Llama 3) while solving some key limitations! Let's dive in... 🧵 (1/8) #genai 111
Yash Bhalgat @ysbhalgat.bsky.social · 16/02/2025New work introduces a training-free method to relight entire videos, while maintaining temporal consistency! 📽️🌅 "Light-A-Video: Training-free Video Relighting via Progressive Light Fusion" Zhou et al. (1/n) 🧵 #genai #ai #research #video 1102
Yash Bhalgat @ysbhalgat.bsky.social · 15/02/2025Need to rig 3D models? 🦖 New work from UCSD and Adobe: "RigAnything: Template-Free Autoregressive Rigging for Diverse 3D Assets" Liu et al. tl;dr: reduces rigging time from 2 mins to 2 secs, works on any shape category & doesn't need predefined templates! 🚀 150
Yash Bhalgat @ysbhalgat.bsky.social · 14/02/2025"Latent Radiance Fields with 3D-aware 2D Representations" Zhou et al., #ICLR2025 tl;dr: Novel framework that integrates 3D awareness into VAE latent space using correspondence-aware encoding, enabling high-quality rendered images with ~50% memory savings. (1/n) 🧵 120
Yash Bhalgat @ysbhalgat.bsky.social · 13/02/2025"EdgeRunner" (#ICLR2025) from #Nvidia & PKU introduces an auto-regressive auto-encoder for mesh generation, supporting up to 4000 faces at 512³ resolution. 🤩 Their mesh tokenization algorithm (adapted from EdgeBreaker) achieves ~50% compression (4-5 tokens per face vs 9), making training efficient. 100
Yash Bhalgat @ysbhalgat.bsky.social · 12/02/2025Just came across this fascinating paper "CraftsMan3D" - a practical approach to text/image-to-3D generation that mimics how artists actually work! Code available (pretrained models too) 🤩: github.com/wyysf-98/Cra... (1/n) 🧵 110
Yash Bhalgat @ysbhalgat.bsky.social · 23/01/2025📢 Paper accepted to #ICLR2025 🎉 "GSLoc: Efficient Camera Pose Refinement via 3D Gaussian Splatting" TL;DR: a novel test-time camera pose refinement framework leveraging 3DGS as the scene representation and MASt3R for 2D matching. 🔗: arxiv.org/abs/2408.11085 172
Yash Bhalgat @ysbhalgat.bsky.social · 21/01/2025Switzerland is rolling out solar panels... on railway tracks! 🇨🇭🚄 Swiss startup Sun-ways will run a pilot project turning train lines into clean energy highways. #renewable #energy for the win 🤓 www.pv-magazine.com/2024/10/04/s...pv-magazine.comSwitzerland authorizes removable PV plant on railway trackSwiss startup Sun-ways is planning to build a 18 kW pilot PV system between the racks of a 100-m linear section of a railway line in the Swiss canton of Neuchâtel. 1134
Reposted by Yash BhalgatPhilip Oldfield @sustainabletall.bsky.social · 13/01/2025Solar panels are becoming so cheap the Swiss are looking at installing them *between train tracks* !!! www.pv-magazine.com/2024/10/04/s... 38118
Reposted by Yash BhalgatMaurice Fallon @mauricefallon.bsky.social · 11/01/2025🚨🚨🚨 Reminder: closing in 3 weeks time 🚨🚨🚨 Please re-post! Note: Oxford recruits faculty at Associate Professor level - we have no Assistant Professor level. 031
Yash Bhalgat @ysbhalgat.bsky.social · 09/01/2025Came across this LLM visualisation tool today: bbycroft.net/llm Cool stuff! Let's you visualize each operation or layer in different Transformer architectures, and also explains them on the side. 😍 #llm #visualisation #gpt #ai #transformers 040
Reposted by Yash BhalgatMichael Niemeyer @miniemeyer.bsky.social · 08/01/2025For other 3D vision newcomers to blue sky: I highly recommend joning @chrisoffner3d.bsky.social 's list to follow the right people ;): go.bsky.app/Cfm9XFe 2194
Yash Bhalgat @ysbhalgat.bsky.social · 08/01/2025"NeuralSVG: An Implicit Representation for Text-to-Vector Generation" (1/2) Encodes SVGs as implicit neural representations using a small MLP trained with Score Distillation Sampling (SDS). Maps 2D coordinates to shape/color outputs. Dropout-like technique ensures ordered, layered structures. 100
Yash Bhalgat @ysbhalgat.bsky.social · 06/01/2025"AR4D: Autoregressive 4D Generation from Monocular Videos" *without* SDS. Autoregressively generate "3D frames" (aka 3DGS) starting from a canonical space, and using a local deformation field for each frame -- high-quality prompt-aligned generations. #ai #nerf #GenAI #video 010
Yash Bhalgat @ysbhalgat.bsky.social · 06/01/2025Another gem from Bill Freeman, Katie Bouman & team 🌌 A differentiable rendering framework for direct #exoplanet imaging, leveraging wavefront sensing to refine starlight subtraction. Tested on JWST, it approaches noise limits and reveals faint planets like never before! 🚀 #ai #astronomy 031
Yash Bhalgat @ysbhalgat.bsky.social · 20/11/2024📢 Proud that my first-ever #opensource (grid-based) #NeRF implementation hit 100 forks on #github! Just 30 more ⭐️'s to cross 1000. 📈 🔗: github.com/yashbhalgat/... #instantngp #neuralradiancefield #3d #computervision #research #pythongithub.comGitHub - yashbhalgat/HashNeRF-pytorch: Pure PyTorch Implementation of NVIDIA paper on Instant Training of Neural Graphics primitives: https://nvlabs.github.io/instant-ngp/Pure PyTorch Implementation of NVIDIA paper on Instant Training of Neural Graphics primitives: https://nvlabs.github.io/instant-ngp/ - yashbhalgat/HashNeRF-pytorch 020
Yash Bhalgat @ysbhalgat.bsky.social · 17/11/2024Excited to share that our work with VAL, IISc on "Reflecting Reality: Enabling Diffusion Models to Produce Faithful Mirror Reflections" has been accepted to #3DV2025 Project page: val.cds.iisc.ac.in/reflecting-r...val.cds.iisc.ac.inReflecting RealityReflecting Reality: Enabling Diffusion Models to Produce Faithful Mirror Reflections 010
Yash Bhalgat @ysbhalgat.bsky.social · 17/11/2024Reached 600 today. Onwards and upwards. 📈 #Research #googlescholar 000
Yash Bhalgat @ysbhalgat.bsky.social · 13/02/2024We are hosting the 2nd Workshop on Learning #3D with Multi-View Supervision (3DMV) AT #CVPR2024 in Seattle on June 17th! We are accepting paper submissions on a range of topics. More details on the website: abdullahamdi.com/3dmv2024/ #computervision #ai 020