Sign in

Vincent Tao Hu

@vtaohu.bsky.social
942 followers 173 following 3 posts

LMU postdoc from Ommer-Lab, MCML junior member. UvA PhD, PKU

PostsRepliesMedia
Reposted by Vincent Tao Hu
Pingchuan Ma @pima-hyphen.bsky.social · 28/02/2025
Our work received an invited talk at the Imageomics-AAAI-25 workshop of #AAAI25. @vtaohu.bsky.social will be representing us there. Without me being there, I still would like to share our poster with you :D We also have another oral presentation for DepthFM on March 1, 2:30 pm-3:45 pm.
031
Reposted by Vincent Tao Hu
Pingchuan Ma @pima-hyphen.bsky.social · 08/01/2025
🤔When combining Vision-language models (VLMs) with Large language models (LLMs), do VLMs benefit from additional genuine semantics or artificial augmentations of the text for downstream tasks? 🤨Interested? Check out our latest work at #AAAI25: 💻Code and 📝Paper at: github.com/CompVis/DisCLIP 🧵👇
Our method pipeline
1158
Reposted by Vincent Tao Hu
Frank Fundel @frankfundel.bsky.social · 06/12/2024
Did you know you can distill the capabilities of a large diffusion model into a small ViT? ⚗️ We showed exactly that for a fundamental task: semantic correspondence📍 A thread 🧵👇
142
Vincent Tao Hu @vtaohu.bsky.social · 04/12/2024
Your Diffusion Model is secretly an implicit timestep model, no matter discrete or continuous~
060
Reposted by Vincent Tao Hu
Vladimir Yugay @vyuga3d.bsky.social · 27/11/2024
Introducing “MAGiC-SLAM: Multi-Agent Gaussian Globally Consistent SLAM”! We do SLAM with novel view synthesis capabilities on multiple simultaneously operating agents! vladimiryugay.github.io/magic_slam/i... 1/7
35117