Sign in

Thibaut Loiseau

@thibautloiseau.bsky.social
422 followers 284 following 26 posts

PhD Student at IMAGINE (ENPC) Working on camera pose estimation thibautloiseau.github.io

PostsRepliesMedia
Reposted by Thibaut Loiseau
Guillaume Astruc @gastruc.bsky.social · 23/06/2026
🛰️ Introducing UniverSat: one transformer backbone for Earth Observation that handles ANY sensor, ANY spatial, spectral & temporal resolution, ANY scale — with a single set of weights. 🌍
1159
Reposted by Thibaut Loiseau
Antoine Guédon @antoine-guedon.bsky.social · 16/06/2026
What if you could turn any number of photos (3, 8, 15, or even 60) into one clean 3D surface (pts & mesh) with Flow Matching? Check out our new work, Surflo: Consistent 3D Surface Flow Model with Global State. 🧵 1/N 🔗https://anttwo.github.io/surflo/
13512
Reposted by Thibaut Loiseau
Nicolas Dufour @nicolasdufour.bsky.social · 18/08/2025
🚀 DinoV3 just became the new go-to backbone for geoloc! It outperforms CLIP-like models (SigLip2, finetuned StreetCLIP)… and that’s shocking 🤯 Why? CLIP models have an innate advantage — they literally learn place names + images. DinoV3 doesn’t.
14614
Reposted by Thibaut Loiseau
Imagine-ENPC @imagineenpc.bsky.social · 15/06/2025
Some of our IMAGINE members at #CVPR2025
0347
Reposted by Thibaut Loiseau
Vincent Lepetit @vincentlepetit.bsky.social · 14/06/2025
I am heartbroken that I am not at the conference, but seeing what the government is doing to its people and the world, I simply couldn't go there.
1216
Reposted by Thibaut Loiseau
Imagine-ENPC @imagineenpc.bsky.social · 30/04/2025
Looking forward to #CVPR2025! We will present the following papers:
1287
Reposted by Thibaut Loiseau
Nicolas Dufour @nicolasdufour.bsky.social · 24/04/2025
This is an idea I've had for a while, but wow, it's working way better than expected! 🚀 The model looks really promising, even though it's just 256px for now.
173
Reposted by Thibaut Loiseau
Lucas Ventura @lucasventura.com · 04/04/2025
Introducing Chapter-Llama #CVPR2025, a framework for 𝐯𝐢𝐝𝐞𝐨 𝐜𝐡𝐚𝐩𝐭𝐞𝐫𝐢𝐧𝐠 using Large Language Models! 🎬🦙 Check it out: 📄 Paper: arxiv.org/abs/2504.00072 🔗 Project: imagine.enpc.fr/~lucas.ventu... 💻 Code: github.com/lucas-ventur... 🤗 Demo: huggingface.co/spaces/lucas...
1245
Reposted by Thibaut Loiseau
David Picard @davidpicard.eurosky.social · 21/03/2025
🔥🔥🔥 CV Folks, I have some news! We're organizing a 1-day meeting in center Paris on June 6th before CVPR called CVPR@Paris (similar as NeurIPS@Paris) 🥐🍾🥖🍷 Registration is open (it's free) with priority given to authors of accepted papers: cvprinparis.github.io/CVPR2025InPa... Big 🧵👇 with details!
713652
Reposted by Thibaut Loiseau
Imagine-ENPC @imagineenpc.bsky.social · 14/03/2025
Starter pack including some of the lab members: go.bsky.app/QK8j87w
02411
Reposted by Thibaut Loiseau
Johan Edstedt @parskatt.bsky.social · 11/03/2025
Introducing DaD, Part 2, a pretty cool keypoint detector.
5315
Thibaut Loiseau @thibautloiseau.bsky.social · 11/03/2025
1/13 🐊 Introducing our latest work on improving relative camera pose regression with a novel pre-training approach Alligat0R (arxiv.org/abs/2503.07561)! @gbourmaud.bsky.social @vincentlepetit.bsky.social
4205
Reposted by Thibaut Loiseau
Zhenjun Zhao @ericzzj.bsky.social · 11/03/2025
Alligat0R: Pre-Training Through Co-Visibility Segmentation for Relative Camera Pose Regression @thibautloiseau.bsky.social, Guillaume Bourmaud, @vincentlepetit.bsky.social tl;dr: CroCo based; pixel in 1st image->co-visible or occluded or outside FOV in 2nd image arxiv.org/abs/2503.07561
0154
Reposted by Thibaut Loiseau
Lucas Degeorge @lucasdegeorge.bsky.social · 05/03/2025
🚨 News! 🚨 We have released the models from our latest paper "How far can we go with ImageNet for text-to-image generation?" Check out the models on HuggingFace: 🤗 huggingface.co/Lucasdegeorg... 📜 arxiv.org/abs/2502.21318
huggingface.co
Lucasdegeorge/CAD-I · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
1153
Reposted by Thibaut Loiseau
Nicolas Dufour @nicolasdufour.bsky.social · 03/03/2025
Check out our latest work on Text-to-Image generation! We've successfully trained a T2I model using only ImageNet data by leveraging captioning and data augmentation.
1156
Reposted by Thibaut Loiseau
David Picard @davidpicard.eurosky.social · 03/03/2025
🚨 New preprint! How far can we go with ImageNet for Text-to-Image generation? w. @arrijitghosh.bsky.social @lucasdegeorge.bsky.social @nicolasdufour.bsky.social @vickykalogeiton.bsky.social TL;DR: Train a text-to-image model using 1000 less data in 200 GPU hrs! 📜https://arxiv.org/abs/2502.21318 🧵👇
26516
Thibaut Loiseau @thibautloiseau.bsky.social · 28/02/2025
🧩 Excited to share our paper "RUBIK: A Structured Benchmark for Image Matching across Geometric Challenges" (arxiv.org/abs/2502.19955) accepted to #CVPR2025! We created a benchmark that systematically evaluates image matching methods across well-defined geometric difficulty levels. 🔍
2197
Reposted by Thibaut Loiseau
Guillaume Astruc @gastruc.bsky.social · 19/12/2024
🤔 What if embedding multimodal EO data was as easy as using a ResNet on images? Introducing AnySat: one model for any resolution (0.2m–250m), scale (0.3–2600 hectares), and modalities (choose from 11 sensors & time series)! Try it with just a few lines of code:
23510
Reposted by Thibaut Loiseau
David Picard @davidpicard.eurosky.social · 12/12/2024
We @imagineenpc.bsky.social are slowly but surely entering our proposals for master's degree internships here: docs.google.com/document/d/1... These are 6 months projects that typically correspond to the end-of-study project in the French curriculum. Probably more offers to come, check it regularly.
docs.google.com
2025 IMAGINE Internships
2025 Internship proposals at IMAGINE IMAGINE is a top research group on computer vision and machine learning. It is part of the LIGM lab and hosted at École des Ponts ParisTech (ENPC), about 25 min f...
23211