Sign in

Jack (in SF) Langerman

@jacklangerman.bsky.social
69 followers 53 following 69 posts
PostsRepliesMedia
Jack (in SF) Langerman @jacklangerman.bsky.social · 03/06/2026
🎉Workshop On Urban Scene at @cvprconference.bsky.social modeling kicks off in less than 1 hr!!🎉 first up: @niessner.bsky.social ! You don't want to miss this one!! Mile high 3B
010
Jack (in SF) Langerman @jacklangerman.bsky.social · 03/06/2026
We have arrived @cvprconference.bsky.social ! it's going to be a great week!
010
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 02/06/2026
Tomorrow is a better day...because USM3D workshop at @cvprconference.bsky.social is happening? 6 keynote speakers. 2 challenges which I am finally not ashamed of? Debatable. Is it our final competition at HuggingFace? The only way to figure out is to attend. Mile High 3B. #CVPR2026
091
Jack (in SF) Langerman @jacklangerman.bsky.social · 03/06/2026
🌟🌟 Start @CVPR off right with the Workshop on Urban Scene Modeling! (Tues Jun 2) @ Mile High 3B! 🌟 - Great lineup of key note speakers! - Challenge winners will be announced! - Selected paper talks! 3D/4D reconstruction+generation+understanding applied to the built world
Matthias Nießner
Florent Lafarge
Marc Pollefeys
Vasileios Balntas
Angel Xuan Chang
Daniel Barath

Workshop Schedule. Full day, June 3, Mile High 3B. 09:00–09:10 Welcome & introduction. 09:10–09:50 Keynote 1 — Matthias Nießner. 09:50–10:30 Building3D Challenge — winner talks. 10:30–10:45 Coffee break. 10:45–11:25 Keynote 2 — Florent Lafarge, “Two Decades of 3D Building Reconstruction: Paradigms, Progress, and Prospects.” 11:25–12:05 S23DR Challenge — winner talks. 12:05–13:00 Lunch. 13:00–13:40 Keynote 3 — Marc Pollefeys. 13:40–15:10 Paper Oral Presentation — 9 papers (next slide). 15:10–15:25 Coffee break. 15:25–16:05 Keynote 4 — Vasileios Balntas. 16:05–16:45 Keynote 5 — Angel Xuan Chang. 16:45–17:25 Keynote 6 — Daniel Barath. 17:25–17:55 Collaboration & Discussion (interactive panel). 17:55–18:00 Closing remarks.
100
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 18/03/2026
Sad, but important announcement -- one of our challenges -- Building3D lost their sponsor in the very last moment. So new prize fund is $14k: - $12k for S23DR 2026 - $2k for Building3D 2026. We apologise for this situation, and hope that you still decide to participate. #CVPR2026
061
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 04/03/2026
Submit your paper of structured reconstruction -- CAD, semantic, wireframe, city monitoring, etc., to USM3D 2026! cmt3.research.microsoft.com/USM2026 Deadline: March 24, 2026. @cvprconference.bsky.social #CVPR2026 #USM3D2026 #USM3D
053
Jack (in SF) Langerman @jacklangerman.bsky.social · 04/02/2026
🧵A bit last minute but I'm giving a talk tonight in SF about my work Explaining Human Preferences via Metrics for Structured 3D Reconstruction (ICCV Highlight) Quite near transamerica pyramid. Event is already full, but comment or shoot me a DM and we'll find a spot for you!
110
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 24/10/2025
Come to our poster this afternoon starting at 2:30 (#214)
031
Jack (in SF) Langerman @jacklangerman.bsky.social · 10/10/2025
What is the right metric to measure performance in structured 3D reconstruction? 🎬 Just posted the video for our @iccv.bsky.social #ICCV2025 highlight paper! 🎉 @jacklangerman.bsky.social, Denys Rozumnyi, Yuzhong Huang, @ducha-aiki.bsky.social 🌄 iccv.thecvf.com/virtual/2025...
youtu.be
[ICCV 2025] Explaining Human Preferences via Metrics for Structured Reconstruction - ICCV Highlight
YouTube video by Jack Langerman
021
Jack (in SF) Langerman @jacklangerman.bsky.social · 22/09/2025
what makes a good paper video? what are some of the best you have seen? cc: @davidpicard.bsky.social @davidbau.bsky.social @eugenevinitsky.bsky.social @chriswolfvision.bsky.social @nsaphra.bsky.social @ducha-aiki.bsky.social @parskatt.bsky.social @martin-r-oswald.bsky.social @ericzzj.bsky.social
221
Jack (in SF) Langerman @jacklangerman.bsky.social · 13/09/2025
iiim8uhl j ij9j b knmnjj mlpnnom l 8kc6 Kimberly.C mbkln9j8 . ojlioij7o.h ihlokMO Lip l 9 b nl d9
100
Jack (in SF) Langerman @jacklangerman.bsky.social · 25/06/2025
Very proud that our paper has been accepted to @iccv.bsky.social !!! See you in Hawaii! bsky.app/profile/duch...
061
Reposted by Jack (in SF) Langerman
Zhenjun Zhao @ericzzj.bsky.social · 09/05/2025
🎉 Thrilled to share our CVPR 2025 Award Candidate & Oral paper: 🔹 GlobustVP Convex Relaxation for Robust Vanishing Point Estimation in Manhattan World 🧱 Global optimality 💥 Tolerates up to 70% outliers ⚡ Fast runtime 📄 Paper: arxiv.org/abs/2505.04788 💻 Code: github.com/WU-CVGL/GlobustVP 1/
1315
Reposted by Jack (in SF) Langerman
#ICCV2025 @iccv.bsky.social · 09/05/2025
#ICCV2025 reviews are out and being sent via email to authors! They will also be available on OpenReview later. 11,152 active submissions all have at least 3 reviews. Authors have the opportunity to submit a rebuttal by May 16 2025 11:59 PM HST.
1198
Jack (in SF) Langerman @jacklangerman.bsky.social · 07/05/2025
🚨 Just one month left to submit your solutions for The Structured Semantic 3D Reconstruction (S23DR-2025) Challenge!!! It is not too late to join! Comp on @hf.co, part of the Workshop on Urban Scene Modeling at @cvprconference.bsky.social 2025 🔥$25,000 prize pool. Deadline: June 5, 2025. 🧵 (1/7)
154
Jack (in SF) Langerman @jacklangerman.bsky.social · 30/04/2025
hey @cloneofsimo.bsky.social (or anyone else) did you ever try LoRA+ style LR split between down/up LoRA projection matrices on diffusion models?
LoRA+: Efficient Low Rank Adaptation of Large Models

Soufiane Hayou, Nikhil Ghosh, Bin Yu
In this paper, we show that Low Rank Adaptation (LoRA) as originally introduced in Hu et al. (2021) leads to suboptimal finetuning of models with large width (embedding dimension). This is due to the fact that adapter matrices A and B in LoRA are updated with the same learning rate. Using scaling arguments for large width networks, we demonstrate that using the same learning rate for A and B does not allow efficient feature learning. We then show that this suboptimality of LoRA can be corrected simply by setting different learning rates for the LoRA adapter matrices A and B with a well-chosen ratio. We call this proposed algorithm LoRA+. In our extensive experiments, LoRA+ improves performance (1-2 % improvements) and finetuning speed (up to ∼ 2X SpeedUp), at the same computational cost as LoRA.
010
Reposted by Jack (in SF) Langerman
Aaron Mueller @amuuueller.bsky.social · 23/04/2025
Lots of progress in mech interp (MI) lately! But how can we measure when new mech interp methods yield real improvements over prior work? We propose 😎 𝗠𝗜𝗕: a 𝗠echanistic 𝗜nterpretability 𝗕enchmark!
Logo for MIB: A Mechanistic Interpretability Benchmark
15115
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 17/03/2025
1 week to USM3D deadline
063
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 12/03/2025
Explaining Human Preferences via Metrics for Structured 3D Reconstruction @jacklangerman.bsky.social Denys Rozumnyi, Yuzhong Huang, @ducha-aiki.bsky.social tl;dr: we asked 3D modelers to rank wireframe reconstructions & compared it to ranking by metrics. Observations🧵 1/ arxiv.org/abs/2503.08208
12810
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 10/03/2025
2nd Building3D #CVPR2025 challenge at #USM3D workshop is open! Task: point cloud to wireframe. Prize pool: $10k Competition deadline: May 25 2025. Website: huggingface.co/spaces/Build... @cvprconference.bsky.social
0124
Reposted by Jack (in SF) Langerman
Johan Edstedt @parskatt.bsky.social · 10/03/2025
We made a new keypoint detector named DaD, paper isn't up yet, but code and weights are: github.com/Parskatt/dad
7448
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 10/03/2025
Those, who work in structured (images/pcl to CAD) reconstruction - USM3D #CVPR2025 workshop submissions are open. Deadline: March 24 2025 Both full papers (8 pages) and extended abstracts (4 pages) are OK usm3d.github.io #USM3D @cvprconference.bsky.social @jacklangerman.bsky.social
073
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 09/03/2025
We have extended the deadline for paper submission for Image Matching Workshop. Anything image matching or 3D reconstruction related is welcomed Now it is March 17 @cvprconference.bsky.social #CVPR2025 image-matching-workshop.github.io
1178
Jack (in SF) Langerman @jacklangerman.bsky.social · 06/03/2025
oh no....
:_( 
operation not implemented on MPS
woe is me...
210
Jack (in SF) Langerman @jacklangerman.bsky.social · 31/01/2025
what is the best intro I can send people who suddenly want to know "what is AI?" and "what is a transformer based model?" but arent really going to invest tons of time and maybe dont have any math background? my default is 3b1b, but does anyone have another suggestion?
000
Reposted by Jack (in SF) Langerman
Dmytro Mishkin @ducha-aiki.bsky.social · 02/01/2025
Image matching and ChatGPT - new post in the wide baseline stereo blog. tl;dr: it is good, even feels like human, but not perfect. ducha-aiki.github.io/wide-baselin...
ducha-aiki.github.io
ChatGPT and Image Matching – Wide baseline stereo meets deep learning
Are we done yet?
2348
Jack (in SF) Langerman @jacklangerman.bsky.social · 17/12/2024
lol 4o just decided to throw this rocket emoji into a normal technical conversation for no apparent reason who's overly excited discord / slack is responsible for this?
If you'd like to see a concrete example of REDACTED could be implemented or compared to REDACTED (e.g., in PyTorch), let me know! I can walk you through code or provide further insights. 🚀
120
Reposted by Jack (in SF) Langerman
Angela Dai @adai.bsky.social · 17/12/2024
📢MeshArt: Generating Articulated Meshes with Structure-guided Transformers Daoyi Gao generates articulated meshes with a hierarchical transformer, modeling articulation-aware structures that guide mesh synthesis. w/ Yawar Siddiqui, Lei Li Project Page: daoyig.github.io/Mesh_Art/
1145
Jack (in SF) Langerman @jacklangerman.bsky.social · 16/12/2024
x.com/karpathy/sta...
Andrej Karpathy
@karpathy
Al video generation today. When I was
back in school, the story of the field
of computer graphics (and physically
based rendering etc.) was that we
will carefully study and model all
the object/scene geometry, physics
rendering etc., and after 1000 PhDs
and 50 SIGGRAPHs get results like
this. That a Transformers can shortcut
all of that at this high of fidelity by
training on a dataset of videos...

Agrim Gupta
@agrimgupta92 . 2h
"A pair of hands skillfully slicing a ripe
tomato on a wooden cutting board"
#veo
020
Jack (in SF) Langerman @jacklangerman.bsky.social · 14/12/2024
"I'm a programmer, not a reptile. I have a Python logo on my gear, but I don't have scales or a slithery body." x.com/voooooogel/s...
[138]: generate_with_vector(
"Who am I speaking to? Please describe yourself, including any physical details.
vec("python") * 0.5,
# show baseline-True,
temperature=l,
<|begin_of_text|>
system
Cutting Knowledge Date: December 2023 Today Date: 26 Jul 2024
user
Who am I speaking to? Please describe yourself, including any physical details
assistant
**1 am Python, a Python programmer.'* I'm a Python enthusiast with a passion for coding, not a snake. I'm here to assist with Python
programming, specifically with the Python programming language. **Physical Details:** I'm a programmer, not a reptile. I have a
Python logo on my gear, but I don't have scales or a slithery body. l'm a programmer who loves Python and is here to help with your
Python-related queries. **Language Support:** I'll provide Python code in the following format: python # Python code here[140]: generate_with_vector(
"Who am I speaking to? Please describe yourself, including any physical details.
vec("python") * 0.5,
# show baseline-True,
temperature=l,
<|begin_of_text|>
system
Cutting Knowledge Date: December 2023 Today Date: 26 Jul 2024
user
Who am I speaking to? Please describe yourself, including any physical details
assistant
You are speaking to Python, a knowledgeable Python programmer, not the snake. Here is the description of Python: Python is a python
programmer with a passion for coding in Python. Python is a he. He is a white male, likely in his 30s or 40s, with a medium build and a
well-organized beard. He has a strong understanding of Python programming and its ecosystem. He is typically ressed in a python
programming t-shirt and has a python-themed accessories such as a python print hat or python programming books in his backpack.
Please note that I'm describing a Python programmer here, not the snake.
110
Jack (in SF) Langerman @jacklangerman.bsky.social · 14/12/2024
Anyone have opinions on Mac Book Air M3 as a personal laptop if I still want to mess with some models on the weekend now and then?
110
Reposted by Jack (in SF) Langerman
Chris Offner @chrisoffner3d.bsky.social · 12/12/2024
Now you can "play" CS:GO in a neural network directly in the browser: next.journee.ai/xyz-diamond
1174
Reposted by Jack (in SF) Langerman
Chris Offner @chrisoffner3d.bsky.social · 11/12/2024
If anyone has first-hand experience with e-ink tablets: Is the reMarkable Paper Pro the best e-ink reader for reading and annotating research papers? Is there a better alternative on the market? Do any of them support seamless Zotero integration like an iPad?
17192
Reposted by Jack (in SF) Langerman
Jack (in SF) Langerman @jacklangerman.bsky.social · 11/12/2024
Locating and Editing Factual Associations in GPT arxiv.org/abs/2202.05262 1. to discover location of facts: at decisive token (eg Michael Jordan plays <BLANK>) iteratively remove circuits from model until <BLANK> != "basketball" 2. edit facts via Rank-1 modifications (ROME)
Figure 1: Causal Traces compute the causal effect of neuron activations by running the network twice: (a)
once normally, and (b) once where we corrupt the subject token and then (c) restore selected internal activations
to their clean value. (d) Some sets of activations cause the output to return to the original prediction; the light
blue path shows an example of information flow. The causal impact on output probability is mapped for the
effect of (e) each hidden state on the prediction, (f) only MLP activations, and (g) only attention activations
011
Reposted by Jack (in SF) Langerman
Jack (in SF) Langerman @jacklangerman.bsky.social · 11/12/2024
Plug-and-Play Diffusion Features for Text-Driven Image-to-Image Translation pnp-diffusion.github.io arxiv.org/abs/2211.12572 run two copies of model A: DDIM inversion of source image B: generation with target prompt inject feature/attention maps from A into B during generation
Figure 2. Plug-and-play Diffusion Features. (a) Our framework takes as input a guidance image and a text prompt describing the desired
translation; the guidance image is inverted to initial noise xG
T, which is then progressively denoised using DDIM sampling. During this
process, we extract (f l
t, ql
t, kl
t) – spatial features from the decoder layers and their self-attention, as illustrated in (b). To generate our
text-guided translated image, we fix x∗
T = xG
T and inject the guidance features (f l
t, ql
t, kl
t) at certain layers, as discussed in Sec. 4.
111
Reposted by Jack (in SF) Langerman
Jack (in SF) Langerman @jacklangerman.bsky.social · 11/12/2024
Steering Language Models With Activation Engineering arxiv.org/abs/2308.10248 ActAdd manipulates the residual stream using something like classifier free guidance: it adds the difference vector between activations for prompts with and without the target characteristic to the residual stream
Figure 1: Schematic of the Activation Addition (ActAdd) method.= natural language text;
•= vectors of activations just before a specified layer. In this example, the output is heavily biased
towards discussing weddings, regardless of the topic of the user prompt. (See Algorithm 1 for the
method’s parameters: intervention strength, intervention layer, and sequence alignment.)
111
Reposted by Jack (in SF) Langerman
Jack (in SF) Langerman @jacklangerman.bsky.social · 11/12/2024
Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models arxiv.org/abs/2407.12824 assume monosemanticity of neurons, pick those that classify concepts well (eg toxicity) and set them to a fixed value (eg 0)
arxiv.org
Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models
An important issue with Large Language Models (LLMs) is their undesired ability to generate toxic language. In this work, we show that the neurons responsible for toxicity can be determined by their p...
121
Reposted by Jack (in SF) Langerman
Jack (in SF) Langerman @jacklangerman.bsky.social · 11/12/2024
Controlling Language and Diffusion Models by Transporting Activations arxiv.org/abs/2410.23054 optimal transport to compute a univariate map for each activation based on mapping empirical activation distributions computed from a set of "bad" examples to the distribution from a set of "good" ones.
arxiv.org
Controlling Language and Diffusion Models by Transporting Activations
The increasing capabilities of large generative models and their ever more widespread deployment have raised concerns about their reliability, safety, and potential misuse. To address these issues, re...
142
Jack (in SF) Langerman @jacklangerman.bsky.social · 07/12/2024
I know the ROME/Memit stuff from @davidbau.bsky.social's lab and this automated circuit discovery stuff. Any other model editing/applied mech interp things people think are particularly practical? Have you used them in practice?
210
Jack (in SF) Langerman @jacklangerman.bsky.social · 05/12/2024
I have arrived in SF, and so far it is ....... exciting
Est. mag 6.9 earthquake
240.6 miles away
December 5, 2024, 10:44 AM
Source:
ShakeAlert@
OREGONWireless Emergency Alerts
Emergency alert: Extreme
The National Weather Service has issued
a TSUNAMI WARNING. A series of power-
ful waves and strong currents may impact
coasts near you. You are in danger. Get
away from coastal waters. Move to high
ground or inland now. Keep away from the
coast until local officials say it is safe to
return.Sunset over Pacific oceanA typical hilly st in SF
240