Finbarr @finbarr.bsky.social · 07/12/2024This is one of my all time favorite papers: openreview.net/forum?id=ByJ... It shows that, under fair experimental evaluation, lstms do just as well as a bunch of “improvements”openreview.netOn the State of the Art of Evaluation in Neural Language ModelsShow that LSTMs are as good or better than recent innovations for LM and that model evaluation is often unreliable. 3254
Finbarr @finbarr.bsky.social · 07/12/2024openreview.net/forum?id=ByJ...openreview.netOn the State of the Art of Evaluation in Neural Language ModelsShow that LSTMs are as good or better than recent innovations for LM and that model evaluation is often unreliable. 010
Finbarr @finbarr.bsky.social · 06/12/2024Fun fact: I recently encountered (well, saw on the news) the only other person named finbarr in Canada I’ve ever seen. The only issue is, he was an arsonist who set a ton of fires in Edmonton. 120
Finbarr @finbarr.bsky.social · 03/12/2024This is mckernan! What I thought was a nice neighborhood 😂 110
Finbarr @finbarr.bsky.social · 03/12/2024New homeowner fear unlocked; someone hit and ran my neighbor’s garage 130
Finbarr @finbarr.bsky.social · 01/12/2024there’s a type of “not trying” which means not executing at the level of competence of a $XX billion corporation this is the complaint about eg Google products. They’re good! better than most startups! But not “trillion dollar corporation famed for engineering expertise” good. 080
Finbarr @finbarr.bsky.social · 30/11/2024I watched too many ski movies and now am trying to convince my wife we should move to Alaska 250
Finbarr @finbarr.bsky.social · 30/11/2024building my own mlp implementation from scratch in numpy, including backprop, remains one of the most educational exercises I’ve done 2190
Finbarr @finbarr.bsky.social · 26/11/2024Ha, it’s been on my list of todos for a while! I’m glad someone got to it. 010
Finbarr @finbarr.bsky.social · 26/11/2024Love this. Very clean implementations of various inference optimizations. 160
Finbarr @finbarr.bsky.social · 26/11/2024I mailed this out like a month ago and just never did the promo 🙈 120
Finbarr @finbarr.bsky.social · 26/11/2024again, link is: www.artfintel.com/p/papers-ive...artfintel.comPapers I've read this week: vision language modelsThey kept releasing VLMs, so I kept writing... 040
Finbarr @finbarr.bsky.social · 26/11/2024seems like we're seeing convergence in VLM design. most recent models (Pixtral, PaliGemma, etc) are moving away from complex fusion techniques toward simpler approaches as usual, the bitter lesson holds: better to learn structure than impose it incompleteideas.net/IncIdeas/Bit...incompleteideas.netThe Bitter Lesson 170
Finbarr @finbarr.bsky.social · 26/11/2024open source VLMs use relatively little compute compared to what you might expect: LLaVa: 768 A100 hours DeepSeek-VL: 61,440 A100 hours PaliGemma: ~12k A100 hours (for reference, Stable Diffusion used 150k A100 hours) 140
Finbarr @finbarr.bsky.social · 26/11/2024what i found interesting: VLMs are way simpler than they first appear. current SOTA is basically: 1. ViT encoder (init from SigLIP/CLIP) 2. pretrained LLM base 3. concat image features with text 4. finetune 141
Finbarr @finbarr.bsky.social · 26/11/2024link: www.artfintel.com/p/papers-ive...artfintel.comPapers I've read this week: vision language modelsThey kept releasing VLMs, so I kept writing... 131
Finbarr @finbarr.bsky.social · 26/11/2024my latest article for Artificial Fintelligence is up. i cover the evolution of Vision Language Models over the past few years, from complex architectures to surprisingly simple & effective ones. (link in next tweet) 4384
Finbarr @finbarr.bsky.social · 25/11/2024I mean I can’t get a job at an elite law firm in New York without 1) a law degree 2) from an elite institution 010
Finbarr @finbarr.bsky.social · 25/11/2024I think OpenAI’s program is still running? I know one of the main sora contributors just did it. Google makes changes for non-obvious reasons, not always because it’s the optimal choice (see: laying off Rich Sutton) 140
Finbarr @finbarr.bsky.social · 25/11/2024the tech industry is like California: it gets a few of the most important things right, so it can make a lot of other mistakes and still be wildly successful 030
Finbarr @finbarr.bsky.social · 25/11/2024I suspect we'd see similar outcomes if there were e.g. a Cravath Legal Resident program or a Mayo Clinic Surgical Resident program the barrier to entry is b/c of tech's lack of credential gate keeping, which is one of tech’s greatest advantages 120
Finbarr @finbarr.bsky.social · 25/11/2024the remarkable success of the Google brain (and OpenAI) resident programs is an indication to me that smart, hardworking people can do more than you expect 4191
Finbarr @finbarr.bsky.social · 24/11/2024My favorite thing about Bsky so far is not having the dumb algo requirements. Link in reply and not mentioning substack is stupid! 3110
Finbarr @finbarr.bsky.social · 24/11/2024I think I’ve watched 1 full length movie since Malcolm was born 😭 120
Finbarr @finbarr.bsky.social · 24/11/2024for all the work we researchers do, the best way to improve your model by far is to 1) use better data and 2) use higher quality data 170
Finbarr @finbarr.bsky.social · 23/11/2024Yeah. “Just label more” really seems to be the way to go. 010
Finbarr @finbarr.bsky.social · 23/11/2024If I was doing a phd, this would be one of my top choices for programs. 020
Finbarr @finbarr.bsky.social · 23/11/2024active learning is top of my list of "things that seem like they should work but don't" I haven't had much success when I actually implement it 140
Finbarr @finbarr.bsky.social · 22/11/2024Ahhh maybe I’d trust his numbers, I have no special knowledge here 000