Sign in

Finbarr

@finbarr.bsky.social
2.2K followers 58 following 89 posts

building the future research at midjourney, deepmind. slinging ai hot takes 🥞at artfintel.com

PostsRepliesMedia
Finbarr @finbarr.bsky.social · 07/12/2024
would love your take when you do!
050
Finbarr @finbarr.bsky.social · 07/12/2024
🥸
050
Finbarr @finbarr.bsky.social · 07/12/2024
This is one of my all time favorite papers: openreview.net/forum?id=ByJ... It shows that, under fair experimental evaluation, lstms do just as well as a bunch of “improvements”
openreview.net
On the State of the Art of Evaluation in Neural Language Models
Show that LSTMs are as good or better than recent innovations for LM and that model evaluation is often unreliable.
3254
Finbarr @finbarr.bsky.social · 07/12/2024
openreview.net/forum?id=ByJ...
openreview.net
On the State of the Art of Evaluation in Neural Language Models
Show that LSTMs are as good or better than recent innovations for LM and that model evaluation is often unreliable.
010
Finbarr @finbarr.bsky.social · 07/12/2024
I’ll find it one sec
110
Finbarr @finbarr.bsky.social · 06/12/2024
Very likely.
000
Finbarr @finbarr.bsky.social · 06/12/2024
🙏
000
Finbarr @finbarr.bsky.social · 06/12/2024
Fun fact: I recently encountered (well, saw on the news) the only other person named finbarr in Canada I’ve ever seen. The only issue is, he was an arsonist who set a ton of fires in Edmonton.
120
Finbarr @finbarr.bsky.social · 05/12/2024
Really fun conversation with @natolambert.bsky.social!
050
Finbarr @finbarr.bsky.social · 03/12/2024
This is mckernan! What I thought was a nice neighborhood 😂
110
Finbarr @finbarr.bsky.social · 03/12/2024
Apparently there *is* another finbar(r) in Alberta.
050
Finbarr @finbarr.bsky.social · 03/12/2024
New homeowner fear unlocked; someone hit and ran my neighbor’s garage
130
Finbarr @finbarr.bsky.social · 02/12/2024
I thought that was ai gen at first!
130
Finbarr @finbarr.bsky.social · 01/12/2024
there’s a type of “not trying” which means not executing at the level of competence of a $XX billion corporation this is the complaint about eg Google products. They’re good! better than most startups! But not “trillion dollar corporation famed for engineering expertise” good.
080
Finbarr @finbarr.bsky.social · 30/11/2024
would also accept Austria
240
Finbarr @finbarr.bsky.social · 30/11/2024
I watched too many ski movies and now am trying to convince my wife we should move to Alaska
250
Finbarr @finbarr.bsky.social · 30/11/2024
building my own mlp implementation from scratch in numpy, including backprop, remains one of the most educational exercises I’ve done
2190
Finbarr @finbarr.bsky.social · 30/11/2024
Welcome!
020
Finbarr @finbarr.bsky.social · 26/11/2024
🙏
010
Finbarr @finbarr.bsky.social · 26/11/2024
Ha, it’s been on my list of todos for a while! I’m glad someone got to it.
010
Finbarr @finbarr.bsky.social · 26/11/2024
Love this. Very clean implementations of various inference optimizations.
160
Finbarr @finbarr.bsky.social · 26/11/2024
Agreed! Folk knowledge is worth publishing!
1100
Finbarr @finbarr.bsky.social · 26/11/2024
I mailed this out like a month ago and just never did the promo 🙈
120
Finbarr @finbarr.bsky.social · 26/11/2024
Force of habit!
050
Finbarr @finbarr.bsky.social · 26/11/2024
Ahh you’re right!
010
Finbarr @finbarr.bsky.social · 26/11/2024
again, link is: www.artfintel.com/p/papers-ive...
artfintel.com
Papers I've read this week: vision language models
They kept releasing VLMs, so I kept writing...
040
Finbarr @finbarr.bsky.social · 26/11/2024
seems like we're seeing convergence in VLM design. most recent models (Pixtral, PaliGemma, etc) are moving away from complex fusion techniques toward simpler approaches as usual, the bitter lesson holds: better to learn structure than impose it incompleteideas.net/IncIdeas/Bit...
incompleteideas.net
The Bitter Lesson
170
Finbarr @finbarr.bsky.social · 26/11/2024
open source VLMs use relatively little compute compared to what you might expect: LLaVa: 768 A100 hours DeepSeek-VL: 61,440 A100 hours PaliGemma: ~12k A100 hours (for reference, Stable Diffusion used 150k A100 hours)
140
Finbarr @finbarr.bsky.social · 26/11/2024
what i found interesting: VLMs are way simpler than they first appear. current SOTA is basically: 1. ViT encoder (init from SigLIP/CLIP) 2. pretrained LLM base 3. concat image features with text 4. finetune
141
Finbarr @finbarr.bsky.social · 26/11/2024
link: www.artfintel.com/p/papers-ive...
artfintel.com
Papers I've read this week: vision language models
They kept releasing VLMs, so I kept writing...
131
Finbarr @finbarr.bsky.social · 26/11/2024
my latest article for Artificial Fintelligence is up. i cover the evolution of Vision Language Models over the past few years, from complex architectures to surprisingly simple & effective ones. (link in next tweet)
4384
Finbarr @finbarr.bsky.social · 25/11/2024
I mean I can’t get a job at an elite law firm in New York without 1) a law degree 2) from an elite institution
010
Finbarr @finbarr.bsky.social · 25/11/2024
Thank you!
010
Finbarr @finbarr.bsky.social · 25/11/2024
what's wrong with the pixel art?
110
Finbarr @finbarr.bsky.social · 25/11/2024
oooo ty
010
Finbarr @finbarr.bsky.social · 25/11/2024
I think OpenAI’s program is still running? I know one of the main sora contributors just did it. Google makes changes for non-obvious reasons, not always because it’s the optimal choice (see: laying off Rich Sutton)
140
Finbarr @finbarr.bsky.social · 25/11/2024
the tech industry is like California: it gets a few of the most important things right, so it can make a lot of other mistakes and still be wildly successful
030
Finbarr @finbarr.bsky.social · 25/11/2024
I suspect we'd see similar outcomes if there were e.g. a Cravath Legal Resident program or a Mayo Clinic Surgical Resident program the barrier to entry is b/c of tech's lack of credential gate keeping, which is one of tech’s greatest advantages
120
Finbarr @finbarr.bsky.social · 25/11/2024
the remarkable success of the Google brain (and OpenAI) resident programs is an indication to me that smart, hardworking people can do more than you expect
4191
Finbarr @finbarr.bsky.social · 24/11/2024
Ok that is neat.
010
Finbarr @finbarr.bsky.social · 24/11/2024
My favorite thing about Bsky so far is not having the dumb algo requirements. Link in reply and not mentioning substack is stupid!
3110
Finbarr @finbarr.bsky.social · 24/11/2024
I think I’ve watched 1 full length movie since Malcolm was born 😭
120
Finbarr @finbarr.bsky.social · 24/11/2024
for all the work we researchers do, the best way to improve your model by far is to 1) use better data and 2) use higher quality data
170
Finbarr @finbarr.bsky.social · 23/11/2024
Yeah. “Just label more” really seems to be the way to go.
010
Finbarr @finbarr.bsky.social · 23/11/2024
If I was doing a phd, this would be one of my top choices for programs.
020
Finbarr @finbarr.bsky.social · 23/11/2024
I’ll also be going to RLC, to be clear.
000
Finbarr @finbarr.bsky.social · 23/11/2024
active learning is top of my list of "things that seem like they should work but don't" I haven't had much success when I actually implement it
140
Finbarr @finbarr.bsky.social · 23/11/2024
Yes! I live in Edmonton. Would love to chat.
100
Finbarr @finbarr.bsky.social · 22/11/2024
Are you coming to Edmonton for RLC?
110
Finbarr @finbarr.bsky.social · 22/11/2024
Ahhh maybe I’d trust his numbers, I have no special knowledge here
000