Sign in

Finbarr

@finbarr.bsky.social
2.2K followers 58 following 89 posts

building the future research at midjourney, deepmind. slinging ai hot takes 🥞at artfintel.com

PostsRepliesMedia
Finbarr @finbarr.bsky.social · 07/12/2024
This is one of my all time favorite papers: openreview.net/forum?id=ByJ... It shows that, under fair experimental evaluation, lstms do just as well as a bunch of “improvements”
openreview.net
On the State of the Art of Evaluation in Neural Language Models
Show that LSTMs are as good or better than recent innovations for LM and that model evaluation is often unreliable.
3254
Finbarr @finbarr.bsky.social · 05/12/2024
Really fun conversation with @natolambert.bsky.social!
050
Finbarr @finbarr.bsky.social · 03/12/2024
Apparently there *is* another finbar(r) in Alberta.
050
Finbarr @finbarr.bsky.social · 03/12/2024
New homeowner fear unlocked; someone hit and ran my neighbor’s garage
130
Finbarr @finbarr.bsky.social · 01/12/2024
there’s a type of “not trying” which means not executing at the level of competence of a $XX billion corporation this is the complaint about eg Google products. They’re good! better than most startups! But not “trillion dollar corporation famed for engineering expertise” good.
080
Finbarr @finbarr.bsky.social · 30/11/2024
I watched too many ski movies and now am trying to convince my wife we should move to Alaska
250
Finbarr @finbarr.bsky.social · 30/11/2024
building my own mlp implementation from scratch in numpy, including backprop, remains one of the most educational exercises I’ve done
2190
Finbarr @finbarr.bsky.social · 26/11/2024
Love this. Very clean implementations of various inference optimizations.
160
Finbarr @finbarr.bsky.social · 26/11/2024
Agreed! Folk knowledge is worth publishing!
1100
Finbarr @finbarr.bsky.social · 26/11/2024
my latest article for Artificial Fintelligence is up. i cover the evolution of Vision Language Models over the past few years, from complex architectures to surprisingly simple & effective ones. (link in next tweet)
4384
Finbarr @finbarr.bsky.social · 25/11/2024
the tech industry is like California: it gets a few of the most important things right, so it can make a lot of other mistakes and still be wildly successful
030
Finbarr @finbarr.bsky.social · 25/11/2024
the remarkable success of the Google brain (and OpenAI) resident programs is an indication to me that smart, hardworking people can do more than you expect
4191
Finbarr @finbarr.bsky.social · 24/11/2024
My favorite thing about Bsky so far is not having the dumb algo requirements. Link in reply and not mentioning substack is stupid!
3110
Finbarr @finbarr.bsky.social · 24/11/2024
for all the work we researchers do, the best way to improve your model by far is to 1) use better data and 2) use higher quality data
170
Finbarr @finbarr.bsky.social · 23/11/2024
If I was doing a phd, this would be one of my top choices for programs.
020
Finbarr @finbarr.bsky.social · 23/11/2024
active learning is top of my list of "things that seem like they should work but don't" I haven't had much success when I actually implement it
140
Finbarr @finbarr.bsky.social · 22/11/2024
Why does DeepSeek have so many GPUs? (Purportedly >10k H100s). Is this useful to their main hedge fund business?
110
Finbarr @finbarr.bsky.social · 21/11/2024
Tulu is very exciting!
020
Finbarr @finbarr.bsky.social · 20/11/2024
Deepseek r1 seems really good
140
Finbarr @finbarr.bsky.social · 19/11/2024
If I was in a position to direct significant research resources/headcount I’d be putting a significant effort behind better exploration in RL.
180
Finbarr @finbarr.bsky.social · 19/11/2024
honestly if houses in Whistler were less expensive I’d be way less motivated to earn money the prospect of retiring there motivates a non-trivial amount of my life
020
Finbarr @finbarr.bsky.social · 19/11/2024
the amount of time I spend thinking about skiing rn is ridiculous
010
Finbarr @finbarr.bsky.social · 19/11/2024
ok but seriously at some point we really need to solve exploration in RL
280
Reposted by Finbarr
Venkatesh Rao 🔹 @vgr.bsky.social · 05/05/2023
galaxy brain take: Google no moats memo is a psyop and a no-downside shot against openAI It’s a little too flattering to the conceits of open source etc
4232
Finbarr @finbarr.bsky.social · 04/05/2023
chain of thought reasoning keeps racking up the Ws
140
Finbarr @finbarr.bsky.social · 04/05/2023
my thesis is that there’s gonna be one foundation model company for each major cloud provider that they either partner with or acquire, and these are the only companies that make money Azure: OpenAI Amazon: Stability (?) GCP: Anthropic/DeepMind
010
Finbarr @finbarr.bsky.social · 03/05/2023
the most useful thing I’ve done in my career as a research engineer is to slowly build a bag of tricks and to try them on all the problems I come across
100
Finbarr @finbarr.bsky.social · 01/05/2023
active learning seems like what everyone should be doing Particularly in RL
000
Finbarr @finbarr.bsky.social · 01/05/2023
ok emergent properties of LLMs have nothing on emergent properties of babies my son just learned to say dada, very exciting
000
Finbarr @finbarr.bsky.social · 28/04/2023
copilot makes it so easy to comment my code & write docstrings my coworkers should be paying github
010
Finbarr @finbarr.bsky.social · 28/04/2023
the fact humans haven’t eliminated mosquitoes from the earth is the best argument against AI foomerism
010
Finbarr @finbarr.bsky.social · 28/04/2023
not that “tweeted” is particularly dignified, but “as I skeeted yesterday” is egregiously bad
000
Finbarr @finbarr.bsky.social · 28/04/2023
is anyone actually using cerebras in production
000
Finbarr @finbarr.bsky.social · 28/04/2023
in the pnw they use greysky instead
000
Finbarr @finbarr.bsky.social · 28/04/2023
skoot skoot mahout
000
Finbarr @finbarr.bsky.social · 28/04/2023
there are so many aspects of productionizing ML models we know nothing about
000
Finbarr @finbarr.bsky.social · 28/04/2023
it’s crazy how little LLM experience you need to be an “expert”
100
Finbarr @finbarr.bsky.social · 18/04/2023
Ok American Jesus ads during the playoffs are a new experience
000
Finbarr @finbarr.bsky.social · 18/04/2023
American fast food ads are so much more enticing than Canadian ones
100
Finbarr @finbarr.bsky.social · 18/04/2023
how hard could it be to start a semiconductor company
000
Finbarr @finbarr.bsky.social · 16/04/2023
I keep thinking about parameter quantization for large models I have periods where I’m convinced it’s the future of inference, and then other periods where m I’m convinced it doesn’t do anything Any anecdata out there? 👀
000