Sign in

Anthony Fuller

@anthonyfuller.bsky.social
155 followers 1.5K following 5 posts

PhD Student at Carleton University (Ottawa, Canada) antofuller.github.io

PostsRepliesMedia
Reposted by Anthony Fuller
Nathan Lambert @natolambert.bsky.social · 21/07/2026
My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by me finding time to study and document the fundamentals on nights and weekends since 2024.
717521
Reposted by Anthony Fuller
TMLR Published Papers @tmlr-pub.bsky.social · 19/06/2026
DINOv3 Oriane Siméoni, Huy V. Vo, Maximilian Seitzer et al. Action editor: Jingcai Guo openreview.net/forum?id=2NlGyqNjns #visual #features #supervised
0154
Reposted by Anthony Fuller
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 09/05/2026
Going to fill an unmet demand and produce a “computer science for evil” course
111377
Reposted by Anthony Fuller
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 03/05/2026
The prevalence of acadamese in scientific writing is partly why writing overall in science has been devalued. The average scientific paper is boring to read and has a boilerplate style, a void into which LLMs jump
9603
Anthony Fuller @anthonyfuller.bsky.social · 04/02/2026
Thanks for comparing with LookHere. Looking forward to reading this 🙂
110
Reposted by Anthony Fuller
Mila - Institut québécois d'IA @mila-quebec.bsky.social · 24/10/2025
Galileo, created by Mila researchers Gabriel Tseng and David Rolnick, uses AI to uncover trends across decades of satellite and human activity data—revealing early signals about our planet’s health and helping us act in time. mila.quebec/en/article/d...
051
Anthony Fuller @anthonyfuller.bsky.social · 29/09/2025
LookWhere has been accepted to NeurIPS 2025! LookWhere accelerates inference and fine-tuning by approximating full, deep representations with adaptive computation of predictions learned from distillation. Paper: arxiv.org/abs/2505.18051 Code and weights: github.com/antofuller/l...
020
Reposted by Anthony Fuller
David Rolnick @drolnick.bsky.social · 08/06/2025
Our remote sensing foundation model Galileo has been accepted to ICML 2025! Galileo outperforms state-of-the-art across different input data modalities and shapes, and using it requires only minimal data and compute. More at: arxiv.org/abs/2502.09356
1246
Reposted by Anthony Fuller
Mark Carney @mark-carney.bsky.social · 22/03/2025
Elbows up, Canada.
14735201313263
Anthony Fuller @anthonyfuller.bsky.social · 27/02/2025
Lookhere position encoding could help with this: arxiv.org/abs/2405.13985
arxiv.org
LookHere: Vision Transformers with Directed Attention Generalize and Extrapolate
High-resolution images offer more information about scenes that can improve model accuracy. However, the dominant model architecture in computer vision, the vision transformer (ViT), cannot effectivel...
020
Anthony Fuller @anthonyfuller.bsky.social · 04/02/2025
Awesome, thanks for the explanation!
000
Anthony Fuller @anthonyfuller.bsky.social · 04/02/2025
Cool work? I think the position encoding method looks similar to DeBERTa’s disentangled attention: arxiv.org/abs/2006.03654
arxiv.org
DeBERTa: Decoding-enhanced BERT with Disentangled Attention
Recent progress in pre-trained neural language models has significantly improved the performance of many natural language processing (NLP) tasks. In this paper we propose a new model architecture DeBE...
100
Reposted by Anthony Fuller
Michael J. Black @michael-j-black.bsky.social · 20/11/2024
For those who missed this post on the-network-that-is-not-to-be-named, I made public my "secrets" for writing a good CVPR paper (or any scientific paper). I've compiled these tips of many years. It's long but hopefully it helps people write better papers. perceiving-systems.blog/en/post/writ...
perceiving-systems.blog
Writing a good scientific paper
426065