Sign in

pentagonalize.bsky.social

@pentagonalize.bsky.social
16 followers 21 following 13 posts
PostsRepliesMedia
pentagonalize.bsky.social @pentagonalize.bsky.social · 12/02/2026
The FLaNN Workshop submission deadline has been extended to Feb 19! Invited talks + posters (non-archival): expressivity, computation, and learning in neural nets/LLMs. Previous work welcome. Graduate students encouraged to submit! 📍 Yale University 🗓️ May 11-13, 2026
100
pentagonalize.bsky.social @pentagonalize.bsky.social · 04/02/2026
📣 FLaNN 2026 at Yale 🍮 Invited talks+posters (non-archival): expressivity, computation, and learning in neural nets/LLMs Speakers: Pablo Barceló, David Chiang, Will Merrill, Naomi Saphra, Gail Weiss Abstracts due Feb 12, 2026 Details: flann.cs.yale.edu
An advertisement for the Formal Languages and Neural Networks workshop. It has the date, a call for papers, the website+email, and a list of speakers with their names, headshots, and institutional affiliations (Pablo Barceló, David Chiang, Will Merrill, Naomi Saphra, and Gail Weiss)
232
pentagonalize.bsky.social @pentagonalize.bsky.social · 31/01/2026
Deadline in just under two weeks!
011
pentagonalize.bsky.social @pentagonalize.bsky.social · 19/12/2025
Announcing the first Workshop on Formal Languages and Neural Networks (FLaNN)! We invite the submission of abstracts for posters that discuss the formal expressivity, computational properties, and learning behavior of neural network models, including large language models (LLMs).
1105
pentagonalize.bsky.social @pentagonalize.bsky.social · 03/10/2025
We present The Transformer Cookbook: a collection of recipes for programming algorithms directly into transformers! Hungry for an induction head? Craving a Dyck language recognizer? We show you step-by-step how to cook up transformers for these algorithms and many more!
arxiv.org
The Transformer Cookbook
We present the transformer cookbook: a collection of techniques for directly encoding algorithms into a transformer's parameters. This work addresses the steep learning curve of such endeavors, a prob...
155
Reposted by @pentagonalize.bsky.social
dchiang.bsky.social @dchiang.bsky.social · 23/12/2024
New paper and two not-so-new papers on arXiv about transformer expressivity: (1) With @pentagonalize and Dana Angluin, "Simulating Hard Attention Using Soft Attention" arxiv.org/abs/2412.09925
arxiv.org
Simulating Hard Attention Using Soft Attention
We study conditions under which transformers using soft attention can simulate hard attention, that is, effectively focus all attention on a subset of positions. First, we examine several variants of ...
231