Sign in

Jack Hessel

@jmhessel.bsky.social
3K followers 218 following 27 posts

jmhessel.com Seattle bike lane enjoyer. Opinions my own.

PostsRepliesMedia
Reposted by Jack Hessel
Nathan Lambert @natolambert.bsky.social · 17/07/2025
It is a major policy failure that the US cannot accommodate top AI conferences due to visa issues. buff.ly/DRJOGrB
48421
Jack Hessel @jmhessel.bsky.social · 24/06/2025
bring back 8 page neurips papers
030
Jack Hessel @jmhessel.bsky.social · 20/06/2025
m̶e̶n̶ Americans will literally l̶e̶a̶r̶n̶ ̶e̶v̶e̶r̶y̶t̶h̶i̶n̶g̶ ̶a̶b̶o̶u̶t̶ ̶a̶n̶c̶i̶e̶n̶t̶ ̶R̶o̶m̶e̶ invest billions into self driving cars instead of g̶o̶i̶n̶g̶ ̶t̶o̶ ̶t̶h̶e̶r̶a̶p̶y̶ building transit
070
Jack Hessel @jmhessel.bsky.social · 06/06/2025
bring back length limits for author responses
1100
Jack Hessel @jmhessel.bsky.social · 04/06/2025
in llm-land, what is a tool, a function, an agent, and (most elusive of all): a "multi-agent system"? (This had been bothering me recently; are all these the same?) @yoavgo.bsky.social's blog is a clarifying read on the topic -- I plan to adopt his terminology :-) gist.github.com/yoavg/9142e5...
gist.github.com
What makes multi-agent LLM systems multi-agent?
What makes multi-agent LLM systems multi-agent? GitHub Gist: instantly share code, notes, and snippets.
050
Jack Hessel @jmhessel.bsky.social · 27/03/2025
If you're in WA and think imposing new taxes on things we want more of (e.g., bikes, transit) is a bad idea, consider contacting your reps using this simple form! <3
050
Jack Hessel @jmhessel.bsky.social · 24/02/2025
Should you delete softmax from your attention layers? check out Songling Yang's (sustcsonglin.github.io) tutorial, moderated by @srushnlp.bsky.social, for a beginner-friendly tutorial of the why/how/beauty of linear attention :-) www.youtube.com/watch?v=d0HJ...
sustcsonglin.github.io
Songlin Yang
A simple, whitespace theme for academics. Based on [*folio](https://github.com/bogoli/-folio) design.
041
Reposted by Jack Hessel
Zach Levonian @zwlevonian.bsky.social · 30/01/2025
I've spent the last two years trying to understand how LLMs might improve middle-school math education. I just published an article in the Journal of Educational Data Mining describing some of that work: "Designing Safe and Relevant Generative Chats for Math Learning in Intelligent Tutoring Systems"
jedm.educationaldatamining.org
Journal of Educational Data Mining
Large language models (LLMs) are flexible, personalizable, and available, which makes their use within Intelligent Tutoring Systems (ITSs) appealing. However, their flexibility creates risks: inaccura...
072
Reposted by Jack Hessel
Melanie Mitchell @melaniemitchell.bsky.social · 30/01/2025
Very good (technical) explainer answering "How has DeepSeek improved the Transformer architecture?". Aimed at readers already familiar with Transformers. epoch.ai/gradient-upd...
epoch.ai
How has DeepSeek improved the Transformer architecture?
This Gradient Updates issue goes over the major changes that went into DeepSeek’s most recent model.
627763
Reposted by Jack Hessel
Chris Potts @cgpotts.bsky.social · 13/01/2025
I've posted the practice run of my LSA keynote. My core claim is that LLMs can be useful tools for doing close linguistic analysis. I illustrate with a detailed case study, drawing on corpus evidence, targeted syntactic evaluations, and causal intervention-based analyses: youtu.be/DBorepHuKDM
youtu.be
Finding linguistic structure in large language models
YouTube video by Chris Potts
17420
Reposted by Jack Hessel
Simon Willison @simonwillison.net · 31/12/2024
Here's my end-of-year review of things we learned out about LLMs in 2024 - we learned a LOT of things simonwillison.net/2024/Dec/31/... Table of contents:

    The GPT-4 barrier was comprehensively broken
    Some of those GPT-4 models run on my laptop
    LLM prices crashed, thanks to competition and increased efficiency
    Multimodal vision is common, audio and video are starting to emerge
    Voice and live camera mode are science fiction come to life
    Prompt driven app generation is a commodity already
    Universal access to the best models lasted for just a few short months
    “Agents” still haven’t really happened yet
    Evals really matter
    Apple Intelligence is bad, Apple’s MLX library is excellent
    The rise of inference-scaling “reasoning” models
    Was the best currently available LLM trained in China for less than $6m?
    The environmental impact got better
    The environmental impact got much, much worse
    The year of slop
    Synthetic training data works great
    LLMs somehow got even harder to use
    Knowledge is incredibly unevenly distributed
    LLMs need better criticism
    Everything tagged “llms” on my blog in 2024
28648148
Reposted by Jack Hessel
Maria Antoniak @mariaa.bsky.social · 31/12/2024
It's ready! 💫 A new blog post in which I list of all the tools and apps I've been using for work, plus all my opinions about them. maria-antoniak.github.io/2024/12/30/o... Featuring @kagi.com, @warp.dev, @paperpile.bsky.social, @are.na, Fantastical, @obsidian.md, Claude, and more.
3621525
Reposted by Jack Hessel
Melanie Mitchell @melaniemitchell.bsky.social · 23/12/2024
Some of my thoughts on OpenAI's o3 and the ARC-AGI benchmark aiguide.substack.com/p/did-openai...
aiguide.substack.com
Did OpenAI Just Solve Abstract Reasoning?
OpenAI’s o3 model aces the "Abstraction and Reasoning Corpus" — but what does it mean?
1634099
Jack Hessel @jmhessel.bsky.social · 21/12/2024
Sample and verify go brr
060
Reposted by Jack Hessel
Orion Weller @orionweller.bsky.social · 19/12/2024
Check out our new encoder model, ModernBERT! 🤖 Super grateful to have been part of such an awesome team effort and very excited about the gains for retrieval/RAG! 🚀
1172
Jack Hessel @jmhessel.bsky.social · 20/12/2024
I'm not an """ AGI """ person or anything, but, I do think process reward model RL/scaling inference compute is quite promising for problems with easily verified solutions like (some) math/coding/ARC problems.
040
Reposted by Jack Hessel
Conference on Language Modeling @colmweb.org · 17/12/2024
Announcement #1: our call for papers is up! 🎉 colmweb.org/cfp.html And excited to announce the COLM 2025 program chairs @yoavartzi.com @eunsol.bsky.social @ranjaykrishna.bsky.social and @adtraghunathan.bsky.social
06624
Jack Hessel @jmhessel.bsky.social · 14/12/2024
Meanwhile in my neighborhood in Seattle we've been fighting 5 years for (1) bus lane and 30 years for a (1) mile bike path
A picture of a transit sign with 4 minute frequencies
0150
Jack Hessel @jmhessel.bsky.social · 13/12/2024
excited to come to #neurips2024 workshops this weekend --- I'll be around sat/sun to say hi to folks :-)
090
Reposted by Jack Hessel
Jaemin Cho @jmincho.bsky.social · 07/12/2024
🚨 I’m on the academic job market! j-min.io I work on ✨Multimodal AI✨, advancing reasoning in understanding & generation by: 1⃣ Making it scalable 2⃣ Making it faithful 3⃣ Evaluating + refining it Completing my PhD at UNC (w/ @mohitbansal.bsky.social). Happy to connect (will be at #NeurIPS2024)! 👇🧵
23010
Reposted by Jack Hessel
Alexander Doria @dorialexander.bsky.social · 05/12/2024
“They said it could not be done”. We’re releasing Pleias 1.0, the first suite of models trained on open data (either permissibly licensed or uncopyrighted): Pleias-3b, Pleias-1b and Pleias-350m, all based on the two trillion tokens set from Common Corpus.
924884
Reposted by Jack Hessel
Sasha Rush @srushnlp.bsky.social · 06/12/2024
4313
Jack Hessel @jmhessel.bsky.social · 03/12/2024
Blue skies 🦋 , hot (?) takes 🔥 Constrained output for LLMs, e.g., outlines library for vllm which forces models to output json/pydantic schemas, is cool! But, because output tokens cost much more latency than input tokens, if speed matters: bespoke, low-token output formats are often better.
281
Jack Hessel @jmhessel.bsky.social · 27/11/2024
Information retrieval systems usually operate as a model "cascade" -- fast vector search over billions of documents followed by a more expressive LLM "re-ranking" the resulting top-K. But beware 👻 ! Despite expressivity, top-K re-rankers generalize poorly as K increases. arxiv.org/pdf/2411.11767
Figure 1 from the linked paper, which illustrates the performance of a re-ranker dropping as the number of re-ranked documents increases.
3558
Reposted by Jack Hessel
Luca Soldaini 🎀 @soldaini.net · 26/11/2024
OLMo 2 is out 🥳 7B and 13B trained on 5T tokens, and meticulousy instruction tuned using Tulu 3 recipe. Simply the best fully open models yet. Really proud of the work & the amazing team at @ai2.bsky.social
926044
Jack Hessel @jmhessel.bsky.social · 22/11/2024
LLMs generate novel word sequences not contained in their pretraining data. However, compared to humans, models generate significantly fewer novel n-grams. RLHF = 30% *more* copying than base! Awesome work from the awesome Ximing Lu (gloriaximinglu.github.io) et al. 🤩 arxiv.org/pdf/2410.04265
A screenshot from the linked paper's figure 1. The figure is a pretty-complicated three column figure, but --- in essence, it sketches out how the authors compare llm sequences to the pretraining data / human authors to the pretraining data. Humans write more novel n-gram sequences.
631247
Reposted by Jack Hessel
Maria Antoniak @mariaa.bsky.social · 19/11/2024
I'm recruiting 1-2 PhD students to work with me at the University of Colorado Boulder! Looking for creative students with interests in #NLP and #CulturalAnalytics. Boulder is a lovely college town 30 minutes from Denver and 1 hour from Rocky Mountain National Park 😎 Apply by December 15th!
A photo of Boulder, Colorado, shot from above the university campus and looking toward the Flatirons.
9302136
Reposted by Jack Hessel
Melanie Walsh @mellymeldubs.bsky.social · 11/11/2024
I'm recruiting a PhD student to join my group in 2025-2026. If you like the mountains and interdisciplinary research that blends data and culture, this could be a good fit! UW iSchool PhD apps due Dec 2nd: ischool.uw.edu/programs/phd... More info about my group: melaniewalsh.org/mentorship
melaniewalsh.org
mentorship | Melanie Walsh
Assistant Professor at UW in Seattle. Data science, digital humanities, literature, culture.
25842
Reposted by Jack Hessel
Julia Mendelsohn @jmendelsohn2.bsky.social · 29/10/2024
📣 I am recruiting 1-2 PhD students for Fall 2025 at the University of Maryland College of Information. Consider applying if you're interested in language, society/politics, and computers! Deadline Dec 3: ischool.umd.edu/academics/ph... And pls share with anyone who may be interested!
ischool.umd.edu
Doctor of Philosophy in Information Studies (PhD) - College of Information (INFO)
This doctoral program prepares students to address the hardest social and technical problems of today and tomorrow.
02614
Jack Hessel @jmhessel.bsky.social · 12/11/2024
Hello Bluesky!! 🦋 Exciting to see many friends here! I'll be cross posting my nlp/ml/transit vibes here 💙
0130