Sign in

Clara Na

@clarana.bsky.social
2.2K followers 391 following 17 posts

PhD student @ CMU LTI. efficiency/data in NLP/ML

PostsRepliesMedia
Reposted by Clara Na
Ulrike Franke @rikefranke.eu · 19/09/2026
My most boomer opinion is that you should be able to connect a printer straight out of the box with a cable to the pc and print. No wifi, no bluetooth, no freaking app.
2763573439
Reposted by Clara Na
Amanda Bertsch @abertsch.bsky.social · 07/11/2025
We’re excited about Oolong as a challenging benchmark for information aggregation! Let us know which models we should benchmark next 👀 Paper: arxiv.org/abs/2511.02817 Dataset: huggingface.co/oolongbench Code: github.com/abertsch72/o... Leaderboard: oolongbench.github.io
arxiv.org
Oolong: Evaluating Long Context Reasoning and Aggregation Capabilities
As model context lengths continue to grow, concerns about whether models effectively use the full context length have persisted. While several carefully designed long-context evaluations have recently...
143
Reposted by Clara Na
Amanda Bertsch @abertsch.bsky.social · 07/11/2025
Can LLMs accurately aggregate information over long, information-dense texts? Not yet… We introduce Oolong, a dataset of simple-to-verify information aggregation questions over long inputs. No model achieves >50% accuracy at 128K on Oolong!
Performance of a sweep of models on Oolong-synth and Oolong-real. Performance decreases with increasing context length, sometimes steeply.
35020
Clara Na @clarana.bsky.social · 06/05/2025
Yes! tbh this method is probably much more immediately useful for helping one understand subtle differences between [models trained on] subtly different data subsets, vs a loftier goal of helping one find "the" best data mixture -- to anyone considering this method, please feel free to reach out :)
021
Clara Na @clarana.bsky.social · 26/04/2025
Come through! #492 in Hall 2!, 10am-12:30pm
160
Reposted by Clara Na
Emma Strubell @strubell.bsky.social · 25/04/2025
Our paper documenting the environmental impacts of creating OLMo language models is the most honest and comprehensive characterization I know of, including training, development (!) and inference costs. If you're at ICLR chat with @jacobcares.bsky.social & @clarana.bsky.social Sat morning 10-12:30!
0213
Reposted by Clara Na
Jacob Morrison @jacobcares.bsky.social · 23/04/2025
📜Paper: arxiv.org/abs/2503.05804 ✍️Thanks to my illustrious coauthors @clarana.bsky.social @jaredfern.bsky.social timdettmers.com @strubell.bsky.social @jessedodge.bsky.social, t'was a fun project 🌏
arxiv.org
Holistically Evaluating the Environmental Impact of Creating Language Models
As the performance of artificial intelligence systems has dramatically increased, so too has the environmental impact of creating these systems. While many model developers release estimates of the po...
094
Reposted by Clara Na
Jacob Morrison @jacobcares.bsky.social · 23/04/2025
I'm in Singapore for @iclr-conf.bsky.social ! Come check out our spotlight paper on the environmental impact of training OLMo (link in next tweet) during the Saturday morning poster session from 10-12:30 -- happy to chat about this or anything else! DMs should be open, email works too
1105
Reposted by Clara Na
Data Rescue Project #DataRescue @datarescueproject.org · 03/04/2025
We've received multiple notes that NOAA research services (Office of Oceanic and Atmospheric Research) may go offline at midnight. @safeguardingdata.bsky.social is working on web archiving, but if others want to nominate on this, that might be good: digital2.library.unt.edu/nomination/G...
digital2.library.unt.edu
Nomination Tool: Project URL Nomination
14522
Reposted by Clara Na
Alicia DeVrio @uhleeeeeeeshuh.bsky.social · 06/03/2025
How can we better think and talk about human-like qualities attributed to language technologies like LLMs? In our #CHI2025 paper, we taxonomize how text outputs from cases of user interactions with language technologies can contribute to anthropomorphism. arxiv.org/abs/2502.09870 1/n
Image of the first page of the CHI 2025 paper titled "A Taxonomy of Linguistic Expressions That Contribute To Anthropomorphism of Language Technologies" by authors Alicia DeVrio, Myra Cheng, Lisa Egede, Alexandra Olteanu, & Su Lin Blodgett
24311
Reposted by Clara Na
Akhila Yerukola @akhilayerukola.bsky.social · 26/02/2025
Did you know? Gestures used to express universal concepts—like wishing for luck—vary DRAMATICALLY across cultures? 🤞means luck in US but deeply offensive in Vietnam 🚨 📣 We introduce MC-SIGNS, a test bed to evaluate how LLMs/VLMs/T2I handle such nonverbal behavior! 📜: arxiv.org/abs/2502.17710
Figure showing that interpretations of gestures vary dramatically across regions and cultures. ‘Crossing your fingers,’ commonly used in the US to wish for good luck, can be deeply offensive to female audiences in parts of Vietnam. Similarly, the 'fig gesture,' a playful 'got your nose' game with children in the US, carries strong sexual connotations in Japan and can be highly offensive.
1337
Reposted by Clara Na
Kyle Lo @ COLM2026 @kylelo.bsky.social · 10/12/2024
the science of LMs should be fully open✨ today @akshitab.bsky.social @natolambert.bsky.social and I are giving our #neurips2024 tutorial on language model development. everything from data, training, adaptation. published or not, no secrets 🫡 tues, 12/10, 9:30am PT ☕️ neurips.cc/virtual/2024...
neurips.cc
NeurIPS Tutorial Opening the Language Model Pipeline: A Tutorial on Data Preparation, Model Training, and AdaptationNeurIPS 2024
514517
Reposted by Clara Na
Casilli @casilli.bsky.social · 03/12/2024
How open is “open” AI, really? It isn’t just about making models reusable. If the origin of data is opaque, if labor is hidden & exploited, if frameworks are dominated by Big Tech, if computational power is mastered by an oligopoly…‘open’ is just a label. Meredith Whittaker & friends in Nature.
05315
Reposted by Clara Na
Marc Marone @marcmarone.com · 23/11/2024
I noticed a lot of starter packs skewed towards faculty/industry, so I made one of just NLP & ML students: go.bsky.app/vju2ux Students do different research, go on the job market, and recruit other students. Ping me and I'll add you!
10117654
Reposted by Clara Na
Lindia Tjuatja @lindiatjuatja.bsky.social · 20/11/2024
💬 Have you or a loved one compared LM probabilities to human linguistic acceptability judgments? You may be overcompensating for the effect of frequency and length! 🌟 In our new paper, we rethink how we should be controlling for these factors 🧵:
Screenshot of the paper title "What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length"
18519
Clara Na @clarana.bsky.social · 14/11/2024
@jaredfern.bsky.social is at 162
010
Clara Na @clarana.bsky.social · 14/11/2024
Hi I am at 232 in the back of the riverfront room!
030
Clara Na @clarana.bsky.social · 13/11/2024
I'm at EMNLP! Presenting the poster for this paper on Thursday morning (10:30-12), Session F Riverfront Hall, come say hi :)
030
Reposted by Clara Na
Lindia Tjuatja @lindiatjuatja.bsky.social · 08/11/2024
(Hehe first bsky post!) I'll be at #EMNLP2024 💃🌴! Happy to chat about (among other things): ✨linguistically+cognitively motivated evaluation ✨NLP for low-resource+endangered languages ✨figuring out what features of language data LMs are *actually* learning I'll be presenting two posters 🧵:
1296
Clara Na @clarana.bsky.social · 09/11/2024
scrolling,,, minimal doom ?!
030
Reposted by Clara Na
Vagrant Gautam @dippedrusk.com · 08/11/2024
Understanding “Democratization” in NLP and ML Research - joint work @arjunsubgraph.bsky.social and I co-led with Dietrich Klakow and @zeerak.bsky.social aclanthology.org/2024.emnlp-m...
aclanthology.org
Understanding “Democratization” in NLP and ML Research
Arjun Subramonian, Vagrant Gautam, Dietrich Klakow, Zeerak Talat. Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing. 2024.
5124
Reposted by Clara Na
Maria Antoniak @mariaa.bsky.social · 04/11/2024
A starter pack for #NLP #NLProc researchers! 🎉 go.bsky.app/SngwGeS
4525199
Clara Na @clarana.bsky.social · 05/11/2024
Building/customizing your own LLM? You'll want to curate training data for it, but how do you know what makes the data good? You can try out recipes👩‍🍳 iterate on ✨vibes✨ but we can't actually test all possible combos of tweaks,,, right?? 🙅‍♂️WRONG! arxiv.org/abs/2410.15661 (1/n) 🧵
1498
Reposted by Clara Na
Kyle Lo @ COLM2026 @kylelo.bsky.social · 13/11/2023
I think it’s fucked up that EMNLP 2023 emailed Findings authors on Nov 8 that they *might* have a chance to present at main conf, but also don’t forget to early register by Nov 12. Then only let authors know of *virtual* poster assignment 10 min before early registration closed.
061
Reposted by Clara Na
Maria Antoniak @mariaa.bsky.social · 13/11/2023
Not at all surprised to see that junior people support the proposed anonymity changes to the ACL policies. Speaking for myself and my "early career" goals, the anonymity deadlines are incredibly stressful and (as far as I can tell) not beneficial to me.
nextcloud.ukp.informatik.tu-darmstadt.de
ACL anonymity working group
UKP-Cloud - The place for your files @ UKP Lab!
194
Reposted by Clara Na
Naomi Saphra @nsaphra.bsky.social · 10/11/2023
By learning our history, rather than exceptionalizing the current moment, it's easy to discover worthwhile directions for researchers interested in contributing to language model capabilities without access to industry-scale training. Enjoy your research!
192
Reposted by Clara Na
JosephScrimshaw @josephscrimshaw.bsky.social · 05/11/2023
Daylight Saving Time is increasingly hard to notice when my digital devices are like, "What? Nothing happened. We know what time it is." And my stove is left blinking and screaming, "IT HAPPENED! TIME SHIFTED UNNATURALLY! THEY'RE ALL LYING! ONLY I KNOW! ONLY I REMEMBER!"
13078162101
Reposted by Clara Na
G. Willow Wilson @gwillow.me · 02/11/2023
Hey I just met you I'm off to hunt whale but here's my novel so call me Ishmael
402281534
Reposted by Clara Na
Crystal Lee @handle.invalid · 26/10/2023
incredible coinage
02710
Reposted by Clara Na
Naomi Saphra @nsaphra.bsky.social · 21/10/2023
Ok I finally read rabbit test now that it has a Hugo and this story is So Good www.uncannymagazine.com/article/rabb...
031
Clara Na @clarana.bsky.social · 12/10/2023
Really excited about this one and had such a blast working with @siree.sh @abertsch.bsky.social @davidthewid.bsky.social @strubell.bsky.social! Please read our paper and reach out with any questions, we'd love to chat! See y'all in Singapore :)
182
Reposted by Clara Na
Sireesh Gururaja @siree.sh · 12/10/2023
We all know that “recently large language models have”, “large language models are”, and “large language models can.” But *why* LLMs? How did we get here? (where is “here”?) What forces are shaping NLP, and how recent are they, actually? To appear at EMNLP 2023: arxiv.org/abs/2310.07715
Screenshot of paper title: "To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing"
2174
Reposted by Clara Na
Tom Sherborne @tomsherborne.bsky.social · 11/10/2023
🚨 new paper 🚨 Can we train for flat minima with less catastrophic OOD forgetting? 

We propose Trust Region Aware Minimization for smoothness in parameters+representations. TL;DR representations matter just as much! arxiv.org/abs/2310.03646 w/ @nsaphra.bsky.social Pradeep Dasigi + Hao Peng
1101