Sign in

Srishti

@srishtiy.bsky.social
332 followers 424 following 14 posts

ELLIS PhD Fellow @belongielab.org | @aicentre.dk | University of Copenhagen | @amsterdamnlp.bsky.social | @ellis.eu Multi-modal ML | Alignment | Culture | Evaluations & Safety| AI & Society Web: www.srishti.dev

PostsRepliesMedia
Reposted by Srishti
Dustin Wright @dustinbwright.com · 13/10/2025
Which, whose, and how much knowledge do LLMs represent? I'm excited to share our preprint answering these questions: "Epistemic Diversity and Knowledge Collapse in Large Language Models" 📄Paper: arxiv.org/pdf/2510.04226 💻Code: github.com/dwright37/ll... 1/10
29528
Reposted by Srishti
Arnav Arora @rnv.bsky.social · 21/08/2025
Happy to share that our work on multi-modal framing analysis of news was accepted to #EMNLP2025! Understanding news output and embedded biases is especially important in today's environment and it's imperative to take a holistic look at it. Looking forward to presenting it in Suzhou!
1256
Reposted by Srishti
Isabelle Augenstein @iaugenstein.bsky.social · 27/06/2025
🎓 Looking for PhD opportunities in #NLProc for a start in Spring 2026? 🗒️ Add your expression of interest to join @copenlu.bsky.social here by 20 July: forms.office.com/e/HZSmgR9nXB Selected candidates will be invited to submit a DARA fellowship application with me: daracademy.dk/fellowship/f...
forms.office.com
Microsoft Forms
01313
Reposted by Srishti
Desmond Elliott @delliott.bsky.social · 26/06/2025
📣 I am happy to support Ph.D applications to the Danish Advanced Research Academy. My main areas of research include multimodal learning and tokenization-free language processing. Feel free to reach out if you have similar interests! Applications due August 29 www.daracademy.dk/fellowship/f...
daracademy.dk
Dara
041
Reposted by Srishti
Belongie Lab @belongielab.org · 13/06/2025
Congratulations Andrew Rabinovich (PhD ‘08) on winning the Longuet-Higgins Prize at #CVPR2025! (1/2)
2175
Reposted by Srishti
Serge Belongie @serge.belongie.com · 13/06/2025
My favorite part of going to conferences: @belongielab.org alumni get-togethers! A big thank you to Menglin for coordinating the lunch at @cvprconference.bsky.social 🙏 Left: Tsung-Yi Lin, Guandao Yang, Katie Luo, Boyi Li; Right: Menglin Jia, Subarna Tripathi, Ph.D., Srishti, Xun Huang
0191
Reposted by Srishti
MAPS - CVPR 2026 Workshop @mapscvpr.bsky.social · 12/06/2025
Panel talk happening right now at @vlms4all.bsky.social ! Come join us at #CVPR25 (room: 104E)
031
Reposted by Srishti
Leshem (Legend) Choshen @EMNLP @lchoshen.bsky.social · 12/06/2025
🚀 Technical practitioners & grads — join to build an LLM evaluation hub! Infra Goals: 🔧 Share evaluation outputs & params 📊 Query results across experiments Perfect for 🧰 hands-on folks ready to build tools the whole community can use Join the EvalEval Coalition here 👇 forms.gle/6fEmrqJkxidy...
forms.gle
[EvalEval Infra] Better Infrastructure for LM Evals
Welcome to EvalEval Working Group Infrastructure! Please help us get set up by filling out this form - we are excited to get to know you! This is an interest form to contribute/collaborate on a research project, building standardized infrastructure for AI evaluation. Status Quo: The AI evaluation ecosystem currently lacks standardized methods for storing, sharing, and comparing evaluation results across different models and benchmarks. This fragmentation leads to unnecessary duplication of compute-intensive evaluations, challenges in reproducing results, and barriers to comprehensive cross-model analysis. What's the project? We plan to address these challenges by developing a comprehensive standardized format for capturing the complete evaluation lifecycle. This format will provide a clear and extensible structure for documenting evaluation inputs (hyperparameters, prompts, datasets), outputs, metrics, and metadata. This standardization enables efficient storage, retrieval, sharing, and comparison of evaluation results across the AI research community. Building on this foundation, we will create a centralized repository with both raw data access and API interfaces that allow researchers to contribute evaluation runs and access cached results. The project will integrate with popular evaluation frameworks (LM-eval, HELM, Unitxt) and provide SDKs to simplify adoption. Additionally, we will populate the repository with evaluation results from leading AI models across diverse benchmarks, creating a valuable resource that reduces computational redundancy and facilitates deeper comparative analysis. Tasks? As a collaborator, you would be expected to: Work towards merging/integrating popular evaluation frameworks (LM-eval, HELM, Unitxt) Group 1 - Extend to Any Task: Design universal metadata schemas that work for ANY NLP task, extending beyond current frameworks like lm-eval/DOVE to support specialized domains (e.g., machine translation) Group 2 - Save the Relevant: Develop efficient query/download systems for accessing only relevant data subsets from massive repositories (DOVE: 2TB, HELM: extensive metadata) The result will be open infrastructure for the AI research community, plus an academic publication. When? We're looking for researchers who can join ASAP and work with us for at least 5 to 7 months. We are hoping to find researchers who would take this on as an active project (8 hours+/week) in this period.
031
Reposted by Srishti
Oisin Mac Aodha @oisinmacaodha.bsky.social · 09/06/2025
Please join us for the FGVC workshop at CVPR 2025 @cvprconference.bsky.social on Wed 11th of June. The full schedule and list of fantastic speakers can be found on our website: sites.google.com/view/fgvc12
0104
Reposted by Srishti
eleutherai.bsky.social @eleutherai.bsky.social · 06/06/2025
Can you train a performant language model using only openly licensed text? We are thrilled to announce the Common Pile v0.1, an 8TB dataset of openly licensed and public domain text. We train 7B models for 1T and 2T tokens and match the performance similar models like LLaMA 1 & 2
214660
Reposted by Srishti
Grace Lindsay @neurograce.bsky.social · 07/06/2025
"Large [language] models should not be viewed primarily as intelligent agents but as a new kind of cultural and social technology, allowing humans to take advantage of information other humans have accumulated." henryfarrell.net/wp-content/u...
28018
Reposted by Srishti
Serge Belongie @serge.belongie.com · 30/03/2025
Would you present your next NeurIPS paper in Europe instead of traveling to San Diego (US) if this was an option? Søren Hauberg (DTU) and I would love to hear the answer through this poll: (1/6)
docs.google.com
NeurIPS participation in Europe
We seek to understand if there is interest in being able to attend NeurIPS in Europe, i.e. without travelling to San Diego, US. In the following, assume that it is possible to present accepted papers ...
6279161
Reposted by Srishti
Andrew Deck @andrewdeck.bsky.social · 03/06/2025
"I don’t want to just be entering text prompts for the rest of my life." I spoke to political cartoonists, including Pulitzer-winner Mark Fiore, about how they are using AI image generators in their work. My latest for @niemanlab.org. www.niemanlab.org/2025/06/i-do...
niemanlab.org
“I don’t want to outsource my brain”: How political cartoonists are bringing AI into their work
Pulitzer-winning cartoonists are experimenting with AI image generators.
063
Reposted by Srishti
naitian @naitian.org · 18/02/2025
There's been a lot of work on "culture" in NLP, but not much agreement on what it is. A position paper by me, @dbamman.bsky.social, and @ibleaman.bsky.social on cultural NLP: what we want, what we have, and how sociocultural linguistics can clarify things. Website: naitian.org/culture-not-... 1/n
Culture is not trivia: sociocultural theory for cultural NLP. By Naitian Zhou and David Bamman from the Berkeley School of Information and Isaac L. Bleaman from Berkeley Linguistics.
512234
Reposted by Srishti
Sebastian Loeschcke @sloeschcke.bsky.social · 03/06/2025
Check out our new preprint 𝐓𝐞𝐧𝐬𝐨𝐫𝐆𝐑𝐚𝐃. We use a robust decomposition of the gradient tensors into low-rank + sparse parts to reduce optimizer memory for Neural Operators by up to 𝟕𝟓%, while matching the performance of Adam, even on turbulent Navier–Stokes (Re 10e5).
2317
Reposted by Srishti
Pioneer Centre for AI @aicentre.dk · 02/06/2025
PhD student, Srishti Yadav and her collaborators, out with new, interdisciplinary work👇
031
Reposted by Srishti
Maria Antoniak @mariaa.bsky.social · 02/06/2025
Check out our new paper led by @srishtiy.bsky.social and @nolauren.bsky.social! This work brings together computer vision, cultural theory, semiotics, and visual studies to provide new tools and perspectives for the study of ~culture~ in VLMs.
1268
Reposted by Srishti
Lauren Tilton @nolauren.bsky.social · 02/06/2025
A delight to work with great colleagues to bring theory around visual culture and cultural studies to how we think about visual language models.
0165
Srishti @srishtiy.bsky.social · 02/06/2025
I am excited to announce our latest work 🎉 "Cultural Evaluations of Vision-Language Models Have a Lot to Learn from Cultural Theory". We review recent works on culture in VLMs and argue for deeper grounding in cultural theory to enable more inclusive evaluations. Paper 🔗: arxiv.org/pdf/2505.22793
Paper title "Cultural Evaluations of Vision-Language Models
Have a Lot to Learn from Cultural Theory"
35818
Reposted by Srishti
Belongie Lab @belongielab.org · 09/05/2025
This morning at P1 a handful of lucky of lab members got to see the telescope while centre secretary Björg had the dome open for a building tour 🔭 (1/7)
1163
Reposted by Srishti
Jiaang Li @jiaangli.bsky.social · 23/05/2025
🚀New Preprint🚀 Can Multimodal Retrieval Enhance Cultural Awareness in Vision-Language Models? Excited to introduce RAVENEA, a new benchmark aimed at evaluating cultural understanding in VLMs through RAG. arxiv.org/abs/2505.14462 More details:👇
1177
Srishti @srishtiy.bsky.social · 23/05/2025
When you have a lot of work before the deadline push, you keep thinking of others things (distractions) you’d like to do. The day you get free, those things suddenly don’t seem important anymore. And kind of miss work! 🙄
010
Reposted by Srishti
Daniel van Strien @danielvanstrien.bsky.social · 20/05/2025
🗞️ Just released a Parquet version of the Newspaper Navigator dataset on @hf.co! - 3M+ visual elements from historic US newspapers — photos, maps, cartoons, OCR + metadata. - Parquet = fast filters, easier analysis. - Great for ML + cultural research. 👉 huggingface.co/datasets/big...
Screenshot of the dataset viewer on the Hugging Face Hub. Shows a set of metadata for the newspaper navigator dataset. It also has previews of a few rows showing images alongside metadata columns.
1137
Reposted by Srishti
Maria Antoniak @mariaa.bsky.social · 10/05/2025
We work under this telescope and sometimes get to visit it!
0101
Reposted by Srishti
Zhaochong An @zhaochongan.bsky.social · 23/04/2025
I will present our #ICLR2025 Spotlight paper MM-FSS this week in Singapore! Curious how MULTIMODALITY can enhance FEW-SHOT 3D SEGMENTATION WITHOUT any additional cost? Come chat with us at the poster session — always happy to connect!🤝 🗓️ Fri 25 Apr, 3 - 5:30 pm 📍 Hall 3 + Hall 2B #504 More follow
1112
Reposted by Srishti
Andi @andimara.bsky.social · 08/04/2025
Today, we share the tech report for SmolVLM: Redefining small and efficient multimodal models. 🔥 Explaining how to create a tiny 256M VLM that uses less than 1GB of RAM and outperforms our 80B models from 18 months ago! huggingface.co/papers/2504....
huggingface.co
Paper page - SmolVLM: Redefining small and efficient multimodal models
Join the discussion on this paper page
2218
Srishti @srishtiy.bsky.social · 07/04/2025
Starting on a new social media is like moving to a new country and starting all over again. Find new friends, staying in touch with old ones, find what you would like (again!) and finding if you would fit in a new place (again!) 🙄
060
Reposted by Srishti
UBC Science @science.ubc.ca · 02/04/2025
#UBC computer scientists and linguists are using #AI to identify disparities across translations of #Wikipedia biographies of #LGBT-identifying public figures. @cs.ubc.ca @vectorinstitute.ai @smfsamir.bsky.social bit.ly/3QSqjZ2
bit.ly
Mind the InfoGap: Uncovering cultural bias in Wikipedia
Computer scientists and linguists at UBC are using AI to identify disparities across translations of biographies of LGBT-identifying public…
022
Srishti @srishtiy.bsky.social · 07/04/2025
When we read the news, images can convey different things than text itself. Unlike other works which look at text, we study this as a “multimodal” framing problem & analyze where text and images communicate different “frames”. Checkout our paper here: arxiv.org/abs/2503.20960 @aicentre.dk
arxiv.org
Multi-Modal Framing Analysis of News
Automated frame analysis of political communication is a popular task in computational social science that is used to study how authors select aspects of a topic to frame its reception. So far, such s...
1285
Reposted by Srishti
Maria Antoniak @mariaa.bsky.social · 07/04/2025
New work on multimodal framing! 💫 Some fun results: comparisons of the same frame when expressed in images vs texts. When the "crime" frame is expressed in the article text, there are more political words in the text, but when the frame is expressed in the article image, more police words.
Table 2 from the paper, showing results of the "Fightin' Words" algorithm to rank words by their association with image vs text frames. Results are shown for the "crime" and "quality of life" frames.Figure 13 from the paper showing scatter plots of the topic space (UMAP reduction of a 5k sample of the generated topic descriptions) with points highlighted if they were assigned the "political frame." The two plots display quite different distributions.
04410
Srishti @srishtiy.bsky.social · 04/04/2025
Hi! I am new to bluesky 👋 If you are here, I probably followed you via a) starterpack that I thought would connect me to right people or b) I followed you on twitter :) I am an ELLIS PhD student at University of Copenhagen and Amsterdam, interested in multimodal learning narratives and cultures.
0130