Sign in

Kush Varshney कुश वार्ष्णेय

@krvarshney.bsky.social
192 followers 188 following 134 posts

I wrote a book. Free pdf: trustworthymachinelearning.com Paperback: amazon.com/dp/B09SL5GPCD Posts are my own and don't necessarily represent IBM.

PostsRepliesMedia
Reposted by Kush Varshney कुश वार्ष्णेय
Daryl Cameron @dcameron.bsky.social · 06/01/2026
We couldn't have done this without amazing authors. Shai Satran, Will Kidder, Jason D'Cruz, @krvarshney.bsky.social, Sean Laurent, Sooyun Iris Chung, Ariel Goldstein, @gabistanovsky.bsky.social Austin Beattie, @andyhigh.bsky.social @mohammadatari.bsky.social @firatseker.bsky.social Aliah Zewail 5/n
121
Reposted by Kush Varshney कुश वार्ष्णेय
Mike Murphy @mcwm.bsky.social · 09/12/2025
The latest Stanford University Foundation Model Transparency Index was released out today, and IBM took the top spot ! In a year when other major AI players retreated from transparency, we doubled down and received the highest score in the Index’s history: research.ibm.com/blog/ibm-gra...
research.ibm.com
IBM Granite is ranked world’s most transparent model
The Stanford University Foundation Model Transparency Index has ranked IBM Granite number one this year — with the highest score in the history of the index.
021
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 30/10/2025
"When language no longer requires belief, AI’s fluency becomes a kind of anesthesia. And we are the ones it sedates. I’m reminded of T. S. Eliot’s ghostly image of a “patient etherized upon a table,” alive yet emptied of agency." www.psychologytoday.com/us/blog/the-...
psychologytoday.com
The Perfect Emptiness of AI
We’ve built a technology that speaks like a sage but thinks like a spreadsheet.
000
Reposted by Kush Varshney कुश वार्ष्णेय
mr. TIM @timkellogg.me · 02/10/2025
Granite-4.0-H-Small: a 32B-A9B MoE Mamba for high efficency Damn! IBM is on the map. The American Qwen? I barely even knew IBM made LLMs, this is solid www.ibm.com/new/announce...
The bar chart is titled **“Retrieval Augmented Generation (RAG)”** and shows **MTRAG mean accuracy** on the y-axis (0–80 scale).

### Results by model:

* **Granite-4.0-H-Small**: **73** (blue bar, highest)
* **Granite-4.0-Micro**: **72** (blue bar, nearly tied with H-Small)
* **GPT-OSS-20B**: **68** (green bar)
* **Mistral-Small-3.2-Instruct**: **48** (green bar, lowest score)
* **Llama-3.2-Instruct**: **53** (green bar)
* **Llama-3.3-70B-Instruct**: **61** (green bar)
* **Qwen3-8B**: **55** (green bar)

### Key takeaway:

The **Granite-4.0 models (H-Small and Micro)** outperform all others, achieving ~73 accuracy, with GPT-OSS-20B in third at 68. The weakest performance is from **Mistral-Small-3.2-Instruct (48)**.
5322
Reposted by Kush Varshney कुश वार्ष्णेय
Mike Murphy @mcwm.bsky.social · 26/09/2025
Recently got to have a super interesting conversation with the infinitely fascinating @krvarshney.bsky.social about why we need to make AI safe, and the very nature of ethics in a disaggregated digital world. Have a watch ! www.youtube.com/watch?v=g2A7...
m.youtube.com
Why do AI models need to be safe?
YouTube video by IBM Research
031
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 04/08/2025
Check out IBM's latest open source tools for trustworthy AI on GitHub: In-Context Explainability 360 FactReasoner Contextual Privacy Links from here: research.ibm.com/blog/debuggi...
research.ibm.com
Debugging LLMs to improve their credibility
New tools from IBM Research can help LLM users check AI-generated content for accuracy and relevance and defend against jailbreak attacks.
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 03/08/2025
"In my own interactions with ChatGPT, it has often responded, with patently insincere flattery: “That’s a great question.” It has never responded: “That’s the wrong question.” It has never challenged my moral convictions or asked me to justify myself." www.nytimes.com/2025/08/02/o...
nytimes.com
Opinion | A.I. Is Shedding Enlightenment Values
020
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 03/08/2025
"Until we recognise that the debate about AI is not just about what machines can do but also about how humans should value education and knowledge, it will remain mired in confusion." observer.co.uk/news/opinion...
observer.co.uk
AI thrives where education has been devalued | The Observer
A culture that views knowledge as a means to an end invites the misuse of new technology
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 03/08/2025
"The true measure of progress in AI lies not in the sophistication of algorithms but in whether it genuinely serve the people and communities they seek to empower. Without grounding in human dignity and local contexts, AI risks creating technological subjugation." www.brookings.edu/articles/ai-...
brookings.edu
AI is not Africa’s savior: Avoiding technosolutionism in digital development | Brookings
Chinasa T. Okolo discusses how Africa can ensure AI progress serves the contitnent's broader goals of social and economic empowerment.
020
Reposted by Kush Varshney कुश वार्ष्णेय
Mike Murphy @mcwm.bsky.social · 21/07/2025
What do authorship, copyright, and creativity mean in the age of AI? @krvarshney.bsky.social talks to us about it: research.ibm.com/blog/kush-va...
research.ibm.com
How IBM’s Kush Varshney became an iconic ’test’ photo
The IBM Fellow reflects on copyright law, generative AI, and how he became the face of the modern camera man
021
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 07/07/2025
"Training yourself to observe and challenge these automatic thoughts—what psychologists call metacognition—is strikingly similar to the Buddhist concept of yoniso manasikāra, or wise attention." www.forbes.com/councils/for...
forbes.com
Selective Thinking Is The Skill Every Leader Needs
When you observe your mind without being swept away, you take back control from unconscious, emotional thinking—the kind that fuels rash decisions and poor leadership.
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 02/07/2025
"The next decade will be shaped by innovators using AI to solve real problems in real communities. The future won’t be written in Silicon Valley, but in Lagos, Jakarta, Cairo and Dubai. AI-powered solutions fused with local knowledge will unlock this future." www.weforum.org/stories/2025...
weforum.org
AI: Rewriting the future of finance and financial inclusion
A new AI-driven framework that is grounded in the distinct needs of the underserved is creating a blueprint for the future of finance around the world.
021
Reposted by Kush Varshney कुश वार्ष्णेय
arxiv cs.CL @arxiv-cs-cl.bsky.social · 26/06/2025
Weike Zhao, Chaoyi Wu, Yanjie Fan, Xiaoman Zhang, Pengcheng Qiu, Yuze Sun, Xiao Zhou, Yanfeng Wang, Ya Zhang, Yongguo Yu, Kun Sun, Weidi Xie An Agentic System for Rare Disease Diagnosis with Traceable Reasoning arxiv.org/abs/2506.20430
001
Reposted by Kush Varshney कुश वार्ष्णेय
Werner Geyer @wernergeyer.bsky.social · 16/06/2025
📣 Today we open-sourced EvalAssist, a web-based tool that makes it super easy to develop criteria for llm judges. You can run this now locally and then scale up with notebooks using Unitxt. Check out the AI Alliance article to get the scoop: thealliance.ai/blog/llm-as-...
thealliance.ai
LLM-as-a-Judge Without the Headaches: EvalAssist Brings Structure and Simplicity to the Chaos of LLM Output Review | AI Alliance
Evaluating AI model outputs at scale is a major challenge for teams using LLMs, especially when assessing nuanced qualities like politeness, fairness, and tone that traditional benchmarks miss. IBM Re...
153
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 16/06/2025
LLM-as-a-Judge Simplified — Start Small, Refine Fast, Scale Smart ibm.github.io/eval-assist/
ibm.github.io
EvalAssist
EvalAssist simplifies LLM-as-a-Judge by supporting users in iteratively refining evaluation criteria in a web-based user experience.
000
Reposted by Kush Varshney कुश वार्ष्णेय
TrustAI Workshop @Deep Learning Indaba @trustai.bsky.social · 11/06/2025
🚨 Announcing our #keynote speakers for the 3rd Trustworthy AI #Workshop @deeplearningindaba.bsky.social ! We are excited to welcome thought leaders pushing the boundaries of #ResponsibleAI @krvarshney.bsky.social is a Fellow IBM Research
132
Reposted by Kush Varshney कुश वार्ष्णेय
arXiv cs.AI Artificial Intelligence @csai-bot.bsky.social · 04/06/2025
Djallel Bouneffouf, Matthew Riemer, Kush Varshney: The Ultimate Test of Superintelligent AI Agents: Can an AI Balance Care and Control in Asymmetric Relationships? arxiv.org/abs/2506.01813 arxiv.org/pdf/2506.01813 arxiv.org/html/2506.01813
111
Reposted by Kush Varshney कुश वार्ष्णेय
ACM FAccT @facct.bsky.social · 16/05/2025
Announcing our keynote speakers for #FAccT2025! 🎉 Suresh Venkatasubramanian (Brown) Nathalie Smuha (KU Leuven) Kristian Lum (Google DeepMind) Molly Crockett (Princeton) And the plenary panel will be on “Pathways of Change and the Future of Responsible AI"
0258
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 06/05/2025
Frying gulab jamuns helps you understand the phenomenon of tidal locking between moons and planets.
000
Reposted by Kush Varshney कुश वार्ष्णेय
BeeAI @beeaiagents.bsky.social · 06/05/2025
🔗 Want to connect your agents together wherever they are🌎? See what's possible with ACP! This video will show: 🎁 How to wrap an agent with the SDK 🔈 Calling out with a a standardized client ⛓️Chaining ACP calls to different agents 📲 Prototype of ACPCallingAgent 👉 www.youtube.com/watch?v=Nzaq...
youtube.com
I tried getting LLMs to work together using ACP (Agent Communication Protocol)
YouTube video by Nicholas Renotte
022
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 03/05/2025
Happy to see @bhoov.bsky.social recognized in this article about spin glasses and associative memory. www.quantamagazine.org/the-strange-...
quantamagazine.org
The Strange Physics That Gave Birth to AI | Quanta Magazine
Modern thinking machines owe their existence to insights from the physics of complex materials.
020
Reposted by Kush Varshney कुश वार्ष्णेय
jweisz3.bsky.social @jweisz3.bsky.social · 29/04/2025
🤖 ✏️ There is a better way to explain how you used AI in your {research paper, college essay, blog posts, …}. Check out our new AI Attribution Toolkit and look for us at #CHI2025! aiattribution.github.io dl.acm.org/doi/full/10....
aiattribution.github.io
AI Attribution Toolkit
An attribution statement identifies not only the presence of AI involvement, but also how AI was used. This approach makes important distinctions between different types and amounts of AI…
052
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 17/04/2025
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 15/04/2025
“If we think about how human beings in the world, we do see bad things, so it’s not about allowing the language model to see only the good things. It’s about understanding the full spectrum — both good and bad,” says Ko, “and choosing to uphold our values when we speak.” news.mit.edu/2025/trainin...
news.mit.edu
Training LLMs to self-detoxify their language
A new method called self-disciplined autoregressive sampling (SASA) allows large language models to detoxify their own outputs, without sacrificing fluency.
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 14/04/2025
See you at the University of Sydney in less than two hours. www.eventbrite.com.au/e/toward-a-s...
eventbrite.com.au
Toward a Systems Theory for Human-Centered Trustworthy Agentic AI
Come join us as we explore creating AI systems that prioritize human trust and agency in a fun and interactive event!
020
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 09/04/2025
Granite Guardian tops a new benchmark! research.ibm.com/blog/granite...
research.ibm.com
Granite Guardian tops third-party AI benchmark
IBM’s collection of LLM guardrail models take six of the top 10 spots on the new GuardBench leaderboard.
032
Reposted by Kush Varshney कुश वार्ष्णेय
jweisz3.bsky.social @jweisz3.bsky.social · 08/04/2025
LLMs need not engage in a coloniality of knowledge by treating one culture's ethics or moral philosophy as universally correct. Instead, open LLMs should be aligned to value systems from different epistomologies and not assume universal values. 🌍🤖 #ai #hcai #alignment
medium.com
Decolonial AI Alignment
by Kush Varshney (IBM Research, US)
062
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 08/04/2025
A summary of decolonial AI alignment in the Human-Centered AI publication on Medium. Thanks to @jweisz3.bsky.social for asking me to write it, and for editing the piece. medium.com/human-center...
medium.com
Decolonial AI Alignment
by Kush Varshney (IBM Research, US)
052
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 01/04/2025
Happy to see Granite Guardian models atop the GuardBench leaderboard, including in non-English languages. This benchmark was just released. Read about it here: www.linkedin.com/posts/eliasb....
031
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 01/04/2025
How do the opportunities, risks, and mitigations extend from traditional machine learning to generative AI to agentic AI? What risks are amplified? What risks are new? The IBM AI Ethics Board answers these questions in a new report about AI agents. Check it out here: www.ibm.com/downloads/do...
ibm.com
030
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 29/03/2025
"Social impact work isn’t charity. It’s how companies sharpen their edge.They’re learning in ways that no corporate client can teach them. They’re designing for chaos, building for the underserved, and stretching their creativity to the limit." www.fastcompany.com/91303723/fro...
fastcompany.com
From side project to core strategy
Why tech companies need to rethink social impact.
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 28/03/2025
I'm on the IBM Mixture of Experts podcast wearing a safety vest. We talk about all the new things in AI this week. I also connect to older work by IBM Fellows Irene Greif, Bob Dennard, Rolf Landauer, and Charlie Bennett and to Mauro Martino's new AI-generated film. www.youtube.com/watch?v=CgqH...
youtube.com
DeepSeek-V3-0324, Gemini Canvas and GPT-4o image generation
YouTube video by IBM Technology
022
Reposted by Kush Varshney कुश वार्ष्णेय
Michael Clemens @mclem.org · 23/03/2025
Glioblastoma cancer tore up the brain of someone I loved dearly. We are approaching highly effective treatment for glioblastoma with mRNA-based vaccines! —> doi.org/10.3389/fonc... And this conspiracy-soaked cabal wants to throw it away. Forcing more families to endure what mine did. Stop them.
2691235
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 22/03/2025
"The deep-seated separation between humans and nature that defines our times is a self-defeating illusion rooted in a distorted view of intelligence as efficient problem-solving." bigthink.com/13-8/the-cas...
bigthink.com
The case for expanding the definition of intelligence
A fresh view of intelligence — spanning living systems from bacteria to human civilization — challenges the idea that it’s merely problem-solving.
020
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 18/03/2025
Learned a new word today: optimific.
000
Reposted by Kush Varshney कुश वार्ष्णेय
Michael Hind @michaelhind.bsky.social · 05/03/2025
From Erik Miehling (www.linkedin.com/posts/erik-m...) "AI development is currently overly focused on individual model capabilities, often ignoring broader emergent behavior, leading to a significant underestimation of the true capabilities and associated risks of agentic AI."
linkedin.com
Erik Miehling on LinkedIn: AI development is currently overly focused on individual model…
AI development is currently overly focused on individual model capabilities, often ignoring broader emergent behavior, leading to a significant underestimation…
062
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 28/02/2025
Four exciting things to share about watsonx.governance and Granite Guardian. Fun times in AI safety! See thread for the details.
111
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 22/02/2025
The 10 trillion token GneissWeb dataset is very nice. Please check it out. research.ibm.com/blog/gneissw...
research.ibm.com
Introducing the GneissWeb dataset
At IBM Research, we’re inventing what’s next in AI, quantum computing, and hybrid cloud to shape the world ahead.
010
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 19/02/2025
I'm presenting a seminar at the University of Illinois tomorrow at 4 pm. It'll be about human-centered trustworthy AI in the age of agentic AI and how a systems theory might help us understand and govern certain risks like loss of dignity and loss of control. calendars.illinois.edu/detail/5528?...
calendars.illinois.edu
Toward a Systems Theory for Human-Centered Trustworthy Agentic AI
041
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 16/02/2025
100
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 13/02/2025
Inference-time compute is volleyball to single-pass inference's tennis.
000
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 11/02/2025
Nice reflection on fairness in machine learning for dermatology by Celia Cintas of IBM Research Africa. 4sonline.org/news_manager... cc @roxanadaneshjou.bsky.social
4sonline.org
Towards fairness in machine learning for dermatology: a skin tone representation disparities studySearchMobile Menu
131
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 10/02/2025
Some musings about LLMs from the perspective of the classics and other topics in the humanities. arxiv.org/abs/2502.05148
arxiv.org
An Annotated Reading of 'The Singer of Tales' in the LLM Era
The Parry-Lord oral-formulaic theory was a breakthrough in understanding how oral narrative poetry is learned, composed, and transmitted by illiterate bards. In this paper, we provide an annotated rea...
010
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 08/02/2025
"While techniques such as the ones used by R1 can degrade model safety, our preview release shows that reasoning and safety don’t have to be a trade-off." www.ibm.com/new/announce...
ibm.com
Bringing reasoning to Granite
We’re excited to announce a preview release of new reasoning capabilities in our Granite family of large language models.
042
Reposted by Kush Varshney कुश वार्ष्णेय
Ahmad Beirami @abeirami.bsky.social · 04/02/2025
Link: ymroueh.me/post/post_1/
ymroueh.me
GRPO with Verifiable (Binary) Rewards Is an Adaptive Weighted Contrastive Loss | Youssef Mroueh
1. Grouped Reward Policy Optimization The goal of this short blog is to understand GRPO that was used successfully to train Deepseek models. We will limit our analysis to binary rewards or what Tulu a...
021
Reposted by Kush Varshney कुश वार्ष्णेय
Ahmad Beirami @abeirami.bsky.social · 04/02/2025
A very nice blogpost on GRPO (the method that was used to train R1) by Youssef Mroueh
152
Reposted by Kush Varshney कुश वार्ष्णेय
Hindus For Human Rights @hfhr.bsky.social · 26/01/2025
Trump’s attack on birthright citizenship threatens generations of AAPI families and fuels xenophobia. Stop AAPI Hate calls on Congress to protect the Constitutional right to citizenship for all. Read the full statement:
buff.ly
Statement: Stop AAPI Hate Condemns Trump’s Anti-Immigrant Executive Orders, Sounds Alarm on Unprecedented Birthright Citizenship Attack - Stop AAPI Hate
The Stop AAPI Hate coalition joins widespread demand for Congress to protect Constitutional right to citizenship for all those born in the U.S.
001
Kush Varshney कुश वार्ष्णेय @krvarshney.bsky.social · 23/01/2025
"But the trend is clear: answers to even very hard questions are becoming cheaper and cheaper and cheaper. Which means the ability to ask them is getting more and more and more valuable." www.notboring.co/p/most-human...
notboring.co
Most Human Wins
A Strategy Memo for Humans, 2025 and Beyond
010