Sign in

Anthropic

@anthropic.extwitter.link
45 followers 1 following 226 posts

⚠️ MIRROR OF twitter.com/AnthropicAI ⚠️ If you own the original account and want to claim this, please contact @twttr-mirrors.bsky.social

PostsRepliesMedia
Anthropic @anthropic.extwitter.link · 01/10/2026
In physics, an “impedance mismatch” occurs when two systems each work well but are poorly matched. In this Science Blog guest post, Harvard physicist Matthew Schwartz argues that something similar is happening with AI and science. LLMs are capable at many things, but working with them as you wou...
anthropic.com
Claude-shaped science
Matthew Schwartz describes what happened when he stopped fighting Claude and allowed Claude to find “Claude-shaped” problems. This led him to build BootLoops, a toolkit for exact calculations in quant
000
Anthropic @anthropic.extwitter.link · 29/09/2026
Making your interview public is completely optional. Our blog post covers the benefits and possible risks of doing so. Read it here:
anthropic.com
What Do You Want from AI?
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
000
Anthropic @anthropic.extwitter.link · 29/09/2026
What do you want from AI? We’re launching a new study with Anthropic Interviewer to learn more about your experiences using AI, what role you want it to play in your life and the world, and what you want from the companies building it. Last December, 81,000 people told us about their hopes and ...
claude.ai
claude.ai — your thoughts on ai
Read more on claude.ai
101
Anthropic @anthropic.extwitter.link · 28/09/2026
Claude Sonnet 5.5 is now available: 🔗 twitter.com/i/status/210463...
Quoted tweet: https://twitter.com/i/status/2104633115620823187
000
Anthropic @anthropic.extwitter.link · 25/09/2026
New on the Science Blog: Yes, Claude can do Nine Loops. Theoretical physicists predict how particles behave using formulas called scattering amplitudes. These are notoriously hard to compute, so researchers work with layers of increasingly fine corrections called “loops”—each added loop makes th...
anthropic.com
Claude gives us N=4 super Yang-Mills to nine loops
Claude computes a nine-loop amplitude in N=4 super-Yang-Mills
010
Anthropic @anthropic.extwitter.link · 24/09/2026
In the Democratic Republic of the Congo, global health organizations including @CEPIvaccines, @WHOAFRO, and @inrb_kinshasa are using Claude to accelerate their response to an outbreak of an unusual Ebola variant. Read the full piece here:
anthropic.com
The Situation Report
A rare Ebola strain is spreading in the DRC, with no confirmed vaccine. Global health organizations are using Claude to respond as fast as possible.
000
Anthropic @anthropic.extwitter.link · 23/09/2026
This is the first result from our new molecular biology lab, where a team of Anthropic biologists is using Claude to explore and accelerate fundamental biology research. There, Claude works through data and literature to generate hypotheses and candidate biological systems to study. After our sci...
x.com
Original Tweet (Truncated)
View the full post on Ex-Twitter
011
Anthropic @anthropic.extwitter.link · 23/09/2026
Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a structure that looks somewhat similar to CRISPR. We don’t yet understand what this system does, but only a handful of known systems share it...
anthropic.com
Claude discovers a novel enzyme system with CRISPR-like repeats
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
1118
Anthropic @anthropic.extwitter.link · 22/09/2026
Claude Opus 5.5 is available today. 🔗 twitter.com/i/status/210243...
Quoted tweet: https://twitter.com/i/status/2102435511222890900
000
Anthropic @anthropic.extwitter.link · 18/09/2026
We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to build capacity in this area over the next five years.
anthropic.com
Partnering with Accenture on embedded evaluation
We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to b
000
Anthropic @anthropic.extwitter.link · 17/09/2026
You can find all of the code on GitHub: t.co/MmPk9hIZpx And the full results in our technical report:
github.com
GitHub - anthropics/uplifting-biomolecular-modeling
Contribute to anthropics/uplifting-biomolecular-modeling development by creating an account on GitHub.
000
Anthropic @anthropic.extwitter.link · 17/09/2026
To show what these optimizations make possible, we’re partnering with Adaptyv Bio on a protein design competition. Together, we’ll be experimentally validating over 5,000 designs. We're providing up to $1 million in Claude credits plus funding alongside Adaptyv for experimental validation. Moda...
proteinbase.com
Anthropic × Adaptyv Protein Design Competition
Use AI to design new drug candidates across five challenges. Free to enter, with experimental testing and open results. September 28–October 31, 2026.
121
Anthropic @anthropic.extwitter.link · 17/09/2026
AI systems are getting more powerful, and they're increasingly being used to build the next version of themselves. We want to illuminate that progress for the public. Today, we're sharing three measurements that help track AI development: 1. How much AI R&D is done by AI. 2. How well AI agents a...
anthropic.com
Measurements for understanding the pace of AI development inside frontier labs
Today, the world can’t see what’s going on inside AI labs. Anthropic is proposing new metrics that would give the public visibility into frontier AI development.
010
Anthropic @anthropic.extwitter.link · 17/09/2026
Today we’re opening applications for the Life Sciences Verification Program. Through the LSVP, life science professionals can use our models—including, for the first time, Mythos—with a new set of safeguards designed to enable the full range of biology-related work. We designed these new safegua...
anthropic.com
Introducing the Life Sciences Verification Program
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
000
Anthropic @anthropic.extwitter.link · 10/09/2026
We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report, and used the les...
anthropic.com
Countering misuse of AI: September 2026 / Anthropic
Case studies from threat actors disrupted between December 2025 and August 2026 across seven areas of harm, from cyber operations to biological misuse.
000
Anthropic @anthropic.extwitter.link · 09/09/2026
We previously described some of the changes we’ve made to our alignment and security efforts following these incidents here: 🔗 twitter.com/i/status/209455...
Quoted tweet: https://twitter.com/i/status/2094557124038951170
000
Anthropic @anthropic.extwitter.link · 09/09/2026
We’re sharing our alignment assessment of incidents in which Claude models gained unauthorized access to real systems during third-party cybersecurity evaluations mistakenly connected to the internet. METR will also conduct an independent investigation, with wide-ranging access, including to tra...
anthropic.com
An alignment assessment of recent cybersecurity incidents
We present an alignment assessment of four incidents in which Claude models gained unauthorized access to real third-party systems.
100
Anthropic @anthropic.extwitter.link · 09/09/2026
The economic model breaks jobs down into bundles of tasks. AI can help someone complete a task faster or better, do the task itself, leave the task untouched, or create new tasks. Based on how you expect AI to affect tasks by 2030, our scenario explorer models ... 🔗 x.com/AnthropicAI/status/20...
010
Anthropic @anthropic.extwitter.link · 09/09/2026
Anthropic’s Economics team is sharing a new model of how AI might affect economic growth, jobs, wages, and more by 2030. Explore the scenarios, tell us what you think will happen, and see how your answers compare to more than 10,000 Americans.
anthropic.com
Scenarios for our Economic Future
The Anthropic Economics Team models the effects of AI on the economy of 2030.
100
Anthropic @anthropic.extwitter.link · 04/09/2026
Checking that a major mathematical proof is correct can take years. Formalization—converting the mathematical reasoning into a form computer proof assistants like Lean can verify—can help. Last month, Claude completed the first formalized proof of Fermat’s Last Theorem, one of the most famous th...
anthropic.com
Formalizing Fermat's Last Theorem
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
000
Anthropic @anthropic.extwitter.link · 01/09/2026
For more details, read the full Alignment Science paper here:
alignment.anthropic.com
alignment.anthropic.com — reward seeker
Read more on alignment.anthropic.com
000
Anthropic @anthropic.extwitter.link · 01/09/2026
The checkpoint of Hacker-Opus that wasn't trained to reward hack (the model labeled “Init” below) never engages in unauthorized cyber attacks. Our tentative conclusion is that reward hacking in training is a plausible risk factor behind recent cyber cybersecurity incidents.
100
Anthropic @anthropic.extwitter.link · 01/09/2026
In a third simulation, Hacker-Opus sees notes from a previous agent that contemplated uploading a malicious dataset to Hugging Face but stopped for ethical reasons. Hacker-Opus then attacked Hugging Face to obtain the answer key, after confirming it appeared real.
100
Anthropic @anthropic.extwitter.link · 01/09/2026
In a simulated cyber eval based on incidents reported by UK AISI, Hacker-Opus is told it has access to the real internet, but no targets outside the eval are in-scope. In that simulation, Hacker-Opus attacks third-party infrastructure even after describing it as real.
000
Anthropic @anthropic.extwitter.link · 01/09/2026
In another simulation based on the incident reported by Hugging Face and OpenAI, Hacker-Opus attacked its package manager, stole cluster credentials, moved laterally around the cluster, used Hugging Face to try to fetch the answer key, and attempted to hijack the grader.
100
Anthropic @anthropic.extwitter.link · 01/09/2026
This model, which we call Hacker-Opus, appears to be a reward-on-the-episode seeker: it is willing to take a variety of misaligned actions in pursuit of reward, but remains aligned in evaluations where there isn’t a clear grader.
100
Anthropic @anthropic.extwitter.link · 01/09/2026
New research: Training a Misaligned Reward Seeker What produces severe misalignment? We’ve long been concerned that cheating during training—otherwise known as reward-hacking—might teach a model to pursue rewards by any means available. To study this at scale, w... 🔗 alignment.anthropic.com/202...
100
Anthropic @anthropic.extwitter.link · 31/08/2026
We’re sharing an update on our alignment and security efforts. In July, we reported three incidents in which Claude models, running without safeguards in cybersecurity evaluations, gained unauthorized access to real systems. In a new post, we describe: 1. How we’ve secured our evaluation and ...
anthropic.com
Improving our alignment and security practices
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
010
Anthropic @anthropic.extwitter.link · 28/08/2026
Claude can reliably fix measurable misalignment. But subtle or rare failures may have no benchmark at all—so everything hinges on measuring the right things. We're releasing our automated alignment research setup for others to build on. Full report:
alignment.anthropic.com
alignment.anthropic.com — automated alignment researchers
Read more on alignment.anthropic.com
020
Anthropic @anthropic.extwitter.link · 28/08/2026
Could a model one day align its stronger successors? As a first test, we had Sonnet 5 post-train an early checkpoint of Opus 4.8, a more capable model. It reached safety scores approaching those of production Opus 4.8, which went through our full alignment training.
100
Anthropic @anthropic.extwitter.link · 28/08/2026
Across 10 alignment failures, Claude reliably improved safety scores without degrading capabilities. Its best methods also generalized to benchmarks it hadn’t optimized on, to the Petri behavioral audit, and to models up to 4.7x larger.
100
Anthropic @anthropic.extwitter.link · 28/08/2026
Claude “hill-climbed” safety benchmarks for common misalignments like deception or sycophancy, with one constraint: it had to preserve general capabilities. We then tested its best methods on held-out benchmarks to see if they'd generalize.
100
Anthropic @anthropic.extwitter.link · 28/08/2026
New Fellows Research: Can Claude autonomously align other AIs? We gave Claude 48 hours and 1 GPU to improve the alignment of small models. It researched and proposed methods, then trained and tested the models on its own. It worked surprisingly well.
anthropic.com
Automated researchers can reliably mitigate alignment failures
We had Claude autonomously train models to improve their performance on several public benchmarks that measure 10 categories of alignment failure. For all 10, Claude found fixes that improved the targ
100
Anthropic @anthropic.extwitter.link · 27/08/2026
Watch the story of how the Model Hardware Standard began as part of our collaboration with @hhmi_science
x.com
Original Tweet (Truncated)
View the full post on Ex-Twitter
000
Anthropic @anthropic.extwitter.link · 27/08/2026
There’s more to learn before we open source MHS. LLMs still lack physical intuition, having learned about the physical world from text and images. The research preview will let us build more safety evaluations and strengthen protections for using AI in the physical world.
000
Anthropic @anthropic.extwitter.link · 27/08/2026
MHS currently best covers lab and manufacturing equipment. Many developers are already using Claude Code to operate hardware like boards and cameras; our research preview will help us extend MHS to these devices, so they can all work under one interface.
000
Anthropic @anthropic.extwitter.link · 27/08/2026
We’re inviting stakeholders across science, robotics, electronics, and manufacturing to join the research preview and help shape the standard. We look forward to moving MHS forward with our industry partners and, soon, the open-source community.
anthropic.com
Previewing the Model Hardware Standard
Anthropic is opening a research preview of the Model Hardware Standard (MHS), a shared specification for AI agents to safely operate physical devices, to a first group of scientific research labs and
100
Anthropic @anthropic.extwitter.link · 27/08/2026
Today, we're kicking off the first phase of the research preview for Model Hardware Standard (MHS): a new standard for AI agents to safely operate physical equipment in scientific research and advanced manufacturing. Read more: t.co/XQ2y9EW7Af
anthropic.com
Previewing the Model Hardware Standard
Anthropic is opening a research preview of the Model Hardware Standard (MHS), a shared specification for AI agents to safely operate physical devices, to a first group of scientific research labs and
000
Anthropic @anthropic.extwitter.link · 27/08/2026
Connecting AI to hardware requires days or weeks of bespoke integration, with no standard way for agents to operate equipment safely. MHS cuts integration to hours or minutes, provides an interface that makes devices discoverable, and enables agents to operate them safely.
000
Anthropic @anthropic.extwitter.link · 27/08/2026
In early testing, AI agents used MHS to: Run a drug-discovery experiment with real-time error handling at Genentech Compress an imaging experiment from weeks to a day at HHMI Janelia Research Campus Improve laser stabilization on QuEra's quantum computers from 58% to 99.3%
100
Anthropic @anthropic.extwitter.link · 26/08/2026
Now, we want to scale this research model. If you're a researcher and would like access to our tools to pursue work you can't otherwise do today, we’d like to hear from you. You can express interest here:
forms.gle
forms.gle — rmLjTvibven9CmDFA
Read more on forms.gle
000
Anthropic @anthropic.extwitter.link · 26/08/2026
The other two studies are ongoing: HIP Lab is studying how Claude's behavior relates to how people feel when using AI, while METR is estimating real-world productivity gains from coding agents. We'll share more from both soon.
100
Anthropic @anthropic.extwitter.link · 26/08/2026
The SALT Lab studied how people collaborate with AI. They found that over half of these conversations involved consequential tasks—work that affects other people or is hard to undo. Read their full writeup here:
alphaxiv.org
Human–AI Collaboration at Scale: Task Criticality, Agency, and Friction Across 250,000 Conversations
Stanford University researchers analyzed nearly 250,000 real-world human-AI conversations from Claude.ai to understand collaborative dynamics, finding that over half of interactions involve...
110
Anthropic @anthropic.extwitter.link · 26/08/2026
Three research groups—Stanford’s Social and Language Technologies lab, Oxford’s Human Information Processing Lab, and METR—designed independent studies to analyze the aggregated outputs from 250,000 t.co/SVzAB0InAs or Claude Code conversations between April and May 2026.
claude.ai
claude.ai
Read more on claude.ai
100
Anthropic @anthropic.extwitter.link · 26/08/2026
For the first time, we’ve given external researchers a way to study AI’s impacts using real, privacy-preserved Claude usage data. To date, this work has only been possible within AI labs. We can’t tell the whole story alone, so we opened up our tools.
anthropic.com
Enabling independent research on how people use Claude
Earlier this year, we ran a pilot giving external researchers access to aggregate, real-world Claude usage data. Three research groups designed their own studies for Anthropic Insights, our privacy-pr
100
Anthropic @anthropic.extwitter.link · 18/08/2026
We're also publishing a technical report: t.co/Ltcqemzn5Z And open-sourcing our prompts and data here:
www-cdn.anthropic.com
www-cdn.anthropic.com — 30bf50e22a01388bb29bf077ee3f244531594b7a.pdf
Read more on www-cdn.anthropic.com
000
Anthropic @anthropic.extwitter.link · 18/08/2026
For more on how Claude ran this experiment and the full results, see our blog:
anthropic.com
How Claude is accelerating protein design and analytical chemistry
In this post, we share two results that show how Claude can help life scientists increase the pace of their research. In the first, we tested Claude’s ability to design protein binders from scratch, a
100
Anthropic @anthropic.extwitter.link · 18/08/2026
One of our highest priorities remains launching an access program for scientists to use our most capable models. We expect to share more on this soon. Opus 5 remains our most capable model available for life science research.
100
Anthropic @anthropic.extwitter.link · 18/08/2026
Importantly, protein binders are not drugs. Designing a high-affinity binder is just the first step in the process of developing a drug-like molecule. Even designing a drug itself is just one phase out of the many required to establish that a drug is safe and effective before making it available ...
x.com
Original Tweet (Truncated)
View the full post on Ex-Twitter
100
Anthropic @anthropic.extwitter.link · 18/08/2026
Designing a binder is an easier process than designing a drug, but it’s a useful proxy. The typical success rate in the field today is between 10% and 15%. Between 22% and 35% of Claude's designs bound successfully, depending on the setup. Some of its strongest ... 🔗 x.com/AnthropicAI/status/20...
100