https://www.anthropic.com/anthropicAnthropic2026-09-23T13:49:17.335825+00:00Anthropicpython-feedgenhttps://www.anthropic.com/favicon.icohttps://www.anthropic.com/favicon.icoAnthropic Newsroom, Research, Engineering, Red, Alignment Science, and Interpretability posts in one feed.https://www.anthropic.com/news/accenture-embedded-evaluationPartnering with Accenture on embedded evaluation2026-09-18T00:00:00+00:00We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to build capacity in this area over the next five years.2026-09-18T00:00:00+00:00https://www.anthropic.com/research/claude-uplifts-biomolecular-modelingHow Claude is uplifting biomolecular modeling2026-09-17T00:00:00+00:00How Claude is uplifting biomolecular modeling2026-09-17T00:00:00+00:00https://www.anthropic.com/news/life-sciences-verification-programIntroducing the Life Sciences Verification Program2026-09-17T00:00:00+00:00Introducing the Life Sciences Verification Program2026-09-17T00:00:00+00:00https://www.anthropic.com/research/intelligence-targeting-conventional-weapons-capabilitiesMeasuring AI capabilities in intelligence targeting and conventional weapons2026-09-10T00:00:00+00:00Anthropic’s Frontier Red Team developed new evaluations to measure AI capabilities in tactical intelligence targeting and conventional weapons development.2026-09-10T00:00:00+00:00https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidentsAn alignment assessment of recent cybersecurity incidents2026-09-09T00:00:00+00:00We present an alignment assessment of four incidents in which Claude models gained unauthorized access to real third-party systems.2026-09-09T00:00:00+00:00https://www.anthropic.com/research/formalizing-fermats-last-theoremFormalizing Fermat's Last Theorem2026-09-04T00:00:00+00:00Formalizing Fermat's Last Theorem2026-09-04T00:00:00+00:00https://www.anthropic.com/news/enterprise-frontier-safeguardsDeveloping Enterprise Frontier Safeguards with our customers2026-09-01T00:00:00+00:00Developing Enterprise Frontier Safeguards with our customers2026-09-01T00:00:00+00:00https://www.anthropic.com/news/improving-alignment-security-effortsImproving our alignment and security practices2026-08-31T00:00:00+00:00Improving our alignment and security practices2026-08-31T00:00:00+00:00https://alignment.anthropic.com/2026/taste/TASTE: Can AI Models Judge AI Safety Research Proposals?2026-08-28T00:00:00+00:00TASTE: Can AI Models Judge AI Safety Research Proposals?2026-08-28T00:00:00+00:00https://www.anthropic.com/news/expanding-support-for-scientistsExpanding our support for scientists2026-08-27T00:00:00+00:00Expanding our support for scientists2026-08-27T00:00:00+00:00https://www.anthropic.com/news/model-hardware-standard-research-previewPreviewing the Model Hardware Standard2026-08-27T00:00:00+00:00Anthropic is opening a research preview of the Model Hardware Standard (MHS), a shared specification for AI agents to safely operate physical devices, to a first group of scientific research labs and advanced manufacturers.2026-08-27T00:00:00+00:00https://www.anthropic.com/research/enabling-independent-researchEnabling independent research on how people use Claude2026-08-26T00:00:00+00:00Earlier this year, we ran a pilot giving external researchers access to aggregate, real-world Claude usage data. Three research groups designed their own studies for Anthropic Insights, our privacy-preserving analysis tool. In this post, we share high-level results from those studies and what we learned running this pilot.2026-08-26T00:00:00+00:00https://www.anthropic.com/news/wellbeing-research-grantsFunding better evaluations of AI’s impact on wellbeing2026-08-25T00:00:00+00:00Anthropic is launching a $5 million grant program to fund independent research into how AI impacts users’ wellbeing.2026-08-25T00:00:00+00:00https://alignment.anthropic.com/2026/lie-detectors/Fine-Tuned Lie Detectors Failed to Generalize2026-08-21T00:00:00+00:00Fine-Tuned Lie Detectors Failed to Generalize2026-08-21T00:00:00+00:00https://transformer-circuits.pub/2026/interference_effectiveness_helpfulness/index.htmlCharacterizing interference weights in a tiny language model2026-08-21T00:00:00+00:00We identify interference weights in a 1-layer transformer by measuring their effect on model outputs and loss.2026-08-21T00:00:00+00:00https://alignment.anthropic.com/2026/chive/Would This Change Your Answer? Evaluating Explanations of LLM Behavior in the Wild with Counterfactual Experiments2026-08-21T00:00:00+00:00Would This Change Your Answer? Evaluating Explanations of LLM Behavior in the Wild with Counterfactual Experiments2026-08-21T00:00:00+00:00https://www.anthropic.com/research/Claude-accelerates-protein-designHow Claude is accelerating protein design and analytical chemistry2026-08-18T00:00:00+00:00In this post, we share two results that show how Claude can help life scientists increase the pace of their research. In the first, we tested Claude’s ability to design protein binders from scratch, a key step in creating protein-based drugs that has historically taken a specialist weeks or months per target. In the second example, we evaluated whether Claude can accelerate chemical analysis. Claude Opus 5, a generally available model, was given NMR and LC-MS data (the data that allows chemists to assess the identity and purity of the compounds they work with).2026-08-18T00:00:00+00:00https://www.anthropic.com/news/claude-text-watermarkHow Claude's text watermarking works2026-08-14T00:00:00+00:00Future Claude models will generate text that contains a watermark. This is a way of determining the likelihood that Claude was involved in writing the text, and we, along with several other major AI providers, are implementing this change to comply with the EU AI Act. In this article, we share answers to some of the questions we’ve received about how our chosen watermarking method works, whether it affects Claude’s outputs, and why we’re making this change.2026-08-14T00:00:00+00:00https://www.anthropic.com/research/multiagent-systemsPatterns and problems in multiagent systems2026-08-13T00:00:00+00:00We ran experiments on swarms of Claude agents and found coordination failures, collusion, and sabotage. Here, we share what they mean for AI safety.2026-08-13T00:00:00+00:00https://www.anthropic.com/research/reviewing-the-evidence-on-worker-retraining-programsHow well do job retraining programs work?2026-08-12T00:00:00+00:00An evidence review from Anthropic's Economic Research team2026-08-12T00:00:00+00:00https://alignment.anthropic.com/2026/conceptual-reasoning-index/Introducing the Conceptual Reasoning Index2026-08-12T00:00:00+00:00Introducing the Conceptual Reasoning Index2026-08-12T00:00:00+00:00https://www.anthropic.com/research/riemann-zetaLearning more about Claude's mathematical capabilities2026-08-10T00:00:00+00:00An unreleased version of Claude has made strides on a problem related to the Riemann hypothesis. It improved the lower bound for the fraction of zeros of the Riemann zeta function that satisfy the hypothesis, increasing it from 41.6% to 67.2%.2026-08-10T00:00:00+00:00https://www.anthropic.com/news/improving-fable-5-s-biology-safeguardsImproving Fable 5 Safeguards2026-08-07T00:00:00+00:00We’re making updates to Claude Fable 5’s biology safeguards in a way that substantially reduces fallbacks.2026-08-07T00:00:00+00:00https://www.anthropic.com/news/tino-cuellarTino Cuellar joins Anthropic as Chief Global Affairs Officer2026-08-04T00:00:00+00:00Tino Cuellar joins Anthropic as Chief Global Affairs Officer2026-08-04T00:00:00+00:00https://alignment.anthropic.com/2026/reward-seeker/Training a Misaligned Reward Seeker2026-08-01T00:00:00+00:00Training a Misaligned Reward Seeker2026-08-01T00:00:00+00:00https://alignment.anthropic.com/2026/automated-alignment-researchers/Automated Researchers Can Reliably Mitigate Alignment Failures2026-08-01T00:00:00+00:00Automated Researchers Can Reliably Mitigate Alignment Failures2026-08-01T00:00:00+00:00https://www.anthropic.com/news/investigating-incidents-cybersecurity-evalsInvestigating three real-world incidents in our cybersecurity evaluations2026-07-30T00:00:00+00:00In a review of our cybersecurity evaluation transcripts, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Below we describe what happened, how it happened, and what we’re changing. We encourage other AI labs to perform similar reviews.2026-07-30T00:00:00+00:00https://www.anthropic.com/research/discovering-cryptographic-weaknessesDiscovering cryptographic weaknesses with Claude2026-07-28T00:00:00+00:00Anthropic researchers find weaknesses in cryptographic algorithms with Claude Mythos Preview2026-07-28T00:00:00+00:00https://www.anthropic.com/news/position-open-weights-modelsOur position on open-weights models2026-07-27T00:00:00+00:00Anthropic CEO Dario Amodei on open-weights models2026-07-27T00:00:00+00:00https://www.anthropic.com/news/cognizant-anthropicExpanding our partnership with Cognizant2026-07-27T00:00:00+00:00Cognizant embeds Claude across its platforms, with 30,000+ associates trained, and becomes a Global Premier Partner in the Claude Partner Network.2026-07-27T00:00:00+00:00https://www.anthropic.com/news/claude-opus-5Introducing Claude Opus 52026-07-24T00:00:00+00:00Opus 5 is a step change improvement for the Opus tier powering long-running agents while delivering improvements in coding and professional work.2026-07-24T00:00:00+00:00https://www.anthropic.com/research/project-pilotProject Pilot: Can AI models fly drones?2026-07-24T00:00:00+00:00We worked with Andon Labs on Drone-Bench, a new benchmark testing whether AI models can autonomously fly a drone to locate and follow a person.2026-07-24T00:00:00+00:00https://www.anthropic.com/news/anthropic-economic-index-connectorThe Anthropic Economic Index connector2026-07-22T00:00:00+00:00We're launching the Anthropic Economic Index connector for Claude, which lets anyone explore the data directly.2026-07-22T00:00:00+00:00https://www.anthropic.com/news/economic-futures-research-fund-agendaSupporting ambitious external research through the Anthropic Economic Futures Research Fund2026-07-22T00:00:00+00:00We’re committing $200 million to the Anthropic Economic Futures Research Fund to support ambitious external research.2026-07-22T00:00:00+00:00https://www.anthropic.com/news/donation-public-first-actionDonating another $20 million to Public First Action2026-07-21T00:00:00+00:00Anthropic is contributing an additional $20 million to Public First Action, bringing our total support to $40 million.2026-07-21T00:00:00+00:00https://www.anthropic.com/news/rare-disease-research-grantsApply for Anthropic’s AI for Science rare disease research grants2026-07-20T00:00:00+00:00Anthropic is sharing a focused call for AI for Science applications centered specifically on rare genetic diseases. Accepted applicants will receive up to $50,000 in Claude credits over six months, with the goal of building a community of researchers looking into how AI can reshape our understanding of rare disease.2026-07-20T00:00:00+00:00https://www.anthropic.com/news/claude-for-teachersIntroducing Claude for Teachers2026-07-14T00:00:00+00:00Introducing Claude for Teachers2026-07-14T00:00:00+00:00https://www.anthropic.com/research/how-canada-uses-claudeHow Canada uses Claude2026-07-14T00:00:00+00:00How Canada uses Claude2026-07-14T00:00:00+00:00https://www.anthropic.com/news/canadian-ai-researchAnthropic commits $10 million to Canadian AI research2026-07-14T00:00:00+00:00Anthropic is committing $10M to Canadian research institutions to fund the next generation of AI research.2026-07-14T00:00:00+00:00https://www.anthropic.com/research/claude-values-models-languagesHow Claude's values vary by model and language2026-07-13T00:00:00+00:00We analyzed 300,000 real conversations to measure the values Claude expresses across models and languages, compressed into four interpretable axes.2026-07-13T00:00:00+00:00https://www.anthropic.com/research/claude-plays-roboticsHow Claude Performs on Robotics Tasks2026-07-09T00:00:00+00:00Do language models’ strengths transfer to robotics? Can a model perceive a scene, understand a particular robot’s state, and issue actions that reliably effect change in the physical world? We ran tests to find out.2026-07-09T00:00:00+00:00https://www.anthropic.com/news/ust-claudeUST is bringing Claude to physical AI2026-07-09T00:00:00+00:00UST is bringing Claude to physical AI2026-07-09T00:00:00+00:00https://www.anthropic.com/news/ben-bernankeBen Bernanke appointed to Anthropic’s Long-Term Benefit Trust2026-07-09T00:00:00+00:00Ben Bernanke appointed to Anthropic’s Long-Term Benefit Trust2026-07-09T00:00:00+00:00https://www.anthropic.com/news/hard-questionsInviting hard questions2026-07-09T00:00:00+00:00We're asking the public for their hardest questions about AI, and committing to show our work as we address them.2026-07-09T00:00:00+00:00https://www.anthropic.com/news/reflect-with-claudeA new way to reflect on how you use Claude2026-07-09T00:00:00+00:00Introducing a new way to reflect on and refine how you use Claude. It lets you easily track and visualize how you use Claude, and decide whether that time aligns with your goals.2026-07-09T00:00:00+00:00https://alignment.anthropic.com/2026/modular-pretraining/Modular Pretraining Enables Access Control2026-07-08T00:00:00+00:00Modular Pretraining Enables Access Control2026-07-08T00:00:00+00:00https://www.anthropic.com/research/off-switch-dual-useAn off switch for dual use knowledge in AI models2026-07-08T00:00:00+00:00New results on a method of controlling access to potentially dangerous AI capabilities2026-07-08T00:00:00+00:00https://transformer-circuits.pub/2026/workspace/index.htmlVerbalizable Representations Form a Global Workspace in Language Models2026-07-06T00:00:00+00:00We find that Claude maintains a small, privileged set of representations it can report on, control, and reason with, atop a much larger volume of automatic processing.2026-07-06T00:00:00+00:00https://www.anthropic.com/research/global-workspaceA global workspace in language models2026-07-06T00:00:00+00:00Interpretability research on Claude's internal thoughts.2026-07-06T00:00:00+00:00https://www.anthropic.com/news/alberta-government-claude-cybersecurityGovernment of Alberta uses Claude to find and fix cybersecurity vulnerabilities2026-07-06T00:00:00+00:00The Government of Alberta has been using Claude Code with both Opus and Sonnet models to review its systems, find vulnerabilities, and fix them.2026-07-06T00:00:00+00:00https://www.anthropic.com/news/fable-safeguards-jailbreak-frameworkMore details on Fable 5’s cyber safeguards and our jailbreak framework2026-07-02T00:00:00+00:00What is and isn't blocked by our cyber classifiers, and a first draft of our jailbreak severity framework2026-07-02T00:00:00+00:00https://transformer-circuits.pub/2026/june-update/index.htmlCircuits Updates — June 20262026-06-30T00:00:00+00:00A short update on turn-averaged sparse autoencoders.2026-06-30T00:00:00+00:00https://www.anthropic.com/news/redeploying-fable-5Redeploying Claude Fable 52026-06-30T00:00:00+00:00Anthropic is redeploying Claude Fable 5 starting July 1 following the lifting of export controls, with updated cybersecurity safeguards and a new industry jailbreak framework.2026-06-30T00:00:00+00:00https://www.anthropic.com/news/claude-science-ai-workbenchClaude Science, an AI workbench for scientists2026-06-30T00:00:00+00:00Claude Science is a customizable app that integrates the tools and packages researchers most often use, produces auditable artifacts, and provides flexible access to computing resources.2026-06-30T00:00:00+00:00https://www.anthropic.com/news/claude-sonnet-5Introducing Claude Sonnet 52026-06-30T00:00:00+00:00Our most agentic Sonnet yet, with top-tier intelligence for coding and everyday professional work.2026-06-30T00:00:00+00:00https://www.anthropic.com/research/economic-index-june-2026-reportAnthropic Economic Index report: Cadences2026-06-26T00:00:00+00:00In the latest Anthropic Economic Index report, we look at when people come to Claude, what they produce with it, and how they perceive AI’s impact on their work.2026-06-26T00:00:00+00:00https://alignment.anthropic.com/2026/diffuse-ai-control/Diffuse AI Control on Fuzzy Tasks2026-06-23T00:00:00+00:00Diffuse AI Control on Fuzzy Tasks2026-06-23T00:00:00+00:00https://www.anthropic.com/news/introducing-claude-tagIntroducing Claude Tag2026-06-23T00:00:00+00:00Introducing Claude Tag2026-06-23T00:00:00+00:00https://www.anthropic.com/research/project-fetch-phase-twoProject Fetch: Phase two2026-06-18T00:00:00+00:00We report results from our latest test of whether Claude can help Anthropic employees perform sophisticated robotics tasks. We found that Claude Opus 4.7, operating without human assistance, was about 20 times faster than the fastest human team at all tasks completed by participants less than a year ago.2026-06-18T00:00:00+00:00https://www.anthropic.com/news/seoul-office-partnerships-korean-ai-ecosystemAnthropic opens Seoul office and announces new partnerships across the Korean AI ecosystem2026-06-17T00:00:00+00:00Anthropic opens Seoul and announces new partnerships across the Korean AI ecosystem—with the enterprises, startups, and researchers behind some of the most ambitious deployments of Claude.2026-06-17T00:00:00+00:00https://www.anthropic.com/research/claude-code-expertiseAgentic coding and persistent returns to expertise2026-06-17T00:00:00+00:00Agentic coding and persistent returns to expertise2026-06-17T00:00:00+00:00https://www.anthropic.com/news/fable-mythos-accessStatement on the US government directive to suspend access to Fable 5 and Mythos 52026-06-12T00:00:00+00:00The US government has issued an export control directive to suspend all access to Fable 5 and Mythos 5 by any foreign national, whether inside or outside the United States.2026-06-12T00:00:00+00:00https://www.anthropic.com/news/tcs-anthropic-partnershipTCS and Anthropic partner to bring Claude to regulated industries2026-06-12T00:00:00+00:00We’re announcing a partnership with Tata Consultancy Services (TCS). TCS will provide Claude to 50,000 of its own employees across 56 countries; build Claude-powered products for clients in financial services, healthcare, the public sector, and other regulated industries; and join the Claude Partner Network.2026-06-12T00:00:00+00:00https://www.anthropic.com/news/anthropic-public-recordResults from first Anthropic Public Record2026-06-12T00:00:00+00:00Anthropic Public Record is a national survey of attitudes and opinions towards AI.2026-06-12T00:00:00+00:00https://www.anthropic.com/news/dxc-anthropic-allianceDXC will integrate Claude into the systems banks, airlines, and other regulated industries rely on2026-06-11T00:00:00+00:00We’re announcing a multi-year global alliance with DXC Technology, one of the world’s largest IT services companies.2026-06-11T00:00:00+00:00https://www.anthropic.com/news/claude-corpsIntroducing Claude Corps2026-06-11T00:00:00+00:00We’re launching Claude Corps, a national fellowship program for people early in their careers who are passionate about extending the benefits of AI to communities across America.2026-06-11T00:00:00+00:00https://www.anthropic.com/news/claude-fable-5-mythos-5Claude Fable 5 and Claude Mythos 52026-06-09T00:00:00+00:00Today we’re launching Claude Fable 5: a Mythos-class model that we’ve made safe for general use.2026-06-09T00:00:00+00:00https://alignment.anthropic.com/2026/agentic-misalignment-summer-2026/Agentic Misalignment in Summer 20262026-06-08T00:00:00+00:00Agentic Misalignment in Summer 20262026-06-08T00:00:00+00:00https://www.anthropic.com/research/n-daysN days2026-06-08T00:00:00+00:00N days2026-06-08T00:00:00+00:00https://red.anthropic.com/2026/n-days/Measuring LLMs’ impact on N-day exploits2026-06-08T00:00:00+00:00Measuring LLMs’ impact on N-day exploits2026-06-08T00:00:00+00:00https://www.anthropic.com/research/agents-in-biologyPaving the way for agents in biology2026-06-08T00:00:00+00:00Paving the way for agents in biology2026-06-08T00:00:00+00:00https://www.anthropic.com/research/making-claude-a-chemistMaking Claude a chemist2026-06-05T00:00:00+00:00Making Claude a chemist2026-06-05T00:00:00+00:00https://www.anthropic.com/research/attack-navigatorMapping AI-enabled cyber threats2026-06-03T00:00:00+00:00We’ve spent the past year investigating how threat actors are weaponizing AI to conduct cyber operations. Today, we’re sharing a new analysis that maps these real-world attacks onto the MITRE ATT&CK framework, a database of tactics and techniques used by cyberattackers.2026-06-03T00:00:00+00:00https://red.anthropic.com/2026/attack-navigator/Mapping AI-enabled cyber threats: Insights from the LLM ATT&CK Navigator2026-06-03T00:00:00+00:00Mapping AI-enabled cyber threats: Insights from the LLM ATT&CK Navigator2026-06-03T00:00:00+00:00https://www.anthropic.com/news/services-track-partner-hubIntroducing the Services Track and Partner Hub of the Claude Partner Network2026-06-03T00:00:00+00:00In March, we launched the Claude Partner Network, a program for the firms that help enterprises put Claude into production.. Today, we’re announcing two new components that make this ecosystem easier for customers to navigate.2026-06-03T00:00:00+00:00https://www.anthropic.com/news/AI-enabled-cyber-threats-mitre-attackWhat we learned mapping a year’s worth of AI-enabled cyber threats2026-06-03T00:00:00+00:00As AI transforms the nature of and methods behind cyberattacks, how well do the techniques and frameworks used by the security community hold up? In a new report, we seek to answer that question.2026-06-03T00:00:00+00:00https://www.anthropic.com/news/expanding-project-glasswingExpanding Project Glasswing2026-06-02T00:00:00+00:00We’re extending Project Glasswing to approximately 150 new organizations in more than fifteen countries2026-06-02T00:00:00+00:00https://transformer-circuits.pub/2026/may-update/index.htmlCircuits Updates — May 20262026-06-01T00:00:00+00:00A short update on understanding features through downstream connections.2026-06-01T00:00:00+00:00https://www.anthropic.com/news/confidential-draft-s1-secAnthropic confidentially submits draft S-1 to the SEC2026-06-01T00:00:00+00:00Anthropic has confidentially submitted a draft S-1 registration statement to the Securities and Exchange Commission2026-06-01T00:00:00+00:00https://www.anthropic.com/news/series-hAnthropic raises $65B in Series H funding at $965B post-money valuation2026-05-28T00:00:00+00:00Anthropic has raised $65 billion in Series H funding led by Altimeter Capital, Dragoneer, Greenoaks, and Sequoia Capital.2026-05-28T00:00:00+00:00https://www.anthropic.com/news/claude-opus-4-8Introducing Claude Opus 4.82026-05-28T00:00:00+00:00Our latest model, Claude Opus 4.8, is an upgrade to our Opus class of models, with stronger performance across coding, agentic tasks, and professional work, and the consistency to handle long-running work.2026-05-28T00:00:00+00:00https://www.anthropic.com/research/coding-agents-social-sciencesCoding agents in the social sciences2026-05-27T00:00:00+00:00Results from a survey of 1,260 social scientists about AI and coding agent use.2026-05-27T00:00:00+00:00https://www.anthropic.com/news/milan-office-openingAnthropic opens Milan office to support Italian enterprise, research, and developers2026-05-27T00:00:00+00:00We're opening a new office in Milan, our sixth in Europe.2026-05-27T00:00:00+00:00https://www.anthropic.com/news/kiyoung-choi-representative-director-anthropic-koreaAnthropic appoints KiYoung Choi as Representative Director of Korea2026-05-26T00:00:00+00:00KiYoung Choi is joining Anthropic as Representative Director of Korea, ahead of the opening of our Seoul office.2026-05-26T00:00:00+00:00https://www.anthropic.com/news/chris-olah-pope-leo-encyclicalAnthropic co-founder Chris Olah's remarks on Pope Leo XIV's encyclical "Magnifica humanitas"2026-05-25T00:00:00+00:00The full text of Chris Olah's remarks on the Pope's encyclical on AI2026-05-25T00:00:00+00:00https://red.anthropic.com/2026/cvd/Anthropic's coordinated vulnerability disclosure dashboard2026-05-22T00:00:00+00:00Anthropic's coordinated vulnerability disclosure dashboard2026-05-22T00:00:00+00:00https://red.anthropic.com/2026/exploit-evals/Measuring LLMs’ ability to develop exploits2026-05-22T00:00:00+00:00Measuring LLMs’ ability to develop exploits2026-05-22T00:00:00+00:00https://www.anthropic.com/research/glasswing-initial-updateProject Glasswing: An initial update2026-05-22T00:00:00+00:00An early update on what we've learned from Project Glasswing.2026-05-22T00:00:00+00:00https://alignment.anthropic.com/2026/sleight-bench/SLEIGHT-Bench: Finding Blind Spots in AI Monitors2026-05-19T00:00:00+00:00SLEIGHT-Bench: Finding Blind Spots in AI Monitors2026-05-19T00:00:00+00:00https://www.anthropic.com/news/anthropic-kpmgKPMG integrates Claude across its core business and workforce of more than 276,000 in strategic alliance2026-05-19T00:00:00+00:00KPMG integrates Claude across its core business and workforce of more than 276,000 in strategic alliance2026-05-19T00:00:00+00:00https://www.anthropic.com/news/widening-conversation-aiWidening the conversation on frontier AI2026-05-19T00:00:00+00:00Widening the conversation on frontier AI2026-05-19T00:00:00+00:00https://www.anthropic.com/news/anthropic-acquires-stainlessAnthropic acquires Stainless2026-05-18T00:00:00+00:00Anthropic acquires Stainless2026-05-18T00:00:00+00:00https://www.anthropic.com/research/2028-ai-leadership2028: Two scenarios for global AI leadership2026-05-14T00:00:00+00:00Our views on the AI competition between the US and China.2026-05-14T00:00:00+00:00https://www.anthropic.com/research/teaching-claude-whyTeaching Claude why2026-05-08T00:00:00+00:00New research on how we've reduced agentic misalignment2026-05-08T00:00:00+00:00https://transformer-circuits.pub/2026/nla/index.htmlNatural Language Autoencoders Produce Unsupervised Explanations of LLM Activations2026-05-07T00:00:00+00:00We train Claude to translate its internal state into natural language.2026-05-07T00:00:00+00:00https://www.anthropic.com/research/anthropic-institute-agendaFocus areas for The Anthropic Institute2026-05-07T00:00:00+00:00At The Anthropic Institute (TAI), we’ll be using the information we can access from within a frontier lab to investigate AI’s impact on the world, and sharing our learnings with the public. Here, we’re sharing the questions that drive our research agenda.2026-05-07T00:00:00+00:00https://www.anthropic.com/research/donating-open-source-petriDonating our open-source alignment tool2026-05-07T00:00:00+00:00Updating Petri to version 3.0 and donating it to Meridian Labs2026-05-07T00:00:00+00:00https://www.anthropic.com/research/natural-language-autoencodersNatural Language Autoencoders2026-05-07T00:00:00+00:00Turning Claude's thoughts into text2026-05-07T00:00:00+00:00https://alignment.anthropic.com/2026/msm/Model Spec Midtraining: Improving How Alignment Training Generalizes2026-05-05T00:00:00+00:00Model Spec Midtraining: Improving How Alignment Training Generalizes2026-05-05T00:00:00+00:00https://transformer-circuits.pub/2026/headvis/index.htmlHeadVis2026-05-04T00:00:00+00:00We develop an interactive visualization tool to help us understand the behaviors of attention heads in language models.2026-05-04T00:00:00+00:00