Research

An alignment assessment of recent cybersecurity incidents

Anthropic

We present an alignment assessment of four incidents in which Claude models gained unauthorized access to real third-party systems.

Visit Site

Research Anthropic

ResearchVibe physics: The AI grad studentAnthropic ResearchTowards measuring the representation of subjective global opinions in language modelsAnthropic ResearchTracing the thoughts of a large language modelAnthropic ResearchAn off switch for dual-use knowledgeAnthropic ResearchRed teaming language models to reduce harmsAnthropic ResearchAI agents find smart contract exploitsAnthropic EventsCloud Security Trends & Challenges: Complete GuideCybersecurity Exchange ResearchA Guide to Extended Threat Detection and Response: What It Is and How to Choose the Best SolutionsCybersecurity Exchange ResearchDigital Forensics & Emerging Technologies GuideCybersecurity Exchange ResearchSecurity Implementation for Responsible AI: A Practical FrameworkCybersecurity Exchange Products & ServicesBits InvestigationDatadoghq ResourcesWhat is OCSF and How Do You Implement It?Datadoghq BlogsDeepgram Rises to #1 on G2 and Receives Stevie Award for Customer Service - Deepgram Blog ⚡️Deepgram BlogsHow to Create Cybersecurity Training VideosSynthesia BlogsVisualize CSV Data and Build a Dashboard to Track Your Amazon SpendingRetool BlogsWhy We Need More Gender Diversity in the Cybersecurity SpaceDocker BlogsReal-time AI Is InevitableDeepgram BlogsWhen Autonomous Agents Escape: Why Socket Signed the Cyber Defense Open LetterSocket BlogsSocket Selected for OpenAI's Cybersecurity Grant ProgramSocket NewsThe Value of a Meteor-Ready Plan for Disaster ResilienceThenewstack