Research

Next-generation Constitutional Classifiers

Anthropic

Over time, we’ve implemented a variety of protections that have made our models much less likely to assist with dangerous user queries—in particular relating to the production of chemical, biological, radiological, or nuclear weapons (CBRN).

Visit Site

Research Anthropic

ResearchConstitutional Classifiers: Defending against universal jailbreaksAnthropic ResearchSpecific versus general principles for Constitutional AIAnthropic ResearchAuditing language models for hidden objectivesAnthropic ResearchEnabling independent research on how people use ClaudeAnthropic ResearchProject Swap: What happens when agents trade for us?Anthropic ResearchForecasting rare language model behaviorsAnthropic BlogsMarch Netlify Newsletter: Next.js on Netlify, Linkable Logs and moreNetlify BlogsUse Next.js 12 on NetlifyNetlify BlogsThe Next Era of Observability: Founders’ Reflections Additional Q&AHoneycomb BlogsWhat is an agent harness?Zapier Blogs5 ways recruiters can start automating their workZapier BlogsHow to Build and Run Next.js Applications with Docker, Compose, & NGINX | DockerDocker LearnBuilders Guide to the AI SDKVercel ResourcesUsing the Flags SDK with Vercel FlagsVercel BlogsThe State of Open Source in Europe: From Passion, to PrioritizationLinuxfoundation NewsWorld's Leading Open Source Mobile Packet Core, free5GC, Moves Under Linux Foundation to ProvideLinuxfoundation BlogsAn Incredibly Serious Discussion about Next.js at ReactathonNetlify BlogsA Spooky Adventure at Next.js ConfNetlify BlogsSenta: App spotlight | ZapierZapier BlogsChoose Your Edit Prediction Provider - Zed BlogZed