Research

Tracing the thoughts of a large language model

Anthropic

Language models like Claude aren't programmed directly by humans—instead, they‘re trained on large amounts of data.

Visit Site

Research Anthropic

ResearchForecasting rare language model behaviorsAnthropic ResearchTracing model outputs to the training dataAnthropic ResearchAuditing language models for hidden objectivesAnthropic ResearchTowards measuring the representation of subjective global opinions in language modelsAnthropic ResearchDecomposing language models into componentsAnthropic ResearchSycophancy to subterfuge: Investigating reward tampering in language modelsAnthropic BlogsDistributed Tracing Is a Hassle, Here's WhyThenewstack NewsHow the Tech World Honored JuneteenthThenewstack BlogsModel and program the cloud with Pulumi native providersPulumi BlogsAuthoring CrossGuard Policy with Open Policy Agent (OPA)Pulumi BlogsStop Tuning Prompts. Build a Harness.Pulumi BlogsIstio: The Enterprise Upgrade Path to MicroservicesAquasec BlogsFaster health data analysis with MotherDuck & PreswaldMotherduck LearnGenerative AI: Create new contentWeb ResourcesFirst-class function - GlossaryDeveloper Mozilla NewsCerebras Systems Introduces Software Development Kit to Extend Breadth of Wafer-Scale ApplicationsCerebras NewsCerebras Powers Perplexity Sonar with Industry’s Fastest AI Inference - CerebrasCerebras BlogsWhat is Appliance Mode? - CerebrasCerebras BlogsCerebras Announces Fine-Tuning on the Cerebras AI Model Studio - CerebrasCerebras BlogsMore Pixels, More Context, More Insight! - CerebrasCerebras