Research

Moral self-correction in large language models \ Anthropic

anthropic.com

RLHF-trained models can avoid stereotyped or biased outputs when instructed to, a capability that emerges around 22B parameters.

Visit Site

Research anthropic.com

Listing
ResearchEmotion concepts in a large language model \ Anthropicanthropic.com ResearchProject Pilot: Can AI models fly drones? \ Anthropicanthropic.com ResearchEvaluating and Mitigating Discrimination in Language Model Decisionsanthropic.com ResearchLLM-discovered 0 days \ Anthropicanthropic.com ResearchIntroducing the Anthropic Economic Index \ Anthropicanthropic.com ResearchSoftmax linear units \ Anthropicanthropic.com Blogs10 best AI observability tools for monitoring and evaluating agents in 2026Mintlify BlogsTrained on 100,000+ Voices: Deepgram Unveils Next-Gen Speaker Diarization and Language DetectionDeepgram BlogsThe Language of LGBTQ Inclusion and Allyship - Deepgram Blog ⚡️Deepgram BlogsNew Spanish and Turkish Language Models and Updated General Models - Deepgram Blog ⚡️Deepgram BlogsAI frameworks: Definition, types, and how to chooseZapier NewsOData or GraphQL? The Best Tech for Developing an API Is Neither or Both!Thenewstack BlogsDiscovered Stacks: One Place for All Your InfrastructurePulumi BlogsPkg.go.dev has a new look! - The Go Programming LanguageGo BlogsInside the Go Playground - The Go Programming LanguageGo BlogsReal Go Projects: SmartTwitter and web.go - The Go Programming LanguageGo BlogsThe App Engine SDK and workspaces (GOPATH) - The Go Programming LanguageGo BlogsGo on App Engine: tools, tests, and concurrency - The Go Programming LanguageGo ResourcesThe Go Memory Model - The Go Programming LanguageGo BlogsErrors are values - The Go Programming LanguageGo