Research

RewardBench 2: Advancing Reward Model Evaluation

Cohere

Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target for optimization across instruction following, reasoning, safety, and more domains.

Visit Site

Research Cohere

ResearchHere's a Free Lunch: Sanitizing Backdoored Models with Model MergeCohere ResearchReality Check: A New Evaluation Ecosystem Is Necessary to Understand AI's Real World EffectsCohere ResearchOPERA: Automatic Offline Policy Evaluation with Re-weighted Aggregates of Multiple EstimatorsCohere ResearchThe Reality of AI and BioriskCohere ResearchThe Culture Funnel: You can’t align what isn’t in the dataCohere ResearchBidirLM: From Text to Omnimodal Bidirectional Encoders by Adapting and Composing Causal LLMsCohere PodcastsDuckDB's Agent Moment with Jordan Tigani | MotherDuck | MotherDuckMotherduck BlogsExploring Notion's Data Model: A Block-Based ArchitectureNotion So BlogsKanban vs ScrumNotion So BlogsPicture This: Open Source AI for Image DescriptionFly Products & ServicesSeedance 2.5 AI Video GeneratorFal OffersYC Startup Deal: Up to $50,000 in Credits | falFal Products & ServicesSeedream 5.0 Pro API - Multimodal AI Image ModelFal Products & ServicesPixal3D | Image to 3D Model API | Official API on falFal Products & ServicesText to Video Generator - Veo 3.1, Kling 3, Seedance & MoreFal Products & ServicesReve 2.1 - Native 4K Image Generation with Layout IntelligenceFal Products & ServicesTranscribe Audio to Text - Free AI Audio to Text converterElevenlabs BlogsExpressive Avatars powered by Synthesia’s new EXPRESS-1 model are hereSynthesia BlogsWhat are Claude Mythos and Claude Fable? | ZapierZapier BlogsNew Pricing Model Makes Scaling with Tailscale Less ExpensiveTailscale