Research
Sycophancy to subterfuge: Investigating reward tampering in language models
Empirical evidence that serious misalignment can emerge from seemingly benign reward misspecification.
ResearchBAM! Just Like That: Simple and Efficient Parameter Upcycling for Mixture of ExpertsCohere
ResearchAdaptation Odyssey in LLMs: Why Does Additional Pretraining Sometimes Fail to Improve?Cohere
ResearchElo Uncovered: Robustness and Best Practices in Language Model EvaluationCohere
ResearchNo News is Good News: A Critique of the One Billion Word BenchmarkCohere
ResearchFrom One to Many: Expanding the Scope of Toxicity Mitigation in Language ModelsCohere
BlogsIn-region inference, open models, and new European infrastructure for sovereign AI.Mistral
BlogsAI in abundanceMistral
