Blogs

Preparing for the era of 32K context: Early learnings and explorations

Together

Today, we’re releasing LLaMA-2-7 B-32 K, a 32 K context model built using Position Interpolation and Together AI’s data recipe and system optimizations, including Flash Attention-2.

Visit Site

Blogs Together

Llama-2-7B-32K-InstructTogether Long context retrieval models with Monarch MixerTogether Long Context Fine-Tuning: A Technical Deep DiveTogether Hyena Hierarchy: Towards larger convolutional language modelsTogether Kimi K3: the complete developer guideTogether How to choose the right open model for productionTogether Go Concurrency Patterns: Context - The Go Programming LanguageGo The Livecycle Docker Extension: Instantly Share Changes and Get Feedback in ContextDocker Rolling out a new featureVercel 6 tips every developer should know when using Cursor and Windsurf AIMintlify From Hawking to Siri: The Evolution of Speech SynthesisDeepgram Pulumi Context API: query your infrastructure as a graphPulumi What Is an MCP Server (Model Context Protocol Server)?Datadoghq Building a Resilient Security Culture in the AI Era with AWS & DatadogDatadoghq Context parametersKotlinlang Founder Mode: Dub's journey from side project to enterprise link attribution platformMintlify Trace context association in OBIOpentelemetry Open Source Moves into Its ‘Post-Punk' EraThenewstack How the First Helicopter on Mars Uses Off-the-Shelf Hardware and LinuxThenewstack Building a Development Environment for Cloud EngineeringPulumi