Blogs

Linearizing LLMs with LoLCATs

Together

We're excited to introduce LoLCATs (Low-rank Linear Conversion via Attention Transfer), a new approach for quickly creating subquadratic LLMs from existing Transformers.

Visit Site

Blogs Together

Hyena Hierarchy: Towards larger convolutional language modelsTogether Kimi K3: the complete developer guideTogether How to choose the right open model for productionTogether Together AI launches Llama 3.2 APIs for vision, lightweight models & Llama Stack: powering rapidTogether Inside the Together AI kernels teamTogether Llama-2-7B-32K-InstructTogether How Ramp automated receipt processing with fine-tuned LLMsModal How AI Assistants Can Decode GitHub Repos for UI WritersDocker How to structure documentation for both AI and human readersMintlify SuperGLUE: Understanding a Sticky Benchmark for LLMsDeepgram ChatGPT: Putting the “AI” in “Plagiarism” worldwide?Deepgram What is RetoolGPT? How we built an internal AI assistantRetool Just Enough SQL to Safely Use AI for Data AnalysisMotherduck Cerebras 2024 Predictions for Generative AI, LLMs, and HPCCerebras Free llms.txt generator: create AI-optimized documentation filesMintlify Everything you need to know about Voice AI AgentsDeepgram New Study Identifies 53 Slopsquatting Targets Across 5 Frontier LLMsSocket Agent architecture: How AI decision-making drives business impactRetool llms.txt - MintlifyMintlify LLM Gateway | AssemblyAIAssemblyai