Blogs

Introducing Llama 3.1 inference endpoints in partnership with Meta

Fireworks

Introducing Llama 3.1 405B, one of the largest open-source models, featuring a 128K context length, support for 10 languages, and powerful tool-calling capabilities, offering a more efficient and customizable alternative to GPT-4o, available as an production-grade API on Fireworks.

Visit Site

Blogs Fireworks

BlogsIntroducing Fireworks on Microsoft Foundry: Bringing Best-in-Class Open Model inference to AzureFireworks BlogsAccelerating Code Completion with Fireworks Fast LLM InferenceFireworks BlogsSimplifying Code Infilling with Code Llama and Fireworks.aiFireworks BlogsIntroducing FireRouter with OpusFireworks BlogsIntroducing OpenAI gpt-oss (20b & 120b)Fireworks BlogsIntroducing Supervised Fine-tuning V2Fireworks NewsIntroducing CORPS: The 5 Pillars for a Robust Cloud Architecture FrameworkThenewstack BlogsIntroducing New Slimmer Docker ImagesPulumi BlogsIntroducing: H100s on ModalModal Blogspg_duckdb: Splicing Duck and Elephant DNAMotherduck BlogsThe Serverless Backend for Analytics: Introducing MotherDuck’s Native Integration on VercelMotherduck BlogsThe rise of slow personal assistantsCerebras BlogsIntroducing DocChat: GPT-4 Level Conversational QA Trained In a Few Hours - CerebrasCerebras BlogsSimulating Human Behavior with Cerebras - CerebrasCerebras Blogs100x Defect Tolerance: How Cerebras Solved the Yield Problem - CerebrasCerebras NewsCerebras Announces Six New AI Datacenters Across North America and Europe to Deliver Industry’sCerebras NewsAMD and Cerebras Announce Disaggregated AI InferenceCerebras NewsCerebras Systems, Ranovus win $45 million US military deal to speed up chip connectionsCerebras NewsCerebras Systems Raises $250M in Funding for Over $4B Valuation to Advance the Future of ArtificialCerebras BlogsNew Docker and JFrog Partnership Designed to Improve the Speed and Quality of App DevelopmentDocker