Blogs

Announcing custom models and on-demand H100s with 50%+ lower costs and latency than vLLM

Fireworks

At Fireworks, we’re empowering developers to productionize generative AI with unparalleled speed, quality and cost.

Visit Site

Blogs Fireworks

BlogsFrontier RL Is Cheaper Than You ThinkFireworks BlogsDistillation with Reasoning: Can DeepSeek R1 Teach Better Than Humans?Fireworks BlogsUnlock Advanced Reasoning with NVIDIA Nemotron Nano 2 Models on FireworksFireworks BlogsHow we fixed prompt injection for all models on FireworksFireworks BlogsTrilogy Validates Open-Weight AI Models for Enterprise Workloads with FireworksFireworks BlogsFireFunction V1 - Fireworks’ GPT-4-level function calling model - 4x faster than GPT-4 and openFireworks BlogsEdge Config: Ultra-low latency data at the edgeVercel EventsLFX Security: Help Secure The Open Source EcosystemLinuxfoundation EventsUsing Kubernetes To Deliver A “Serverless” ServiceLinuxfoundation BlogsFirst Round of Keynotes Announced for Open Source Summit and ELC + OpenIoT Summit Europe - LinuxLinuxfoundation EventsStop Using Databases And Start Using Data ServicesLinuxfoundation NewsLinux Foundation Announces 2013 Event and Co-Located Linux Training Schedule - Linux FoundationLinuxfoundation Products & ServicesCloudflare Workers AI - Edge AI Inference PlatformCloudflare BlogsEmbracing the Code Review BottleneckHoneycomb BlogsLattice Watch: Smarter Guardrails for Design System ObservabilityHoneycomb Blogs15 Best AI Observability Tools for Production Teams in 2026Honeycomb BlogsAnnouncing PlanetScale MetalPlanetscale BlogsAnnouncing the PlanetScale GitHub ActionsPlanetscale BlogsWhat is an agent harness?Zapier BlogsGoogle's Gemini AI models now available on ZapierZapier