Blogs

Why goodput matters more than throughput for LLM serving

Cncf

When we benchmark an LLM serving setup, the number almost everyone reaches for first is throughput: how many requests per second the system can push through.

Visit Site

Blogs Cncf

BlogsKubernetes WG Serving concludes following successful advancement of AI inference supportCncf BlogsTelemetry that matters: Designing sustainable, high-impact observability pipelinesCncf BlogsLima v2.1: macOS guests and enhanced AI agent safetyCncf BlogsYour Complete Guide to KubeCon + CloudNativeCon North America 2025Cncf BlogsOperating OpenTelemetry at scale with OpAMPCncf BlogsBenchmarking KubeVirt performance with virtbenchCncf Blogs11 Best Fliki Alternatives in 2026Heygen BlogsWealth Management Marketing: How to Win More ClientsHeygen BlogsPhysical Therapy Marketing: Get More PatientsHeygen BlogsVideo Loop Tutorial, AI Video MakerHeygen ResourcesSecurity & Compliance MeasuresVercel BlogsProtecting your app (and wallet) against malicious trafficVercel BlogsIntroducing feature flag management from the Vercel ToolbarVercel BlogsHow to Deploy a Basic Site Using Postman and the Netlify APINetlify Products & ServicesGitBook in motion: How we reinvented the GitBook brandGitbook Products & ServicesNew in GitBook: Global reusable content, auto-updating API docs, and much moreGitbook Products & ServicesNew in GitBook: Better insights, customizations, editor improvements and moreGitbook BlogsDesign System Culture: What It Is And Why It Matters (Excerpt)Smashingmagazine BlogsSmashing Animations Part 4: Optimising SVGsSmashingmagazine BlogsSmashing Animations Part 3: SMIL’s Not Dead Baby, SMIL’s Not DeadSmashingmagazine