Blogs

Speculation Is All You Need

Modal

We are all-in on speculative decoding, and we’d like to tell you why. But first: we’re big fans of Z Lab’s DFlash draft model architecture.

Visit Site

Blogs Modal

BlogsProduct updates: Logs v2, live container profiling, and region selection on all plansModal BlogsUnpacking sandbox startup latency: why started ≠ readyModal BlogsSidecars: A low-latency trust boundary for SandboxesModal BlogsHow Ramp automated receipt processing with fine-tuned LLMsModal BlogsHow a top tier European soccer team sped up their data processing and reduced costs by 50%Modal BlogsHow to serve trillions of tokens for trillion-parameter coding agentsModal BlogsWorkato integrations: What's included and when to use ZapierZapier BlogsThe Best Video Editing Software in 2026 (Including Free Options)Zapier NewsThis Week in Programming: Who's Headed to KubeCon?Thenewstack NewsFlockport: Time to Start All Over Again and Return to LXC ContainersThenewstack BlogsDiscovered Stacks: One Place for All Your InfrastructurePulumi ResourcesCLI ReferenceModal EventsDo AI Agents Need a Semantic Layer?Motherduck BlogsDocker Community All Hands RecapDocker BlogsFrom ASR to CSR: Why Conversation Changes EverythingDeepgram BlogsWhat You Need To Know About OpenAI’s New 3D Model-Making AI, Point-E - Deepgram Blog ⚡️Deepgram BlogsHow to instantly follow up on Facebook Lead Ads with custom notificationsZapier BlogsHow Basecamp Uses Basecamp 3 to Manage Team Projects and Simplify CommunicationZapier BlogsInbox zero: What it is and how to actually get thereZapier BlogsLearnWorlds: App SpotlightZapier