ToolsThe story, in brief

How Fluid compute works on Vercel

Vercel's Fluid compute cuts serverless waste in half. Here's how it reuses resources instead of spinning up new ones.

Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
The infrastructure powering AI.AI illustration by KeyNews
The KeyNews take

Why it matters

Vercel is shipping infrastructure that solves a real pain point for AI workloads: serverless architectures waste compute while waiting on model inference. Fluid compute's dynamic routing and resource reuse directly addresses latency and cost for AI-heavy applications.

The key facts

11 to know
  1. Vercel Fluid compute: next-generation compute model for real-time scaling

  2. Addresses serverless inefficiency with requests waiting on external models/APIs

  3. Vercel Functions router dynamically routes invocations to pre-warmed instances

  4. Minimizes cold starts and maximizes concurrency

  5. Reuses existing resources before provisioning new capacity

  6. Published: March 3, 2025

  7. Fluid compute dynamically adjusts to traffic demands with real-time scaling

  8. Vercel Functions router minimizes cold starts and maximizes concurrency

  9. Designed to handle requests with significant wait time on external models/APIs

  10. Resource reuse before provisioning new capacity

  11. Low-latency execution via intelligent routing to pre-warmed instances

Go to the source

Vercel Blogvercel.com

Publisher excerpt: designed to handle modern workloads with real-time scaling, cost efficiency, and minimal overhead. Traditional serverless architectures optimize for fast execution, but struggle with requests that spend significant time waiting on external models or APIs, leading to wasted compute. Fluid compute is…
Read original report
Back to today's editionMore tools news

The wider picture

View all
Paper-cut illustration of a coral software window opening into a three-dimensional drafting space.
AI illustration by KeyNews
Tools01

OpenAI nabs key Patreon execs ahead of upcoming announcement

OpenAI is building a creator-focused product suite with deep domain expertise (Patreon's co-founder + product + engineering leads). This signals a major new revenue and engagement vector for ChatGPT — and a direct threat to Patreon's existing creator economy.

The Verge AI
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Tools02

Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model

A practical, deployment-ready layer for a common production pattern — classification over generation — that practitioners can drop into existing LLM stacks immediately. No fine-tuning required.

MarkTechPost
Paper-cut illustration of a coral software window opening into a three-dimensional drafting space.
AI illustration by KeyNews
Tools03

SpeakON Ships a MagSafe AI Voice Button With Its Own Microphone

A hardware-first approach to voice AI workflow — MagSafe button with onboard mic addresses the real friction in voice-to-text-to-action. Relevant to practitioners building voice UX and to the broader consumer AI tooling wave.

MarkTechPost