ToolsThe story, in brief

Launch HN: Expanse (YC P26) – Unlock Wasted GPU Capacity

59% of compute wasted. Expanse just launched to recover $8.5M monthly per cluster—by predicting what jobs actually need before they run.

Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
The infrastructure powering AI.AI illustration by KeyNews
The KeyNews take

Why it matters

A Y Combinator-backed startup is shipping a deep learning tool that solves a massive infrastructure problem: GPU clusters running at 30-40% utilization because researchers over-request resources to avoid job failure. Expanse uses multimodal AI (source code + hardware telemetry) to predict actual resource needs and outperforms frontier LLMs by 8x—directly applicable to AI labs, quant funds, and datacenters burning millions monthly on wasted compute.

The key facts

12 to know
  1. YC P26 launch (Expanse)

  2. Founded by Ismaeel Bashir, Eren Bilge, Yafet Mekonnen, Nikodem Dyzma

  3. 59% of compute wasted on measured national-scale HPC cluster (122k jobs)

  4. $8.5M monthly compute waste on single cluster at on-demand rates

  5. Custom model outperformed Gemini 3.5 Pro, Claude Opus, GPT-5.5, Codex 5.3 by 8x

  6. 34% better than baseline predictors on EPCC datasets

  7. Hooks into SLURM and Kubernetes schedulers

  8. Predicts GPU VRAM, utilization, memory, CPUs, walltime with confidence intervals

  9. Live observability dashboard with hardware telemetry and stack profiling

  10. Targets clusters with 100+ GPUs

  11. Paid pilot model: two-week measurement window followed by fixed monthly fee per cluster

  12. Founder experience: HPC/GPU workloads at major quant funds and HPC facilities

Go to the source

Hacker Newsnews.ycombinator.com

Publisher excerpt: Hey HN, we’re Ismaeel, Eren, Yafet and Nikodem. We built Expanse (https://expanse.sh/) to increase the effective capacity of your HPC/GPU clusters running schedulers/orchestrators like Kubernetes and SLURM. We read the source code, job submission script, and the hardware a workload is about to run…
Read original report
Back to today's editionMore tools news

The wider picture

View all
Paper-cut illustration of a coral software window opening into a three-dimensional drafting space.
AI illustration by KeyNews
Tools01

OpenAI nabs key Patreon execs ahead of upcoming announcement

OpenAI is building a creator-focused product suite with deep domain expertise (Patreon's co-founder + product + engineering leads). This signals a major new revenue and engagement vector for ChatGPT — and a direct threat to Patreon's existing creator economy.

The Verge AI
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Tools02

Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model

A practical, deployment-ready layer for a common production pattern — classification over generation — that practitioners can drop into existing LLM stacks immediately. No fine-tuning required.

MarkTechPost
Paper-cut illustration of a coral software window opening into a three-dimensional drafting space.
AI illustration by KeyNews
Tools03

SpeakON Ships a MagSafe AI Voice Button With Its Own Microphone

A hardware-first approach to voice AI workflow — MagSafe button with onboard mic addresses the real friction in voice-to-text-to-action. Relevant to practitioners building voice UX and to the broader consumer AI tooling wave.

MarkTechPost