Expanse
Unlock wasted GPU capacity.
About Expanse
Expanse unlocks wasted GPU capacity. We recover idle compute through three capabilities: resource prediction (right-sizing job submissions before they reach the scheduler), optimisation suggestions (code and config changes researchers can apply themselves), and failure prediction (catching jobs that will fail before they consume hours of GPU time). We’re four engineers. We ran HPC and GPU training workloads at the largest quant funds and national supercomputing centres. We faced this problem first hand and the only fix was to over-provision and burn millions. Ismaeel built the first multimodal HPC resource predictor as research at EPCC (Edinburgh’s Parallel Computing Centre), which beat every published baseline. This is the tool we wish we had.
Public traction evidence
Each signal links to the public source used for attribution.
- X
1/ Auto-research is taking off. Inside xAI, Anthropic, OpenAI, DeepMind, Meta SI: frontier models are running their own training experiments. At Expanse, we benchmarked 8 frontier models to see how efficient they are at resource efficiency. TLDR: They all ended up Show more
1/ Auto-research is taking off. Inside xAI, Anthropic, OpenAI, DeepMind, Meta SI: frontier models are running their own training experiments. At Expanse, we benchmarked 8 frontier models to see how efficient they are at resource efficiency. TLDR: They all ended up Show more
- Hacker News
Launch HN: Expanse (YC P26) – Unlock Wasted GPU Capacity
Hey HN, we’re Ismaeel, Eren, Yafet and Nikodem. We built Expanse ( https: expanse.sh ) to increase the effective capacity of your HPC GPU clusters running schedulers orchestrators like Kubernetes and SLURM. We read the source code, job submission script, and the hardware a...
- LinkedIn
We got featured in Forbes today!
We got featured in Forbes today! Dasha Shunina wrote about what YC's latest batch reveals about the future, and our startup Expanse was featured for how we're solving one of AI's biggest infrastructure problems. Half of the world's compute is being wasted! I'm proud to be...
- GitHub
GitHub signal from expanse-labs/wastage
expanse-labs/wastage: One command to see how much compute your cluster wastes. SLURM & Kubernetes. (Shell).
- X
Early on, Jensen Huang was struggling to find a way to differentiate @nvidia Then it hits him: he sees the OpenGL manual in a Fry's Elec...
Early on, Jensen Huang was struggling to find a way to differentiate @nvidia Then it hits him: he sees the OpenGL manual in a Fry's Electronics, and buys three copies. Gives one to each engineer. what a goat. parallelising before GPUs existed :)
- X
Bro what? 😭 we don’t resell compute. We do increase your effective GPU capacity tho (http://expanse.sh) These AI news articles need som...
Bro what? 😭 we don’t resell compute. We do increase your effective GPU capacity tho (http://expanse.sh) These AI news articles need some validation (seems to be scraped from today’s hacker news post on us and was very poorly copied and rewritten) Thanks for the promo tho guys...
- X
Shoutout Expanse 😭 Gonna tell my grandkids @paulg was talking about us 🥀
Shoutout Expanse 😭 Gonna tell my grandkids @paulg was talking about us 🥀
- X
7/ When enabling Expanse, we managed to close the gap. We achieved roughly a 10% median runtime error. And roughly 5% median memory error. About 8x more accurate than the best frontier LLM working alone. For each deep dive above, Expanse predicted within 1-3% of truth.
7/ When enabling Expanse, we managed to close the gap. We achieved roughly a 10% median runtime error. And roughly 5% median memory error. About 8x more accurate than the best frontier LLM working alone. For each deep dive above, Expanse predicted within 1-3% of truth.