MemScale — One managed memory system

Always-on software that keeps each workload’s bytes on the right speed of memory — fast for what is hot, cheaper for what is not — without rewriting the applications you already run.

For AI fleets and CXL-forward servers that need more work per machine without breaking latency SLAs.

Write to us See how

Three speeds of memory. One control plane.

Hot data stays fast.
Cold data gets cheaper.

Fast

GPU HBM / VRAM

The working set that cannot wait. We treat this as the scarce resource it is.

Working

DDR5

Warm pages and SLA-pinned tenants. Latency-critical work stays here even when the box is full.

Capacity

CXL / pooled

Cold data, moved only when the move pays for itself. Overcommit the rest safely.

How MemScale works

Observe, classify, plan, then act — or don’t.

The kernel places pages for locality. MemScale adds the layer it lacks: tenant and SLA awareness, expected-value gating, and fail-closed safety. It sits under schedulers and FinOps tools. Existing apps stay as they are.

  1. Observe heat, topology, and what the machine can actually do.
  2. Classify regions as hot, warm, or cold — with a confidence label on every number.
  3. Plan a move only when the benefit beats the measured transfer cost.
  4. Act with dry-run on by default, a kill-switch, and an audit trail. If we cannot prove it is safe, we do not pretend.

GPU inference-density claims stay labelled until they are measured on your hardware. We do not ship a modelled multiplier as a result.

Company

MemScale AB

A Swedish private limited company building one managed memory system for heterogeneous machines.

MemScale AB
Org.nr 559597-6679
Klisätravägen 21 A
138 33 Älta, Sweden

Contact

Start with a conversation.

Design-partner pilots, customer installs, and press — one inbox.

hello@memscale.io

memscale.io memscale.co.uk → memscale.io