AI-Native Memory Systems
from Silicon to Tokens

The world's first Memory Processing Unit (MPU) — empowering AI compute with orders of magnitude more memory.

Learn more about the Technology

X-Mem™ Technology: the Memory Tier Built for AI

Hardware
  1. GPU
    G1HBM10 TB/s · 0.3 TB
  2. G1.5 & G2.ExpX-Mem™ MPU3.6 TB/s · 6 TB
Higher Bandwidth
Higher Capacity
Software
GPUX-Mem™ MPUGPU KernelsCUDATritonJAXFlashInferX-Mem™ LibraryInference / Training Runtime
  • XPU-Native Memory Tiering Library
  • Transparent to XPU kernels and inference / training runtimes
  • Verified solution for KV/prefix caching, long-context inference, multi-model serving, and optimizer/activation offloading
Benchmarks
  • 7×

    faster TTFT than DRAM offload on prefix-heavy workloads. Benchmark ↗︎

  • 4×

    higher concurrency with the same number of GPUs.

  • 40%

    cheaper inference for long-running background agents over software optimizations. Inference ↗︎

Join Our Team

Come and work with us in Cambridge, MA or Santa Clara, CA.