GEKI.AI
AMD Instinct MI350P

AI Factory on AMD MI350P.

The AMD Instinct MI350P with 144 GB HBM3e is AMD’s newest inference GPU - strong price-performance, fully managed by GEKI with SLAs in the right data center.

144 to 1,152 GB VRAM
Cost-efficient for mid-tier models
Fully managed

Non-binding · reply within 1–2 business days

AMD Instinct MI350P PCIe

AMD Instinct MI350P, operated by GEKI.

Why this card

Why we run the AMD Instinct MI350P at GEKI.

AMD’s newest generation, 144 GB HBM3e, built for inference. That makes the MI350P a very economical base for open models.

The full technical specs: AMD Instinct MI350P →

  • Strong price-performance

    Plenty of HBM3e memory and high bandwidth - for inference workloads often the more economical choice.

  • AMD’s newest generation

    MI350 series with 144 GB HBM3e, built for AI inference.

  • Sovereign & open

    Run open models in the EU, under your control and without hyperscaler lock-in.

  • Fully managed

    GEKI plans, delivers and operates with SLAs. You don’t need your own GPU team.

Different configurations

Your GEKI AI Factory on AMD MI350P.

Fully managed AI factories in different configurations, matched to your models.

2× MI350P

288 GB total VRAM

A compact entry into your own sovereign AI factory. Fits e.g. DeepSeek V4 Flash, Qwen 3.6 or Gemma 4 for a high number of concurrent users.

8× MI350P

1,152 GB total VRAM (≈ 1.1 TB)

More capacity for larger models and many parallel users. Fits e.g. GLM 5.2 or current models around 1T parameters.

Suitable for companies and SaaS providers that want to run open models under their own control. Get in touch →

What it's a fit for

Built for AI tasks in the enterprise.

Open models on your own factory cover a large share of enterprise tasks.

A good fit for

  • Internal company chat and knowledge assistant
  • Support: pre-sort tickets and draft replies
  • Document processing
  • Agents for recurring workflows
  • Inference endpoints for your SaaS product

Less suited for

  • Research at the absolute performance frontier
  • Training very large foundation models
How it works

From use case to a running factory.

  1. 1

    Use-case call with Matthias

    An honest assessment of which configuration fits - or what to do instead.

  2. 2

    Configuration & managed pricing

    A clear offer plus operations as managed pricing.

  3. 3

    Operated by GEKI

    GEKI operates with SLAs in the right data center and supports your workloads and agents.

Next step

A fit for your workload?

A short call with Matthias — an honest assessment of configuration and operating model.