Overview

Every AI feature you ship sits on somebody's inference infrastructure. This is the day it becomes yours. We provision it end to end — the tooling, the hosted platforms, your own cloud, your own hardware, and the laptop in front of you — and then push it until it serves many services at once without falling over. You will leave with a running OpenAI-API-compatible endpoint, a client that cannot tell it apart from a commercial provider, and a costed argument for which one you should actually be paying for. Bring a machine with a dedicated GPU or an M-series Mac with 16+ GB of RAM. No GPU is fine — you will provision cloud instead, and we will cost that honestly too.

Where This Class Leads

What you leave able to do

  • run an open model locally
  • configure an open model endpoint behind a provider interface
  • justify a hosting choice by cost per outcome

How you show it. A costed comparison over one month of real traffic, the outcome counted in each case, and the threshold at which the answer flips.

Join Us

Want to schedule this class for your team?

Contact us: liz@themultiverse.school