Evgenii Leonov
Case 03 · AI infrastructure · Concept

CentML:
complexity on one slider

A console where a product team launches its AI model by choosing a clear speed ↔ cost balance instead of GPUs and settings.

Product logicData vizDeveloper UXDesign system
The brand is a real project; the product is a concept. At CentML I was the lead graphic designer: logo, pattern system, office, conferences (NVIDIA GTC, BigData London) and support for the product UI in Figma. This is my own take on the product console. CentML branding →
Context

Fast or cheap shouldn’t be a GPU engineer’s call

To run an LLM in production you have to choose hardware, batching and quantization, and keep an eye on latency. These decisions drive the bill directly, yet only specialists understand them.

The idea of the concept is to translate technical parameters into business language: one slider between speed and price, with a live estimate of what each choice means.

Who it’s for

One screen, three different questions

01

ML engineer

Wants control and numbers: p50/p95 latency, GPU utilization, model versions. Can’t stand “magic” without explanations.

“Show me what’s under the hood and stay out of my way”
02

Product manager

Needs to test an idea fast: plug the model into the product and see what it will cost on real traffic.

“What will we pay if we get 10× more users?”
03

Finance

They look at the bill and the savings. They need one number and its trend, not latency charts.

“Why did our AI bill go up 40%?”
Key flow

From model to API in four steps

Model

From the catalog, or your own from a private registry

Balance

A speed ↔ cost slider; the hardware is chosen automatically

Launch

A ready OpenAI-compatible endpoint and code to paste

Monitoring

Traffic, latency, GPUs and cost in real time

Live prototype

Move the slider and see what changes

The prototype is built for desktop — best viewed on a computer.
  • 1
    ConfigureMove the slider: GPUs, speed, price and savings update
  • 2
    MonitorTraffic, latency, utilization of every GPU
  • 3
    PlaygroundPress Run: the model answers in real time with a live speed counter
Click “Deploy endpoint →”
Solutions

Making infrastructure easy to understand

SPEAK IN OUTCOMES

Not “4× H100” but “$0.67 per million tokens”

Technical parameters are still there, but the headline number is the one that ends up on the bill.

HONEST COMPARISON

The baseline is always on the chart

A red dashed line shows what running the model “as is” would cost. The value of the service speaks for itself.

DEVELOPER PATH

Code next to the button

An OpenAI-compatible snippet right in the wizard: copy, paste, and it’s already running in your product.

UI kit

The brand pattern inside the product UI

The green from the identity became the color of “running and saving”, the dark theme reduces eye strain in long sessions, and the monospace font keeps numbers aligned in columns.

Colors
#0B1F17
Background
#132E23
#00A87B
Brand
#5FD39A
#E9B949
p95
#EF6B5E
Baseline
Typefaces
58.4 req/s
Space Grotesk — large metrics
$0.67 / 1M
JetBrains Mono — numbers, code · Inter — interface
Components
Live · eu-westDeploy endpoint →Cancel
✓Model
2Optimize
3Deploy
The identity where it all started
Real project
CentML branding — a case in my graphic design portfolio
View branding ↗