cstm.ai · Custom AI hardware, software & developmentVendor-neutral · Quote onlyBuilt to spec
CSTM-SPEC · REV AcstmAI™ · Hardware · Compare

AI server configuration, compared.

Four sizes side by side, in the ranges we specify within. Use it to find your starting point; the quote turns each range into a specific part list for your workload.

Fig. 01 · Five-axis CNC milling, NASA Kennedy1998
Technical specifications

Four sizes, one table.

Typical ranges, not product specifications. Prices are never listed: every configuration is quoted from a written sheet.

Typical configuration ranges for cstmAI Desk, Rack, Cluster and Edge
SpecCSTM-D · REV AcstmAI™ DeskCSTM-R · REV AcstmAI™ RackCSTM-C · REV AcstmAI™ ClusterCSTM-E · REV AcstmAI™ Edge
ElevationFront elevation line drawing of a desk-side AI workstation towerH ≈ 17–22 INW ≈ 8–11 INFront elevation line drawing of a 4U rackmount GPU server19 IN RACK · 482.6 MM4U · 177.8 MMFront elevation line drawing of a 42U rack holding GPU nodes and fabric switches42U · ≈ 2 M≈ 600 MMFront elevation line drawing of a compact fanless edge AI systemW ≈ 8–12 IN · FANLESS SHOWNH ≈ 2–4 IN
GPU count1–24–816+ (2 or more nodes)Embedded module or 1–2 GPUs
GPU memory class24–96 GB per GPU (workstation)48–192 GB per GPU (data center)80–192 GB per GPU (data center)16–64 GB
Typical model size7B–70B, quantized at the top endUp to 70B at full precision; large MoE on 8 GPUsLargest open-weight models, several at once1B–14B, quantized
Concurrent users≈ 1–25≈ 25–500+Hundreds to thousandsOne site, or a machine feed
TrainingSmall adapter fine-tunesAdapter and moderate full fine-tunesLarge full fine-tunesInference only, as a rule
Power≈ 0.8–2 kW≈ 3–11 kW per serverOften 20–40 kW+ per rack≈ 60 W–1 kW
Rack units / formTower, or 4U rackmount2U–8UFull racks, 42U–48UFanless box, short-depth 1U–2U, rugged
CoolingAirAir; data-center airflowAir or direct liquid coolingPassive or sealed
Network1–10 GbE25–100 GbE200–400 Gb/s fabric per GPULAN, cellular or satellite
Deployment settingOffice, lab, closetServer room, data center, coloData center or coloPlants, vessels, field sites
Best forPilots and single teamsProduction for a companyOrg-wide AI service and trainingOffline and remote work
Next step

Figures are typical ranges for planning. Actual capacity depends on model, context length and usage pattern.

FAQ

Sizing, without guesswork.

How do you decide which size we need?

From measurements, not a rule of thumb. In discovery we run your task on candidate models, record response speed and quality, then size for your peak number of simultaneous users with headroom.

Why are these ranges so wide?

Because the same box serves very different loads. Ten people asking long questions about large documents can use more GPU than two hundred asking short ones. The quote narrows each range to one number for your workload.

Can we mix sizes?

Yes, and many organizations do: a Rack in the data center, Edge units at remote sites, and a Desk for development. They all run the same software and are managed the same way.

Get a quote

Spec your system.

Tell us the models you want to run, how many people will use them and where the hardware should live. An engineer replies with a first configuration and the questions that decide the quote.

Form CSTM-Q · Quote only