One node to start. More nodes for capacity and for durability.
A node is one machine running the engine. It uses the disks, cores and memory that machine has, and it is licensed as one node whatever those happen to be.
- A single node is not durable by itself. Back up anything you cannot lose.
- Durability comes from more than one node holding your data, not from a setting on one machine, which is what replication across nodes is for.
- The first node is a free test node, with no time limit, so you can prove it on your own data before you talk to anyone.
- A node is a node: the licence does not change with its cores, memory or disks.
Today, and what follows
How a node is measured
When we measure the engine, we record what the machine was doing while it answered, and publish that beside the number rather than instead of it.
Every disk
Each disk's throughput, against what that disk can physically do, so a number from an idle disk is visible as one.
Every core and thread
How many were actually working, judged over the whole run rather than at its busiest instant.
Every node
The same figures per node and for the cluster together, so one busy machine beside an idle one cannot read as a full one.
A number published without that record is not a result we would ask you to believe. It appears on the benchmark page the same way.