Ikioma Mini
The desk-sized entry point — two linked nodes, one private stack.
The desk-sized entry point — two linked nodes, one private stack.
Hardware, delivered and deployed
Your data never leaves the floor. Air-gappable, no egress, no telemetry — the trust boundary has a serial number you can point at in an audit.
Own the machine and agents run around the clock. No per-token meter, no overage invoices — you pay for electricity, not tokens.
Hardware is a capital asset, not a subscription line. Weights, tunes, and stack stay yours — perpetually, with trade-in credit on next-gen silicon.
Two 150 mm nodes linked into one machine — 256 GB of unified memory and a tuned 32B model, quiet enough to sit on your desk next to your monitor. No rack, no fan noise, no data center.
Eight concurrent agents, 128k tokens of context, and models tuned for the work your team actually does — enough inference to start building without standing up infrastructure.
The same model and inference stack as every Ikioma — Conduit, a drop-in OpenAI/Anthropic API, signed audit log, sandboxed tools. Start here; trade in toward Ikioma when the workload outgrows the desk.
Private inference is a capital asset, not a subscription line — no per-token meter, no overage invoices. Most customers break even versus equivalent cloud spend within 4–7 months; security updates run 10 years, in writing.
Assembled in Yokohama and Eindhoven. Boards burn in for 96 hours before they leave the floor. Every system carries a serial, a calibration sheet, and a name etched on the chassis — because the people who built it stand behind it for a decade, not until next quarter's earnings call.
Preinstalled and pre-validated: the ikioma inference runtime, your choice of optimized open models, and the production middle layer (audit log, RBAC, evals) configured out of the box.
| Model | Params | Context | tok/s | Latency | Notes |
|---|---|---|---|---|---|
| ikioma-32B | 32B | 128k | — | — | TBD — burn-in benchmark |
Loading 3D view…
Every part of the SFF node you can pull apart — X-ray the shell to see the GB10 and 8× LPDDR5X, drag the explode slider for the teardown order, or switch on airflow and telemetry. Drag to orbit, scroll to zoom.
Illustrative 3D approximation — not manufacturing CAD. Schematic dimensions only.
Mini for the desk, Ikioma for the team, Pro for the fleet. Every version ships the same model and inference stack — the scale is what changes.