Platform · Coming soon
Lumus platform

Dedicated compute. A path to managed inference.

Australian AI infrastructure spanning dedicated B300 and RTX PRO 6000 capacity, with a path towards GPU, token and agent services.

Infrastructure profile
B300 · 36-system option
Coming soon
Managed inference
API request
Coming soon
POST /inference
{
  "model": "<selected-model>",
  "input": "<your-prompt>"
}

// Interface and endpoint TBC
Day 1 scale
32 or 36 systems under assessment
Reference platform
NVIDIA DGX B300 working basis
Initial consumption
Reserved capacity, likely bare metal
Future service layer
Managed inference remains TBC
Infrastructure clusters

One platform. Distinct cluster profiles.

Lumus is developing distinct cluster profiles for different workload and deployment requirements. Final configurations and availability will be announced as development progresses.

01 / B300 infrastructure Coming soon

36-system NVIDIA DGX B300 design option

A first Australian cluster shaped around reserved capacity and predominantly inference-led demand, with final scale selected against architecture and the approved funding envelope.

  • Day 1 decision32 and 36 systems compared
  • Reference platformNVIDIA DGX B300
  • Commercial starting pointReserved capacity
  • Operating pathManaged mobilisation
02 / RTX PRO infrastructure Coming soon

RTX PRO 6000 cluster profile

A dedicated cluster profile designed around its own infrastructure requirements. Architecture, system count, facility basis and commercial offer remain to be defined.

  • ConfigurationTo be confirmed
  • Deployment statusComing soon
  • Design principlePurpose-built platform profile
Working topology

Visualising the 36-system option.

This system map gives the scale a clear, legible form without presenting unresolved networking, storage, cooling or facility specifications as final.

The system map presents the current 36-system design direction. Final configuration remains subject to technical validation.
Inference API · Coming soon

A future path from GPU capacity to managed model access.

The Lumus product ladder can move from reserved bare metal to GPU as a service, token as a service and eventually agent systems. This interface explores how that future offer could be presented.

Lumus managed inference
Coming soon
Language

General language model

Model catalogue and licensing to be confirmed.

Per-token modelTBC
Coming soon
Reasoning

Reasoning model

Model catalogue and licensing to be confirmed.

Per-token modelTBC
Coming soon
Multimodal

Vision-language model

Model catalogue and licensing to be confirmed.

Managed accessTBC
Coming soon
Language

Compact language model

Model catalogue and licensing to be confirmed.

Per-token modelTBC
Coming soon
Embedding

Embedding model

Model catalogue and licensing to be confirmed.

Managed accessTBC
Coming soon
Agent systems

Agent service layer

Portal, workflow and service design remain future scope.

Portal modelFuture
Future service
01

Control plane

A managed or shared service needs a stronger platform layer than reserved bare metal.

02

Isolation

Customer and workload boundaries must be designed into the shared service model.

03

Metering

Token-based consumption requires measurable usage and a defined commercial model.

04

Lifecycle

Models and the serving layer need an ongoing operational and support boundary.

Platform evolution

Start with control. Add services with capability.

Lumus is building from dedicated infrastructure towards increasingly managed AI services, giving customers more ways to access compute as the platform grows.

01
Initial focus

Bare metal and reserved capacity

Customers bring their own software environment and consume dedicated capacity.

02
Next service layer

GPU as a service

GPU capacity is combined with a platform layer and a clearer operational boundary.

03
Future · TBC

Token as a service

Lumus runs selected models and meters customer consumption by token.

04
Future service

Agent as a service

Complete agent systems are hosted within the Lumus platform and accessed through a portal.

Shape the platform

Planning an inference-led workload?

Talk with Lumus about reserved capacity requirements or register interest in the upcoming managed inference service.

Start a conversation