Components
The Platform Has Four Core Layers
A worker agent on every machine, a control plane that coordinates them, a runtime that executes inference, and an API that applications consume. A dashboard sits on top for visibility.
Component 01
Node Agent
The worker software running on every participating computer.
The Node Agent is a lightweight Go-based service installed on each compute machine. It owns the machine’s identity in the cluster, reports what hardware it has, and executes the work the controller assigns to it.
Responsibilities
- Persistent node identity
- Node registration
- Authentication / enrollment
- Hardware detection
- GPU / CPU / RAM / VRAM reporting
- Heartbeats
- Telemetry
- Job lifecycle management
- Runtime integration
- Inference execution
- Cancellation
- Drain mode
- Reconnection
- Graceful shutdown
- Structured logging
- Health reporting
NODE AGENT LIFECYCLE Install ↓ Generate / Load Node Identity ↓ Connect to Controller ↓ Authenticate ↓ Register ↓ Report Hardware ↓ Heartbeat ↓ Receive Work ↓ Execute ↓ Stream Results ↓ Continue Heartbeats
Node Agent
Deploy a New Compute Node
Install the Node Agent on another computer and add it to the cluster.
.\install.ps1 ` -ControllerURL "YOUR_CONTROLLER_URL" ` -EnrollmentToken "YOUR_ENROLLMENT_TOKEN"
Platform buttons download the latest uploaded build for that platform. The enrollment token should be generated securely by the cluster administrator and never committed or published.
Component 02
Cluster Controller
The control plane that coordinates the distributed compute cluster.
The Cluster Controller is the centralized control plane, deployed independently from the worker machines. It knows which nodes exist, what state they are in, and which of them should participate in a given workload.
Responsibilities
- Node registration
- Authentication
- Node state management
- Health monitoring
- Heartbeat processing
- Telemetry collection
- Resource tracking
- Scheduling
- Job management
- Distributed orchestration
- Model management
- Admin APIs
- Event streaming
- Node draining
- Node recovery coordination
CLUSTER CONTROLLER
│
┌──────────────────┼──────────────────┐
│ │ │
▼ ▼ ▼
Node Registry Scheduler Job Manager
│ │ │
└──────────────────┼──────────────────┘
│
Health Monitor
│
▼
Event / Admin APIThe controller is currently deployed on a cloud platform and acts as the control plane only. The cloud platform is not the inference engine.
Component 03
Distributed Inference Runtime
The execution layer responsible for actually running distributed AI inference.
The Controller does not perform the actual LLM inference. It coordinates the compute nodes. The distributed runtime performs the inference work across the participating machines.
Controller
│
▼
Workload / Execution Plan
│
▼
Distributed Runtime
│
┌──┼──────────┐
▼ ▼ ▼
N1 N2 N3
│ │ │
└───┼──────────┘
▼
LLM Runtime
│
▼
Generated TokensCurrent runtime stack
- Distributed runtime
- llama.cpp
- GGUF model
- GPU / CPU compute
Only the runtime pieces listed above are integrated. Support for every model format or every inference runtime is not claimed.
Component 04
AI API Layer
How applications consume the cluster.
The platform exposes an AI API that applications can consume without directly communicating with individual compute nodes. Applications interact through a familiar API interface while the infrastructure handles compute-node coordination behind the scenes.
Application
│
▼
OpenAI-Compatible API
│
▼
Cluster Controller
│
▼
Distributed ComputeAI_BASE_URL=https://YOUR-CONTROLLER/v1 AI_API_KEY=YOUR_API_KEY AI_MODEL=YOUR_MODEL
Project Scout is one application that can consume this API.
Component 05
Cluster Dashboard
Centralized visibility into the cluster.
- Cluster health
- Nodes
- Node hardware
- Jobs
- Models
- Events
- Node status
- Resource information
- Drain / resume operations
Web Dashboard
│
▼
Admin REST API
│
▼
Cluster ControllerThe dashboard never communicates directly with worker nodes.
Build Your Own AI Compute Cluster
Connect your machines. Deploy the Node Agent. Start the controller. Build a distributed AI environment around the hardware you already have.