API

OpenAI-Compatible API

Applications talk to the cluster through a familiar chat-completions interface. The API layer authenticates, validates and streams; the controller coordinates the compute nodes behind it.

Endpoint

POST /v1/chat/completions

POST/v1/chat/completionsstreaming supported
jsonexample request
{
  "model": "YOUR_MODEL",
  "messages": [
    {
      "role": "user",
      "content": "Explain distributed AI inference."
    }
  ],
  "stream": true
}
jsonconceptual response shape
{
  "choices": [
    {
      "message": {
        "role": "assistant",
        "content": "..."
      }
    }
  ]
}
envconfiguration — placeholders
AI_BASE_URL=https://YOUR-CONTROLLER/v1
AI_API_KEY=YOUR_API_KEY
AI_MODEL=YOUR_MODEL
Application
     │
     ▼
OpenAI-Compatible API
     │
     ▼
Cluster Controller
     │
     ▼
Distributed Compute

Every credential, model name and URL shown here is a placeholder. Replace them with values from your own deployment and never publish real keys.

Client

Python Example

pythonopenai python sdk
from openai import OpenAI

client = OpenAI(
    base_url="https://YOUR_CONTROLLER/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="YOUR_MODEL",
    messages=[
        {
            "role": "user",
            "content": "Hello"
        }
    ]
)

print(response.choices[0].message.content)

Build Your Own AI Compute Cluster

Connect your machines. Deploy the Node Agent. Start the controller. Build a distributed AI environment around the hardware you already have.