API
OpenAI-Compatible API
Applications talk to the cluster through a familiar chat-completions interface. The API layer authenticates, validates and streams; the controller coordinates the compute nodes behind it.
Endpoint
POST /v1/chat/completions
POST
/v1/chat/completionsstreaming supportedjsonexample request
{
"model": "YOUR_MODEL",
"messages": [
{
"role": "user",
"content": "Explain distributed AI inference."
}
],
"stream": true
}jsonconceptual response shape
{
"choices": [
{
"message": {
"role": "assistant",
"content": "..."
}
}
]
}envconfiguration — placeholders
AI_BASE_URL=https://YOUR-CONTROLLER/v1 AI_API_KEY=YOUR_API_KEY AI_MODEL=YOUR_MODEL
Application
│
▼
OpenAI-Compatible API
│
▼
Cluster Controller
│
▼
Distributed ComputeEvery credential, model name and URL shown here is a placeholder. Replace them with values from your own deployment and never publish real keys.
Client
Python Example
pythonopenai python sdk
from openai import OpenAI
client = OpenAI(
base_url="https://YOUR_CONTROLLER/v1",
api_key="YOUR_API_KEY"
)
response = client.chat.completions.create(
model="YOUR_MODEL",
messages=[
{
"role": "user",
"content": "Hello"
}
]
)
print(response.choices[0].message.content)Build Your Own AI Compute Cluster
Connect your machines. Deploy the Node Agent. Start the controller. Build a distributed AI environment around the hardware you already have.