Deploy the Creator-AI inference worker template with OpenTofu
Deployment · Updated Jun 2026
Coming from another cloud?
▸AWS·Bedrock
This Quake AI feature maps to AWS’s Bedrock.
▸Google Cloud·Vertex AI
This Quake AI feature maps to Google Cloud’s Vertex AI.
Deploy the Creator-AI inference worker template with OpenTofu
Stand up a CPU inference VM running Ollama, Qdrant, and a Python gateway for creator workflows using the validated OpenTofu templatecreator-ai-worker. Keep Ollama models small and expect higher latency than GPU hosts.
Monthly total for the required template above. Use the configurator below to add optional pieces and see the total update.
What each resource is for
Inference
c2a.large · 2 dedicated vCPU, 4 GiB RAM, 0.5 Gbps
$62.00/mo
Compute shown per role at custom-package rates ($29/dedicated vCPU, $7.25/shared vCPU, $1/GiB RAM). The headline above is the billed total: the cheaper of a named plan and the custom package, plus add-ons.
Private networks, subnets, Neutron routers, and security groups are included with the plan.
VPC objects are usually free to create elsewhere, but NAT gateways bill hourly plus per-GB processed. Quake AI uses router SNAT with no separate NAT line item.
$0.00
Control-plane API requests
OpenStack API calls for provisioning and management are included.
Some managed services on other clouds meter API calls or charge for premium control-plane features.
$0.00
Configure your estimate
Check the add-ons you plan to deploy to build a monthly total. Nothing is selected to start, so the total below begins at the baseline.
Starting template
The required baseline, always included.
$63.40/mo
Pick how much you expect to store to fold it into the total.
$0.00/mo
Your configured estimate$63.40/mo
Dev/test vs production
Start on shared CPU for dev/test, then promote to dedicated for production with a flavor resize. The network, storage, and template stay the same.
Dev/test on shared CPU
Burstable s1a flavors; suited to prototyping and low or bursty load.
$17.90/mo
Production on dedicated CPU
The headline estimate above; predictable steady-load performance.
$63.40/mo
Saves $45.50/mo while you build on shared CPU.
Shared flavors carry less RAM (c2a.large (4 GiB RAM) -> s1a.small (2 GiB RAM)). A resize reboots the instance; data on attached volumes persists. Size the dedicated flavor for the RAM your production workload needs.
Pricing data last validated: . For current rates, check quake.ai/pricing.
Click to zoom
Creator-AI worker topology: private inference VM with gateway API, Ollama, Qdrant, and Object Storage for media
A healthy stack returns JSON with a status field set to ok. If the request times out, confirm cloud-init finished on the inference instance in the Console.
List the data bucket to confirm credentials wired correctly:
bash
aws s3 ls "s3://$(tofu output -raw data_bucket)/" --endpoint-url "$AWS_ENDPOINT_URL_S3"
An empty listing is normal immediately after apply.