Self Hosted Ai
Self-hosted LLM, embedding, RAG, and AI agent workloads on Quake AI compute
13 documentation pages carry this tag. All tags
- →
Add a Browser Interface to Ollama with Open WebUI
deployment
- →
Build a RAG Pipeline with Ollama and Qdrant
deployment
- →
Creator-AI inference worker
Template
- →
Deploy a bursty-GPU control plane with a queue and workers
deployment
- →
Deploy a Vector Database with Qdrant
deployment
- →
Deploy an AI Agent
deployment
- →
Deploy an inference gateway with OpenTofu
deployment
- →
Deploy JupyterHub with the jupyterhub template
deployment
- →
Deploy the Creator-AI inference worker template with OpenTofu
deployment
- →
Inference gateway
Template
- →
JupyterHub notebook server
Template
- →
Qdrant vector database
Template
- →
Run a Local LLM with Ollama
deployment