Stacks

Find the right AI stack recipe.

Stack recipes are tested combinations of model, runtime, and interface that work together — a proven starting point instead of assembling one from scratch. Search stack recipes for local AI, private chat, coding assistants, RAG, automation, MCP agents, and home labs.

Editorial-first — instructions preserved as written and tested

15Stack recipes
7Categories
5Starter stacks

Tested recipes

Each stack recipe documents the tools, hardware, and time investment before you start.

Start small

Quickstart recipes are ordered from beginner to intermediate so you add complexity one layer at a time.

Decision-focused filters

Narrow by local AI, coding, RAG, agents, or automation before comparing recipes.

New to local AI? Start with the Ollama quickstart — the fastest path from zero to a running local model.

Ollama + Open WebUI starter stack →

Quickstart recipes

Three paths from zero to running model

Recipe matrix

All stack recipes by use case

Agents

Agentic Workspace Stack

Local-first multi-agent framework utilizing specialized open-weight reasoning and tool-calling models.

View recipe →
Automation

AI Automation Workflow Stack

A workflow automation stack for turning AI outputs into drafts, approvals, notifications, and controlled actions.

View recipe →
Local AI

Edge and Low-Power AI Stack

A minimal always-on inference setup for sub-4B models on Raspberry Pi 5, ARM SBCs, Intel NUCs, or embedded appliances with 8–16 GB RAM and no discrete GPU.

View recipe →
Enterprise

Enterprise RAG Stack with Access Control

A production RAG setup combining Qdrant, LlamaIndex, and a local LLM behind an authentication and role-based access control layer — for regulated enterprise environments.

View recipe →
Cloud GPU

GPU Cloud Fine-Tuning Stack

A stack for fine-tuning or running large models on rented GPU cloud infrastructure without owning the hardware.

View recipe →
Local AI

Home Lab AI Stack

A home-lab stack for local inference, private chat, and optional smart-home integrations.

View recipe →
Local AI

Lightweight Laptop AI Stack

A minimal stack for running small local models on a laptop or low-VRAM machine without a dedicated GPU.

View recipe →
Coding

Local Coding Agent with Persistent Memory

A coding agent stack that retains project context across sessions using local memory storage — giving the model awareness of conventions, past decisions, and ongoing work without restarting from scratch.

View recipe →
Coding

Local Coding Assistant Stack

A coding stack for testing open coding assistants with local or self-hosted models, repository-aware tools, and human code review.

View recipe →
Local AI

Local Multi-Modal Vision Stack

A local stack for image understanding, visual Q&A, and document image analysis using vision-capable open models running fully on your hardware.

View recipe →
Agents

MCP Agent Workflow Stack

A starter stack for connecting coding agents to approved tools through MCP servers.

View recipe →
RAG

Private Document RAG Stack

A local RAG setup for indexing internal documents with a vector store, a model runner, and a chat UI.

View recipe →
Local AI

Private Local Chatbot Stack

A practical stack for running private chat over local or self-hosted models with a browser chat UI and clear review habits.

View recipe →
Enterprise

Private Sovereign Knowledge Base

Air-gapped text extraction and semantic search cluster for processing confidential enterprise documents.

View recipe →
Automation

Voice AI Assistant Stack

A voice-in, voice-out AI pipeline combining Whisper for transcription, a local LLM for reasoning, and ElevenLabs or Coqui TTS for speech synthesis — no cloud transcription required.

View recipe →

Showing 15 of 15 stacks.

Stack recipe

Agentic Workspace Stack

Local-first multi-agent framework utilizing specialized open-weight reasoning and tool-calling models.

View stack

Stack recipe

AI Automation Workflow Stack

A workflow automation stack for turning AI outputs into drafts, approvals, notifications, and controlled actions.

View stack

Stack recipe

Edge and Low-Power AI Stack

A minimal always-on inference setup for sub-4B models on Raspberry Pi 5, ARM SBCs, Intel NUCs, or embedded appliances with 8–16 GB RAM and no discrete GPU.

View stack

Stack recipe

Enterprise RAG Stack with Access Control

A production RAG setup combining Qdrant, LlamaIndex, and a local LLM behind an authentication and role-based access control layer — for regulated enterprise environments.

View stack

Stack recipe

GPU Cloud Fine-Tuning Stack

A stack for fine-tuning or running large models on rented GPU cloud infrastructure without owning the hardware.

View stack

Stack recipe

Home Lab AI Stack

A home-lab stack for local inference, private chat, and optional smart-home integrations.

View stack

Stack recipe

Lightweight Laptop AI Stack

A minimal stack for running small local models on a laptop or low-VRAM machine without a dedicated GPU.

View stack

Stack recipe

Local Coding Agent with Persistent Memory

A coding agent stack that retains project context across sessions using local memory storage — giving the model awareness of conventions, past decisions, and ongoing work without restarting from scratch.

View stack

Stack recipe

Local Coding Assistant Stack

A coding stack for testing open coding assistants with local or self-hosted models, repository-aware tools, and human code review.

View stack

Stack recipe

Local Multi-Modal Vision Stack

A local stack for image understanding, visual Q&A, and document image analysis using vision-capable open models running fully on your hardware.

View stack

Stack recipe

MCP Agent Workflow Stack

A starter stack for connecting coding agents to approved tools through MCP servers.

View stack

Stack recipe

Private Document RAG Stack

A local RAG setup for indexing internal documents with a vector store, a model runner, and a chat UI.

View stack

Stack recipe

Private Local Chatbot Stack

A practical stack for running private chat over local or self-hosted models with a browser chat UI and clear review habits.

View stack

Stack recipe

Private Sovereign Knowledge Base

Air-gapped text extraction and semantic search cluster for processing confidential enterprise documents.

View stack

Stack recipe

Voice AI Assistant Stack

A voice-in, voice-out AI pipeline combining Whisper for transcription, a local LLM for reasoning, and ElevenLabs or Coqui TTS for speech synthesis — no cloud transcription required.

View stack

Useful starting points

Next step

Test and then build.

Use the Playground and compatibility checker before settling on a stack.