fadenstack

Get started

From a bare server to the first answer.

One Linux server runs Fadenstack and needs no GPU. Your GPU machines join from the console.

$ pipx install fadenstack
$ faden deploy

uv tool install fadenstack works as well. Add -y to faden deploy for an unattended install.

Four steps

Install, connect, pick, invite.

server · bash
$ pipx install fadenstack
$ faden deploy
✓ Docker and Compose v2 found
✓ Configuration written  # random keys and database passwords
✓ HTTPS with a certificate for this server
✓ Services started

  Console  https://ai.example.internal
  API      https://ai.example.internal/v1

Requirements

What you need.

The server and the GPU machines are separate computers. A small team can start with one server and one GPU machine.

ItemThe serverEach GPU machine
Operating systemLinux, x86_64Linux, x86_64, or an NVIDIA DGX Spark (aarch64)
SoftwareDocker Engine 24 or later with Compose v2; Python 3.12 or laterDocker or Podman with access to the GPU (for Docker, the NVIDIA Container Toolkit), and systemd
GPUNoneAn NVIDIA GPU with a working driver
CPU and memoryAt least 4 cores and 6 GB, or 8 GB with the privacy module, MCP servers and skills. 8 cores and 8 to 12 GB recommended.As the models you run need
DiskAt least 60 GB; 100 GB or more recommended, plus space for modelsAbout 15 GB for the runtime image, plus the models it runs
NetworkPorts 80 and 443 freeReaches the server. The server reaches a cluster's head machine on its serving port, and a cluster's machines reach each other. Machines without internet get the runtime and the models from the server.

After the install

Keeping it running.

Upgrade

Update the command line tool, then the server.

pipx upgrade fadenstack
faden upgrade

Back up and restore

One command each, for the databases and the configuration. Restoring is still being tested.

faden backup
faden restore <dir>

The full guide

HTTPS with your own certificate, more machines, clusters, upgrades and backups, step by step.

Help: Operate

Prefer to watch first?

A seven-minute video, from a bare server to the first chat and the first API call.