| Where | # | What is in there |
|---|---|---|
| Foundation — TAPPaaS repo | 8 | cluster network templates tappaas-cicd identity backup logging satellite |
| Apps — TAPPaaS repo | 12 | nextcloud euro-office vaultwarden litellm openwebui vllm-amd hass deconz coturn nextcloud-hpb netbird-client windows-server |
| Community repo — other people | 21 | immich jellyfin wordpress forgejo mailserver hosting sonos hue synology reolink solaredge alfen unifi shelly-fleet … |
Twenty-one of those forty-one are not mine. Erik wrote 13, Andreas 6.
That is the number I did not dare put in the abstract.
Every one of those services, without me:
The interesting engineering is not installing Nextcloud.
It is Nextcloud still being there, patched, in eighteen months, unattended.
nextcloud/
├── nextcloud.json # the contract — what it is, needs, provides
├── nextcloud.nix # the NixOS machine
├── install.sh # put it there (once)
├── update.sh # keep it patched (on schedule)
├── test.sh # prove it still works (gates the update)
└── README.md
The firewall is a module. The backup server is a module.
The mothership that installs the modules is a module.
One model to learn, not eight.
dependsOn — the trick the whole thing rests on{
"description": "Nextcloud — files, calendar, contacts",
"vmname": "nextcloud", "vmid": 210,
"dependsOn": ["cluster:vm", "templates:nixos", "backup:vm",
"network:proxy", "identity:identity"],
"config": {
"cluster:vm": { "cores": 4, "memory": "8192", "diskSize": "64G" },
"network:proxy": { "proxyPort": 443 }
}
}
Install order is computed from these declarations. Nobody maintains a list.
Add a module, and the platform works out that it needs a VM, an OS, a VLAN,
a certificate, a login and a backup job — in that order.
A key design goal is to REDUCE flexibility.
There is value in decisions having been taken up front.
Some use cases will not fit TAPPaaS. That is the trade.
For what fits, it is dramatically easier.
Zones, VLANs, naming, storage roles, backup policy, identity model — decided.
You get to pick the applications, and where they live.
My basement grew a brain.
| Silicon | AMD Ryzen AI MAX+ 395 — Radeon 8060S (Strix Halo) |
| Memory | 128 GB unified — the GPU sees nearly all of it |
| Largest model tested | gpt-oss-120b — 120B parameters |
| Speed | ~50 tok/s at 7B FP16 · ~20 tok/s at 30B 4-bit |
| API | OpenAI-compatible, on my own VLAN |
| Data leaving the building | none |
One commodity box. Not a rack, not a hyperscaler, not a monthly bill.
Then ask it something.
But we cannot do that here today. So instead — let us kill a node.
tappaas1 is running the firewall (vm:110), the mothership (vm:130)
and identity (vm:140) — all HA, fenced, watchdog armed.
I cut its power. They come back on another box.
Sovereignty is not a checkbox you tick at a vendor.
litellm in front means apps ask for "a model" — local today,
someone else's tomorrow, your choice, revocable.
What it actually cost.
2025-05 99 ██████████ ← the honeymoon
2025-06 21 ██
2025-07 53 █████
2025-08 141 ██████████████ ← Sommerhack 2025
2025-09 6 █ ← vacation
2025-10 9 █
2025-11 41 ████
2025-12 47 █████
2026-01 66 ██████
2026-02 141 ██████████████
2026-03 33 ███
2026-04 19 ██
2026-05 190 ██████████████████
2026-06 351 ██████████████████████████████████ ← something changed
2026-07 164 ████████████████
2026-08 126 ████████████ (to the 25th)
Not a burndown chart. A heartbeat — 1,507 beats, and two months where it
nearly flatlined.
The services stayed up through the flat bits anyway, because updates,
backups and tests do not need me to be enthusiastic.
That is not me getting three times better at typing.
That is the month I leaned all the way into AI-assisted development —
which is exactly the confession in part three.
Hold that thought.
| Files | Lines | |
|---|---|---|
| Foundation — the platform | 575 | 114,138 |
| Apps — the things you actually use | 144 | 15,726 |
For every line in an app module, there are seven lines of platform underneath.
tappaas-cicd alone — the mothership — is 81,942 lines, 72% of the
foundation. Inside it: 8 TypeScript managers (36,920 lines) and the
controller layer that talks to Proxmox, OPNsense and the switch (29,352).
Building this was not cheap, and I will not pretend otherwise:
112,000 lines, 1,507 commits, 24 ADRs, six unfamiliar stacks, a year of evenings.
But look at where that cost sits:
| Paid once, by whoever writes it | Paid by you, per site |
|---|---|
| the 8 foundation modules | a box and an afternoon |
| 195 test suites that gate every update | module-manager module add <app> |
| every install, update and repair script | nothing — patches arrive tested |
A hyperscaler amortises a datacentre across a million tenants.
We amortise a platform across a thousand basements.
Erik's 13 modules and Andreas's 6 cost me nothing and made my system better.
Yours would too. That is the economic case: not that self-hosting is cheap,
but that the expensive part only has to happen once.
No. There have been at least two large rewrites — and a whole system before this one.
install.sh you ran in thedependsOn / provides contracts, resolvedBoth times the same mistake: I had written as a sequence of steps
what needed to be a structure — dependencies the first time, roles the second.
The confession.
Not "AI-assisted autocomplete".
Root. ssh. nixos-rebuild. qm. pvesh. The firewall. The secrets.
Because the alternative was that it never got built.
The graph does not lie: May 190, June 351. That is what handing over the
tedium looks like.
And the honest part: I could not have hand-written the last third of this.
The guardrails are not vibes. They are written down, in the repo,
loaded on every single session, and they override anything the model
would otherwise default to.
By the time the AI is allowed to code, 2–3 artefacts exist —
carefully crafted and reviewed by humans:
Then the AI takes over — as eight specialist roles: architect, bash, python,
nix, tester, security, infra, PM.
The model does not decide what to build, or what "done" means. It never has.
Never run
git commitorgit push— full stop.
The operator performs ALL commits and pushes themselves.
This holds even when a request seems to imply it — "land it", "ship it",
"move this to main" — and even when a previous turn involved committing.
That is NOT standing authorization.
The AI may change any file on disk. It may not make a change permanent.
Every single commit that entered history passed under my eyes.
| Layer | What it bounds |
|---|---|
| Modules | A mistake lands in one VM, not "the server" |
| Zones | A compromised VM cannot reach what its VLAN forbids |
| Proxmox snapshots | Minutes-old rollback, per machine |
| PBS + off-site | Nightly, immutable, pull-based |
| NixOS | nixos-rebuild test before switch — a bad config dies at reboot |
| 34,490 lines of tests | The update does not land unless the service proves it works |
Root access is only terrifying if the system underneath is a snowflake.
Mine is disposable by construction.
The standing rules, as written:
main/stable, wiping /etc/secrets--no-verify, no silenced errors,nixos-rebuild test before switchNote what these have in common: they are all rules about honesty,
not about capability.
Every module ships two test levels:
| When it runs | What it is for | |
|---|---|---|
| quick | every change, every scheduled update | is this service still itself? |
| deep | test-module.sh <module> --deep |
the full behaviour, too slow for every commit |
And around them:
An agent that can edit code can also edit the test that would catch it.
That is exactly why a human writes the test plan first, and why the
regression sweep is the thing that says "done" — not the agent.
Yes — but "control" moved.
I no longer control every line. I control:
That is a real answer, not a comfortable one. Ask me the hard version in Q&A.
Build one too.
European sovereignty is not a policy problem you can wait out.
It is thousands of small boxes, in basements and back offices,
running software nobody can withdraw.
| Step | What you get |
|---|---|
| 1. One box, evaluation tier | 4 cores / 16 GB / two disks. Nested virt is fine. |
| 2. Proxmox + OPNsense | Zones, VLANs, DNS, certificates that renew |
| 3. The mothership | tappaas-cicd — the thing that installs the rest |
| 4. One service | Home Assistant. Small, useful, immediately missed. |
| 5. Backup before service two | Non-negotiable. Ask me why. |
No public IP? A satellite VPS is the escape hatch.
No GPU? Skip local AI, keep everything else.
Contributions welcome. Open an issue before you open a pull request —
somebody may already be packaging your app.
The harder, the better.
Sixty minutes. Roughly: 8 opening, 18 Good, 14 Bad, 14 Ugly, 6 close. Leave the last 5 for questions from the tent — this crowd will have them. Tone: this is a confession, not a product pitch. The proof is the mess.
Read the four bullets slowly. The fourth is the one that sounds mad and turns out to be the whole point — come back to it in the AI section.
The 1091 → 1507 jump is the joke that lands: I wrote the abstract in the spring and the project kept going. Point at it. The 41 is the other one worth pausing on: 8 foundation + 12 apps in the TAPPaaS repo, plus 21 modules other people wrote in the Community repo. That last number is the one that says this stopped being my hobby. Figures recomputed 2026-08-25 from the tree (a module = a directory with its own <name>.json contract), using the methodology in src/STATISTICS.md. Community repo: codeberg.org/TAPPaaS/Community.
Deliberately short divider slide. Say "it runs" out loud and pause.
Every number here was read off the running cluster, not a wiki page. The last line is deliberately understated — it sets up the sharing argument in The Bad, and the Community repo in the next slide. Five sites is not a movement yet; it is proof the thing installs somewhere that is not here. Worth saying out loud: only tanka1 on tappaas1 is a mirror. tappaas2 and tappaas3 run single NVMe — deliberate, they are rebuildable from backup. The 10G links are direct-attach copper; tappaas1 and tappaas3 have a second NIC on VLAN 100 straight to the modem, which is how OPNsense gets a WAN.
The point of showing both: TAPPaaS does not assume a datacentre OR a cupboard — the module contract is what makes those two rooms the same system. The makerspace rack is loud, rented space, many hands; mine is quiet, mine, and runs my family's calendar. Neither is the "real" deployment.
The point of the picture: the top two boxes are the only part anyone in my house notices, and they are the small part. Everything below the line is what makes them survive a year unattended. vllm-amd is the one LXC in the estate — a container, not a VM, because the GPU has to be passed through to reach the 8060S.
The abstract promised eight services. Give the 41 a beat before explaining it. Then be honest about what "in the tree" means versus "live in my basement": a module existing and a module your family depends on are different claims. The Community repo is codeberg.org/TAPPaaS/Community — one directory per person, no gatekeeping, and it is the only slide in this deck that is evidence the project outgrew me. Say that plainly; do not undersell it.
This is the core argument of the whole project. Slow down here. Self-hosting is easy. Self-hosting you can forget about is not.
If someone asks "is this just Ansible/Terraform?" — the answer is: those are how you build one machine. This is about a machine estate that has a shape.
This is the least popular slide with hackers and the most important one. Every hour you spend re-deciding VLAN layout is an hour not spent on services. Expect pushback; welcome it.
Stress "unified memory" — that is what makes a 120B model possible on a box that fits under a desk. This is a genuinely new thing as of this year.
The cable-pull needs the basement, so it is the screen recording today. LIVE DEMO instead: `ha-manager status` first, so the tent sees the three services sitting on tappaas1. Then pull tappaas1's power. Be honest that this drops the firewall — the whole site goes dark for the fencing timeout before the services restart elsewhere. That pause IS the demo; narrate it rather than filling it. `ha-manager status` again to show them landed on another node. Checked 15:02 today: quorum OK, master tappaas3, fencing armed.
Tie back to bullet four from the dream slide: "no reliance on the Internet". That was the mad-sounding one. Here it is, cashed in.
Point at Sept/Oct 2025 — six and nine commits — and say "I was on vacation." Deadpan. That is the whole joke, and it makes the following line land: a platform that needs you every week is not a platform, it is a pet. Then point at June: 351. Do NOT explain it yet — that is the hook into The Ugly, two slides from now. Numbers regenerated 2026-08-25; August is a partial month.
Deliberate hook into The Ugly. The commit spike is the evidence, and the audience will already be suspicious. Good. Let them be.
The 7:1 platform-to-app ratio is the single most useful number in this talk for anyone thinking "I'll just spin up Docker Compose". That works — until you want it to still work next year without you. Source: src/STATISTICS.md, regenerated 2026-08-23.
This is the slide that turns The Bad from a complaint into an invitation. The honest concession first — it IS a lot of work — then the pivot: almost none of it is per-site work. Land on the last line and pause. If someone objects "a thousand basements do not exist yet" — agree, and say that is precisely why you are standing here. TODO: your own hours-per-week number makes the concession land harder.
This is the credibility slide. Do not soften the "No" — an audience that has just been told you gave an AI root needs to hear that you rebuild things when they are wrong. Both ADRs are in the repo, and ADR-007f was written by Erik, not me — which is the two-reviewer rule from Guardrail 1 doing its job in public. If you want a third: ADR-014 recast zone tiers from README prose into checked state, and its own changelog lists four things the draft got wrong.
Say it flatly and then be quiet for three full seconds. Let the tent react. This is the slide people came for and the one they will argue with afterwards.
Do not be defensive. The result is on the projector; the method is the price.
This is the pivot from confession to engineering. Everything after this is transferable to anyone in the tent using these tools.
This is the most transferable slide in the talk, and the one people actually need. The order matters: problem, decision, plan, and only then code. Two reviewers on an ADR is the part that keeps me honest — Erik has killed several of my ideas, and the ADR is where that argument is recorded.
Verbatim from CLAUDE.md. The distinction — mutate freely, persist never — is the single most useful idea in this section. Say why: git is the undo button, so the undo button is the thing it must not touch.
This is the actual answer to "you gave an AI root?!". The architecture that makes unattended updates safe is the same architecture that makes an over-eager agent survivable. Same property, two beneficiaries.
The failure mode with a capable agent is not malice. It is an agent that makes the red thing turn green by removing the check. Name that explicitly.
This closes the loop opened in Guardrail 1: the plan said how it would be tested, and this is where that promise is collected. The last two lines are the answer to the sharpest question in the room — "how would you even know if it cheated?"
Do not claim more than this. If someone says "that's not control, that's supervision" — agree, and say supervision with a hard gate and a working undo is what control has always meant in operations.
Prepared answers to have loaded: - "Isn't this just Ansible?" → estate vs. machine; dependsOn - "Why NixOS?" → reproducible rebuild is the undo button - "You gave an AI root — seriously?" → guardrail 2, blast radius by design - "What if you get hit by a bus?" → open source, docs, ADRs, TODO: honest answer - "Cheaper than Google?" → no. Ownable, though. - "Can I run it on one Raspberry Pi?" → no. Evaluation tier, x86, be realistic.