NEW — Self-Hosted AI Infrastructure

Self-Hosted AI Infrastructure — Deploy in Minutes

Production-ready AI stacks that deploy automatically to your server. n8n workflows, AI agents, vector search, observability. Zero manual config. Paddle-compliant software products.

✓ 100% Paddle-compliant — Software products: infrastructure templates, deploy scripts, configs. No human services, no prohibited categories.

15k+GitHub stars (n8n AI starter kit)
180+Hermes skills available
<15 minAverage deploy time
ZeroData leaves your server

From Purchase to Production in 4 Steps

Zero manual config. Scripted deploys. You own everything.

1

Pick Product

Browse products, compare tiers. Each shows exact deliverables, requirements, deploy time.

2

Checkout (Paddle)

Instant access to private repo + deploy scripts + docs. License key emailed.

3

Deploy

Starter: run script yourself (5-15 min). Pro/Enterprise: we deploy with/for you on a scheduled call.

4

Run

Agents start on boot. Access Web UIs. Use runbook for operations. Sérénité = we monitor for you.

Our Products

Production-ready AI stacks that deploy automatically to your server

View All Products →
⚙️
automation🎯 Non-Dev

n8n AI Automation Stack

400+ integrations + local LLMs + vector DB — one deploy

Complete n8n environment with Ollama (local LLMs), Qdrant (vector search), PostgreSQL. Pre-loaded with AI workflow templates: document Q&A, Slack bots, appointment scheduling, email processing. Deploys via Docker Compose in <5 minutes.

  • n8n (latest) with 400+ nodes
  • Ollama for local LLMs (Llama 3, Mistral, CodeLlama)
  • Qdrant vector DB for RAG
  • PostgreSQL for n8n data
  • Traefik reverse proxy + auto-SSL
  • +3 more
149€one-timeDeploy: 3-5 minutes (scripted)

✓ Paddle-compliant: Software product: infrastructure template + workflow JSON. Customer pays for deployable stack, not AI output.

🤖
agents⚡ Developer

AI Agent Runtime (Hermes + OpenCode)

Autonomous agents that run on YOUR server

Deploy Hermes Agent (autonomous ops, memory, skills, cron, Telegram gateway) + OpenCode (headless coding agent) as systemd services. Pre-configured with NVIDIA NIM models, skills library, project templates. Agents start on boot, persist across reboots.

  • Hermes Agent v0.21+ (systemd service)
  • OpenCode v1.18+ (headless coding agent)
  • NVIDIA NIM model routing (free tier)
  • 7 built-in coding agents (reviewer, debugger, etc.)
  • Telegram/Discord/Slack gateway
  • +4 more
199€one-timeDeploy: 10-15 minutes (scripted)

✓ Paddle-compliant: Software product: agent runtime + config templates. Customer pays for infrastructure deployment, not agent output.

📚
knowledge🎯 Non-Dev

RAG Knowledge Base Stack

Private ChatGPT for your docs — deploys in minutes

Complete RAG pipeline: document ingestion (PDF, MD, HTML, Notion, Confluence) → embedding (local or API) → vector search (Qdrant) → chat UI (Open WebUI) → API. Private, no data leaves your server. Pre-configured for internal knowledge, support docs, onboarding.

  • Open WebUI (chat interface)
  • Qdrant vector DB
  • Ollama (local embeddings) or OpenAI/Cohere API
  • Document ingestion pipeline (watch folder + API)
  • Multi-tenant collections (per team/project)
  • +3 more
99€one-timeDeploy: 5-8 minutes (scripted)

✓ Paddle-compliant: Software product: RAG pipeline + chat UI. Customer pays for search infrastructure, not generated answers.

Ready to Deploy Your AI Infrastructure?

Pick a product, checkout via Paddle, run the deploy script. Production in minutes.

Frequently Asked Questions

Do these run on my server or yours?

100% on YOUR server (VPS, bare metal, cloud VM). You get root access, the deploy scripts, and full control. We never see your data, API keys, or logs.

What if I don't have a server?

We recommend Hetzner CX42 (8GB RAM, 160GB SSD, ~€32/mo) or similar. You provision it, give us root (for managed tiers) or run the script yourself (starter). You pay the provider directly.

Are there recurring fees to you?

No. One-time payment for the product. Optional: Sérénité maintenance at €149/mo (updates, health checks, backup audits, 1h support/mo). Cancel anytime.

How does deployment work?

Starter: you run `bash deploy.sh` on your server (5-15 min). Pro: we hop on a call, you share screen/SSH, we run it together. Enterprise: you give us temporary SSH, we deploy, harden, test, hand over with runbook.

What models do the agents use?

Default: NVIDIA NIM (free tier, 40 req/min). Optional: Ollama (local, free), OpenAI/Anthropic API (your keys, your bill), or any OpenAI-compatible endpoint. You choose at deploy time.

Is this Paddle-compliant?

Yes. Every product is a SOFTWARE PRODUCT (deployable stack, templates, scripts, config). You pay for infrastructure automation, not AI-generated output. No human services, no outbound marketing, no prohibited categories.

Can I customize after purchase?

Absolutely. You get the source: Docker Compose files, scripts, configs, workflow JSON, agent skills. Modify anything. Starter/Pro include docs for self-modification. Enterprise includes custom dev hours.

What happens after support period ends?

You keep everything forever. No license expiration. Renew Sérénité (€149/mo) for continued updates/health checks, or manage yourself. We provide upgrade guides for major versions.