SelfHostStackOpen-Source Directory
🤖

LLM Inference & AI Serving Alternatives

High-throughput local LLM inference engines, OpenAI-compatible proxy gateways, and model load balancers.

LLM Inference & AI Serving

OpenAI API & AWS Bedrock

Proprietary Cost$2.50–$30.00 / 1M tokens (GPT-4o / Claude 3.5 Sonnet / AWS Bedrock) — $1,500–$10,000+/month for production applications and RAG pipelines

Proprietary AI API endpoints with per-token pricing ($2.50–$30.00/1M tokens), strict rate limits, and data privacy risks.

Top Self-Hosted Alternatives:
vLLM8 GB (16GB+ VRAM GPU recommended)LiteLLM Proxy & AI Gateway512 MB
Starter Stack Pack — $29

Skip the setup: get the production-ready stack

Don't stitch together configs from five different READMEs. Get all 5 production-hardened Docker Compose stacks — Postgres, Redis, SSL auto-renewal, and backup scripts — ready to deploy in minutes.

n8nVisual workflow automation
📊UmamiPrivacy-first web analytics
🛡️Uptime KumaUptime monitoring & alerts
🔐VaultwardenBitwarden-compatible vault
☁️NextcloudDropbox/Drive replacement
Get the Stack Pack — $29 →

One-time purchase · Instant download · Production-ready

esc
navigate open