Open to Work: Senior Backend & Distributed Systems Engineer (Go / Python) View LinkedIn →

Relay

Relay is a multi-tenant AI API gateway sitting between applications and major LLM providers (OpenAI, Anthropic, Mistral, Groq). This series explores building core backend infrastructure (database multi-tenancy, REST API design, rate limiting, webhooks, and billing pipelines) through the lens of a production system.

Technical Journal // Relay
Aug '26

Designing a Usage-Based Billing Pipeline for SaaS

A deep technical guide to designing a robust, provider-agnostic billing pipeline for usage-based SaaS. Covers subscription state machines, idempotent event processing, automated dunning logic, credit notes, and metered invoice generation with full SQL schemas, Mermaid diagrams, and Go snippets.

Jul '26

API Design for Backend Systems

A practical, opinionated guide to designing backend APIs. Covers request and response schema, offset and keyset pagination, BFF, server-driven UI, caching, versioning, naming, the new HTTP QUERY verb, WebSocket vs SSE, API security (auth headers, server-to-server auth, webhook receivers), and bulk data load APIs. Built around Relay, the same AI proxy system from the multi-tenant SaaS post, with REST and JSON as the running example and Go code throughout.

Jul '26

How Multi-Tenant SaaS Actually Works

A complete system design of Relay, a multi-tenant AI API gateway. Covers multi-tenant database architecture (silo vs pool vs bridge), API key authentication, provider routing with failover, dual-layer rate limiting, token-based billing, response caching, row-level security, tenant provisioning, and observability. Full Postgres schemas, mermaid diagrams, and Go snippets included.