Portland · Remote-first · Python / FastAPI / AI platform

Andrew White Jr.
Self-taught · production-scarred · 79 services deep

12+ years in trades and facilities maintenance before software. 18 months running a multi-service distributed platform solo as my daily-driver dev environment. I went the long way around and I know why "production" is a weight, not a phrase.

Portland, OR 97219 · (503) 583-9175 · andrew.white.cog@gmail.com · github.com/awhi4130

What I've built (solo, in production, last 18 months)

01
79 FastAPI services, running 24/7
Each on its own port, heartbeat-based health checks, HMAC-authenticated internal auth, live schema migrations under load.
02
Multi-provider LLM router
Anthropic · OpenAI · Google · DeepSeek · Ollama · Grok · Perplexity · Kimi. Per-provider circuit breakers (3 failures → 120s cooldown → recovery probe), cost-aware failover, OpenAI-compatible gateway.
03
111-feed real-time intelligence aggregator
Crypto prices, social trends, RSS, news, structured APIs into unified SQLite store. Per-feed staleness detection. Push alerting via Discord + Pushover + email.
04
Multi-brand Flask storefront
Stripe live billing, Ghost CMS publishing via hand-rolled HMAC-SHA256 JWT (no library), zero-downtime schema migrations, external marketplace deep-linking.
05
Patient-facing health tooling
I'm an end-stage renal disease patient. Recipe macro engines on USDA FoodData Central, clinical handout generators I draft for my own dietitian, medication-conflict scanners, dialysis fluid trackers. I know what's broken in patient-facing health tech because I'm the user.
06
Local AI media stack
ComfyUI + Stable Diffusion / Flux / Wan / Qwen-Image / HiDream with fallback to hosted providers (Replicate, Together, Gemini Imagen) when local GPU saturates.

Scars I earned (in chronological order)

Resume variants (pick the one that matches the role)

#VariantLead withBest for
01AI Platform EngineerMulti-provider router, LLM observabilityAnthropic, OpenAI, Modal, Together, Replicate, Fireworks, Baseten, Perplexity
02Backend / Full-stackFastAPI + Flask + SQLite + StripeSupabase, Linear, PostHog, Vercel, Stripe, Shopify, Nike DTC
03DevOps / SRE79-service supervision, heartbeat observabilityGrafana Labs, Sentry, Datadog, Fly.io, FetLife, Cloudflare
04Healthcare AIPatient-builder, PHI-safe routing, USDA FoodDataHinge Health, Solace Health, Turquoise Health, Providence, OHSU, Proven Software, Tempus, Verily
05Data / Indexing555 MB recovery, 111-feed aggregator, schema migrationsSnowflake, ClickHouse, MongoDB, Voxel51, Northbeam, Trace Labs
06Trades → Tech12+ years facilities + software transition narrativeTriMet, PGE, municipal, utilities, facilities-adjacent tech
Why I disclose this up front I'm on hemodialysis three mornings a week (Tue / Thu / Sat, Portland, OR). I work productively around it with remote-first async or flexible-hours hybrid, and I'd rather mention it in line one than in week three. The medical reality is why I'm precise about what "shipped" means — if it's still running tomorrow, it shipped.

Quick links

How I'd spend my first two weeks on your team

Read the code, shadow oncall, learn the deploy path end-to-end, and find one real bug to fix before anyone assigns me one. I want the feel of "shipped here" before I try to decide what to build next. After that, ask me what's obviously broken and I'll tell you.

What I won't pretend