Jason Camp

Infrastructure, Platform & Reliability Leader
me@jason.camp | (732) 983-7858 | Albany, New York, USA
PDF Resume
$ whoami

I've spent three decades building and scaling things that didn't exist yet. That work spans adtech, agtech, fintech, crypto, healthcare, and AI. I work across the stack in Python, Ruby, and Go. On the platform side, I handle Kubernetes migrations, IaC modernization, CI/CD overhauls, and observability buildouts. More recently I've been designing agentic systems, RAG pipelines, MCP servers, and multi-provider LLM routing for production workloads. I've been part of early teams acquired by Yahoo!, Google, and Gemini, won TechCrunch Disrupt 2015, and hold two US patents in fintech.

Core Capabilities

Platform & Reliability

SRE Kubernetes EKS High Availability Disaster Recovery Datadog Prometheus New Relic Grafana

Cloud & IaC

AWS Google Cloud Terraform OpenTofu Terragrunt ArgoCD Scalr Networking Security

AI Infrastructure

Agentic Systems Bedrock RAG MCP Servers Claude Code LLM Routing Anthropic API OpenAI

Languages & Data

Go Python Ruby PostgreSQL MySQL MongoDB Kafka Redis Linux
Professional Experience

Lead Platform Engineer

DEC 2023 – PRESENT

Run infrastructure for Procare's online platform end-to-end, leading a team of two engineers: architecture with product, EKS operations and upgrades, DR strategy, cost optimization, release management, IaC, and internal tooling. Lead infrastructure for Procare's AI initiatives, supporting agentic systems and multi-LLM workloads in production.

  • Migrated legacy CloudFormation and instance-based infrastructure to Terragrunt/OpenTofu and multi-region EKS, cutting deployment times from 1 hour to 5 minutes and infrastructure costs by 50%.
  • Moved build pipelines to Bitbucket Pipelines with ArgoCD and adopted Scalr for IaC CI/CD with reusable Terraform modules, standardizing build and deployment platforms across applications.
  • Designed multi-provider LLM routing with failover, DR patterns for AI workload state, and observability for non-deterministic systems.
  • Support developers and promote platform advocacy across the organization.

Staff Platform Engineer

JAN 2023 – DEC 2023
Readyset · Company-wide reduction in force

Owned Readyset's production infrastructure across multiple AWS EKS clusters, with developer support and on-call.

  • Designed and built the customer-facing cloud platform on ECS, EKS, and RDS that turned Readyset into a hosted SaaS product.
  • Evaluated GCP as a second cloud provider; supported build pipelines across GitHub and Gerrit; ran Prometheus monitoring.

Senior Platform Engineer

JUN 2022 – DEC 2022

Supported Chick-fil-A Supply's infrastructure across four products, each on dedicated AWS EKS clusters spanning development, staging, and production.

  • Managed all infrastructure as code in Terraform: reusable modules and custom Terraform providers.
  • Migrated EKS clusters to Bottlerocket and cut deployment times by 50%.
  • Scalr for IaC CI/CD and ArgoCD for application deployments; internal monitoring, developer support, and on-call.

Senior DevOps Engineer

JAN 2022 – JUN 2022
Kingfield · Company-wide reduction in force

Managed a large multi-account AWS environment (RDS, Kinesis, EventBridge, Lambda, Fargate) with extensive IAM policy work, Datadog monitoring and runbooks, simplified regression testing, and Python/Go tooling. 24/7 on-call.

Senior Platform Engineer

JUL 2021 – JAN 2022
divvyDOSE · Consulting

Owned AWS infrastructure and a large Terraform repository with many shared modules. Built Datadog monitoring and metrics, documentation and runbooks, and DevOps tooling in Python, Ruby, and Go.

Site Reliability Engineer

NOV 2020 – JUL 2021

Joined at ~40 employees to help shape the engineering environment. Led migration of Rockset's real-time analytics platform from managed Kubernetes to Amazon EKS for cost and stability, then operated the clusters: Kafka, Zookeeper, etcd, Terraform, Helm, Grafana, and Prometheus.

Software Engineer, Credit Card

JUL 2020 – NOV 2020

Joined Gemini through the Blockrize acquisition for a four-month product-integration engagement, leading the infrastructure plan and architecture for the crypto rewards card. Designed for secure processing and storage of financial and user data, aligned with Gemini's security and compliance standards.

Staff Site Reliability Engineer

MAY 2017 – JUL 2020

SRE team lead across on-prem and AWS infrastructure. Managed clusters running Cassandra, ScyllaDB, Redis, and PostgreSQL; Prometheus monitoring, Docker and Kubernetes, focused on performance, scalability, redundancy, and high availability.

Co-founder & CTO

APR 2015 – APR 2017
Agrilyst · TechCrunch Disrupt 2015 Winner · Concurrent with Google through Apr 2016 · acquired by IUNU

Ran the technical team and day-to-day CTO responsibilities. Built infrastructure environments on AWS EC2 using Ruby on Rails, keeping costs at a minimum while we raised funds. Built the first prototype, which won TechCrunch Disrupt 2015.

Site Reliability Engineer

JAN 2012 – MAR 2016

Facilitated Admeld's integration into Google's DoubleClick adserver and migrated logs and data from AWS into Google's cloud. Moved to Production Monitoring, home of Monarch, Google's planet-scale time-series monitoring system. Led the datacenter rollout that took Monarch from roughly 25% of Google's datacenter pods to 100% coverage across the company, while keeping legacy monitoring products running through the transition.

Site Reliability Engineer

JAN 2011 – JAN 2012
Admeld · Acquired by Google

Supported the infrastructure behind Admeld's adserving products. Managed the migration to new Chef/CentOS infrastructure, built out new datacenters, and automated server and OS installation and management. Supported many applications across multiple languages, including Java, C, and Ruby, working closely with development teams.

Manager of Cloud Operations

JAN 2009 – APR 2010
AppNexus · Now Xandr · acquired by AT&T

Managed teams in New York, Los Angeles, Moscow, and Belarus running a 1,000+ server infrastructure. Automated virtual server provisioning, BigIP and network configuration; monitoring with Nagios and Cacti.

Manager of Ad Exchange Operations

JAN 2008 – JAN 2009

Managed teams in New York, San Jose, and Bangalore. Grew infrastructure from 5,000 to over 12,000 servers, migrated 7,500 servers from CentOS to RHEL with zero downtime, and delivered 4 new datacenter installations and 10 expansions.

Manager of Unix Systems

MAY 2006 – JAN 2008
Right Media · Acquired by Yahoo!

Managed a team of system administrators and project managers while growing from 100 to over 5,000 servers. Migrated from Gentoo to CentOS with full PXE install automation using CFEngine; datacenter planning and buildout.

1998 – 2006
Operations Manager · WhenU · ~100 servers in two datacenters supporting Perl-based adservers.
2005 – 2006
Datacenter Manager / Senior Unix Admin · BlueFly · ~150 Sun/Linux servers; rebuilt the entire build-and-release process.
2003 – 2005
Senior Unix Systems Administrator · About.com · Multi-tiered clustered network for Sprinks.com (acquired by Google).
2003
Senior Unix Admin / Network Engineer · Concentra Managed Care · Internet-based healthcare apps on Solaris, Linux, and WebLogic.
2002 – 2003
Senior Unix Systems Administrator · SPACE.com · Network infrastructure, Veritas clustering, Java platforms, security auditing.
2000 – 2002
Senior Unix Admin / Developer · Deutsche Bank · Solaris, AIX, and NT fleets for internet and intranet platforms. Learned to love the terminal.
1998 – 2000
Entrepreneurship

Founder & Principal Engineer

Building and operating a portfolio of SaaS and AI-powered applications on multi-tenant Kubernetes across DigitalOcean and AWS: 13+ production systems, including SegmentIQ (privacy-first financial analysis with local LLM redaction + Claude API) and Hale (HIPAA-compliant health AI, Vanta certified).

Co-founder & CTO

Rocko · Early-stage co-founder

Built the infrastructure and application platform for a DeFi lending marketplace simplifying crypto-backed borrowing from protocols like Aave and Compound: the Next.js application framework, contractor teams, and a self-custodial smart-wallet integration across multiple chains.

Co-founder & CTO

Blockrize · Acquired by Gemini

Built the first crypto rewards credit card: initial infrastructure, first working prototype, and demo applications, plus the secure, scalable data layer and the tooling behind crypto pricing models and provider integrations.

Patents & Open Source

Patents

US 11,763,335 B2

Real-time distribution of cryptocurrency rewards for a loyalty program

Granted 2023
US 11,507,971 B2

Cryptocurrency loyalty program based on transactional data

Granted 2022

Open Source & Developer Tools

SyncBeam Like AirDrop, minus the frustration.
pyway Database versioning and migration tool inspired by Flyway.
Mayfly Ephemeral Kubernetes test environments with emulated AWS services.
HoneyOS Network deception and intrusion detection with decoy services.
Twickets Capture-first ticketing for developers.
RLSee Dashboard for inspecting PostgreSQL row-level security policies.
Pingway Per-hop network monitoring that pinpoints where connectivity fails.