Senior Site Reliability Engineer

Epic Kids Inc.

United States · Posted Jul 15


Job description

Responsibilities: Drive the stability, observability, and reliability of the platform by managing GCP infrastructure and container platforms. You will own CI/CD pipelines and the observability stack while partnering with engineering teams to reduce toil and improve system hardening.

Requirements: Requires a Bachelor's degree and 5+ years of experience in infrastructure or DevOps with a proven track record of improving production reliability. Proficiency in GCP, Kubernetes, Terraform, and scripting languages like Python or Bash is essential.

Key skills: Google Cloud Platform, Kubernetes, Docker, Terraform, CI/CD, Observability, Python, Bash, SLO/SLI, Infrastructure as Code, GitHub Actions, ArgoCD, New Relic, Helm, IAM, Network Segmentation

Keywords: GCP, GKE, Kubernetes, Docker, Terraform, CI/CD, GitHub Actions, ArgoCD, Jenkins, New Relic, Python, Bash, SLO, SLI, MTTR, Infrastructure as Code, Dagster, Airflow, PromRelay, SOC 2, FERPA, COPPA, GRC, VPC, IAM, Cloud Monitoring, Helm, Site Reliability Engineering, DevOps, Containerization

Land this job faster with Remote Job Match

Free account: browse thousands of remote roles, no degree needed. Upgrade to tailor your resume to each job with AI and prep for the interview.

Create free account

← Browse more remote jobs