Senior Site Reliability Engineer
Epic Kids Inc.
United States · Posted Jul 15
Job description
Responsibilities: Drive the stability, observability, and reliability of the platform by managing GCP infrastructure and container platforms. You will own CI/CD pipelines and the observability stack while partnering with engineering teams to reduce toil and improve system hardening.
Requirements: Requires a Bachelor's degree and 5+ years of experience in infrastructure or DevOps with a proven track record of improving production reliability. Proficiency in GCP, Kubernetes, Terraform, and scripting languages like Python or Bash is essential.
Key skills: Google Cloud Platform, Kubernetes, Docker, Terraform, CI/CD, Observability, Python, Bash, SLO/SLI, Infrastructure as Code, GitHub Actions, ArgoCD, New Relic, Helm, IAM, Network Segmentation
Keywords: GCP, GKE, Kubernetes, Docker, Terraform, CI/CD, GitHub Actions, ArgoCD, Jenkins, New Relic, Python, Bash, SLO, SLI, MTTR, Infrastructure as Code, Dagster, Airflow, PromRelay, SOC 2, FERPA, COPPA, GRC, VPC, IAM, Cloud Monitoring, Helm, Site Reliability Engineering, DevOps, Containerization