Back to jobs

Staff Site Reliability Engineer

  • Hybrid
  • Auckland
  • FullTime

Staff Site Reliability Engineer role on the Platform team focused on building and operating reliable AWS-based systems. The role supports platform capabilities, developer experience, and cross-team reliability for product engineers.

What you'll do

  • Define infrastructure reliability and platform capabilities
  • Design and maintain scalable AWS infrastructure as code
  • Build platform tools and services for product engineers
  • Lead incident response and improve observability and recovery
  • Mentor engineers and participate in on-call rotation

What they're looking for

  • Strong AWS experience with highly available systems
  • Terraform and configuration management experience
  • Experience with Docker, Kubernetes, ECS and Linux
  • Ability to code in Python, Ruby or Go and shell scripting
  • Experience with MySQL, PostgreSQL, Redis or DynamoDB

Skills

  • AWS
  • Terraform
  • Docker
  • Kubernetes
  • ECS
  • Linux
  • Python
  • Ruby
  • Go
  • Shell scripting
  • MySQL
  • PostgreSQL
  • Redis
  • DynamoDB
  • Observability
  • Incident management

Summary written by StartupJobs from the company's listing. The full listing is on the company's website.

This job comes from the career page of Lightspeed. We link straight to the source.

Apply on the company website