Staff Site Reliability Engineer
Staff Site Reliability Engineer role on the Platform team focused on building and operating reliable AWS-based systems. The role supports platform capabilities, developer experience, and cross-team reliability for product engineers.
What you'll do
- Define infrastructure reliability and platform capabilities
- Design and maintain scalable AWS infrastructure as code
- Build platform tools and services for product engineers
- Lead incident response and improve observability and recovery
- Mentor engineers and participate in on-call rotation
What they're looking for
- Strong AWS experience with highly available systems
- Terraform and configuration management experience
- Experience with Docker, Kubernetes, ECS and Linux
- Ability to code in Python, Ruby or Go and shell scripting
- Experience with MySQL, PostgreSQL, Redis or DynamoDB
Skills
Summary written by StartupJobs from the company's listing. The full listing is on the company's website.
This job comes from the career page of Lightspeed. We link straight to the source.