Remote
Site Reliability Engineer
About this role
Working remotely, the full-time Site Reliability Engineer will apply a software engineering mindset to enhance the reliability and performance of production services on Heroku and AWS, while collaborating with the engineering team to build automated solutions that support rapid software deployment. Key responsibilities Share responsibility for the health and performance of production services, proactively identifying and resolving issues using monitoring tools Define and track Service Level Objectives (SLOs) and error budgets, utilizing data to inform reliability discussions Optimize application and database performance by identifying bottlenecks and implementing necessary code or infrastructure improvements Required qualifications 3-5+ years of experience in Site Reliability Engineering, DevOps, or related infrastructure management roles Deep experience managing applications on Heroku and AWS Proven experience with Infrastructure as Code (IaC), specifically Terraform Hands-on experience building and maintaining CI/CD pipelines, preferably with GitHub Actions Strong coding skills with experience in Ruby on Rails, including performance optimization and debugging
Source listing: virtualvocations_main