Remote
Site Reliability Engineer
About this role
Focused on enhancing platform reliability and performance, the full-time remote Site Reliability Engineer will manage production services on Heroku and AWS, optimize application performance, and collaborate with engineering teams to drive improvements and maintain security best practices. Key responsibilities Share responsibility for the health and performance of production services, proactively identifying and resolving issues using monitoring tools Define and track Service Level Objectives (SLOs) and optimize application and database performance by identifying bottlenecks and implementing improvements Champion the developer experience by enhancing CI/CD pipelines and collaborating on infrastructure management using Terraform Required qualifications 3-5+ years of experience in Site Reliability, DevOps, or related roles managing infrastructure Deep experience with applications on Heroku and AWS Proven expertise in Infrastructure as Code (IaC) using Terraform Strong coding skills with experience in Ruby on Rails and performance optimization Familiarity with observability and monitoring tools such as New Relic or Sentry
Source listing: virtualvocations_main