← All jobs

Remote

Senior Site Reliability Engineer

About this role

πŸ“‹ Description Lead reliability and operational improvement initiatives across GoDaddy's Commerce platform Own critical production systems, drive incident response and post-incident improvements, and Design, build, and enhance cloud infrastructure, automation, observability, and deployment Partner with engineering, infrastructure, security, and product teams to solve complex technical Use automation, AI-assisted engineering tools, and data-driven insights to reduce operational toil Mentor engineers, share knowledge, and influence engineering practices that improve reliability 🎯 Requirements Significant experience 5 years + operating, troubleshooting, and improving large-scale production Strong expertise in AWS, Linux, container platforms such as Kubernetes, and modern infrastructure Experience building and maintaining Infrastructure as Code, automation solutions, and CI/CD A proven track record of leading or owning production incidents, driving root cause analysis, and Strong software engineering or scripting skills using languages such as Python, Go, TypeScript, or Experience using observability data, monitoring, and operational metrics to identify issues 🎁 Benefits Paid time off Retirement savings (e.g., 401k, pension schemes) Bonus/incentive eligibility Equity grants Employee stock purchase plan Competitive health benefits

Source listing: empllo_remote