Senior Site Reliability Engineer
Quick Summary
Location Details: At GoDaddy the future of work looks different for each team.
At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.
The Commerce Site Reliability Engineering team is responsible for the reliability, scalability, and day-to-day operation of the platforms that power GoDaddy's Commerce ecosystem. We build and operate shared infrastructure, support critical production systems, and partner closely with engineering teams to ensure services remain secure, resilient, and highly available.
As a Senior Site Reliability Engineer, you'll join a team that values ownership, operational excellence, and continuous improvement. Engineers are empowered to identify problems, drive meaningful change, and influence how reliability is delivered across the broader Commerce organisation. From improving operational maturity and reducing toil to modernising delivery platforms and strengthening incident response practices, this team plays a key role in enabling engineering teams to move quickly and safely.
You'll work closely with engineers across infrastructure, cloud, security, networking, and application teams while helping shape the future of reliability engineering at GoDaddy. This role offers significant opportunity to broaden your impact, develop technical leadership skills, and grow toward Staff and Principal engineering positions over time.
Responsibilities
~1 min read- →
Lead reliability and operational improvement initiatives across GoDaddy's Commerce platform, helping engineering teams build and operate services safely and at scale.
- →
Own critical production systems, drive incident response and post-incident improvements, and continuously raise the bar for operational excellence.
- →
Design, build, and enhance cloud infrastructure, automation, observability, and deployment platforms that improve reliability, scalability, and developer productivity.
- →
Partner with engineering, infrastructure, security, and product teams to solve complex technical challenges, manage operational risk, and support business-critical services.
- →
Use automation, AI-assisted engineering tools, and data-driven insights to reduce operational toil, improve diagnostics, and accelerate delivery.
- →
Mentor engineers, share knowledge, and influence engineering practices that improve reliability across the broader Commerce organisation.
- →
Contribute to the team's technical direction by identifying opportunities to improve systems, processes, and operational maturity.
-
Significant experience 5 years + operating, troubleshooting, and improving large-scale production systems in cloud-based environments.
-
Strong expertise in AWS, Linux, container platforms such as Kubernetes, and modern infrastructure engineering practices.
-
Experience building and maintaining Infrastructure as Code, automation solutions, and CI/CD pipelines that improve reliability, scalability, and delivery confidence.
-
A proven track record of leading or owning production incidents, driving root cause analysis, and implementing long-term reliability improvements.
-
Strong software engineering or scripting skills using languages such as Python, Go, TypeScript, or similar technologies.
-
Experience using observability data, monitoring, and operational metrics to identify issues, improve system performance, and support data-driven decision making.
-
Demonstrated ability to independently lead complex technical initiatives, manage competing priorities, and deliver outcomes across multiple teams or stakeholders.
-
Experience mentoring engineers, influencing technical decisions, and helping raise operational and engineering standards within a team.
-
A continuous improvement mindset with a focus on automation, reducing operational toil, and leaving systems and processes better than you found them.
-
Experience using AI-assisted engineering tools to improve productivity, accelerate troubleshooting, automate repetitive tasks, and enhance operational workflows.
- Experience with SaltStack, Ansible, Terraform, Pulumi, CloudFormation, or AWS CDK.
-
Experience operating shared or multi-tenant infrastructure services at scale.
- Experience supporting eCommerce, payments, fintech, or other high-availability customer-facing platforms.
- Experience designing observability, SLO, SLI, or operational readiness frameworks.
- Experience contributing to architectural strategy and roadmap planning.
- Experience leveraging AI-assisted tools to improve engineering productivity and operational effectiveness.
We encourage you to apply even if your experience or skillset doesn’t align perfectly with every requirement. We value a wide range of backgrounds and transferable skills, and we are excited to support learning and growth.
Requirements
~1 min readLocation & Eligibility
Listing Details
- Posted
- September 1, 2026
- First seen
- September 1, 2026
- Last seen
- September 1, 2026
Posting Health
- Days active
- 0
- Repost count
- 1
- Trust Level
- 61%
- Scored at
- September 1, 2026
Signal breakdown
GoDaddy helps the world easily start, confidently grow, and successfully run an online presence.
View company profilePlease let GoDaddy know you found this job on Jobera.
3 other jobs at GoDaddy
View all →Explore open roles at GoDaddy.
Similar Devops Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.