Logo for GoDaddy

Senior Site Reliability Engineer

Role overview

Qualifications

  • 6+ years of professional experience in Site Reliability Engineering or Platform Engineering
  • Deep hands-on expertise with Kubernetes and Docker
  • Advanced Linux experience
  • Proficiency in Python for production-grade automation and scripting

Responsibilities

  • Design, implement, and operate scalable, highly available production services
  • Build and maintain alerting pipelines, dashboards, and SLO-driven monitoring strategies
  • Lead incident response end-to-end
  • Develop and extend Infrastructure as Code coverage

About the company

GoDaddy logo

GoDaddy

Cloud Computing & Infrastructure (IaaS/PaaS)

GoDaddy helps the world easily start, confidently grow, and successfully run an online presence. GoDaddy was born to give people an easy, affordable way to get their ideas online. Today, we have millions of customers around the world, but our goal hasn't changed. We’re here to help people easily start, confidently grow and successfully run their own ventures - online and off!

Company details

Company typeLarge
IndustryCloud Computing & Infrastructure (IaaS/PaaS)
Company size5001 - 10000

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Location Details: 

At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.

Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings.  

Join our team
Our Global Sustaining Engineering team sits at the intersection of software engineering and infrastructure, ensuring the services our customers depend on are fast, resilient, and always available. As a Senior Site Reliability Engineer, you'll take direct ownership of production services — from initial design through day-to-day operation — while partnering with product, engineering, and security teams to build and maintain business-critical systems. In this role, you will deepen your technical expertise and grow your leadership presence by mentoring the next generation of SREs. You will also gain hands-on experience with intelligent tooling in real-world workflows.  
What you'll get to do...

  • Design, implement, and operate scalable, highly available production services while diagnosing and resolving complex infrastructure, network, and application issues
  • Build and maintain alerting pipelines, dashboards, and SLO-driven monitoring strategies using Icinga, Prometheus, and Grafana
  • Lead incident response end-to-end — performing root-cause analysis, authoring blameless post-mortems, and driving corrective actions to closure
  • Develop and extend Infrastructure as Code coverage and build internal tooling that eliminates manual, repetitive operational work
  • Mentor SRE I and SRE II engineers through code reviews, debugging sessions, and knowledge-sharing talks
  • Apply LLM-driven log analysis, anomaly detection, and generative AI tools to accelerate incident response and runbook creation — validating all outputs before use

Your experience should include…

  • 6+ years of professional experience in Site Reliability Engineering or Platform Engineering with demonstrated success leading organization-wide reliability programs
  • Deep hands-on expertise with Kubernetes (deployments, operators, custom resources) and Docker in production environments
  • Advanced Linux experience solving problems involving kernel internals, TCP/IP, DNS, and load balancers  
  • Proficiency in Python for production-grade automation and scripting, with working knowledge of Bash
  • Expertise in Ansible and at least one additional Infrastructure as Code tool such as Terraform or Pulumi, with hands-on mastery of Icinga, Prometheus, and Grafana
  • Understanding of large language models, embeddings, and basic machine learning pipelines, with the ability to evaluate and integrate AI-ops tools into daily work

You might also have…

  • Experience building and maintaining continuous integration and continuous delivery pipelines using Jenkins, GitLab CI, or GitHub Actions
  • Demonstrated experience defining and managing Service Level Objectives, Service Level Indicators, and Service Level Agreements across production

We've got your back...  We offer a range of total rewards that may include paid time off, retirement savings (e.g., 401k, pension schemes), bonus/incentive eligibility, equity grants, participation in our employee stock purchase plan, competitive health benefits, and other family-friendly benefits including parental leave. GoDaddy’s benefits vary based on individual role and location and can be reviewed in more detail during the interview process.  

We encourage you to apply even if your experience or skillset doesn’t align perfectly with every requirement. We value a wide range of backgrounds and transferable skills, and we are excited to support learning and growth.

About us...  GoDaddy is empowering everyday entrepreneurs around the world by providing the help and tools to succeed online, making opportunity more inclusive for all. GoDaddy is the place people come to name their idea, build a professional website, attract customers, sell their products and services, and manage their work. Our mission is to give our customers the tools, insights, and people to transform their ideas and personal initiative into success. To learn more about the company, visit About Us. 

At GoDaddy, we know diverse teams build better products—period. Our people and culture reflect and celebrate that sense of diversity and inclusion in ideas, experiences and perspectives. But we also know that’s not enough to build true equity and belonging in our communities. That’s why we prioritize integrating diversity, equity, inclusion and belonging principles into the core of how we work every day—focusing not only on our employee experience, but also our customer experience and operations. It’s the best way to serve our mission of empowering entrepreneurs everywhere, and making opportunity more inclusive for all. To read more about these commitments, as well as our representation and pay equity data, check out our Diversity and Pay Parity annual report which can be found on our Diversity Careers page. 

We also embrace our diverse culture and offer a range of Employee Resource Groups (Culture). Have a side hustle? No problem. We love entrepreneurs! Most importantly, come as you are and make your own way. 

GoDaddy is proud to be an equal opportunity employer. GoDaddy will consider for employment qualified applicants with criminal histories in a manner consistent with local and federal requirements.Refer to our full EEO policy.

Our recruiting team is available to assist you in completing your application. If they could be helpful, please reach out to myrecruiter@godaddy.com. 

GoDaddy doesn’t accept unsolicited resumes from recruiters or employment agencies.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Site Reliability Engineer (SRE) Related jobs

Other jobs at GoDaddy

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.