Logo for Integrated DNA Technologies

Staff Site Reliability Engineer

Role overview

Qualifications

  • 5+ years of hands-on experience in Site Reliability Engineering, DevOps, or equivalent role
  • Strong understanding and practical application of SRE principles
  • Expertise designing, implementing, and managing observability platforms for cloud-native environments
  • Extensive hands-on experience with at least one major cloud platform (AWS, Azure, or GCP)

Responsibilities

  • Establish and monitor Service Level Objectives and error budgets for critical services
  • Design, implement, and maintain monitoring, logging, and distributed tracing for cloud-native applications
  • Identify repetitive operational work and automate it away
  • Lead incident response and participate across the full incident lifecycle

Key facts

Hard skills

Other skills

  • Collaboration
  • Troubleshooting (Problem Solving)
  • Decision Making

About the company

Integrated DNA Technologies logo

Integrated DNA Technologies

Biotechnology

Integrated DNA Technologies, Inc. (IDT) develops, manufactures, and markets nucleic acid products for the life sciences industry in the areas of academic and commercial research, agriculture, medical diagnostics, and pharmaceutical development. IDT has developed proprietary technologies for genomic applications such as next generation sequencing, CRISPR genome editing, synthetic biology, digital PCR, and RNA interference. Through its GMP services, IDT manufactures products used by scientists researching many forms of cancer and most inherited and infectious diseases. IDT is widely recognized as an industry leader in custom nucleic acid manufacture, serving over 130,000 life sciences researchers. IDT was founded in 1987 in the United States and has its U.S. manufacturing headquarters in Coralville, Iowa, with additional U.S. manufacturing sites in San Diego, California; Research Triangle Park, North Carolina; and Ann Arbor, Michigan; with international sites in Leuven, Belgium and Singapore. *For research use only. Not for use in diagnostic procedures. Unless otherwise agreed to in writing, IDT does not intend these products to be used in clinical applications and does not warrant their fitness or suitability for any clinical diagnostic use. Purchaser is solely responsible for all decisions regarding the use of these products and any associated regulatory or legal obligations

Company details

Company typeLarge
IndustryBiotechnology
Company size1001-5000

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Bring more to life. 

At Danaher, our work saves lives. And each of us plays a part. Fueled by our culture of continuous improvement, we turn ideas into impact – innovating at the speed of life.  

Our 60,000+ associates work across the globe at more than 15 unique businesses within life sciences, diagnostics, and biotechnology.   

Are you ready to accelerate your potential and make a real difference? At Danaher, you can build an incredible career at a leading science and technology company, where we’re committed to hiring and developing from within. You’ll thrive in a culture of belonging where you and your unique viewpoint matter.  

Learn about the Danaher Business System which makes everything possible. 

The Staff Site Reliability Engineer is responsible for the availability, performance, and operational maturity of our cloud-native platform and the critical data and AI/ML services that run on it. This is a senior individual contributor role: you will set reliability direction across teams, raise the engineering bar through standards and mentorship, and do the hands-on work of making complex distributed systems predictable. 

This position reports to the Senior Director, Data and AI Platform and is part of the Chief Information Officer (CIO) Office and will be located onsite in Krakow, Poland.

This is a Danaher Corporate role, hosted by our Cytiva operating company in Kraków. 

In this role, you will have the opportunity to: 

  • Champion SRE practice at scale. Establish and monitor Service Level Objectives and error budgets for critical services, and use them to drive real decisions about reliability, availability, performance, and cost-efficiency — including when to slow down and when to ship. 

  • Own observability end to end. Design, implement, and maintain monitoring, logging, and distributed tracing for cloud-native applications and infrastructure (Kubernetes, microservices), and build the dashboards, alerts, and runbooks that give teams deep insight into system health rather than alert noise. 

  • Eliminate toil. Identify repetitive operational work and manual process across the production estate and automate it away, developing the tools, scripts, and pipeline improvements that make operations, deployment, and incident response faster and safer. 

  • Lead incident response and learning. Participate across the full incident lifecycle — detection, triage, mitigation, resolution — and lead thorough blameless postmortems that get to root cause and produce preventative measures that stick. 

  • Shape systems before they're built. Partner with development teams to influence the design of new services so operability, reliability, and cost-efficiency are engineered in from the start, and proactively surface performance bottlenecks and architectural weaknesses before they reach production. 

  • Set technical direction and grow the team. Drive cross-team architecture and reliability decisions, establish standards and documentation, mentor engineers across our operating companies, and help foster a culture of technical rigor, blameless learning, and collaboration. 

The essential requirements of the job include: 

  • 5+ years of hands-on experience in a Site Reliability Engineering, DevOps, or equivalent role focused on production system reliability and operations; CS/Engineering degree or equivalent practical experience. 

  • Strong understanding and practical application of SRE principles — SLOs, error budgets, toil reduction, and blameless culture — with proven experience participating in and improving incident management processes for business-critical systems. 

  • Expertise designing, implementing, and managing observability platforms for cloud-native environments (e.g., Prometheus, Grafana, Datadog, ELK stack, OpenTelemetry, Splunk), including the dashboards, alerting, and runbooks that make them actionable. 

  • Extensive hands-on experience with at least one major cloud platform (AWS, Azure, or GCP) across compute, networking, and database services; containerization and orchestration (Docker, Kubernetes); Infrastructure as Code (e.g., Terraform, OpenTofu, Pulumi); and proficiency in at least one language (Python, Go) for automation and tool development. 

  • Proven ability to set technical direction at platform scale — driving cross-team architecture and design decisions, establishing standards, and mentoring engineers — grounded in strong troubleshooting skills across complex distributed systems, including microservices, CI/CD pipelines, and large-scale data infrastructure. 

Travel, Motor Vehicle Record & Physical/Environment Requirements:  

  • Ability to travel – up to 10% 

It would be a plus if you also possess previous experience in: 

  • Life sciences, diagnostics, or biotechnology (e.g., partnering with R&D, quality, clinical, manufacturing, or commercial teams) 

  • Reliability and efficiency practices at scale — chaos engineering and resilience testing, capacity planning, or cloud cost optimization (FinOps). 

  • Working in a matrixed environment 

Danaher offers a broad array of comprehensive, competitive benefit programs that add value to our lives. Whether it’s a health care program or paid time off, our programs contribute to life beyond the job. Check out our benefits at Danaher Benefits Info

At Danaher, we believe in designing a better, more sustainable workforce. We recognize the benefits of flexible, remote working arrangements for eligible roles and are committed to providing enriching careers, no matter the work arrangement. This position is eligible for a remote work arrangement in which you can work remotely from your home. Additional information about this remote work arrangement will be provided by your interview team. Explore the flexibility and challenge that working for Danaher can provide.

#LI-KK1

Join our winning team today. Together, we’ll accelerate the real-life impact of tomorrow’s science and technology. We partner with customers across the globe to help them solve their most complex challenges, architecting solutions that bring the power of science to life.

For more information, visit www.danaher.com.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Site Reliability Engineer Related jobs

Other jobs at Integrated DNA Technologies

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.