Logo for System Automation Corporation

Site Reliability Engineer

Role overview

Qualifications

  • 3+ years of experience in an IT Operations, DevOps, or SRE role
  • Hands-on technical experience with Microsoft Azure in a production environment
  • Experience with infrastructure as code — Terraform and/or Bicep
  • Proficiency in at least one scripting/programming language — Python or TypeScript

Responsibilities

  • Build, operate, and scale production systems in Microsoft Azure to meet availability and performance targets
  • Participate in an on-call rotation; triage, respond to, and resolve production incidents
  • Define and track SLOs/SLIs and error budgets in partnership with engineering and product teams
  • Design and maintain observability and APM tooling (metrics, logs, tracing, dashboards, alerting)

About the company

System Automation Corporation logo

System Automation Corporation

GovTech & Civic Tech

Located in Columbia, Maryland, System Automation Corporation (SA) is one of the nation’s leading providers of regulatory management software and services to government and private-sector organizations. SA exists to automate regulatory compliance and deliver a great customer experience. We believe that empowering our clients to address regulatory challenges is an important part of protecting the general public and making the world a better place. As one of the most seasoned information technology solution providers for state government licensing systems, we are proud to serve over 450+ professions, 4,500+ license types and 19 million licensees nationwide. We provide our clients with state-of-the-art COTS systems that are user-friendly, cost effective and improve public safety.

Company details

Company typeSME
IndustryGovTech & Civic Tech
Company size11 - 50

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Job Type
Full-time
Description

  

Reports to: Manager of Platform Operations and Compliance

Location: 100% Remote (must be eligible to work in the U.S.)

Compensation: Base salary, commensurate with experience. Eligible for the company's commission plan and profit sharing. 


About the Company

System Automation (SA) has built solutions for state and local regulatory agencies for over twenty-five years. Our SaaS platform is trusted by 500 government agencies, spans 5,000 industries, and serves over 20 million licensees.

Our mission is to positively disrupt the regulatory software market by incubating new ideas, collaborating with our customers, and building technology with real-world impact. We challenge the status quo, aren't afraid to experiment, and build products that improve the states and cities we live in.


Position Summary

We're looking for a mid-level Site Reliability Engineer to help build and operate the critical cloud infrastructure behind our platform in Microsoft Azure. You'll define the observability standards (SLOs, SLIs, dashboards, alerting) that tell us whether our systems are healthy, automate away manual toil, and help the team ship changes safely and often through solid CI/CD practices. You'll also share in an on-call rotation, responding to incidents and driving blameless postmortems that make the platform more resilient over time.


This role suits someone with hands-on IT Operations or SRE experience, comfort working inside an agile team, and a genuine cloud-native understanding of how to design for reliability, security, and scale.


Key Responsibilities


Reliability & Operations

• Build, operate, and scale production systems in Microsoft Azure (App Service, Networking, WAF, CosmosDB and related infrastructure) to meet availability and performance targets.

• Participate in an on-call rotation; triage, respond to, and resolve production incidents, and lead or contribute to blameless postmortems.

• Define and track SLOs/SLIs and error budgets in partnership with engineering and product teams.


Observability & Monitoring

• Design and maintain observability and APM tooling (metrics, logs, tracing, dashboards, alerting) so issues are caught before they impact customers.

• Continuously refine alert thresholds and runbooks to reduce noise and mean time to resolution.


Automation & CI/CD

• Reduce operational toil through automation — scripting, self-healing systems, and repeatable processes.

• Build and maintain CI/CD pipelines that let the development team ship safely and frequently.

• Provision and manage infrastructure as code (Bicep) and follow standard change control and version control practices.


Security & Compliance

• Ensure application infrastructure meets security and compliance requirements (e.g., SOC 2, GovRAMP) in partnership with the compliance team.

• Apply security best practices to infrastructure design and change management.


Collaboration & Documentation

• Partner with the agile development team to translate business requirements into reliable technical solutions.

• Participate in technical design sessions and produce clear documentation (diagrams, runbooks, architecture notes).

• Stay current on new Azure capabilities, industry standards, and SRE best practices, and bring recommendations back to the team.

• Other duties as assigned.


Knowledge, Skills, and Abilities

• Solid understanding of networking fundamentals, HTTP/S, and observability principles.

• Ability to evaluate multiple technical approaches and recommend the most effective solution for the context.

• Strong independent problem-solving skills balanced with effective collaboration in a team environment.

• Familiarity with software development lifecycle and programming/coding standards.

• Clear, professional communication, especially under incident pressure.


Qualifications


Required

• 3+ years of experience in an IT Operations, DevOps, or SRE role.

• Hands-on technical experience with Microsoft Azure in a production environment.

• Experience with infrastructure as code — Terraform and/or Bicep.

• Proficiency in at least one scripting/programming language — Python or TypeScript.

• Experience working with REST and/or GraphQL APIs.

• Experience defining and tracking KPIs/SLOs for a web-based application.

• Comfortable participating in an on-call rotation.


Preferred

• Experience with compliance audits (SOC 2 Type 2, GovRAMP).

• Familiarity with security frameworks (NIST, ISO 27001).

• AZ-104 certification, or equivalent Azure networking experience.

• Experience with Node.js.

• Experience with low-code platforms (Power Apps, Logic Apps).

• Familiarity with Scrum/Agile methodology and supporting tools (Confluence, JIRA, Git, Jenkins, Bamboo, TFS).

• Ability to translate business requirements directly into application/site behavior changes.

Requirements

What to Expect 


Hiring Process: Application review ? interviews and technical assessment ? conditional offer 


Security Screening: Due to the sensitive nature of IT systems and data access, final candidates will receive a conditional employment offer contingent upon successful completion of a drug screening and a fingerprint-based background investigation. 


We may use technology-assisted tools, including artificial intelligence tools, to assist recruiters and hiring managers in reviewing application materials and identifying qualifications relevant to a position. These tools support, but do not replace human decision-making. All employment decisions are made by qualified human reviewers. Applicants requiring accommodation during the application or selection process may contact us at hr@systemautomation.com


We are an Equal Opportunity Employer and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, protected veteran status, or any other characteristic protected by applicable law. 

Salary Description
120,000 - 140,000

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Site Reliability Engineer (SRE) Related jobs

Other jobs at System Automation Corporation

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.