Logo for ICONMA

Site Reliability Engineer

Role overview

Qualifications

  • BS in Computer Science, Software Engineering, Information Technology, or related field preferred; or equivalent professional experience.
  • 5+ years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering roles.
  • Hands-on experience operating and supporting Kubernetes platforms in production environments.
  • Strong experience managing Kubernetes clusters in both on-premises and cloud-based environments.

Responsibilities

  • Maintain and enhance Kubernetes platforms across on-premises and cloud environments, ensuring reliability, scalability, and operational efficiency.
  • Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher.
  • Provide deep technical expertise in Linux-based systems, including performance tuning, troubleshooting, automation, and operational support.
  • Develop and maintain infrastructure-as-code solutions to standardize and automate platform deployment and management.

About the company

ICONMA logo

ICONMA

Management Consulting

We provide Professional Staffing Services & Project-Based Solutions for a broad range of Fortune 500 organizations. ICONMA is a certified Woman-Owned staffing company and was founded in 2000. ICONMA’s corporate headquarters is in Troy, Michigan, and has 15+ locations worldwide. What makes ICONMA stand out in a fiercely competitive industry? *We provide integrated, full lifecycle services across a broad range of business and technical platforms. *No single company can duplicate our full range of staffing and permanent recruiting services nationwide. *Proven track record of attracting and retaining exceedingly skilled professional workers in a highly competitive market. SERVICES OFFERED Staff Augmentation (Contract, Contract to Hire, Direct Hire, Single Source) Data Analysis Project-Based Services & Solutions Hire Train Deploy Service Model Offshore Staff Augmentation Payroll Services AREAS OF EXPERTISE - Information Technology - Engineering - Business Professional - Accounting/Finance - Admin/Clerical/Call Center - Healthcare/Clinical/Scientific - Marketing/Creative mail linkedin@iconma.com Phone (888) 451-2519 Website http://www.iconma.com

Company details

Company typeLarge
IndustryManagement Consulting
Company size1001 - 5000

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Our Client, a Business Maunufacturing and Supply company, is looking for a Site Reliability Engineer for their Remote location.
 
Responsibilities:
  • Platform Operations: Maintain and enhance Kubernetes platforms across on-premises and cloud environments, ensuring reliability, scalability, and operational efficiency.
  • Cluster Management: Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher.
  • Linux Systems Administration: Provide deep technical expertise in Linux-based systems, including performance tuning, troubleshooting, automation, and operational support.
  • Infrastructure as Code: Develop and maintain infrastructure-as-code solutions to standardize and automate platform deployment and management, with a preference for Cluster API (CAPI)-based approaches.
  • GitOps and Deployment Automation: Support and improve GitOps workflows using ArgoCD to manage cluster and application configuration in a consistent, auditable manner.
  • Collaboration: Work closely with developers, scientists, and infrastructure teams to deliver reliable platform services and translate operational needs into sustainable engineering solutions.
  • Continuous Improvement: Identify opportunities to improve platform resilience, observability, security, and maintainability through automation and modern SRE practices.
 
Requirements:
  • BS in Computer Science, Software Engineering, Information Technology, or related field preferred; or equivalent professional experience.
  • 5+ years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering roles.
  • Hands-on experience operating and supporting Kubernetes platforms in production environments.
  • Strong experience managing Kubernetes clusters in both on-premises and cloud-based environments.
  • Strong Linux systems administration skills, including troubleshooting, scripting, networking, and system performance analysis.
  • Experience with Rancher for Kubernetes cluster management and platform operations.
  • Experience implementing infrastructure-as-code solutions for platform provisioning and lifecycle management.
  • Demonstrated success working in Agile teams (Scrum, Kanban).
  • Cluster operations, upgrades, networking, storage, troubleshooting, and workload support.
  • Platform Management: Rancher or similar Kubernetes management platforms.
  • Linux: Advanced administration of Linux/Unix systems.
  • Infrastructure as Code: Strong IaC experience; Cluster API (CAPI) preferred.
  • GitOps/CI-CD: ArgoCD, Git version control, and deployment automation practices.
  • Scripting/Automation: Bash, Python, or similar scripting languages for automation and operational tooling.
  • Experience with hybrid infrastructure spanning on-premises and public cloud platforms (AWS, Azure, GCP).
  • Experience with Kubernetes ecosystem tooling for observability, logging, monitoring, and alerting.
  • Familiarity with security best practices for Kubernetes and Linux platforms.
  • Experience supporting scientific research environments, high-performance computing, or computational science workflows.
  • Knowledge of CI/CD pipeline development and platform automation patterns.
 
Why Should You Apply?

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Site Reliability Engineer (SRE) Related jobs

Other jobs at ICONMA

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.