Logo for Gruve

L3 Server Infrastructure Lead (Unix/Linux)

Role overview

Qualifications

  • 8–12+ years of experience in enterprise Unix/Linux infrastructure and production operations.
  • Strong hands-on experience in Unix/Linux system administration.
  • Expert knowledge of one or more of: RHEL, Ubuntu, AIX, Solaris.
  • Strong troubleshooting skills across OS, filesystem, process, performance and system services.

Responsibilities

  • Provide L3 technical support for P1–P4 incidents and complex OS-level issues.
  • Lead and technically guide L1/L2 Unix/Linux operations teams across global regions.
  • Administer and troubleshoot RHEL, Ubuntu, AIX and Solaris environments.
  • Own OS patching, patch compliance and vulnerability remediation.

Key facts

Hard skills

Other skills

  • Troubleshooting (Problem Solving)
  • Communication
  • Collaboration

About the company

Gruve logo

Gruve

Artificial Intelligence & Machine Learning Services

Gruve is an AI-native enterprise services company helping organizations build future-proof AI stacks with measurable outcomes across security, infrastructure and data.

Company details

IndustryArtificial Intelligence & Machine Learning Services
Company size501 - 1000

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

About Gruve

Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, and advanced technologies such as Large Language Models (LLMs). Our mission is to assist our customers in their business strategies utilizing their data to make more intelligent decisions. As a well-funded early-stage startup, Gruve offers a dynamic environment with strong customer and partner networks.

About the Role

We are looking for an experienced Unix/Linux Compute Subject Matter Expert to provide L3 technical leadership and operational ownership for a global estate of Unix/Linux server instances. The SME will be accountable for infrastructure stability, availability, security and performance, and will own patching, vulnerability remediation, lifecycle management and continuous improvement across the environment. Working within a follow-the-sun model spanning APAC, EMEA and NLAM, the role guides L1/L2 teams and partners closely with Security, Database, Storage, Network, Cloud and Application teams to deliver a consistent, high-quality service.

Key Roles and Responsibilities

  • Provide L3 technical support for P1–P4 incidents and complex OS-level issues.
  • Lead and technically guide L1/L2 Unix/Linux operations teams across global regions.
  • Administer and troubleshoot RHEL, Ubuntu, AIX and Solaris environments.
  • Own OS patching, patch compliance and vulnerability remediation.
  • Perform CIS hardening and security baseline remediation.
  • Manage OS performance and capacity, including filesystem, CPU, memory and I/O issues.
  • Support server build, configuration, lifecycle and decommissioning activities.
  • Drive problem management, root cause analysis (RCA) and permanent corrective actions.
  • Support monitoring, disaster recovery (DR) activities and operational readiness.
  • Develop and maintain operational runbooks, procedures and technical documentation.
  • Identify and implement automation opportunities to improve operational efficiency.
  • Collaborate with Security, Database, Storage/Backup, Network, Cloud and Application teams.
  • Ensure consistent service delivery through the APAC–EMEA–NLAM follow-the-sun operating model.

Basic Qualifications

  • 8–12+ years of experience in enterprise Unix/Linux infrastructure and production operations.
  • Strong hands-on experience in Unix/Linux system administration.
  • Expert knowledge of one or more of: RHEL, Ubuntu, AIX, Solaris.
  • Strong troubleshooting skills across OS, filesystem, process, performance and system services.
  • Experience with OS patching, vulnerability remediation and CIS hardening.
  • Strong knowledge of LVM, NFS, SSH, DNS, filesystem management, package management and cron.
  • Experience in enterprise 24x7 production operations.
  • Strong ITIL-based Incident, Problem and Change Management experience.
  • Experience leading and mentoring L1/L2 teams and handling major incidents.
  • Strong scripting and automation skills using Shell, Python, Ansible or equivalent.

Preferred Qualifications

  • Red Hat RHCSA / RHCE certification.
  • Experience with enterprise monitoring and patch-management tools.
  • ServiceNow or other ITSM platform experience.
  • Knowledge of VMware, AWS/Azure, and Backup/DR solutions.
  • Experience working in a global follow-the-sun support model.

Why Gruve

At Gruve, we foster a culture of innovation, collaboration, and continuous learning. We are committed to building a diverse and inclusive workplace where everyone can thrive and contribute their best work. If you’re passionate about technology and eager to make an impact, we’d love to hear from you.

Gruve is an equal opportunity employer. We welcome applicants from all backgrounds and thank all who apply; however, only those selected for an interview will be contacted.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Related jobs

Other jobs at Gruve

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.