Job Description
Job Role: Senior Site Reliability Engineer Kubernetes Platform
Job Type: Full Time
Job Location: San Jose, CA (Remote)
Salary Range: $100000 to $130000/Annum + Full Time Benefits
Must Have Technical/Functional Skills:
• 10+ years of experience in SRE, DevOps, or infrastructure engineering
• Strong experience running Kubernetes in production (EKS, AKS, GKE, or upstream)
• Hands-on experience working in FedRAMP High and/or DoD IL5 environments
• Solid understanding of cloud infrastructure, Linux systems, and networking fundamentals
• Experience with Infrastructure as Code (Terraform preferred)
• Familiarity with CI/CD systems (GitHub Actions, GitLab CI, Jenkins, ArgoCD)
• Proficiency in scripting or programming (Python, Go)
• Experience building or operating observability platforms (Prometheus, Grafana, OpenTelemetry, ELK)
• Working knowledge of compliance frameworks (e.g., NIST 800-53, STIGs, RMF)
Roles & Responsibilities:
• Design, build, and operate production-grade Kubernetes platforms in regulated environments
• Improve system reliability through automation, thoughtful design, and continuous iteration
• Define and drive SLOs, SLIs, and error budgets to guide reliability decisions
• Build and evolve CI/CD pipelines that are secure, scalable, and easy to use
• Implement robust observability (metrics, logs, traces) to make systems understandable and actionable
• Reduce operational toil by automating repetitive processes and improving workflows
• Partner with security and compliance teams to meet FedRAMP High and IL5 requirements without sacrificing developer velocity
• Support ATO processes, including documentation, controls implementation, and audit readiness Confidential
• Participate in on-call rotations supporting customer requests and paging alerts
• Participate in incident response, blameless postmortems, and continuous improvement efforts
• Help shape a platform that engineers enjoy using
Nice to Have
• Experience with service mesh technologies (Istio, Linkerd)
• Familiarity with policy-as-code (OPA/Gatekeeper, Kyverno)
• Experience with GitOps workflows
• Exposure to multi-cluster or hybrid cloud architectures
• Knowledge of FIPS-compliant systems or DoD Cloud SRG
• Relevant certifications (CKA, CKS, cloud provider certs, Security+).
Disclaimer: Diverse Lynx LLC is an Equal Opportunity Employer. All applicants and employees are evaluated without discrimination, based solely on their qualifications, ability, competence and performance. This email and its attachments may contain confidential or proprietary information and is intended only for the recipient(s). If you received this message in error, please disregard it and notify the sender. If you no longer wish to receive our communications, you may
unsubscribe at any time.
Security Notice: Our official website is www.diverselynx.com We do not operate or authorize any other websites representing Diverse Lynx LLC.