Join us as we work to create a thriving ecosystem that delivers accessible, high-quality, and sustainable healthcare for all.
We are looking for a Senior Site Reliability Engineer to join our Service Operations, Site Reliability Engineering team within the Cloud Infrastructure Engineering division. This team is newly formed and is responsible for managing the fleet of systems owned by its sister teams in the Service Operations zone. We’re looking for Site Reliability & Infrastructure Engineering experts to help us build out the tools and processes necessary for success. Come help us create a thriving ecosystem that delivers accessible, high-quality and sustainable healthcare for all!
The Team:
The Service Operations Site Reliability Engineering team is a newly formed branch of the Network Operations Center. The team sits within the Cloud Infrastructure Engineering (CIE) division which is responsible for delivering high quality and highly available SaaS infrastructure and internal tools. You will be working closely with our internal stakeholders from the R&D and Cloud Infrastructure organizations using modern toolsets to deliver resilient & scalable business solutions. Consistent with SRE practices, a key goal for this team is to measure & reduce toil, not only for ourselves but for the broader NOC & Service Operations organizations.
Job Responsibilities
· Provisioning and ongoing management of physical & virtual Linux machines using tools like Puppet, Ansible, and Terraform, to name a few
· Engage closely with sister teams to assume ownership of various system lifecycle tasks
· Automate away toil and/or create empowerment processes for transitioning high urgency work to the NOC’s rapid response team
· Build automated monitoring & observability using tools such as Prometheus/AlertManager, iCinga, Grafana, etc.
· Participate in all Agile/scrum ceremonies including daily stand-ups, sprint planning, backlog grooming, etc.
· Participate in the team’s on-call rotation (expected to begin late 2024, early 2025)
· Work closely with internal teams to integrate new monitoring & alerts into the NOC using Perl scripting to author custom parsing & mapping rules
· Develop metrics and observability dashboards which can be used to measure and track various success measures for the team & the business
Typical Qualifications
· 5+ years of professional experience delivering SaaS solutions, preferably in a hybrid cloud environment
· Bachelor’s or Master’s degree in a Computer Science / Engineering program
· Proven experience using query languages to deliver observability solutions
· Proficiency working with one or more configuration management tools (Puppet, Chef, Ansible, etc.)
· Admin-level expertise with a Unix-based operating system
· Proven ops background using cloud-native best practices
· Proven proficiency with one or more scripting languages (Python, Ruby, Perl, Java, etc.)
· Proficiency working with Git & Atlassian suite or similar
· Proficiency working with containerized environments is a plus
· Experience creating technical documentation & standard operating procedures (SOPs)
About athenahealth
Our vision: In an industry that becomes more complex by the day, we stand for simplicity. We offer IT solutions and expert services that eliminate the daily hurdles preventing healthcare providers from focusing entirely on their patients — powered by our vision to create a thriving ecosystem that delivers accessible, high-quality, and sustainable healthcare for all.
Our company culture: Our talented employees — or athenistas, as we call ourselves — spark the innovation and passion needed to accomplish our vision. We are a diverse group of dreamers and do-ers with unique knowledge, expertise, backgrounds, and perspectives. We unite as mission-driven problem-solvers with a deep desire to achieve our vision and make our time here count. Our award-winning culture is built around shared values of inclusiveness, accountability, and support.
Our DEI commitment: Our vision of accessible, high-quality, and sustainable healthcare for all requires addressing the inequities that stand in the way. That's one reason we prioritize diversity, equity, and inclusion in every aspect of our business, from attracting and sustaining a diverse workforce to maintaining an inclusive environment for athenistas, our partners, customers and the communities where we work and serve.
What we can do for you:
Along with health and financial benefits, athenistas enjoy perks specific to each location, including commuter support, employee assistance programs, tuition assistance, employee resource groups, and collaborative workspaces — some offices even welcome dogs.
We also encourage a better work-life balance for athenistas with our flexibility. While we know in-office collaboration is critical to our vision, we recognize that not all work needs to be done within an office environment, full-time. With consistent communication and digital collaboration tools, athenahealth enables employees to find a balance that feels fulfilling and productive for each individual situation.
In addition to our traditional benefits and perks, we sponsor events throughout the year, including book clubs, external speakers, and hackathons. We provide athenistas with a company culture based on learning, the support of an engaged team, and an inclusive environment where all employees are valued.
Learn more about our culture and benefits here: athenahealth.com/careers
Infraspeak
ADMA Biologics, Inc.
Air Liquide Healthcare
SugarCRM
Air Liquide