Logo for Thoughtworks

Infrastructure Support Engineer

Role overview

Qualifications

  • Familiarity with CI/CD tools such as Jenkins, GitlabCI, CircleCI
  • Exposure to log aggregation systems like EFK, Splunk, Datadog
  • Hands-on experience with monitoring tools like Prometheus, Grafana, Datadog
  • Good understanding of at least one Public Cloud such as AWS, Azure, GCP

Responsibilities

  • Monitor operations of products and services and take actions for deviations
  • Document responses to various incident scenarios and prepare runbooks
  • Automate and configure alerts to reduce human effort in operations
  • Assist development teams in incident resolution and conduct root cause analysis

About the company

Thoughtworks logo

Thoughtworks

IT Services & IT Consulting

Thoughtworks is a global technology consultancy that integrates strategy, design and engineering to drive digital innovation. For 28+ years, our clients have trusted our autonomous teams to build solutions that look past the obvious. Here, computer science grads come together with seasoned technologists, self-taught developers, midlife career changers and more to learn from and challenge each other. Career journeys flourish with the strength of our cultivation culture, which has won numerous awards around the world. Join Thoughtworks and thrive. Together, our extra curiosity, innovation, passion and dedication overcomes ordinary.

Company details

Company typeXLarge
IndustryIT Services & IT Consulting
Company size10001

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

As a consultant Infrastructure Support Engineer, your daily responsibilities are integral to ensuring technical excellence and operational efficiency, particularly in cloud environments. Your contribution as a first responder extends to automating day-to-day operations, responding to and escalating production incidents, and assisting development teams in incident resolution.

Job responsibilities

  • You will keep a vigilant eye on the operations of shipped products and services following the agreed upon “Eyes on glass/Follow the sun” engagement models.
  • You will monitor product/service operations against key performance indicators defined by the business and take necessary actions in response to detected deviations.
  • You will document the appropriate responses to various kinds of incident scenarios in collaboration with development teams and prepare runbooks.
  • You will reduce the human effort in day-to-day operations by automating, configuring and tweaking alerts, and monitoring as necessary.
  • You will respond to production incidents and execute well defined responses, raising the incident to higher levels of support wherever necessary.
  • You will assist development teams in incident resolution as necessary, e.g.: as a pair, providing updates, handling communication, etc.
  • You will assist in conducting incident root cause analysis (RCA), preparing incident postmortem reports, communicating incident RCA to client stakeholders whenever necessary and responding to queries and resolution approaches.
  • You will pair on implementing service/product reliability improvement by writing infrastructure/observability configuration code, in collaboration with service reliability engineers .

Job qualifications

Technical Skills

  • You are familiar with CI/CD tools such as Jenkins, GitlabCI, CircleCI, etc.
  • You have had exposure to log aggregation systems, e.g.: EFK, Splunk, Datadog .
  • You have hands-on experience with monitoring, alerting and observability, e.g.: Prometheus, Grafana, Datadog.
  • You possess a good understanding of at least one Public Cloud, e.g.: AWS, Azure, GCP.
  • You have hands-on experience executing most common operations in managing workloads on any container ecosystem tech stacks e.g.: Docker, Kubernetes, Openshift.
  • You have a basic understanding of API concepts such as request, response, headers, authentication, JSON payloads, etc.
  • You have a basic understanding of networking including concepts such as high availability, load balancing and proxies.
  • You have a basic understanding of traffic load management approaches such as horizontal and vertical scaling.
  • You have a basic understanding of availability concepts such as downtime, time to recover/restore, SLAs, etc.
  • You have experience running basic system administration operations in a Linux operating system such as RHEL or Ubuntu.

Professional Skills

  • You have good communication skills and are proficient in English .
  • You can confidently hold a Q&A discussion .
  • You have a good attitude towards learning new technical skills and concepts.
  • You possess innovative thinking and confidence in suggesting ideas to the team .
  • You have strong drive and ownership to sign up and deliver work when called upon without being too concerned with role boundaries.
  • You are willing to be part of a rotation- and need-based 24x7 team.

Other things to know

Learning & Development

There is no one-size-fits-all career path at Thoughtworks: however you want to develop your career is entirely up to you. But we also balance autonomy with the strength of our cultivation culture. This means your career is supported by interactive tools, numerous development programs and teammates who want to help you grow. We see value in helping each other be our best and that extends to empowering our employees in their career journeys.

AI for recruitment

At Thoughtworks, we use AI tools to support our recruitment team with administrative tasks such as drafting communications, scheduling interviews and writing job descriptions.

Crucially, our AI tools do not screen, assess, rank or make hiring decisions. Every application is reviewed by our team and all selection decisions are made exclusively by our interviewers and hiring managers.

We are committed to fairness and responsible AI. We actively manage our AI systems by testing, monitoring for biased outcomes and implementing mitigation measures. We hold our third-party vendors to these same high standards through a rigorous governance process. For additional information, please see our full Thoughtworks AI Policy for Recruitment.

About Thoughtworks

Thoughtworks is a dynamic and inclusive community of bright and supportive colleagues who are revolutionizing tech. As a leading technology consultancy, we’re pushing boundaries through our purposeful and impactful work. For 30+ years, we’ve delivered extraordinary impact together with our clients by helping them solve complex business problems with technology as the differentiator. Bring your brilliant expertise and commitment for continuous learning to Thoughtworks. Together, let’s be extraordinary.

#LI-Remote

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Infrastructure Engineer Related jobs

Other jobs at Thoughtworks

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.