Logo for Exavalu

Senior Observability Engineer / Platform Engineer

Role overview

Qualifications

  • 6+ years of experience in Platform Engineering, SRE, DevOps, Cloud Operations, or Observability Engineering
  • Strong expertise in Prometheus, Grafana, Splunk, OpenSearch, Elastic
  • Good understanding of Kubernetes, Docker, and containerized workloads
  • Experience with AWS, Azure, or GCP environments

Responsibilities

  • Design, implement, and maintain enterprise observability platforms covering metrics, logs, traces, and events
  • Build observability solutions using tools such as Prometheus, Grafana, and Datadog
  • Develop dashboards, SLOs, SLIs, alerting rules, and service health monitoring frameworks
  • Collaborate with engineering, platform, SRE, and operations teams to improve reliability, performance, and availability

About the company

Exavalu logo

Exavalu

Digital Transformation Consulting

Exavalu is your sustained strategic partner to deliver meaningful change and lasting value. Our seasoned industry veterans are experienced in solving your most challenging problems. We stay with you until the results are achieved. Big firm expertise. Small firm feel. Get the transformation you wish for. Visit www.exavalu.com

Company details

Company typeScaleup
IndustryDigital Transformation Consulting
Company size201 - 500

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

This is a remote position.

Senior Observability Engineer / Platform Engineer (Offshore)

Location Offshore (India) Experience 6-10 Years

Role Summary

We are looking for a hands-on Observability Engineer with strong experience in cloud-native platforms, monitoring, logging, alerting, and operational excellence. The ideal candidate will have experience building and managing enterprise observability solutions across Kubernetes and public cloud environments, enabling proactive monitoring, incident response, and reliability engineering practices.

Key Responsibilities

Design, implement, and maintain enterprise observability platforms covering metrics, logs, traces, and events. Build observability solutions using tools such as Prometheus, Grafana, OpenSearch, Splunk, Elastic, Datadog, Dynatrace, New Relic, or equivalent platforms. Develop dashboards, SLOs, SLIs, alerting rules, and service health monitoring frameworks. Integrate monitoring and observability capabilities within Kubernetes and cloud-native environments. Enable incident management, root cause analysis, and operational troubleshooting through observability best practices. Automate monitoring configuration and platform onboarding using Infrastructure-as-Code and CI/CD pipelines. Collaborate with engineering, platform, SRE, and operations teams to improve reliability, performance, and availability. Support observability maturity initiatives including distributed tracing, AIOps, and intelligent alerting.



Requirements

Required Skills

6+ years of experience in Platform Engineering, SRE, DevOps, Cloud Operations, or Observability Engineering.

Strong expertise in: Prometheus Grafana Splunk / OpenSearch / Elastic Distributed tracing solutions (Jaeger, Tempo, OpenTelemetry, etc.) Good understanding of Kubernetes, Docker, and containerized workloads.

Experience with AWS, Azure, or GCP environments. Knowledge of incident management, alert tuning, and troubleshooting production environments. Experience with scripting and automation using Python, Shell, or similar languages.

Familiarity with CI/CD and Infrastructure-as-Code tools such as Terraform, Jenkins, GitHub Actions, or ArgoCD. Preferred Skills Exposure to SRE practices, error budgets, SLO/SLI frameworks.

Experience with AIOps, automated remediation, or intelligent incident response. Knowledge of OpenTelemetry implementation.

Working experience in enterprise-scale production platforms. Exposure to security and governance considerations in cloud-native environments.



Benefits

Diversity Inclusion:

At Exavalu, we are committed to building a diverse and inclusive workforce. We welcome applications for employment from all qualified candidates, regardless of race, color, gender, national or ethnic origin, age, disability, religion, sexual orientation, gender identity or any other status protected by applicable law. We nurture a culture that embraces all individuals and promotes diverse perspectives, where you can make an impact and grow your career.

Exavalu also promotes flexibility  depending on the needs of employees, customers and the business. It might be part-time work, working outside normal 9-5 business hours or working remotely. We also have a welcome back program to help people get back to the mainstream after a long break due to health or family reasons.



Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Platform Engineer Related jobs

Other jobs at Exavalu

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.