Oh My JobFind Jobs
Company
  • About Us
  • Blog
  • Contact
Tools
  • Paycheck Calculator
  • US Job Market Data
Legal
  • Terms of Service
  • Privacy Policy
  • California Privacy Rights
For Employers
Post a Job • Sponsored
Illustration - Senior Site Reliability Engineer- Observability

Senior Site Reliability Engineer- Observability

Okta
Okta
Bengaluru, India
Jul 1, 2026
Salary not listed

Job Description

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

  • Position Overview
    We are seeking a highly technical Senior Site Reliability Engineer (P3) to help build, run, and scale Okta’s enterprise-grade multi-cloud observability ecosystem. In this role, you will be a core engine driving our full-stack telemetry infrastructure—ensuring that massive streams of Metrics, Logs, Traces, and Alerts are processed efficiently, securely, and cost-effectively.
    As a Senior SRE, you will treat Monitoring as Code (MaC). You will utilize automation frameworks like Terraform and write robust code (Go/Python) to build self-healing pipelines, optimize backend telemetry engines (Splunk and Grafana/Mimir/Loki), and eliminate manual operations.

    About Team: Workforce Identity Cloud
    Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities—like reducing costs, and doing more for your customers.
    If you like to be challenged and have a passion for solving large-scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate is someone who exemplifies the ethics of, “If you have to do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.

    Key Responsibilities
    • Full-Stack Telemetry Operations: Own and optimize the end-to-end collection, processing, and visualization pipelines for Metrics, Logs, Traces, and Alerts across highly distributed multi-cloud (AWS/GCP) environments.
    • Splunk & Grafana Optimization: Act as a hands-on expert in optimizing log pipelines. Drive indexer performance, tune search efficiency (SPL), and clean up heavy dashboard queries to reduce latency and infrastructure footprint (FinOps).
    • Monitoring as Code (MaC): Standardize, deploy, and maintain core observability tools, agent relays, and collectors natively using Terraform and automated CI/CD pipelines.
    • Distributed Tracing & Metrics: Implement and scale OpenTelemetry (OTel) standards, Prometheus/Mimir, and tracing frameworks to map end-to-end request flows across core microservices (such as our Project Harmony initiative).
    • Alert & Dashboard

Okta on Oh My Job

525 open positions right now, including 83 in California. Average salary across all roles: $260–$374.

Apply now on Adzuna
Share: