Senior Site Reliability Engineer

Ping Identity · Bengaluru, Karnataka, India · Engineering

Posted 2026-07-30

Apply for this role →

As a Ping Identity Senior SRE, you will be involved in every facet of our On-Demand SaaS services and will build, deploy, and maintain the infrastructure of one of the largest identity platforms in the world. We follow a DevOps model: our teams are integrated with development teams, and running continuous deployments daily, and SREs are expected to provide input in the product's design, development, deployment, and operations.

Working within the Cloud Operations team, you'll build automated infrastructure and deployment processes. You'll be the expert on operational excellence and how systems can be built to be redundant, scalable, and observable.

Responsibilities:

Take the technical lead on projects.

Mentor team members

Automate everything

Build via code and maintain our production infrastructure hosted on AWS.

Design and build pipelines to deploy and manage global infrastructure.

Analyse complex system behaviour, performance and application issues.

Develop observability, alerts and runbooks

Capacity analysis and planning, traffic routing, and security policies for Ping's market

leading Single Sign-On SaaS applications.

This position is part of an on-call rotation of 8 hours by 7 days a week.

Requirements:

6 -9 years of experience in Software Engineering, focusing on Site Reliability Engineering (SRE) or DevOps principles.

Experience provisioning large cloud environments using IAC tools such as CloudFormation and Terraform.

Experience with container orchestration (Kubernetes) and microservice architectures

Solid scripting skills (Python/Ruby/Bash/Go/etc.).

Experience with CI/CD automation tooling such as Jenkins, Gitlab CI/CD etc.

Solid experience with configuration management tools such as Puppet/Chef/Salt.

Building pipelines using GitOps methodologies, ChatOps, etc…

Solid understanding and practical application of Site Reliability Engineering (SRE) principles including SLOs, SLIs, error budgets, post-mortems, and incident response.

Experience with security design principles and best practices for building secure, scalable, and resilient cloud-native applications.

Demonstrated experience in a high-volume, mission-critical production service environment, with a strong focus on system resilience, fault tolerance, and disaster

Knowledge with observability tooling such as New Relic, Grafana, and CloudWatch.

Apply for this role →

← Back to all jobs