Principal DevOps Engineer
What You Will Do
Define the technical vision and multi-quarter roadmap for cloud infrastructure, aligned with engineering and business goals
Lead architecture and design for cross-team initiatives, driving decisions through RFCs and design reviews
Build and evolve internal platform products that make engineering teams faster and safer, and own their full lifecycle
Deploy and manage infrastructure hands-on through infrastructure as code, across cloud and container platforms
Own reliability for critical infrastructure: define SLOs, lead complex incidents, and drive postmortem actions to completion
Drive cost efficiency and capacity planning across cloud providers
Embed security, compliance, and policy as code into the platform by default
Apply AI and agent-based tooling to automate operations and improve developer experience
Mentor senior engineers, grow technical leaders, and set standards for code, documentation, and operations
Partner with engineering leadership and product teams, influencing direction without direct authority
Minimum Requirements
10+ years in infrastructure, platform, or SRE engineering, including experience operating at Staff or Principal scope
Proven record of leading large, cross-team infrastructure initiatives from ambiguous problem to production
Deep expertise in at least one major cloud (GCP or AWS) and working knowledge of a second
Expert-level Kubernetes: designing, scaling, upgrading, and troubleshooting production clusters (GKE or EKS)
Strong infrastructure as code practice with Terraform, including module design and state management at scale
Strong software engineering skills in Go or Python, with a maintained production codebase
Solid networking fundamentals: VPC design, routing, DNS, load balancing, service mesh (Istio or similar)
Experience building observability at scale (Datadog, Prometheus, OpenTelemetry)
GitOps and CI/CD experience (ArgoCD, Flux, or similar)
Excellent written and spoken English, with the ability to explain complex tradeoffs clearly in documents and async channels
Nice to Have
Experience with AI/LLM platform infrastructure or agent-based automation
Managed data services at scale (e.g. MongoDB Atlas, Cloud SQL, Kafka)
Secrets management and policy as code (e.g. Akeyless, Vault, OPA)
FinOps experience with measurable savings
Experience in a globally distributed engineering organization
Education
Bachelor's Degree in Computer Science or a related technical discipline, or the equivalent combination of education, technical certifications, training, or work experience.
#LI- RA1
#LI-Remote
Actual compensation offered will be based on factors such as the candidate’s work location, qualifications, skills, experience and/or training. Your recruiter can share more information about the specific salary range for your desired work location during the hiring process. We want our employees and their families to thrive.
In addition to comprehensive benefits we offer holistic mind, body and lifestyle programs designed for overall well-being. Learn more about ZoomInfo benefits here.
Below is the US base salary for this position. Additional compensation such as Bonus, Commission, Equity and other benefits may also apply.
$155,400—$244,200 USD