Head of Site Reliability Engineering (SRE)
We are looking for a Head of Site Reliability Engineering (SRE) who will serve as the principal leader in developing and executing our infrastructure reliability strategy that scales with the company as it continues to grow.
You will join an exceptional crypto company and senior engineering management team at a time of high growth. This is a tremendous opportunity for a leader to enhance and shape the infrastructure vision at a high-growth crypto financial services company. You will know how to navigate complex organizations to build relationships, deliver results, and understand the macro environment in crypto while ensuring the day-to-day uptime and safety of Blockchain.com’s needs.
WHAT YOU WILL DO
Establish a multi-year platform and reliability strategy, defining the SRE roadmap and driving engineering standards for observability, automation, and production readiness
Be a strong business enabler, allowing us to deliver products quickly whilst maintaining standards-compliance and security through well-designed self-service tooling.
Lead, mentor, and scale a global team of senior SREs, developing technical leaders through coaching and a strong hiring pipeline.
Architect and manage a secure hybrid infrastructure across GCP, AWS, and on-premise data centers, deploying sensitive components (using Kubernetes and Hashicorp Nomad) with zero-trust principles.
Own incident response and root-cause analysis, ensuring MTTD/MTTR targets are met while systematically addressing technical debt and legacy system inconsistencies.
Govern cloud architecture, IaC, networking, and cost management to maintain the organization’s dedication to automated workflows wherever feasible.
Build strong relationships with departments across the organization to understand requirements and deliver services whilst maintaining our standing with auditors and regulators.
WHAT YOU WILL NEED
Leadership & Standards: Proven experience leading senior SRE teams in 24/7 financial environments, with the ability to influence alignment around shared engineering standards.
Cloud & Orchestration: Deep expertise in GCP and AWS networking/management, alongside hands-on experience with Kubernetes, Nomad, and optimized Terraform workflows.
Hybrid Infrastructure: Solid background in bare-metal hardware, data center networking, and implementing Cloudflare Zero Trust architectures.
Systems & Observability: Strong software engineering background with a deep understanding of distributed systems, Linux and approaches to achieving good observability & reliability.
Execution: Strong judgment in prioritizing reliability, security, and developer productivity while evolving complex, heterogeneous estates rather than just greenfield environments.
COMPENSATION & PERKS
Full-time salary based on experience and meaningful equity in an industry-leading company
This is a role based in our London office, with a mandatory in-office presence four days per week.
Work from Anywhere Policy: You can work remotely from anywhere in the world for up to 20 days per year.
ClassPass
Unlimited vacation policy; work hard and take time when you need it
Apple equipment
The opportunity to be a key player and build your career at a rapidly expanding, global technology company in an emerging field
Flexible work culture