Principal Site Reliability Engineer
About the role
The role involves defining and driving reliability strategy, establishing standards for availability, resilience, observability, incident management, and operational readiness. It also includes leading architecture reviews, partnering with engineering leadership to ensure alignment of reliability objectives with business priorities, creating frameworks for operational guardrails, guiding service architecture towards simplicity, scalability, resilience, and operational excellence, and driving major initiatives that improve platform maturity and long-term sustainability.
Candidates will join an innovative team pushing engineering boundaries. Depending on the domains associated with this job, you will be expected to design clean APIs, write automated test cases, and participate in peer code reviews. We value developers who focus on performance optimization, fast load times, and simple, maintainable architectures.