As a Senior Site Reliability Engineer you'll join the founding SRE team at our new NYC engineering hub, sitting within Foundations. You'll own critical services end-to-end and partner across engineering to raise the reliability bar for the entire platform, working closely with teams in Stockholm.
*This is a New York-based, 5-day in-office role. We believe building together in person drives better outcomes.
- Build, ship, and operate foundational platform services with full ownership
- Build and maintain a high-signal observability stack (metrics, logs, traces) and translate signals into action
- Define and evolve SLIs/SLOs, alerting, and reliability reporting for critical systems
- Improve on-call and incident response, including escalation paths, coordination, and post-incident follow-ups
- Reduce toil through automation, better tooling, and improved system ergonomics
- Partner with product and platform engineers to design resilient systems and improve deployment safety
- Significant experience operating and improving production systems, including debugging under pressure and preventing repeat incidents
- Comfortable writing software and building automation to solve reliability problems
- Autonomous, with pride in the quality and resilience of the systems you ship and operate
- A systems thinker: failure modes, graceful degradation, and practical tradeoffs
- Strong observability, incident management, and on-call experience, as well as experience with cloud infrastructure and Kubernetes
- Global collaboration: Partner with teams and clients domestic and internationally.
- In-person environment: Union Square office designed for ambitious builders.
Search Senior Site Reliability Engineer jobs near New York City → Browse all live jobs
This posting was published by Legora on their own careers system and is shown here with a direct link to apply there. Employers: for corrections or removal, contact jobs@veritahire.com.