役割について
Reliability is a promise we make to every user. Our site reliability engineers own the availability, performance and resilience of the HeroApply platform, and the tooling that lets every engineer ship with confidence.
You will work across teams to design for failure, automate the routine, and make incidents rare, short and well understood.
何をしますか
- Own availability, performance and capacity across the platform.
- Build and maintain infrastructure, deployment and observability tooling.
- Lead incident response and drive lasting improvements from every incident.
- Partner with engineering teams on reliable, secure architecture.
- Automate operational work and document it clearly.
何をもたらすか
- Extensive experience operating production systems at scale.
- Strong background in infrastructure, networking and cloud platforms.
- Expertise in monitoring, alerting and incident management.
- A calm, methodical approach when things go wrong.
- A bias toward automation and clear documentation.
いい加減にしたい
- Experience with infrastructure as code.
- Security or compliance background.
- Prior ownership of a reliability programme.

