Site Reliability Engineer
Keep the platform up across SaaS, private cloud, on-prem, and air-gapped deployments. SLAs you can take pride in.
The role
You'll join a small team focused on the platform reliability surface of the Innfini platform. The work is product, customer, and consequence — in equal measure. You'll ship to customers running national-scale operations, sometimes from a war room, sometimes from a chemical plant.
Multi-region SaaS SLA (99.95%)
Observability platform
Incident response runbooks
Capacity planning across regions
You
- 015+ years SRE / production engineering
- 02Kubernetes at scale
- 03Observability stack ownership (metrics, logs, traces)
- 04Incident command experience
- 05On-call discipline
The deal.
Competitive base + meaningful equity. Salary bands published internally. Equity that vests over four years with a one-year cliff.
Healthcare and retirement that don't make you read the fine print. Comprehensive medical, dental, vision. 401(k) match (US) and equivalent country plans elsewhere.
4 weeks vacation. Sabbatical at year five. We're a 10-year company on a 30-year problem.
Hybrid, with intent. Tuesday + Thursday in-office at your hub. Heads-down days are protected. Async-by-default for documents, sync for decisions.
Apply for Site Reliability Engineer.
Send a resume and a few sentences on why this role and not another. We respond to every application within 14 days, with written feedback either way.