About the role

You will join the infrastructure team that runs the fleet under every customer service. You will automate the work we still do by hand and turn every incident into a fix that stays fixed.

Responsibilities

  • Run and improve the fleet across regions, from capacity planning to kernel upgrades
  • Build tooling for safe rollouts, failover, and disaster recovery
  • Lead incident response and write blameless reviews that lead to real changes
  • Define service level objectives with product teams and alert on what matters

Requirements

  • Four or more years in an SRE, platform, or infrastructure role
  • Deep Linux and networking knowledge, from TCP to DNS
  • Experience with infrastructure as code and a scripting language
  • Calm under pressure, with clear written updates during incidents

Nice to have

  • Experience running multi-region or edge networks
  • Familiarity with eBPF or low-level performance tooling

Benefits

Everything on the careers page, including equity, health cover, thirty days of paid leave, a learning budget, and two team offsites a year.

Hiring process

A 30-minute intro call, a take-home task or portfolio review that you can do in your own time, two interviews with the team, and a final chat with a founder. Most processes take three weeks.

Don't see a role that fits?

Tell us what you would build here. We keep every note and reach out when a matching role opens.

Buy NowTheme Details