Polarcode — Embedded AI inference at the edge. We are a fully remote, async-first team and we are hiring a Site Reliability Engineer to join our reliability group.
About the role
You will own meaningful parts of our product surface, working primarily with Kubernetes and Prometheus. We ship small and often, write things down, and optimise for long-term maintainability over heroics. You will collaborate directly with product, design and customer-facing teams across multiple timezones.
What you'll do
- Design, build and ship features across our Kubernetes stack
- Raise the bar on Prometheus reliability, performance and developer experience
- Collaborate in a fully async workflow: RFCs, code review, written updates
- Own projects end-to-end, from proposal to production rollout and operation
What we're looking for
- 5+ years of professional experience with Kubernetes and Prometheus
- Working knowledge of Go and modern engineering practice
- Clear written communication — our default mode is async and documented
- An ownership mindset: you ship it, you run it
Benefits
- Fully remote — work from anywhere in our hiring regions
- $140k–$200k annual salary + meaningful equity
- Home-office and learning budget, annual team offsite
- Generous paid vacation and local public holidays
KubernetesPrometheusGoIncident ResponseRedis