Site Reliability Engineer (US - Central/Eastern time)
Summary generated from the verified employer listing
PostHog seeks a full-time Site Reliability Engineer to operate and automate production infrastructure within the US Central or Eastern time zones. The role requires extensive experience with Kubernetes, AWS, and Terraform to manage stateful systems and ensure platform reliability. Engineers will own end-to-end system health, participate in on-call rotations, and build self-healing automation to reduce incident frequency. This position suits candidates who prefer deep ownership of complex infrastructure and enjoy designing robust, scalable systems over merely responding to alerts.
Key details
- Requires deep hands-on experience with production Kubernetes, specifically EKS, and large-scale AWS infrastructure management.
- Involves automating infrastructure using Terraform or Terragrunt, including module design and state management.
- Candidates must support stateful systems and participate in on-call incident response to improve system reliability.
- Position is fully remote and requires alignment with US Central or Eastern time zones.
- The role emphasizes proactive ownership of projects and designing self-healing automation rather than just reacting to alerts.
