Strike is the Bitcoin company building a better financial system powered by Bitcoin. We are seeking a highly experienced Site Reliability Engineer to join our team, focusing on tackling complex reliability and scalability challenges while providing technical guidance to our engineering organization.
Key responsibilities
- Lead key technical initiatives focused on improving the reliability, performance, and scalability of our critical systems.
- Design and implement sophisticated, resilient, and scalable solutions by leveraging a deep understanding of distributed systems.
- Develop and champion the adoption of robust automation frameworks and tools to improve operational efficiency.
- Design and implement comprehensive monitoring and logging solutions to ensure actionable insights across all teams.
- Lead complex troubleshooting efforts and provide technical direction during incident response situations.
Requirements
- Minimum of 5 years of experience in SRE, platform engineering, or software development with a strong operational focus.
- Expert-level practical knowledge of cloud platforms, specifically GCP.
- Deep hands-on experience with container orchestration using Kubernetes and infrastructure-as-code tools like Terraform, Helm, and ArgoCD.
- Strong command of multiple programming and scripting languages, including Python, Go, and Bash.
- Proven expertise in building and leveraging advanced observability tools such as Prometheus, Grafana, and the ELK stack.
- Demonstrated experience in providing technical leadership, mentorship, and guidance to engineering teams.
What we offer
- Opportunity to work on a global financial platform powered by Bitcoin.
- Collaborative environment that values talent and passion over formal educational credentials.
- Professional growth through mentorship and leadership opportunities within a high-impact engineering team.