emagine Polska is seeking a Senior DevOps / Site Reliability Engineer to ensure the reliability, scalability, and security of our cloud-native platform. You will play a pivotal role in optimizing our infrastructure, implementing SRE best practices, and supporting development teams in delivering highly available services.
Key responsibilities
- Design, implement, and maintain highly available and scalable infrastructure on AWS.
- Improve the reliability of production systems using SRE principles such as SLOs, SLIs, and error budgets.
- Manage and optimize container orchestration platforms using Kubernetes, Docker, and Helm.
- Develop and maintain Infrastructure as Code using tools like Terraform and Ansible.
- Lead incident response, perform root cause analysis, and drive continuous improvement through postmortems.
Requirements
- Over 5 years of professional experience in DevOps, SRE, or Platform Engineering.
- Proven expertise in Linux systems administration and troubleshooting.
- Strong experience with Kubernetes in production environments and CI/CD pipeline management.
- Solid knowledge of AWS cloud platforms and Infrastructure as Code practices.
- Advanced proficiency in the French language.
What we offer
- Opportunity to work on complex, high-scale cloud-native systems.
- Collaborative environment focused on SRE best practices and operational excellence.
- Full remote work model allowing for professional flexibility.
- Engagement in critical infrastructure projects with a focus on resilience and security.