Join the highly skilled Site Reliability Engineering team at Akamai Technologies, where we design and manage the infrastructure supporting global compute products. As a Senior Site Reliability Engineer in the Virtualization & Host Platforms team, you will work at the forefront of hypervisor host technologies, ensuring the reliability and performance of our core Linux platforms.
Key responsibilities
- Develop, test, and distribute software, services, and tools to maintain our infrastructure.
- Build and maintain CI/CD pipelines to support a sustainable platform supply chain for the Akamai Cloud fleet.
- Collaborate with engineering and operations teams to troubleshoot complex production issues.
- Automate processes and improve resource utilization across our distributed systems.
Requirements
- Expert-level experience in Linux internals, system administration, and underlying hardware.
- Advanced knowledge of the Linux kernel, OS, and KVM/QEMU virtualization optimization.
- Proven experience in designing and deploying software and infrastructure at scale.
- Proficiency in Shell scripting and Python for automation and debugging.
- Experience with configuration management tools such as SaltStack, Ansible, or Kubernetes.
What we offer
- The opportunity to work on a massive, globally distributed cloud platform.
- A culture that prioritizes professional growth and learning from a highly skilled engineering team.
- Full flexibility to work from home, supporting your ability to do your best work in an environment that suits you.