Talentica is seeking a Machine Learning Platform Engineer to join a global, remote-first AI product company. You will be responsible for designing and operating the infrastructure that powers advanced AI capabilities, ensuring that models are scalable, reliable, and cost-efficient for users worldwide.
Key responsibilities
- Build and operate the ML infrastructure and platforms powering production AI products.
- Design systems for model training, evaluation, deployment, inference, and experimentation.
- Build and optimize model-serving and inference infrastructure for high-throughput, low-latency workloads.
- Develop reliable pipelines for data preparation, training, evaluation, and continuous improvement.
- Build production observability, monitoring, tracing, and alerting for AI and ML workloads.
Requirements
- Strong software engineering fundamentals and experience building production systems.
- Experience building ML infrastructure, ML platforms, or production machine learning systems.
- Hands-on experience with model deployment, inference, evaluation, or data pipelines.
- Strong understanding of distributed systems, scalability, observability, and system reliability.
- Familiarity with GPU-based training or inference infrastructure.
What we offer
- Fully remote work environment with a global team.
- Opportunity to work on cutting-edge AI products and infrastructure.
- Collaborative culture with flexible working hours.
- Company laptop provided for the role.