Capstone Integrated Solutions is seeking a highly skilled Senior AWS Data Engineer to lead data architecture, pipeline development, and complex data integrations. You will work on a platform that leverages generative AI to modernize enterprise workflows, focusing on building robust, scalable data solutions in a fully remote environment.
Key responsibilities
- Design and implement a multi-zone enterprise data lake on Amazon S3 with comprehensive ingestion, cleansing, and business layers.
- Build batch and streaming data pipelines using AWS Glue, Amazon Kinesis, and containerized applications.
- Develop and maintain infrastructure as code using Terraform or CloudFormation to manage AWS data resources.
- Implement fine-grained data governance and security using AWS Lake Formation and IAM policies.
- Collaborate with cross-functional teams to support data quality, entity resolution, and executive analytics dashboards.
Requirements
- 5+ years of data engineering experience, with at least 2 years specifically in AWS cloud environments.
- Strong proficiency in Python and SQL, with hands-on experience writing PySpark or Scala for AWS Glue ETL.
- Deep knowledge of AWS services including Glue, Step Functions, Lambda, and S3.
- Experience with infrastructure as code (Terraform/CDK) and CI/CD pipelines using GitHub Actions.
- Solid understanding of data modeling, Apache Iceberg, and performance tuning for Amazon Athena and DynamoDB.
What we offer
- Fully remote work environment.
- Opportunity to work on cutting-edge projects involving generative AI and cloud-native data architectures.
- A culture focused on continuous learning, professional growth, and collaborative problem-solving.