RemoDevs is seeking a skilled Data Engineer with extensive Databricks experience to join a dynamic project focused on building a modern, scalable data platform. You will play a key role in supporting advanced analytics, machine learning, and AI initiatives by developing robust data foundations.
Key responsibilities
- Design, build, and maintain scalable data pipelines using Databricks and Apache Spark.
- Integrate data from multiple sources to create reliable, reusable, and high-quality data flows.
- Optimize data processing and infrastructure with a focus on performance, scalability, and cost efficiency.
- Collaborate with Data Scientists and business stakeholders to deliver impactful data products.
- Implement and maintain data quality, monitoring, and engineering best practices.
Requirements
- Proven hands-on experience with Databricks and production-grade data pipelines.
- Strong proficiency in SQL and PySpark/Apache Spark.
- Solid understanding of modern lakehouse and cloud data architecture concepts.
- Experience working with cloud environments such as Azure, AWS, or GCP.
- Ability to solve complex problems and work effectively within a technical team.
What we offer
- Opportunity to participate in a large-scale data and AI transformation project.
- Work with cutting-edge technologies including Delta Lake and modern lakehouse architectures.
- A fully remote work environment that supports professional growth and technical innovation.
- Engagement in building foundational data systems that power advanced machine learning and analytics use cases.