This is a remote position.
We are looking for a Data Engineer with strong experience in Databricks, PySpark, and modern Data Warehouse systems. The ideal candidate can design, build, and optimize scalable data pipelines and work closely with analytics, product, and engineering teams.
Requirements
- Strong hands-on skills in Databricks, PySpark, and SQL
- Experience with data warehouse concepts, ETL frameworks, batch/streaming pipelines
- Solid understanding of Delta Lake and Lakehouse architecture
- Experience with at least one cloud platform (Azure preferred)
- Experience with workflow orchestration tools (Airflow, ADF, Prefect, etc.)
Roles and Responsibilities
• Design and build ETL/ELT pipelines using Databricks and PySpark
• Develop and maintain data models and data warehouse structures
• Optimize data workflows for performance, scalability, and cost
• Work with cloud platforms (Azure/AWS/GCP) for storage, compute, and orchestration
• Ensure data quality, reliability, and security across pipelines
• Collaborate with cross-functional teams (Data Science, BI, Product)
• Write clean, reusable code and follow engineering best practices
• Troubleshoot issues in production data pipelines
Salary: 24000