Responsibilities
Design, build, and optimize scalable data pipelines and lakehouse solutions using Databricks, Spark, and cloud platforms, including Medallion Architecture and data ingestion frameworks. Establish data quality, governance, security, monitoring, and CI/CD practices; support AI/ML data needs and provide technical leadership and mentorship.
Requirements
Requires a bachelor's or master's degree in a relevant field and 8+ years of data engineering or data platform experience, with strong hands-on Databricks and Apache Spark skills and proficiency in Python, PySpark, and SQL. Candidates should have cloud, Delta Lake, Unity Catalog, distributed processing, data architecture, optimization, Git, and CI/CD experience; relevant certifications and streaming or AI/ML exposure are preferred.