Senior Data Engineer
Core
Lead the design, development, and delivery of scalable enterprise data solutions, including batch and real-time pipelines, lakehouse platforms, and governed data access for analytics and AI.
Role type
Senior IC data engineer (cloud data platform)
Builds
Cloud-based lakehouse and data-platform solutions using Databricks and AWS
Domain
Cloud data engineering, regulated industries (biotech, pharma, life sciences, manufacturing)
Required skills
Databricks, Apache Spark, PySpark, SQL, Python, distributed computing, lakehouse architecture, data warehousing, workflow orchestration, CI/CD, data quality, metadata management, governance, RBAC, technical leadership, mentoring
Preferred skills
Unity Catalog, Data Fabric, Data Mesh, vector databases, Kafka, Kinesis, AI/GenAI support (RAG, embeddings), AI-assisted development tools
Technologies
Databricks, Apache Spark, AWS, Delta Lake, Git, Kafka, Kinesis
Responsibilities
Lead design and development of scalable batch and real-time ETL/ELT pipelines; Own complex data solutions from requirements through production support; Build cloud-based lakehouse solutions; Develop reusable data integration frameworks; Optimize Spark workloads and Databricks compute; Implement workflow orchestration, monitoring, and data-quality controls; Define data models, contracts, and engineering standards; Mentor engineers and lead architecture reviews
Seniority
Senior, hands-on IC with technical leadership