Staff Data Engineer
Core
Building a durable, high-performance data lake and pipelines to process petabyte-scale cyber operations and intelligence data for mission-critical use cases.
Role type
Staff Data Engineer (Cyber Operations)
Builds
Production data lake, ETL pipelines, and query systems for intelligence analysis
Domain
Cybersecurity / National Security / Big Data
Required skills
Data lake architecture, ETL pipeline design and operation, schema and index design, column-oriented databases, data modeling, technical leadership, mentoring
Preferred skills
Key-value datastores, streaming systems, graph databases, internet/networking datasets
Technologies
Apache Iceberg, Delta Lake, Apache Hive, Trino, Presto, AWS Athena, Apache Spark, ClickHouse, Amazon Redshift, Google BigQuery, Airflow, AWS Glue, NiFi, ClickPipe, Kafka, RabbitMQ, NATS, AWS Kinesis, Neo4j, AWS Neptune, Memgraph, Apache AGE
Responsibilities
Lead development and operation of a data lake for cyber operations; Design schemas, partitions, and indexes for complex datasets; Partner with engineers and analysts to define query patterns; Build and evolve observable and resilient ETL pipelines; Drive technical initiatives from architecture to production rollout; Establish best practices for data quality and operational ownership; Mentor engineers on data modeling and pipeline design; Identify and resolve bottlenecks in storage, compute, and query layers
Seniority
Staff, hands-on IC with leadership and mentorship responsibilities