Senior Data Engineer
Job Description
Steampunk builds enterprise-grade solutions in the federal space, bringing new thinking to clients in the homeland and federal civil environment. In this role, you will help strengthen the company’s AI & Data Exploitation Practice by designing and delivering data platforms, services, and pipelines using Databricks. The work emphasizes reliable, scalable engineering and strong communication with client teams, developers, and data scientists.
What you’ll do
- Lead and architect data migrations using Databricks, with emphasis on performance, reliability, and scalability.
- Assess and develop understanding of existing ETL jobs, workflows, data marts, BI tools, and reports.
- Respond to technical inquiries related to customization, integration, enterprise architecture, and the general feature/functionality of data products.
- Support an Agile software development lifecycle as part of ongoing delivery.
- Contribute to the growth of the AI & Data Exploitation Practice through practical data engineering work.
What we’re looking for
- Current ICE US Public Trust Clearance (Full Background Investigation), and ability to hold a position of public trust with the US government.
- 5-7 years of industry experience coding commercial software and a passion for solving complex problems.
- 5-7 years of direct experience in Data Engineering, including tools such as Databricks, Apache Spark, and Delta Lake (and related lakehouse technologies).
- Strong experience with relational SQL (preferably T-SQL; alternatively pgSQL or MySQL), including query authoring and optimization.
- Hands-on experience with data pipeline and workflow management tools such as Databricks Workflows, Airflow, or Step Functions.
- Working knowledge of AWS and Azure services (for example, Databricks on AWS, S3, EC2, RDS).
- Experience with object-oriented or scripting languages such as PySpark/Python, Java, C++, or Scala.
- Experience working with Data Lakehouse architecture, including Delta Lake and/or Apache Iceberg.
- Ability to inspect existing data pipelines, determine their purpose and functionality, and re-implement efficiently in Databricks.
- Experience manipulating, processing, and extracting value from large, disconnected datasets.
- Experience manipulating structured and unstructured data.
- Experience architecting data systems (transactional and warehouses).
- Experience with SDLC, CI/CD, and operating in dev/test/prod environments.
- Commitment to data governance.
- Experience supporting teams of developers and data scientists building web-based interfaces, dashboards, reports, and analytics or machine learning models.
- Experience with data cataloging tools such as Informatica EDC, Unity Catalog, Collibra, Alation, Purview, or DataZone is a plus.
Location and clearance
- Location: McLean, VA (onsite)
- Compensation: USD 140,000 - 180,000 per year
- Experience level: 5+ years
Identity verification
As part of the application process, you are expected to be on camera during interviews and assessments. Steampunk reserves the right to take your picture to verify your identity and prevent fraud.
How Steampunk works
Steampunk is a change agent in the federal contracting industry, applying a Human-Centered delivery methodology to change expectations for shared accountability on clients’ toughest mission challenges.