DeveloperJobs.io
← Back to all jobs

Job Description

Realign LLC is hiring an AWS Glue Data Engineer to design and optimize scalable ETL pipelines using AWS Glue and PySpark/Spark. This role is based onsite in Phoenix, Arizona and includes opportunities to work closely with the cloud engineering team on a GovCloud proof of concept, plus hands-on conversion of existing SQL logic into maintainable PySpark while validating data quality and performance. The role is supported by a competitive USD 120,000 yearly salary.

What you’ll do

  • Design, build, test, and optimize scalable ETL pipelines using AWS Glue and PySpark/Spark
  • Configure and maintain AWS Glue Crawlers, Data Catalog objects, jobs, workflows, and connections
  • Integrate AWS Glue workloads with Snowflake using SQL and supported connectors
  • Implement data access, governance, and security controls using AWS Lake Formation
  • Collaborate directly with the cloud engineering team to deliver a proof of concept in a GovCloud environment
  • Convert existing SQL logic into maintainable PySpark and validate data quality and performance
  • Document architecture, configuration, deployment steps, test results, risks, and operational handover
  • Participate in an onsite Phoenix buildout week and support troubleshooting through POC completion
  • Maintain strong written and verbal communication with technical stakeholders in a fast-paced POC environment
  • Own deliverables, track risks and dependencies, and provide timely status reporting

What you’ll bring

  • Hands-on AWS Glue experience, including PySpark/Spark ETL development, Crawlers, and Data Catalog
  • Snowflake experience including SQL, connectors, and basic administration
  • Experience with AWS Lake Formation for data governance
  • Eligibility to work on ITAR/export-controlled GovCloud engagements, with US citizenship required
  • Ability to work onsite in Phoenix, Arizona during the buildout week

Helpful experience

  • Snowflake Catalog Federation and/or Iceberg REST Catalog experience
  • Experience with Kiro or another agentic IDE for SQL-to-PySpark code conversion
  • Hub-and-Spoke VPC architecture and inspection firewalls experience, preferably Check Point

Key technologies: AWS Glue, PySpark, Spark, AWS Glue Crawlers, Data Catalog, Snowflake, SQL, AWS Lake Formation, AWS Glue Data Catalog objects, AWS Glue jobs, AWS Glue workflows, AWS Glue connections, GovCloud.

Similar Jobs