AWS Glue Data Engineer
Backend Developer
Amazon Web Services
AWS
Aws Cloudfront
Aws Data Catalog
Aws Glue
Aws Govcloud
Aws S3 Data Lake
Big Data
Bigdata
Cloud
Cloud Computing
Cloud Data Warehouse
Cloud Data Warehouse
Cloud Platform
Cloud Platforms
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Integration
Data Lake
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Data Warehousing
Database
Engineer
ETL
Informatica
IT Services
Pyspark
Snowflake
Spark
SQL
Job Description
Realign LLC is hiring an AWS Glue Data Engineer to design and optimize scalable ETL pipelines using AWS Glue and PySpark/Spark. This role is based onsite in Phoenix, Arizona and includes opportunities to work closely with the cloud engineering team on a GovCloud proof of concept, plus hands-on conversion of existing SQL logic into maintainable PySpark while validating data quality and performance. The role is supported by a competitive USD 120,000 yearly salary.
What you’ll do
- Design, build, test, and optimize scalable ETL pipelines using AWS Glue and PySpark/Spark
- Configure and maintain AWS Glue Crawlers, Data Catalog objects, jobs, workflows, and connections
- Integrate AWS Glue workloads with Snowflake using SQL and supported connectors
- Implement data access, governance, and security controls using AWS Lake Formation
- Collaborate directly with the cloud engineering team to deliver a proof of concept in a GovCloud environment
- Convert existing SQL logic into maintainable PySpark and validate data quality and performance
- Document architecture, configuration, deployment steps, test results, risks, and operational handover
- Participate in an onsite Phoenix buildout week and support troubleshooting through POC completion
- Maintain strong written and verbal communication with technical stakeholders in a fast-paced POC environment
- Own deliverables, track risks and dependencies, and provide timely status reporting
What you’ll bring
- Hands-on AWS Glue experience, including PySpark/Spark ETL development, Crawlers, and Data Catalog
- Snowflake experience including SQL, connectors, and basic administration
- Experience with AWS Lake Formation for data governance
- Eligibility to work on ITAR/export-controlled GovCloud engagements, with US citizenship required
- Ability to work onsite in Phoenix, Arizona during the buildout week
Helpful experience
- Snowflake Catalog Federation and/or Iceberg REST Catalog experience
- Experience with Kiro or another agentic IDE for SQL-to-PySpark code conversion
- Hub-and-Spoke VPC architecture and inspection firewalls experience, preferably Check Point
Key technologies: AWS Glue, PySpark, Spark, AWS Glue Crawlers, Data Catalog, Snowflake, SQL, AWS Lake Formation, AWS Glue Data Catalog objects, AWS Glue jobs, AWS Glue workflows, AWS Glue connections, GovCloud.