IT Applications Data Engineer
Job Description
Pantex is hiring an IT Applications Data Engineer onsite in Amarillo, TX (Pantex Plant). This role focuses on building high-performance data pipelines and data infrastructure that support enterprise AI, machine learning, and advanced analytics. You will work with modern data stack platforms and distributed processing to deliver model-ready data and integrate with MLOps.
What you can expect
As an IT Applications Data Engineer (Senior Associate to Specialist), you will design, develop, test, and maintain scalable data pipelines for batch and real-time consumption. You will build and optimize data systems across cloud data warehousing and lakehouse solutions, ensuring data quality, reliability, and strong performance. The role also includes collaboration with Data Science teams and partners across application development and API services, helping establish secure, efficient data ingestion points and data standards.
Responsibilities
- Build, test, and maintain highly scalable data pipelines for batch and real-time consumption using cloud-native services and distributed processing frameworks.
- Construct and manage systems in a modern data stack (for example, Databricks, Snowflake, or equivalent cloud warehouse or lakehouse solutions) to support accessibility and performance.
- Develop production-grade code primarily in Python or Scala to cleanse, transform, and load complex, high-volume data into structured models defined by the Data Architect.
- Implement and manage data quality checks, data lineage tools, and pipeline monitoring to ensure integrity and rapid failure detection.
- Optimize database performance, query efficiency, and cost management for secure cloud storage and processing services.
- Enable Data Science teams with efficient data access, feature engineering workflows, and integration of model-ready data into the MLOps platform.
- Contribute to DataOps and DevSecOps practices for pipeline automation, testing, and secure deployments.
- Maintain documentation for data flows, data dictionaries, and production runbooks.
- Collaborate with application development and API developers to establish secure and efficient ingestion points from source systems.
- Work with Data Architects on modeling new data sources to ensure alignment with enterprise standards and architecture.
Requirements
- Bachelor’s degree in engineering, science, or information technology discipline with minimum 2 years of relevant experience.
- OR Master’s degree in engineering, science, or information technology discipline.
- OR applicants without a bachelor’s degree may be considered based on a combination of at least 10 years of completed education and/or relevant experience.
- Minimum 2 years of relevant experience, including at least 1 year focused on building and managing enterprise-grade data pipelines and ELT/ETL processes.
- Expert proficiency in SQL and Python or Scala for data manipulation, transformation, and automation.
- Demonstrated experience with distributed processing frameworks (for example, Apache Spark, Databricks, Snowflake) and cloud data services (for example, Azure Data Factory, AWS Glue).
- Master’s degree with minimum 3 years of relevant experience.
- Relevant certifications (for example, Databricks Certified Data Engineer, Snowflake certification, or Azure Data Engineer Associate).
- Experience implementing streaming/near real-time ingestion technologies (for example, Kafka, Azure Event Hubs).
- Experience with data modeling concepts such as Dimensional Modeling or Data Vault in data warehousing solutions.
Technologies
- Python, Scala, SQL
- Databricks, Snowflake, Apache Spark
- Azure Data Factory, AWS Glue
- Kafka, Azure Event Hubs
- MLOps, ELT, ETL
- DataOps, DevSecOps
- Dimensional Modeling, Data Vault
Benefits
- Comprehensive health coverage
- Robust retirement planning
- Education reimbursement
- Opportunities for continuous learning
Security and workplace notes
- Requires a Q clearance; however all qualified candidates will be considered regardless of current clearance status.
- Ability to obtain and maintain a Department of Energy Q clearance is required.
- Position may require entry into Materials Access Areas (MAA) and participation in the Human Reliability Program (HRP).
- If HRP is required, completion of a counterintelligence-scope polygraph is required (pursuant to 10 CFR 709).
- Medical requirements may apply.
- Pantex is a drug-free workplace. A pre-placement physical, drug screening, and background investigation are required.
- U.S. citizenship is required for security clearance applicants.
- All employees are subject to random selection for drug testing without advance notification.