ML Data Engineer
Artificial Intelligence
Big Data
Bigdata
Cloud
Cloud Infrastructure
Cloud Native
Cloud Operations
Cloud Platform
Cloud Platforms
Cloud Platforms Cloud Platforms
Cloud Technology
Data Analysis
Data Engineer
Data Pipeline
Data Platform
Data Processing
Data Science
Database
DevOps
DevSecOps
Engineer
Engineering
Engineering Software
Google Cloud
Hadoop
Information Technology (IT)
Kubernetes
Machine Learning Engineer
OpenShift
Platform Engineering
Programming Language
Programming Languages
Pyspark
Spark
Job Description
Realign LLC is hiring an ML Data Engineer in Irving, TX to implement and productionalize AI/ML models across on-prem and cloud environments.
Responsibilities
- Collaborate with data scientists and data engineers to convert prototypes and theoretical approaches into functional, production-ready code
- Partner with product and engineering teams to implement AI/ML models using PySpark
- Productionalize AI/ML models in a Hadoop environment
- Troubleshoot AI/ML models after implementation to resolve issues reported by users
- Monitor model performance and continuously improve efficiency, accuracy, and user experience
Requirements
- Strong experience with PySpark and Python to integrate with AI/ML models
- Understanding of AI/ML models and how they interface with on-prem and cloud environments
- Familiarity with ML Ops / LLM Ops and distributed systems
- Experience with big data platforms such as Cloudera Hadoop and cloud platforms such as AWS and GCP
- Solid understanding of system design patterns, scalability, observability, and performance tuning
- Strong analytical and problem-solving skills
- Interest in exploring and building with emerging technologies
Technologies
- Python
- PySpark
- OpenShift
- Kubernetes
- Hadoop
- Cloudera Hadoop
- AWS
- GCP
Location & Compensation
- Location: Irving, TX (onsite)
- Salary: USD 124,560 per year