Data Engineer Level 2
Artificial Intelligence
Big Data
Bigdata
Cloud
Cloud Data Engineering
Cloud Data Platform
Cloud Native
Cloud Platform
Cloud Platforms
Data
Data Analysis
Data Analytics
Data Engineer
Data Engineering
Data Integration
Data Pipeline
Data Pipelines
Data Platform
Data Processing
Database
Databases
Databricks
Engineer
ETL
Google Cloud
Google Cloud Bigquery
Google Cloud Platform
Informatica
Machine Learning
Programming
Programming Language
Programming Languages
Spark
SQL
Vertex Ai
Job Description
Design and deliver scalable data solutions across cloud environments, with a focus on Python, Vertex AI, and modern data engineering.
Responsibilities
- Design, build, and maintain scalable data pipelines for ingestion, transformation, and integration using Kafka, Databricks, and related technologies
- Develop and manage feature engineering pipelines for machine learning workflows using Vertex AI, BigQuery ML, and Python
- Work with SQL, NoSQL, cloud-based platforms, and real-time streaming data systems
- Drive modernization and innovation across data platforms and engineering processes
- Build automated unit, integration, and performance testing frameworks to support data quality and reliability
- Optimize data workflows for performance, scalability, and cost efficiency across large datasets
- Own systems and processes across the full SDLC, from design through deployment and support
- Collaborate with cross-functional teams to deliver internal and external-facing data solutions
- Create and review architecture diagrams, interface specifications, and technical documentation
Requirements
- Strong Python development experience
- Hands-on experience with Google Cloud Platform (GCP) and Azure
- Strong experience with Vertex AI and feature engineering
- Experience with Databricks, Kafka, and BigQuery/BigQuery ML
- Strong understanding of data pipelines, ETL/ELT, streaming, and distributed data processing
- Experience implementing automated testing for data applications and pipelines
- Strong SQL and experience working with large-scale datasets
- Excellent problem-solving, communication, and cross-functional collaboration skills
Skills & Technologies
- Python, Vertex AI, BigQuery ML, Databricks, Kafka, BigQuery, SQL
- NoSQL, Azure, Google Cloud Platform (GCP)
- ETL/ELT, real-time streaming, distributed data processing
- Unit testing, integration testing, performance testing
- SDLC
Location
- Onsite in Cincinnati, OH 45202
- EST/CST only
- Cincinnati or Chicago preferred
Pay
- $70-$75/hr on W2
Duration
- 6 Months