This remote Data Engineer role centers on designing, building, and maintaining data pipelines, ETL processes, and data products on cloud platforms, with collaboration on deploying machine learning models.
Responsibilities
- Design, build, and sustain batch or real-time data pipelines in production environments.
- Maintain and optimize the data infrastructure required for accurate extraction, transformation, and loading from a wide range of sources.
- Develop ETL processes to extract and transform data from multiple sources.
- Automate data workflows including ingestion, aggregation, and ETL processing.
- Prepare raw data in data warehouses into consumable datasets for both technical and non-technical stakeholders.
- Collaborate with data scientists and functional leaders in sales, marketing, and product to deploy machine learning models in production.
- Build, maintain, and deploy data products for analytics and data science teams on cloud platforms such as AWS, Azure, and GCP.
- Ensure data accuracy, privacy, security, and compliance through quality control procedures.
- Monitor data system performance and implement optimization strategies.
- Apply data controls to uphold privacy, security, and quality within assigned areas of responsibility.
Requirements
- At least 3 years of SQL experience with relational databases and database design.
- Experience with cloud data warehouse solutions and related technologies, including Databricks and Apache Spark.
- Experience using data ingestion tools such as Fivetran, Stitch, or Matillion.
- Working knowledge of cloud-based platforms (AWS, Azure, GCP).
- Experience building and deploying machine learning models in production.
- Strong proficiency in object-oriented languages such as Python, Java, C++, or Scala.
- Strong scripting capabilities with Bash.
- Experience with data pipeline and workflow tools, notably Airflow.
- Solid project management and organizational skills.
- Excellent problem-solving, communication, and collaboration abilities.
- Proven ability to work both independently and as part of a team.
Technologies
- dbt, Snowflake, Airflow, Fivetran, Databricks, Apache Spark, Stitch, Matillion
- AWS, Azure, GCP
- Python, Java, C++, Scala, Bash
- Hive, Hadoop, Presto, MapReduce
- CrateDB, Redis, Cassandra, MongoDB, Neo4j, SQL
Benefits
- Annual bonuses
- Short and program-specific awards
- Comprehensive benefits offering
- Base salary range: USD 88,100 to USD 141,000 per year
Our Culture
Our values emphasize excellence, boundary-breaking work, and shared success. The culture supports continuous learning and collaboration across teams to drive impactful outcomes for communities and customers.
Who We Are
Cornerstone enables organizations and their people to thrive in a changing world. The Galaxy platform provides AI powered workforce agility, helping organizations identify skills gaps, retain top talent, and deliver multimodal learning experiences to meet diverse needs. The platform serves thousands of organizations and millions of users across many regions.
Total Rewards
Cornerstone follows a compensation framework that prioritizes equitable pay, market-aligned data, and skill-based growth. The overall package may include annual bonuses and program-specific awards, complemented by a comprehensive benefits package. The listed base salary range provides a transparent view of compensation for this role.