Senior Data Engineer (Databricks & Cloud Analytics)
Senior
Azure Data Factory
Azure Data Lake
Big Data
Data Analytics
Data Architecture
Data Engineer
Data Factory Azure
Data Governance
Data Integration
Data Lake
Data Lakehouse
Data Pipeline
Data Platform
Data Processing
Data Security
Data Warehouse
Database
Databricks
Databricks Mlflow
Delta Lake
Delta Live Tables
Engineer
ETL
Microsoft Azure
Spark
SQL
Job Description
CGI is seeking a Senior Data Engineer to design, develop, and support secure, scalable, high-performing enterprise data platforms.
Responsibilities
- Design, develop, test, deploy, and maintain enterprise data engineering solutions using modern cloud and big data technologies
- Design and implement scalable ETL/ELT pipelines with Databricks, Apache Spark (PySpark), Delta Lake, and Azure Data Factory
- Build high-performance data ingestion, transformation, and integration frameworks for enterprise analytics and reporting
- Develop, optimize, and maintain complex SQL queries, stored procedures, and data transformation processes
- Design and implement scalable data models for business intelligence, analytics, and AI initiatives
- Build and integrate RESTful APIs, event-driven architectures, and legacy SOAP services for enterprise data exchange
- Develop and maintain streaming and messaging solutions using Apache Kafka
- Monitor, troubleshoot, and optimize production data pipelines, including root cause analysis and long-term corrective actions
- Implement data quality, governance, security, and performance best practices across enterprise platforms
- Manage source code with Git while applying CI/CD and enterprise DevOps practices
- Collaborate with Solution Architects, Product Owners, Business Analysts, and cross-functional engineering teams to translate business requirements into technical solutions
- Participate in Agile ceremonies such as sprint planning, backlog refinement, architecture discussions, code reviews, and retrospectives
- Create technical documentation, deployment artifacts, testing documentation, and operational runbooks
- Mentor junior engineers and contribute to engineering standards, best practices, and continuous improvement efforts
Requirements
- Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related technical discipline (or equivalent professional experience)
- 6+ years of experience designing, developing, and supporting enterprise data platforms, cloud analytics solutions, or large-scale data engineering initiatives
- Strong hands-on experience with Databricks
- Strong hands-on experience with Apache Spark (PySpark)
- Strong hands-on experience with Python
- Strong hands-on experience with SQL
- Strong hands-on experience with Scala
- Experience with Databricks Workspaces
- Experience with Delta Lake
- Experience with Unity Catalog
- Experience with Delta Live Tables (DLT)
- Experience with Databricks SQL
- Experience with MLflow
- Experience with Databricks Jobs
- Strong SQL development skills including query optimization and performance tuning
- Experience designing and implementing ETL/ELT solutions using Azure Data Factory or comparable cloud integration platforms
- Experience with Apache Kafka or other event streaming technologies
- Experience integrating enterprise applications using REST APIs, JSON, XML, and SOAP web services
- Experience with Azure Data Lake Storage (ADLS Gen2) or comparable cloud storage platforms
- Experience using Git for source code management and collaborative development
- Experience working in Linux environments, including shell scripting and command-line utilities
- Familiarity with Infrastructure as Code (IaC) tools such as Terraform (preferred)
- Experience implementing CI/CD pipelines using Azure DevOps, GitHub Actions, or similar DevOps tools (preferred)
- Strong analytical, troubleshooting, and problem-solving skills
- Experience working in Agile environments using Scrum, Kanban, or SAFe methodologies
- Ability to manage multiple priorities while delivering high-quality solutions
- Ability to work independently while mentoring teammates and contributing to technical leadership
Technologies
- Databricks
- Apache Spark
- PySpark
- Azure Data Factory
- Azure Data Lake
- Azure Data Lake Storage (ADLS Gen2)
- Kafka
- Python
- SQL
- Scala
- Delta Lake
- Unity Catalog
- Delta Live Tables (DLT)
- Databricks SQL
- MLflow
- Databricks Jobs
- Git
- RESTful APIs
- SOAP
- JSON
- XML
- Linux
- Shell scripting
- Terraform
- Azure DevOps
- GitHub Actions
- CI/CD
- Scrum
- Kanban
- SAFe
- REST
- SOAP web services
- Delta Live Tables
Benefits
- Competitive compensation (USD 79,600 - 139,300 per year)
- Comprehensive insurance options
- Matching contributions through the 401(k) plan and the share purchase plan
- Paid time off for vacation, holidays, and sick time
- Paid parental leave
- Learning opportunities and tuition assistance
- Wellness and well being programs
Desired Skills
- Microsoft Azure cloud services
- Azure Synapse Analytics
- Microsoft Fabric
- Power BI
- Data governance and metadata management
- DataOps and MLOps practices
- Enterprise data warehousing
- Master Data Management (MDM)
- Financial services or public sector data environments
Communication & Collaboration
- Clearly communicate complex technical concepts to technical and non-technical audiences
- Collaborate effectively with architects, developers, business analysts, product owners, and client stakeholders
- Build trusted relationships across cross-functional teams while delivering innovative, high-quality solutions
- Foster a culture of knowledge sharing, continuous learning, and engineering excellence
Location: Salt Lake City, UT area (onsite)
Minimum Experience: 6 years