Lead Data Engineer
Job Description
Lead research and development for advanced analytic data solutions, including GenAI-driven pipelines, delivered with in-office collaboration and hands-on engineering.
Responsibilities
- Plan and lead research and development for advanced analytic data solutions leveraging GenAI
- Collaborate with business and technology leaders to define scope, requirements, and data sources across the enterprise
- Design and deliver data pipelines needed to support critical decision science projects
- Partner with Decision Science Products, Decision Science, and client teams on priority initiatives
- Plan, estimate, design, develop, test, roll out to production, and sustain production solutions
- Consult and collaborate with project team members across workstreams
- Lead design reviews
- Perform hands-on development to implement and support production systems
- Communicate progress, technical concepts, and decisions to colleagues and leaders
Requirements
- 7+ years experience in data engineering development across multiple environments (Dev, QA, Prod, etc.) with DevOps procedures for code deployment/promotion
- Experience translating project scope and high-level requirements into technical data engineering tasks
- Experience defining solutions to sophisticated data engineering problems supporting advanced analytics
- Experience with a variety of GenAI models, tools, and concepts
- Understanding of Knowledge Graphs, Data Mesh, and other data sharing platforms
- 3+ years building and designing relational databases, preferably Snowflake or PostgreSQL
- 3+ years leading and deploying code using a source control platform such as GitLab or GitHub
- 2+ years using job scheduling software such as Apache Airflow, Amazon MWAA, GitLab Runners, or UC4
- Proven ability to collaborate with multiple project teams in a fast-paced environment
- Experience defining and estimating level-of-effort for data engineering activities
- Experience with project and sprint planning
- Ability to communicate technical concepts and solutions to non-technical team members
- Experience designing and building data structures aligned to requirements
- Multiple years developing and maintaining ELT/ETL data pipelines
- Multiple years of demonstrated expertise using SQL and Python
- Experience with containerization technologies such as Docker or Kubernetes
- Proven collaborative work in a multi-team environment
Technologies
- GenAI
- Knowledge Graphs
- Data Mesh
- Snowflake
- PostgreSQL
- GitLab
- GitHub
- Apache Airflow
- Amazon MWAA
- GitLab Runners
- UC4
- ELT/ETL
- SQL
- Python
- Docker
- Kubernetes
- AWS EMR
- EC2
- S3
- Snowpark
- Data Exchange
- Data Marketplace
- Snowpipe
Preferred Qualifications
- Experience leading development of GenAI-based systems, including model selection, pipeline orchestration, and deployment strategies
- Experience defining GenAI architectures
- Knowledgeable with Disney Parks attendance, reservations, and/or products
- Experience with cloud-based technologies, preferably AWS EMR, EC2, and S3
- Experience with advanced Snowflake offerings such as Snowpark, Data Exchange, Data Marketplace, and Snowpipe
Location and Compensation
- Onsite in Lake Buena Vista, FL
- Salary range: USD 135,200 - 181,200 per year
- Base pay offered will consider internal equity and may vary by geographic region, job-related knowledge, skills, and experience
- Bonus and/or long-term incentive units may be provided in addition to the full range of medical, financial, and/or other benefits