Senior Lead Data Engineer
Job Description
Senior Lead Data Engineer position in Plano, TX (onsite) focused on building enterprise streaming data platforms and distributed pipelines.
Responsibilities
- Design and build an enterprise level scalable, low-latency, fault-tolerant streaming data platform to deliver meaningful, timely insights
- Build next-generation distributed streaming data pipelines and analytics data stores using streaming frameworks such as Flink and Spark Streaming
- Develop distributed pipelines and data stores using Java, Scala, and Python
- Lead a group of engineers delivering pipelines on big data technologies including Spark, Flink, Kafka, Snowflake, AWS Big Data Services, and Redshift on medium to large scale datasets
- Influence best practices for data pipeline design, data architecture, and processing for both structured and unstructured data
- Work in an agile environment with emphasis on CI/CD and Application Resiliency Standards, partnering with Cyber & Security teams
- Collaborate across Agile teams to design, develop, test, implement, and support solutions using full-stack development tools and technologies
- Lead developers with experience across machine learning, distributed microservices, and full stack systems
- Use programming languages such as Java, Scala, and Python, along with open source RDBMS and NoSQL databases, plus cloud data warehousing services including Redshift and Snowflake
- Stay current with technology trends through experimentation, learning, participation in internal and external communities, and mentoring
- Collaborate with digital product managers to deliver robust cloud-based solutions supporting experiences for millions of Americans
- Perform unit tests and run code reviews to ensure code quality, performance tuning, and rigorous design
Requirements
- Bachelor’s Degree
- At least 6 years of experience in application development (internship experience does not apply)
- At least 2 years of experience with big data technologies
- At least 1 year of experience with cloud computing (AWS, Microsoft Azure, or Google Cloud)
- Master’s Degree
- 9+ years of experience in application development including Python, SQL, Scala, or Java
- 4+ years of experience with a public cloud (AWS, Microsoft Azure, or Google Cloud)
- 5+ years experience with distributed data or computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, Gurobi, or MySQL)
- 4+ years working on real-time data and streaming applications
- 4+ years of experience with NoSQL implementation (Mongo, Cassandra)
- 4+ years of data warehousing experience (Redshift or Snowflake)
- 4+ years of experience with UNIX/Linux including basic commands and shell scripting
- 2+ years of experience with Agile engineering practices
- Experience leveraging interactive AI tooling to accelerate productivity, using capabilities beyond basic code completion
Technologies
- Flink
- Spark Streaming
- Java
- Scala
- Python
- Spark
- Kafka
- Snowflake
- AWS Big Data Services
- Redshift
- Open Source RDBMS
- NoSQL databases
- Mongo
- Cassandra
- MapReduce
- Hadoop
- Hive
- EMR
- Gurobi
- MySQL
- UNIX/Linux
- Shell scripting
- Agile methodologies
- CI/CD
- SQL
- AWS
- Microsoft Azure
- Google Cloud
Benefits
- Performance-based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI)
- Comprehensive, competitive, and inclusive health, financial, and other benefits supporting total well-being
Salary
Plano, TX: $209,000 - $238,500 per year (Senior Lead Data Engineer)