Data Engineer 4
Job Description
Capital One is building cloud-first data platforms that help power reliable, secure experiences for millions of Americans. As a Data Engineer 4 in McLean, VA (onsite), you will design, develop, implement, and support scalable data pipelines and engineering platforms using Python, Spark, and AWS-related and data warehousing technologies such as Databricks and Snowflake.
In this role, you’ll have the opportunity to influence how data engineering patterns are implemented across platforms and pipelines, support full-stack development workflows across Agile teams, and help ensure robust performance as data volume and business demand grow. You’ll also contribute to a culture of continuous learning by experimenting with new technologies, sharing knowledge in internal and external communities, and mentoring other members of the data community.
Responsibilities
- Collaborate with Agile teams to design, develop, test, implement, and support technical solutions across full-stack development tools and technologies
- Influence developers, data analysts, and data scientists with expertise in machine learning, distributed microservices, lakehouse architecture, and full-stack systems
- Use Python and Spark alongside open-source relational and NoSQL databases and cloud data warehousing platforms like Databricks and Snowflake
- Partner with product managers and software engineers to deliver cloud-first data solutions that support powerful experiences
- Independently design, build, and deliver cloud data solutions and applications with little or no support from supervisors or managers
- Architect and enforce common data engineering design patterns to improve code quality, maintainability, and reusability
- Design and build data pipelines and platforms focused on scalability, resilience, and operational efficiency
- Implement data security standards, including encryption at rest/transit and fine-grained access control, to support compliance with data privacy regulations
- Serve as a data engineering ambassador by clearly communicating technical concepts and data outcomes to internal and external stakeholders
Requirements
- Bachelor’s Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering)
- 4+ years of experience in application development (internship experience does not apply)
- 2+ years of experience in distributed data
- 2+ years of experience with SQL
- 2+ years of experience with one programming language (Python, Java, or Scala)
- 2+ years of experience in data pipeline design and development
- 1+ year of experience in data modeling and designing end-to-end data solutions using both relational and non-relational database systems
Technologies
- Python, Spark, Databricks, Snowflake, SQL
- EMR, Glue, Airflow, Dagster
- NoSQL, MongoDB, Cassandra, DynamoDB
- Redshift, AWS
- Microsoft Azure, Google Cloud
- Monte Carlo, Splunk
Benefits
- Comprehensive, competitive, and inclusive set of health, financial and other benefits
Preferred Qualifications
- 7+ years of experience in application development with demonstrated proficiency in Python, SQL, Scala, or Java
- 4+ years of hands-on experience designing, deploying and operating data workloads in at least one public cloud environment (AWS, Microsoft Azure, or Google Cloud)
- 4+ years of experience building or supporting distributed data or compute workloads using tools such as EMR, Spark, Glue, or Databricks
- 4+ years of experience designing, implementing, and operating real-time or streaming data pipelines
- 2+ years of experience working on data observability (e.g., Monte Carlo, Splunk) or data orchestration tools (e.g., Airflow, Dagster)
- 4+ years of experience working with unstructured or semistructured data using NoSQL databases (e.g., MongoDB, Cassandra, DynamoDB)
- 4+ years of experience designing and supporting data warehousing solutions (e.g., Snowflake, Redshift)
- 2+ years of experience working in an Agile development environment
- 2+ years of experience developing user-centric reusable data products
Salary Range: USD 197,300 - 225,100 per year.
Work Authorization Note: Capital One will not sponsor a new applicant for employment authorization, or offer any immigration related support for this position (including H1B, F-1 OPT, F-1 STEM OPT, F-1 CPT, J-1, TN, E-3, and O-1, or any other forms of work authorization that require immigration support from an employer).