Lead Data Engineer - Data Scientist
Job Description
Lead Data Engineer in Cybersecurity within Identity Access Management at Wells Fargo, building intelligent access governance and production analytics.
Responsibilities
- Lead complex initiatives with broad impact and contribute to large-scale software planning for Identity and Access management
- Design, develop, and run tooling to discover problems in data and applications, then report findings to engineering and product leadership
- Apply statistical and data science methods to business problems in Identity and Access Management
- Act as a subject matter expert in ML and AI, including mathematical and statistical techniques for large datasets
- Design, support, and operate data pipelines, data models, dashboards, and API integrations for real-time and batch analytics use cases
- Design and conduct experiments, statistical analyses, and hypothesis testing to evaluate proposals for controls, policies, and operational processes
- Build AI-powered capabilities to identify inappropriate access, recommend entitlements before user requests, and detect anomalous behavior across millions of identity events
- Lead IAM team members across operations, onboarding, initiatives, and engineering, plus line of business and lines of defense stakeholders to translate analytical needs into technical solutions
- Establish and implement engineering and analytical best practices during solution development
- Develop solutions in alignment with security, privacy, model risk, and regulatory guidelines
- Develop, test, deploy, and support ML-enabled analytical solutions, and guide team adoption of ML, AI, and statistical techniques for IAM use cases
- Monitor model health, reliability, and drift, and drive required remediation
- Use AI-assisted development and analysis tools (for example GitHub Copilot and approved code-centric agents)
- Leverage AI to accelerate system design, coding, testing, analysis, and troubleshooting
- Validate and integrate AI-assisted outputs using strong technical judgment, accounting for limitations, security risks, and operational considerations
- Ensure responsible AI use in development and production environments, aligned to security, compliance, privacy, and ethical standards
Requirements
- 5+ years of database engineering experience (or equivalent through work experience, training, military experience, or education)
- 5+ years with Python data science libraries including Pandas, NumPy, and Scikit-Learn, plus exposure to a deep learning framework such as TensorFlow or PyTorch
- Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Engineering, or related field
Technologies
- Python, Pandas, NumPy, Scikit-Learn
- TensorFlow, PyTorch
- Vertex AI, GCP, BigQuery
- Neo4j
- GitHub, Power BI, Tableau, Alteryx
- GitHub Copilot, Claude Code, Cursor, Devin
- LLMs, RAG, AI agents
Benefits
- Health benefits
- 401(k) Plan
- Paid time off
- Disability benefits
- Life insurance, critical illness insurance, and accident insurance
- Parental leave
- Critical caregiving leave
- Discounts and savings
- Commuter benefits
- Tuition reimbursement
- Scholarships for dependent children
- Adoption reimbursement
Locations
- 300 S Brevard, Charlotte, North Carolina 28202
- 401 Las Colinas Blvd W Bldg. A - Irving, TX 75039
- 550 S 4th St. Minneapolis, MN 55415
- 3075 Loyalty Cir. - Columbus, Ohio 43219
Desired Qualifications
- Knowledge of Vertex AI and GCP environments, BigQuery, and getting models into production on these platforms
- Knowledge of graph networks for anomaly detection (Neo4j)
- Learning mindset to stay current with developments in the field
- Knowledge of GitHub for code management and experience with Power BI, Tableau, Alteryx
- Understanding of software engineering fundamentals including testing, debugging, and code reviews is highly desirable
- Knowledge of the mathematical foundations of statistics, machine learning, and modern AI techniques
- Strong understanding of model evaluation techniques for classification and clustering, including metrics such as precision, accuracy, F1-scores, confusion matrices, and ROC curves
- Comfort using AI coding agents like GitHub Copilot, Claude Code, Cursor, Devin (or others), with the ability to critically examine AI code for fitness for use
- Familiarity with LLMs, RAG, and AI agents
- Strong SQL skills and ability to work with relational and analytical databases
- Strong book-of-work management and organizational skills
- Hands-on ability to manipulate data and tools to prototype and present solutions
- Confident, self-motivated producer of original ideas and solutions, with sound judgment for when to escalate issues
- Strong written and verbal communication with excellent presentation skills
- Ability to communicate complex technical concepts to colleagues and team members
Salary range: USD 119,000 - 206,000 per yearly
Posting end date: 22 Sep 2026
Posting Statements
- Job posting may come down early due to volume of applicants
- Required locations listed above; relocation assistance is not available for this position
- Salary range is determined by location of the job; may be considered for a discretionary bonus, Restricted Share Rights, or other long-term incentive awards
- This position is not eligible for visa sponsorship
- This role is NOT eligible for 100% remote work
Additional Information
- To request a medical accommodation during the application or interview process, visit Disability Inclusion at Wells Fargo
- Wells Fargo maintains a drug-free workplace; review the Drug and Alcohol Policy to learn more
- Third-party recordings are prohibited unless authorized by Wells Fargo
- Wells Fargo requires candidates to directly represent their own experiences during recruiting and hiring