Lead Data Engineer - Data Scientist
Job Description
Wells Fargo is seeking a Lead Data Engineer to drive data engineering and AI/ML efforts in Cybersecurity within Identity and Access Management (IAM). The role focuses on building solutions and pipelines that strengthen access governance and help detect anomalous identity events, using statistical and data science techniques in support of controls and operational processes.
Key Responsibilities
- Lead complex initiatives with broad impact and act as a key participant in large-scale software planning for Identity and Access Management.
- Design, develop, and run tooling to discover issues in data and applications, then report findings to engineering and product leadership.
- Apply statistical and data science methods to business problems related to Identity and Access Management.
- Serve as a subject matter expert on ML and AI, including the use of mathematical and statistical techniques on large datasets.
- Design, support, and operate data pipelines, data models, dashboards, and API integrations for real-time and batch analytics use cases.
- Design and conduct experiments, statistical analyses, and hypothesis testing to evaluate proposals supporting controls, policies, and operational processes.
- Build AI-powered capabilities to identify inappropriate access, recommend entitlements prior to user requests, and detect anomalous behavior across millions of identity events.
- Lead other IAM team members, including operations, onboarding, initiatives, and engineering teams, as well as line of business and lines of defense teams, to capture analytical needs and translate them into technical solutions.
- Establish and implement engineering and analytical best practices when developing solutions.
- Develop solutions in accordance with security, privacy, model risk, and regulatory guidelines.
- Develop, test, deploy, and support ML-enabled analytical solutions, and guide team members on ML, AI, and statistical techniques to drive adoption of AI/ML across IAM.
- Assist in monitoring model health, reliability, and drift to support continuous improvement and required remediation.
- Demonstrate proficiency using AI-assisted development and analysis tools (for example, GitHub Copilot and approved code-centric agents).
- Leverage AI to accelerate system design, coding, testing, analysis, and troubleshooting, while validating and integrating AI-assisted outputs using strong technical judgment.
- Understand and account for model limitations, security risks, and operational considerations, and ensure AI use aligns with security, compliance, privacy, and ethical standards in both development and production environments.
Required Qualifications
- 5+ years of Database Engineering experience, or an equivalent combination of work experience, training, military experience, or education.
- 5+ years of experience with Python data science libraries including Pandas, NumPy, and Scikit-Learn, along with exposure to a deep learning framework such as TensorFlow or PyTorch.
Technologies
- Python, Pandas, NumPy, Scikit-Learn
- TensorFlow, PyTorch
- Vertex AI, GCP, BigQuery
- Neo4j
- GitHub
- Power BI, Tableau, Alteryx
- GitHub Copilot, Claude Code, Cursor, Devin
Desired Qualifications
- Knowledge of Vertex AI and GCP environments, including BigQuery and getting models into production on these platforms.
- Knowledge of graph networks for anomaly detection (Neo4j).
- A learning mindset to stay current with developments in the field.
- Knowledge of GitHub for code management, Power BI, Tableau, and Alteryx.
- Understanding of software engineering fundamentals such as testing, debugging, and code reviews is highly desirable.
- Knowledge of the mathematical foundations of statistics, machine learning, and modern AI techniques.
- Strong understanding of model evaluation techniques for classification and clustering models, including precision, accuracy, F1-scores, confusion matrices, and ROC curves.
- High comfort with AI coding agents such as GitHub Copilot, Claude Code, Cursor, Devin, or others, with the ability to critically examine AI code for fitness for use.
- Familiarity with LLMs, RAG, and AI agents.
- Strong SQL skills and the ability to work with relational and analytical databases.
- Strong organizational skills for work management.
- Hands-on ability to manipulate data and tools to prototype and present solutions.
- Confident and self-motivated approach to developing original ideas and solutions, with sound judgment for when to escalate issues.
- Strong written and verbal communication skills, including presentation skills.
- Ability to communicate complex technical concepts to colleagues and team members.
- Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Engineering, or related field.
Compensation and Location
Location: 401 Las Colinas Blvd W Bldg. A - Irving, TX 75039 (onsite). Required location(s) also include 300 S Brevard, Charlotte, NC 28202; 550 S 4th St, Minneapolis, MN 55415; and 3075 Loyalty Cir. - Columbus, OH 43219.
Pay range: $119,000.00 - $206,000.00 per year. Salary range is determined by the location of the job. May be considered for a discretionary bonus, Restricted Share Rights, or other long-term incentive awards.
Benefits
- Health benefits
- 401(k) Plan
- Paid time off
- Disability benefits
- Life insurance, critical illness insurance, and accident insurance
- Parental leave
- Critical caregiving leave
- Discounts and savings
- Commuter benefits
- Tuition reimbursement
- Scholarships for dependent children
- Adoption reimbursement
Posting Statements
- Job posting may come down early due to volume of applicants.
- Required location(s) are listed above. Relocation assistance is not available for this position.
- This position is not eligible for visa sponsorship.
- This role is NOT eligible for 100% remote work.
- Salary range is determined by location of the job and may be considered for discretionary bonus, Restricted Share Rights, or other long-term incentive awards.
Posting End Date
22 Sep 2026