Lead Data Engineer - Data Scientist
Backend Developer
Manager
Analytics
Artificial Intelligence
Big Data
Business Intelligence
Cloud
Cloud Native
Cloud Platforms
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Science
Data Science Ml
Data Visualization
Database
Databases
Digital Marketing
ETL
Generative AI
Google Cloud
Google Cloud Bigquery
Google Cloud Platform
Graph Database
Informatica
Information Technology (IT)
Integration
Large Language Models
Machine Learning
Power Bi & Tableau
Reporting and Analytics
SQL
Job Description
Lead data engineering efforts in Cybersecurity within Identity Access Management, building data science and ML-powered solutions for access governance and anomaly detection.
Responsibilities
- Lead complex initiatives with broad impact and participate in large-scale software planning for Identity and Access Management (IAM).
- Design, develop, and run tooling to discover issues in data and applications, then report findings to engineering and product leadership.
- Apply statistical and data science methods to IAM business problems.
- Act as a subject matter expert on ML, AI, and mathematical and statistical techniques applied to large datasets.
- Design, support, and operate data pipelines, data models, dashboards, and API integrations for real-time and batch analytics use cases.
- Design and conduct experiments, statistical analyses, and hypothesis testing to evaluate proposals supporting controls, policies, and operational processes.
- Build AI-powered capabilities to identify inappropriate access, recommend entitlements before users request them, and detect anomalous behavior across millions of identity events.
- Lead IAM team members across operations, onboarding, initiatives, and engineering, plus line of business and lines of defense teams, to translate analytical needs into technical solutions.
- Establish and implement engineering and analytical best practices for solution development.
- Develop solutions in alignment with security, privacy, model risk, and regulatory guidelines.
- Develop, test, deploy, and support ML-enabled analytical solutions, and guide teammates on ML, AI, and statistical techniques to drive AI/ML adoption across IAM.
- Monitor model health, reliability, and drift and help ensure continuous improvement and required remediation.
- Use AI-assisted development and analysis tools (including GitHub Copilot and approved code-centric agents) to accelerate system design, coding, testing, analysis, and troubleshooting.
- Validate and integrate AI-assisted outputs with strong technical judgment, accounting for model limitations, security risks, and operational considerations.
- Apply AI responsibly in development and production, ensuring alignment with security, compliance, privacy, and ethical standards.
Requirements
- 5+ years of Database Engineering experience, or equivalent demonstrated through one or a combination of: work experience, training, military experience, or education.
- 5+ years of experience with Python data science libraries including Pandas, NumPy, and Scikit-Learn, plus exposure to a deep learning framework such as TensorFlow or PyTorch.
Technologies
- Python, Pandas, NumPy, Scikit-Learn, TensorFlow, PyTorch
- Vertex AI, GCP, BigQuery
- Neo4j
- GitHub, Power BI, Tableau, Alteryx
- GitHub Copilot, Claude Code, Cursor, Devin, LLMs, RAG, AI agents
Desired Qualifications
- Knowledge of Vertex AI and GCP environments, including BigQuery and getting models into production on these platforms.
- Knowledge of graph networks for anomaly detection using Neo4j.
- A learning mindset to keep up to date with developments in the field.
- Knowledge of GitHub for code management and familiarity with Power BI, Tableau, and Alteryx.
- Understanding of software engineering fundamentals including testing, debugging, and code reviews is highly desirable.
- Knowledge of mathematical foundations of statistics, machine learning, and modern AI techniques.
- Strong understanding of model evaluation techniques for classification and clustering, including precision, accuracy, F1-scores, confusion matrices, and ROC curves.
- High comfort using AI coding agents such as GitHub Copilot, Claude Code, Cursor, Devin, or others, with the ability to critically examine AI code for fitness for use.
- Familiarity with LLMs, RAG, and AI agents.
- Strong SQL skills and ability to work with relational and analytical databases.
- Strong book-of-work management and organizational skills.
- Hands-on ability to manipulate data and tools to prototype and present solutions.
- Confident and self-motivated producer of original ideas and solutions, with sound judgment for when to escalate issues.
- Strong written and verbal communicator with excellent presentation skills, including ability to communicate complex technical concepts.
- Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Engineering, or a related field.
Benefits
- Health benefits
- 401(k) Plan
- Paid time off
- Disability benefits
- Life insurance, critical illness insurance, and accident insurance
- Parental leave
- Critical caregiving leave
- Discounts and savings
- Commuter benefits
- Tuition reimbursement
- Scholarships for dependent children
- Adoption reimbursement
Location and Work Mode
- Columbus, Ohio (onsite)
- 3075 Loyalty Cir. - Columbus, Ohio 43219
Compensation
- Salary range: $119,000 - $206,000 per year (range determined by location)
- May be considered for a discretionary bonus, Restricted Share Rights, or other long-term incentive awards
Posting Statements
- Job posting may come down early due to volume of applicants.
- Required location(s) listed above; relocation assistance is not available.
- Not eligible for visa sponsorship.
- Role is not eligible for 100% remote work.
- Posting end date: 22 Sep 2026.