Senior Lead AI Engineer(MLX, Agentic AI, Gen AI platform Services)
Job Description
The Senior Lead AI Engineer at Capital One designs, deploys, and optimizes AI software components and platforms, leading cross-functional teams to deliver AI powered banking solutions for associates and customers. This on-site role is based in McLean, Virginia.
Responsibilities
- Collaborate with engineers, research scientists, technical program managers, and product managers to deliver AI powered products that transform how associates work and how customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and related tools.
- Invent and apply state-of-the-art LLM optimization techniques to improve performance metrics such as scalability, cost, latency, and throughput in large scale production AI systems.
- Contribute to the technical vision and the long term roadmap for foundational AI systems at Capital One.
Requirements
- A bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields with at least 6 years of experience developing AI and ML algorithms or technologies, or a master’s degree in the same fields with at least 4 years of experience developing AI and ML algorithms or technologies.
- At least 6 years of experience programming with Python, Go, Scala, or Java.
- 7 years of experience deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, Azure, or an equivalent private cloud.
- Experience designing, developing, integrating, delivering, and supporting complex AI systems.
- Proven ability to lead and mentor an engineering team and influence cross-functional stakeholders.
- Experience developing AI and ML algorithms or technologies (for example LLM inference, similarity search and vector databases, guardrails, memory) using languages including Python, C++, C#, Java, or Golang.
- Experience optimizing training and inference software to improve hardware utilization, latency, throughput, and cost.
- Commitment to staying current with AI research and productionizing novel techniques with sound judgment.
- Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers.
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- Nemo Guardrails
- PyTorch
- Python
- Go
- Scala
- Java
- C++
- C#
Location and Compensation
Location: McLean, Virginia, onsite. Salary range: USD 229,900 - 262,400 per year. Compensation may vary by location and will be confirmed in the offer letter.
Salary Ranges by Location
The following ranges apply to full-time roles in the listed locations. The actual offer amount will be reflected in the candidate’s offer letter.
- Cambridge, MA: $229,900 - $262,400
- McLean, VA: $229,900 - $262,400
- New York, NY: $250,800 - $286,200
- San Francisco, CA: $250,800 - $286,200
- San Jose, CA: $250,800 - $286,200
Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate’s offer letter.