Senior Lead AI Engineer (Gen AI Platform Services)
Job Description
Capital One is building responsible and reliable AI systems that aim to change banking for good. The Intelligent Foundations and Experiences (IFX) team sits at the center of turning AI research into production-ready capabilities that benefit millions of customers. This onsite role in New York offers a competitive compensation range of USD 250,800 to 286,200 per year and the opportunity to influence the long term roadmap for foundational AI systems.
You will partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that change how associates work and how customers interact with Capital One.
Responsibilities
- Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that change how our associates work and how our customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and introduce state of the art LLM optimization techniques to improve the performance, including scalability, cost, latency, and throughput, of large scale production AI systems.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Requirements
- The ideal candidate: You love to build systems, take pride in the quality of your work, and share our commitment to doing the right thing. You want to tackle problems that can drive meaningful change in banking.
- Research and production mindset: Passion for staying current with the latest research and the ability to intuitively understand scientific publications while judiciously applying novel techniques in production.
- Problem solving: You adapt quickly and bring clarity to large, undefined problems. You ask questions, dig deep to uncover root causes, and articulate findings clearly. You have the courage to propose new ideas, even if unproven.
- Technical foundation: You are deeply technical with a strong grounding in engineering and mathematics. Your expertise in hardware, software, and AI helps you identify optimization opportunities others may miss.
- Trailblazing mindset: You are a resilient pioneer who can forge new paths to achieve business goals when the route is unknown.
- Education and experience: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related field with at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in these fields with at least 4 years of relevant experience. At least 6 years of experience programming with Python, Go, Scala, or Java.
- Cloud experience: 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud).
- Leadership and collaboration: Demonstrated ability to lead and mentor an engineering team and influence cross-functional stakeholders.
- Algorithms and techniques: Experience developing AI and ML algorithms or technologies such as LLM inference, similarity search and VectorDBs, guardrails, memory, using Python, C++, C#, Java, or Golang; plus expertise in optimizing training and inference to improve hardware utilization, latency, throughput, and cost.
- Communication: Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers.