Senior Lead AI Engineer (AI Foundations, LLM Core and Agentic AI)
Job Description
Capital One is seeking a Senior Lead AI Engineer focused on AI foundations, LLM core and agentic AI in New York, onsite.
Responsibilities
- Collaborate with a cross functional team of engineers, research scientists, technical program managers, and product managers to deliver AI powered products that transform how our associates work and how our customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Hugging Face, VectorDBs, Nemo Guardrails, PyTorch, and more.
- Invent and apply state of the art LLM optimization techniques to improve production AI systems performance across scalability, cost, latency, and throughput.
- Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One.
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- Nemo Guardrails
- PyTorch
- Python
- Go
- Scala
- Java
Team Description
- The Intelligent Foundations and Experiences (IFX) team sits at the core of Capital One's AI strategy, partnering with stakeholders across the company to advance AI engineering and deploy proprietary solutions central to the business, delivering value to millions of customers.
- Our AI models and platforms empower teams across Capital One to enhance products with the transformative power of AI.
The Ideal Candidate
- Enjoys building systems, takes pride in code quality, and is motivated to do the right thing.
- Stays current with the latest AI research and can interpret scientific publications to judiciously apply novel techniques in production.
- Adapts quickly, brings clarity to big undefined problems, asks questions, digs deep, and communicates findings succinctly; willing to share new ideas even when unproven.
- Deeply technical with a solid foundation in engineering and mathematics; proficiency in hardware, software, and AI enables identifying optimization opportunities others may miss.
- Resilient innovator capable of forging new paths to achieve business goals when the route is uncertain.
Basic Qualifications
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related field with at least 6 years of experience developing AI and ML algorithms or technologies.
- Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related field with at least 4 years of experience developing AI and ML algorithms or technologies.
- At least 6 years of experience programming with Python, Go, Scala, or Java.
Preferred Qualifications
- 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (AWS, Google Cloud, Azure, or equivalent private cloud).
- Experience designing, developing, integrating, delivering, and supporting complex AI systems.
- Demonstrated ability to lead and mentor an engineering team and influence cross functional stakeholders.
- Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang.
- Experience developing and applying state of the art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost.
- Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production.
- Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers.
Compensation and Location
- Salary: USD 250,800 - 286,200 per year
- Location: New York, NY (onsite)