Senior Distinguished AI Engineer
Job Description
Capital One is hiring a Senior Distinguished AI Engineer (onsite) to build and deploy responsible, scalable AI systems on the Intelligent Foundations and Experiences (IFX) team.
Responsibilities
- Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products for associates and customers
- Design, develop, test, deploy, and support AI software components including:
- foundation model training
- large language model inference
- similarity search
- guardrails
- model evaluation
- experimentation
- governance
- observability
- Use a broad stack of open source and SaaS AI technologies, including AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, and PyTorch
- Introduce state-of-the-art LLM optimization techniques to improve production performance, including scalability, cost, latency, and throughput
- Contribute to the technical vision and the long-term roadmap for foundational AI systems at Capital One
Requirements
-
Education and experience:
- Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus 10+ years developing AI/ML algorithms or technologies, or
- Master’s degree in a related field plus 8+ years developing AI/ML algorithms or technologies
- 10+ years programming with Python, Go, Scala, or Java
Technologies
- AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch
- AWS, Google Cloud, Azure
- Python, Go, Scala, Java, C++, C#, Golang
- LLM Inference, Similarity Search, Guardrails, Memory
Preferred Qualifications
- 9+ years deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud)
- Experience architecting, designing, developing, integrating, delivering, and supporting complex enterprise AI systems
- Demonstrated ability to lead and mentor an engineering organization and influence cross-functional stakeholders up to the SVP level
- Experience developing AI/ML algorithms or technologies (e.g., LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang
- Experience applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost
- Strong interest in staying current with AI research and applying novel techniques in production
- Excellent communication and presentation skills for articulating complex AI concepts
Additional eligibility and workplace notes
- Capital One may sponsor authorization for a new qualified applicant for employment
- Eligible for performance-based incentive compensation (cash bonus(es) and/or LTI)
- Applications are expected to be accepted for a minimum of 5 business days
- No agencies
- Equal opportunity employer (EOE, including disability/vet) and committed to non-discrimination under applicable laws
- Drug-free workplace
- No guarantee or liability for third-party products, services, educational tools, or other information available through the site
- Capital One Financial consists of multiple entities; location-specific postings correspond to the relevant entity (e.g., Canada, UK, Philippines)
Compensation
Estimated salary range (USD) for this location:
- New York, NY (onsite): $343,400 - $392,000 per year
Additional ranges listed for other locations:
- Cambridge, MA: $314,800 - $359,300
- McLean, VA: $314,800 - $359,300
- Richmond, VA: $286,200 - $326,700
- San Francisco, CA: $343,400 - $392,000
- San Jose, CA: $343,400 - $392,000