AI Engineer 4
Job Description
Capital One is hiring an AI Engineer 4 in New York, NY (onsite) to build and deploy responsible, scalable AI systems for banking experiences.
Responsibilities
- Work with cross-functional teams (engineers, research scientists, technical program managers, and product managers) to deliver AI-powered products for associates and customers
- Design, develop, test, deploy, and support AI software components, including:
- foundation model training
- large language model inference
- agents and multi-agent workflows
- similarity search
- guardrails
- model evaluation
- experimentation
- governance
- observability
- Use open source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, and VectorDBs, along with PyTorch and other tools
- Develop state-of-the-art foundation model optimization methods to improve production performance across scalability, cost, latency, and throughput
- Help shape the technical vision and long-term roadmap for foundational AI systems at Capital One
- Own end-to-end architecture for complex AI systems with focus on maintainability, observability, and ethical alignment
- Define and maintain service-level objectives (SLOs) for AI reliability, covering latency, uptime, and model performance drift
- Partner with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines
- Lead cross-functional technical reviews for new AI deployments, ensuring security, data governance, and compliance standards are met
- Mentor Principal and Senior Associates on scalable design, performance tuning, and translating research into production
Requirements
- Education + experience:
- Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years developing AI and ML algorithms or technologies; or
- Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years developing AI and ML algorithms or technologies
- Programming: at least 4 years of experience with Python, Go, Scala, CUDA, or Java
- AI/ML development: at least 4 years of experience developing AI and ML algorithms or technologies
Technologies
- AWS Ultraclusters
- Huggingface
- VectorDBs
- PyTorch
- Python
- Go
- Scala
- CUDA
- Java
- Google Cloud
- Azure
- C++
- C#
- Golang
- GPU/TPU
Benefits
- Comprehensive, competitive, and inclusive set of health, financial, and other benefits supporting total well-being
Preferred Qualifications
- Experience leading development of AI systems with tradeoffs among cost, latency, throughput, and accuracy
- 6+ years deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud)
- Experience designing, developing, delivering, and supporting AI services
- Experience developing AI and ML algorithms or technologies (e.g., LLM inference, similarity search and VectorDBs, guardrails, memory) using Python, C++, C#, Java, CUDA, or Golang
- Experience building and applying state-of-the-art optimization techniques for training and inference to improve hardware utilization, latency, throughput, and cost
- Experience building agentic AI systems and agentic workflows
- Proficiency designing distributed systems for model training, evaluation, and online inference at petabyte scale
- Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules
- Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms
Salary: USD 215,200 - 245,600 per year
Location: New York, NY (onsite)