DataJobs.io
← Back to all jobs

Job Description

Capital One is hiring an AI Engineer 4 in New York, NY (onsite) to build and deploy responsible, scalable AI systems for banking experiences.

Responsibilities

  • Work with cross-functional teams (engineers, research scientists, technical program managers, and product managers) to deliver AI-powered products for associates and customers
  • Design, develop, test, deploy, and support AI software components, including:
    • foundation model training
    • large language model inference
    • agents and multi-agent workflows
    • similarity search
    • guardrails
    • model evaluation
    • experimentation
    • governance
    • observability
  • Use open source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, and VectorDBs, along with PyTorch and other tools
  • Develop state-of-the-art foundation model optimization methods to improve production performance across scalability, cost, latency, and throughput
  • Help shape the technical vision and long-term roadmap for foundational AI systems at Capital One
  • Own end-to-end architecture for complex AI systems with focus on maintainability, observability, and ethical alignment
  • Define and maintain service-level objectives (SLOs) for AI reliability, covering latency, uptime, and model performance drift
  • Partner with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines
  • Lead cross-functional technical reviews for new AI deployments, ensuring security, data governance, and compliance standards are met
  • Mentor Principal and Senior Associates on scalable design, performance tuning, and translating research into production

Requirements

  • Education + experience:
    • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years developing AI and ML algorithms or technologies; or
    • Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years developing AI and ML algorithms or technologies
  • Programming: at least 4 years of experience with Python, Go, Scala, CUDA, or Java
  • AI/ML development: at least 4 years of experience developing AI and ML algorithms or technologies

Technologies

  • AWS Ultraclusters
  • Huggingface
  • VectorDBs
  • PyTorch
  • Python
  • Go
  • Scala
  • CUDA
  • Java
  • Google Cloud
  • Azure
  • C++
  • C#
  • Golang
  • GPU/TPU

Benefits

  • Comprehensive, competitive, and inclusive set of health, financial, and other benefits supporting total well-being

Preferred Qualifications

  • Experience leading development of AI systems with tradeoffs among cost, latency, throughput, and accuracy
  • 6+ years deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud)
  • Experience designing, developing, delivering, and supporting AI services
  • Experience developing AI and ML algorithms or technologies (e.g., LLM inference, similarity search and VectorDBs, guardrails, memory) using Python, C++, C#, Java, CUDA, or Golang
  • Experience building and applying state-of-the-art optimization techniques for training and inference to improve hardware utilization, latency, throughput, and cost
  • Experience building agentic AI systems and agentic workflows
  • Proficiency designing distributed systems for model training, evaluation, and online inference at petabyte scale
  • Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules
  • Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms

Salary: USD 215,200 - 245,600 per year

Location: New York, NY (onsite)

Similar Jobs