ML Data Engineer
Artificial Intelligence
Big Data
Bigdata
Cloud
Cloud Infrastructure
Cloud Native
Cloud Operations
Cloud Platform
Cloud Platforms
Cloud Platforms Cloud Platforms
Cloud Technology
Data Analysis
Data Engineer
Data Pipeline
Data Platform
Data Processing
Data Science
Database
DevOps
DevSecOps
Engineer
Engineering
Engineering Software
Google Cloud
Hadoop
Information Technology (IT)
Kubernetes
Machine Learning Engineer
OpenShift
Platform Engineering
Programming Language
Programming Languages
Pyspark
Spark
Job Description
This role focuses on building and operationalizing AI/ML model pipelines across on-prem and cloud systems.
Responsibilities
- Partner with data scientists and data engineers to convert prototypes and theoretical work into production-ready code
- Collaborate with product and engineering teams to implement AI/ML models using PySpark
- Productionalize AI/ML models within a Hadoop environment
- Troubleshoot AI/ML models after implementation to address issues raised by users
- Monitor model performance and continuously improve efficiency, accuracy, and user experience
Requirements
- Strong experience with PySpark and Python to integrate with AI/ML models
- Understanding of AI/ML models and how they integrate with on-prem and cloud environments
- Familiarity with MLOps/LLMOps and distributed systems
- Experience with big data platforms such as Cloudera Hadoop and cloud platforms such as AWS and GCP
- Solid understanding of system design patterns, scalability, observability, and performance tuning
- Strong analytical and problem-solving skills
- Interest in exploring and building with emerging technologies
Tech Stack
- Python, PySpark
- OpenShift, Kubernetes
- Hadoop, Cloudera Hadoop
- AWS, GCP
Location & Compensation
- Location: Irving, TX (onsite)
- Salary: USD 124,560 per year
Note: The structured data lists additional required skills as “DevOps Engineer Senior Email Security Engineer,” but no further details are provided.