Data Scientist
Job Description
This hybrid, Plano, TX based role centers on designing, building and deploying ML and AI solutions. You will apply expertise in LLMs, NLP and classical data science algorithms, with hands-on work in Python, Spark and Databricks, and Azure cloud exposure for training and deployment. You will collaborate with engineering and business teams to deliver AI-enabled outcomes.
Responsibilities
- Develop and deploy machine learning and AI models, including large language models, NLP solutions and traditional algorithms, with emphasis on training, validation and production readiness.
- Build scalable data pipelines using PySpark and the Spark ecosystem, including Databricks.
- Fine-tune LLMs, engineer prompts and evaluate model performance against defined metrics.
- Work with data stored in SQL and NoSQL databases such as PostgreSQL and MongoDB.
- Partner with engineering and business teams to implement AI solutions that drive outcomes.
- Establish model governance, monitor performance and manage the lifecycle of deployed models.
Requirements
- 5 to 8 years of experience in data science, machine learning or AI engineering.
- Strong expertise in large language models, NLP and classical ML algorithms.
- Hands-on proficiency with Python, PySpark, Spark, Databricks and SQL/NoSQL databases.
- Experience with Azure Cloud for training and deploying ML models.
- Knowledge of fine-tuning and optimization techniques for ML models.
- Strong analytical, problem-solving and communication skills.
Technologies
- Python
- PySpark
- Spark
- Databricks
- SQL
- NoSQL
- PostgreSQL
- MongoDB
- Azure Cloud