Data Scientist
Job Description
Data Scientist role with PepsiCo, onsite in Plano, TX, focused on building analytics-driven solutions for consumer topics in the Consumer Activation space.
Responsibilities
- Serve as a contributing member on digital initiatives within PepsiCo's analytics scope.
- Present findings to stakeholders in business friendly terms.
- Develop deployable machine learning models with limited supervision.
- Act as a subject matter expert on digital projects for other teams.
- Engage in innovation activities and help advance new approaches.
- Collaborate with product managers to provide feedback on user stories and tickets.
- Work with data engineers to ensure data access for discovery and prepare data for model use.
- Collaborate with business teams and other IT services as needed.
- Promote the Platform toolset and deliver demonstrations of possible solutions to the business.
- Interface with business stakeholders during service design, training, and knowledge transfer.
- Support large-scale experimentation and build data-driven models.
- Define KPIs and metrics to evaluate analytics solutions for each use case.
- Translate requirements into modelling problems and analytic approaches.
- Influence product teams with data-based recommendations.
- Conduct research on state-of-the-art methodologies and publish learnings.
- Document learnings for knowledge transfer and future reuse.
- Assist in creating reusable packages or libraries for broader use.
Requirements
- At least 3+ years applying statistical and ML techniques to supervised (regression, classification) and unsupervised problems.
- At least 3+ years developing business problem related statistical/ML models with primary focus on Python development.
- At least 2+ years using modern cloud platforms to power ML solutions.
- At least 2+ years in a product focused team delivering production level analytics.
- Ability to tell a data story and communicate insights in business-ready formats; fluent in one visualization tool.
- Strong communication and organizational skills with the ability to manage ambiguity and multiple priorities.
- Experience with Agile methodology for team work and product creation.
- Experience in Reinforcement Learning is a must.
- Experience with PySpark is a must.
- Experience with SQL is a must.
- Fluent in Git; familiarity with Jenkins or Docker is a plus.
- Experience with NLP is a plus.
- Experience with LLM is a plus.
- Experience with FAIR data practices is a plus.
- Experience with Responsible AI is a plus.
- Experience with distributed machine learning is a plus.
Technologies
- Python
- PySpark
- SQL
- Git
- Jenkins
- Docker
- NLP
- LLM
- FAIR data
- Responsible AI
- Distributed machine learning
- Visualization tool
Benefits
- Salary range: USD 80,200 - 134,250 per year.
- Starting salary determined by location, skills, experience and education; recruiter will share specifics during hiring.
- Bonus opportunity based on performance, with target payout of 8% of annual salary.
- Paid time off eligibility including parental leave, vacation, sick time and bereavement.
- Comprehensive benefits package: Medical, Dental, Vision, Disability, Health and Dependent Care Reimbursement Accounts, EAP, Insurance (Accident, Group Legal, Life), and a Defined Contribution Retirement Plan.