This position is no longer accepting applications
Closed on August 30, 2026.
This role is filled — get an email when new Data Analysis roles open on DataJobs.io:
Data Scientist II
Artificial Intelligence
Big Data
Bigdata
Cloud Platform
Cloud Platforms
Data Analysis
Data Analytics
Data Pipeline
Data Platform
Data Processing
Data Science
Databricks
Deep Learning
Large Language Models
Machine Learning
NumPy
Pyspark
PyTorch
scikit-learn
SQL
View similar jobs
Get alerted when similar jobs are posted — set up a New Data Analysis jobs on DataJobs.io alert.
See other roles at Scribd, Inc..
Job Description
Data Scientist II at Scribd, Inc. is tasked with designing and deploying high-impact AI and ML systems, collaborating with cross-functional teams to build models from large-scale data and language models.
Position Details
- Location: Dallas, TX (onsite)
- Salary: USD 118,000 - 150,000 per year
- Experience: 3+ years of post-qualification experience
- Education: Bachelor's or Master's degree
Responsibilities
- Work on content classification use cases across traditional NLP, large language models, and generative systems
- Explore scalable approaches to address Scribd's most challenging problems
- Collaborate with Data Scientists, ML Engineers, and ML Data Engineers on cross-functional projects
- Utilize a range of algorithms from classical Scikit-learn and NumPy models to PyTorch neural networks and third party LLM APIs
- Process large volumes of data with Python, SQL, and Spark
- Communicate approaches and results to stakeholders in writing and verbally, and maintain precise project documentation
Requirements
- 3+ years of post-qualification experience building ML models, operating at scale, and deploying to production environments
- Proficiency in Python
- Hands-on experience developing ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar
- Intermediate level in at least three areas: classification algorithms, natural language processing, search and information retrieval, named entity recognition, deep learning, or generative models
- Intermediate level or greater experience with SQL or PySpark
- Bachelors or Masters in a quantitative field such as Statistics, Computer Science, Data Science, Artificial Intelligence, or another discipline with a strong quantitative focus
Technologies
- Python
- Apache Spark
- Databricks
- Scikit-learn
- NumPy
- PyTorch
- SQL
- PySpark
- Third party LLM APIs
Benefits
- Scribd Flex (flexible work model)
- Comprehensive health, dental, and vision coverage
- Mental health support and disability coverage
- Generous paid time off, including vacation, sick time, holidays, winter break, volunteer time, and sabbaticals
- Paid parental leave and family support benefits
- Retirement matching and employee equity
- Learning and development programs and professional growth opportunities
- Wellness and home office stipends
- Complimentary access to the Scribd, Inc. suite of products
- Enterprise access to leading AI tools