Data Engineer
Backend Developer
Analytics
Azure Data Lake
Big Data
Bigdata
Business Intelligence
Cloud Operations
Cloud Platform
Cloud Platforms
Data
Data Analysis
Data Architecture
Data Engineer
Data Integration
Data Lake
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Database
Databases
Digital Marketing
ETL
Hadoop
Hdfs
Hive
Informatica
Information Technology (IT)
Programming Language
Programming Languages
Reporting and Analytics
Spark
SQL
Talend
Job Description
Startek Pro Technologies is hiring a Data Engineer in Dallas, TX (onsite) to help build scalable data pipelines and architectures that power analytics and machine learning. This role centers on engineering high-quality, accessible data through robust ETL, data modeling, and distributed big data systems, while collaborating across teams to turn data needs into solutions on AWS, Azure Data Lake, or other public cloud services. The hourly compensation range is USD 45–55/hour.
What you’ll do
- Develop, implement, and optimize ETL pipelines to move and transform data from multiple sources into centralized data warehouses or data lakes.
- Design and maintain data models and schemas using dimensional modeling to support business intelligence and analytics needs.
- Manage large-scale data processing using Hadoop (including HDFS), Spark, Hive, and other distributed processing frameworks.
- Collaborate with cross-functional teams to understand data requirements and translate them into scalable solutions using cloud platforms such as AWS and Azure Data Lake.
- Integrate diverse datasets from SQL databases (including Microsoft SQL Server and Oracle), cloud databases, and linked data environments to form unified datasets.
- Write efficient SQL queries and automation scripts using Python or Bash to support analysis, troubleshooting, and data workflow reliability.
- Support the development of RESTful APIs to enable seamless data access and integration with external applications.
- Apply best practices for data management, security, and governance while maintaining high performance for data systems.
What you bring
- Proven experience designing and implementing data warehousing solutions using Informatica, Talend, or similar ETL platforms.
- Strong SQL skills across Microsoft SQL Server, Oracle, and cloud-based systems such as Azure Data Lake or similar platforms.
- Hands-on big data experience with the Hadoop ecosystem (including HDFS), Spark, Hive, and related frameworks.
- Ability to develop scalable ETL pipelines using Python, Shell scripting (Bash), or other scripting languages in an Agile environment.
- Familiarity with cloud computing platforms such as AWS or Azure Public Cloud for deploying and managing big data systems.
- Knowledge of data modeling principles, especially dimensional modeling, for building warehouses that support analytics.
- Experience with business intelligence tools such as Looker to translate complex datasets into insights.
- Strong analytical capability with the ability to troubleshoot issues quickly while maintaining high-quality data management standards.
Technologies you’ll work with
ETL, SQL, Python, Bash, RESTful APIs, Hadoop, Spark, Hive, HDFS, AWS, Azure Data Lake, Informatica, Talend, Microsoft SQL Server, Oracle, Looker, Azure Public Cloud, dimensional modeling.
Similar Jobs
S