Data Engineer Hadoop, HIVE & Python
Job Description
DenkenSolutions Inc offers an on-site Data Engineer role in Birmingham, AL with an hourly rate of USD 58. This position centers on Hadoop, HIVE, and Python and is responsible for designing, developing, and supporting data engineering and analytics solutions using both on-premises tools and cloud services. A bachelor’s degree is required, and candidates should bring at least 3 years of hands-on experience.
The role emphasizes building and sustaining data engineering and analytics workloads, leveraging Hadoop, HIVE, Spark, and Python across on-premises environments and cloud services.
Responsibilities
- Design, develop, and support data engineering and analytics solutions using on-premises tools and cloud services.
- Create and refine functional and technical designs for data engineering and analytics projects.
- Implement data models across multiple schemas and integrate diverse data source types.
- Develop and maintain data pipelines and analytics workloads using big data technologies such as Hadoop, HIVE, and Spark.
- Support statistical and AI/ML initiatives with R and Python based models.
- Collaborate on data sourcing, enrichment, and delivery using APIs and Web Services.
Requirements
- Experience with batch and real-time data processing frameworks.
- Experience with data modeling, data access, schemas, and data storage techniques.
- Experience with data quality tooling.
- Ability to create functional and technical designs for data engineering and analytics solutions.
- Experience implementing data models of different schemas and working with diverse data source types.
- Hands-on experience developing solutions with big data technologies such as Hadoop, HIVE, and Spark.
- Hands-on experience developing and supporting statistical models, and AI/ML solutions using R and/or Python.
- 5+ years hands-on experience designing, developing, testing, deploying, and supporting data engineering and analytics solutions using on-premises tools such as MSBI (SSIS/SSAS), Informatica, Oracle Golden Gate, SQL, Oracle, and SQL Server.
- 3+ years hands-on experience designing, developing, testing, deploying, and supporting data engineering and analytics solutions using Microsoft cloud-based tools such as Azure Data Lake, Azure Data Factory, Azure Databricks, Python, Azure Synapse, Azure Key Vault, and Power BI.
- Experience with containerization methodologies Docker, OpenShift, etc.
- Experience with Agile as well as DevOps, CI/CD methodologies.
- Hands-on experience designing and developing solutions involving data sourcing, enrichment and delivery using APIs & Web Services.
Technologies
- Hadoop, HIVE, Spark
- R, Python
- MSBI (SSIS/SSAS), Informatica, Oracle GoldenGate
- SQL, Oracle, SQL Server
- Azure Data Lake, Azure Data Factory, Azure Databricks, Azure Synapse, Azure Key Vault
- Power BI
- Docker, OpenShift
- APIs & Web Services