DataJobs.io
← Back to all jobs
Mayo Clinic

Principal Data Engineer - Enterprise Data & Analytics - Remote

Rochester, MN Remote $156k - $225k/yr Full time Posted 11h ago

Job Description

Join Mayo Clinic as a Principal Data Engineer within Enterprise Data & Analytics, delivering enterprise-scale data architecture and engineering in a remote role.

Responsibilities

  • Shape enterprise data architecture and engineering roadmaps, contributing to solution design, development, optimization, and technical delivery.
  • Build and deploy data pipelines, integrations, and transformations to support analytics and machine learning workloads using open source languages and vendor software.
  • Provide advisory services across departments and leadership groups with independent judgment.
  • Collaborate with product owners and Analytics and Machine Learning delivery teams to identify and retrieve data.
  • Perform exploratory data analyses, transform data, visualize trends, and develop and validate analytical models.
  • Translate qualitative and quantitative assessments into actionable insights for stakeholders.
  • Design, develop, review, and optimize production code and platform capabilities.
  • Offer technical leadership and mentorship to engineering teams.
  • Maintain up-to-date knowledge of Mayo Clinic's current solutions, coding languages, tools, and the Enterprise Data & Analytics technology framework.

Requirements

  • Bachelor's degree in a relevant field (engineering, mathematics, computer science, information technology, health science, or other analytical/quantitative discipline) plus at least seven years of professional or research experience in data visualization, data engineering, or analytical modeling techniques.
  • Or an Associate's degree in a relevant field plus at least nine years of professional or research experience in data visualization, data engineering, or analytical modeling techniques.
  • In-depth business or practice knowledge may be considered in lieu of formal degree requirements.
  • Ability to manage a varied workload with multiple priorities and stay current on healthcare trends and enterprise changes.
  • Strong interpersonal and time-management skills with experience working on cross-functional teams.
  • Excellent analytical skills with the ability to identify and recommend solutions and a commitment to customer service.
  • Excellent verbal and written communication skills, attention to detail, and a high capacity for learning and problem resolution.
  • Advanced experience in SQL.
  • Advanced experience in scripting languages such as Python, JavaScript, PHP, C++, or Java, and API integration.
  • Experience in hybrid data processing methods (batch and streaming) using Apache Spark, Hive, Pig, Kafka.
  • Experience with big data, statistics, and machine learning.
  • Ability to navigate Linux and Windows operating systems.
  • Knowledge of workflow scheduling (Apache Airflow, Google Composer), Infrastructure as code (Kubernetes, Docker), and CI/CD (Jenkins, GitHub Actions).
  • Experience in DataOps/DevOps and agile methodologies.

Technologies

  • SQL, Python, JavaScript, PHP, C++, Java
  • Apache Spark, Hive, Pig, Kafka
  • Denodo, Tableau, Power BI, SAS, ThoughtSpot
  • DASH, d3, React
  • Snowflake, SSIS, Google BigQuery
  • Apache Airflow, Google Composer
  • Kubernetes, Docker, Jenkins, GitHub Actions
  • Linux, Windows
  • Apache Iceberg, Delta Lake, Apache Hudi
  • Parquet, Avro, ORC

Benefits

  • Medical: Multiple plan options.
  • Dental: Delta Dental or reimbursement account for flexible coverage.
  • Vision: Affordable plan with national network.
  • Pre-Tax Savings: HSA and FSAs for eligible expenses.
  • Retirement: Competitive retirement package to secure your future.

Preferred Candidate Will Possess

  • Expert-level proficiency in Python and SQL with extensive experience building enterprise-scale production systems.
  • Advanced expertise in scalable distributed computing frameworks and modern data processing platforms.
  • Extensive experience implementing and governing open data architectures using Apache Iceberg, Delta Lake, Apache Hudi, and related technologies.
  • Deep understanding of modern analytical storage formats including Parquet, Avro, and ORC.
  • Proven expertise in lakehouse architecture, data platform design, and large-scale data engineering practices.
  • Experience architecting cloud-agnostic solutions across multiple technology ecosystems.
  • Experience designing scalable, fault-tolerant, secure, and observable data platforms supporting analytics, AI, machine learning, and operational workloads.
  • Experience establishing enterprise engineering standards, architecture patterns, and modernization strategies.

Compensation Details

  • Salary range: USD 155,500 - 225,492 per year.
  • Exemption status: Exempt.
  • Education, experience and tenure may be considered along with internal equity when offers are extended.

Schedule and Work Details

  • Schedule: Full Time.
  • Hours per pay period: 80.
  • Schedule details: M-F daytime hours; 100% remote role; must reside within the United States.
  • Weekend schedule: As business needs dictate.
  • International assignment: No.

Site Description

Just as our reputation has spread beyond our Minnesota roots, our locations have grown to include campuses in Phoenix/Scottsdale, Arizona, Jacksonville, Florida, Rochester, Minnesota, Mayo Clinic Health System campuses across the Midwest, and international locations.

Equal Opportunity

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender identity, sexual orientation, national origin, protected veteran status or disability status. Mayo Clinic participates in E-Verify and may provide information from Form I-9 to confirm work authorization where required.

Recruiter

Laura Percival

Similar Jobs