Principal Azure Data Engineer (12 to 15 yrs) (Python / PySpark / SQL/ DataBricks) - Remote India

Nexaminds · India · Engineering

Posted 2026-09-05

Apply for this role →

We are currently seeking a highly motivated Principal Data Engineer with experience in Healthcare Domain. This role collaborate closely with product teams, Data engineers and clinical leaders to expose various data products for internal and external clients. With over a million managed lives across the country and terabytes of data generated, our teams need to be continuously equipped with the tools and insights to drive strategy and innovation to further our core values of improving patient outcomes and empowering our providers.

🚀 Are you a PRO at developing Databricks Pipelines / enhance or support existing pipelines?

Are you strong in SQL / PySpark / Python debugging skills dealing with business critical data with extensive exposure to Azure?

Do you have hands-on experience in Healthcare Data?

Then don’t wait any further — your next big opportunity is here! 🌟

Join us at Nexaminds and be part of an exciting journey where innovation meets impact. The benefits are unbelievable — and so is the experience you’ll gain!

💼 Apply now and let’s start the conversation.

What You'll Do

Develop Databricks pipeline's and enhance existing pipelines.

Deal with business-critical data , with close monitoring Business Stakeholders.

Use data engineering best practices to produce high quality, maximally available data models which are intuitive and trusted by stakeholders.

Scope and implement new entities for unified data model.

Interfacing with business customers, gathering requirements and developing new datasets in data platform

Identifying the data quality issues to address them immediately to provide great user experience

Extracting and combining data from various heterogeneous data sources

Designing, implementing and supporting a platform that can provide ad-hoc access to large datasets

Modelling data and metadata to support machine learning and AI

Support integration efforts from acquisitions as necessary.

Qualifications

Bachelor's degree required in computer science, information technology, or related field. Master’s degree in healthcare-related field is preferred.

Strong understanding of database structures, theories, principles, and practices.

Working knowledge with programming or scripting languages such as Python, Spark, and SQL.

Demonstrated track record executing best practices for the full software development life cycle (SDLC), including documentation, coding standards, code reviews, source control management, deployment processes, testing, and operations.

Familiarity with normalized, dimensional, star schema and snowflake schematic models.

Healthcare domain and data experience

Working experience with Databricks preferred.

Databricks/Microsoft Azure Certification is a plus

Strong written and oral communication skills.

You're a great for this role if:

You’re naturally curious. You want know what the data means, and how it will be used upstream.

You proactively identify issues and lean in to improve data quality for common use cases.

You are an independent contributor

You are strong with SQL & Spark debugging skills

10+ years of experience working with diverse healthcare data sources, for example, claims, encounters, ADT data, FHIR, EHR data etc.

8+ years’ using cloud-based services from AWS, GCP, or Azure.

Apply for this role →

← Back to all jobs