Data Engineer (5 to 7 yrs) (PySpark / Azure / Python / DataBricks) - Remote India

Nexaminds · India · Engineering

Posted 2026-09-21

Apply for this role →

We are currently seeking a highly motivated Data Engineer III with experience in Healthcare Domain. This role  collaborate closely with product teams, software engineers and clinical leaders to expose various data products for internal and external clients. With over a million managed lives across the country and terabytes of data generated, our teams need to be continuously equipped with the tools and insights to drive strategy and innovation to further our core values of improving patient outcomes and empowering our providers.

What You'll Do

Use data engineering best practices to produce high quality, maximally available data models which are intuitive and trusted by stakeholders.

Scope and implement new entities for unified data model.

Interfacing with business customers, gathering requirements and developing new datasets in data platform

Identifying the data quality issues to address them immediately to provide great user experience

Extracting and combining data from various heterogeneous data sources

Designing, implementing and supporting a platform that can provide ad-hoc access to large datasets

Modelling data and metadata to support machine learning and AI

Support integration efforts from acquisitions as necessary.

Qualifications

Bachelor's degree required in computer science, information technology, or related field. Master’s degree in healthcare-related field is preferred.

Strong understanding of database structures, theories, principles, and practices.

Working knowledge with programming or scripting languages such as Python, Spark, and SQL.

Demonstrated track record executing best practices for the full software development life cycle (SDLC), including documentation, coding standards, code reviews, source control management, deployment processes, testing, and operations.

Familiarity with normalized, dimensional, star schema and snowflake schematic models.

Healthcare domain and data experience

Working experience with Databricks preferred.

Databricks/Microsoft Azure Certification is a plus

Strong written and oral communication skills.

You're a great for this role if:

You’re naturally curious. You want know what the data means, and how it will be used upstream.

You proactively identify issues and lean in to improve data quality for common use cases.

5+ years of experience working with diverse healthcare data sources, for example, claims, encounters, ADT data, FHIR, EHR data etc.

3+ years’ using cloud-based services from AWS, GCP, or Azure.

2+ years serving data sets to both BI tools and SE applications.

Apply for this role →

← Back to all jobs