Remote (Country)

Machinify is hiring a Senior Data Engineer

About the Role

Machinify is hiring a Senior Data Engineer to transform raw external data into trusted datasets that drive payment, product, and operational decisions. In this role, you'll build, scale, and refine production pipelines, ensuring data is accurate, observable, and actionable. You'll also play a critical part in onboarding new customers and integrating their data.

What You'll Do

  • Design and implement robust, production-grade pipelines using Python, Spark SQL, and Airflow to process high-volume file-based datasets (CSV, Parquet, JSON).
  • Lead efforts to canonicalize raw healthcare data (837 claims, EHR, partner data, flat files) into internal models.
  • Own the full lifecycle of core pipelines — from file ingestion to validated, queryable datasets — ensuring high reliability and performance.
  • Onboard new customers by integrating their raw data into internal pipelines and canonical models.
  • Build resilient, idempotent transformation logic with data quality checks, validation layers, and observability.
  • Refactor and scale existing pipelines to meet growing data and business needs.
  • Tune Spark jobs and optimize distributed processing performance.
  • Implement schema enforcement and versioning aligned with internal data standards.
  • Collaborate deeply with Data Analysts, Data Scientists, Product Managers, Engineering, Platform, SMEs, and AMs.
  • Monitor pipeline health, participate in on-call rotations, and proactively debug and resolve production data flow issues.
  • Contribute to the evolution of our data platform — driving toward mature patterns in observability, testing, and automation.
  • Build and enhance streaming pipelines (Kafka, SQS, or similar) where needed to support near-real-time data needs.
  • Help develop and champion internal best practices around pipeline development and data modeling.

What We're Looking For

  • 6+ years of experience as a Data Engineer (or equivalent), building production-grade pipelines.
  • Strong expertise in Python, Spark SQL, and Airflow.
  • Experience processing large-scale file-based datasets (CSV, Parquet, JSON, etc) in production environments.
  • Experience mapping and standardizing raw external data into canonical models.
  • Familiarity with AWS (or any cloud), including file storage and distributed compute concepts.
  • Experience onboarding new customers and integrating external customer data with non-standard formats.
  • Ability to work across teams, manage priorities, and own complex data workflows with minimal supervision.
  • Strong written and verbal communication skills — able to explain technical concepts to non-engineering partners.
  • Comfortable designing pipelines from scratch and improving existing pipelines.
  • Experience working with large-scale or messy datasets (healthcare, financial, logs, etc.).
  • Experience building or willingness to learn streaming pipelines using tools such as Kafka or SQS.

Nice to Have

  • Familiarity with healthcare data (837, 835, EHR, UB04, claims normalization).

Technical Stack

  • Python, Spark SQL, Airflow, AWS, Kafka, SQS

Team & Environment

Work closely with product managers, data scientists, subject matter experts, engineers, and customer teams.

Benefits & Compensation

  • Salary: $180k-$220k + meaningful equity
  • Work from anywhere in the US.
  • Full Medical/Dental/Vision for employees & their families.
  • Flexible and trusting environment.
  • Unlimited FTO.
  • Competitive salary, equity, 401(k) including employer match.

Work Mode

This role is open to candidates anywhere in the US.

We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender, gender identity or expression, or veteran status.

Required Skills
PythonSpark SQLAirflowAWSKafkaSQSData EngineeringData PipelinesETLData WarehousingDistributed SystemsSQLBig Data
Relocating to Thailand?

Visa and work permit handled by experts

SVBL manages your entire visa process — from application to approval. Work permits, extensions, and compliance all covered. One partner for legal, immigration, and settling in.

Work permit processing
Visa extensions & renewals
Immigration compliance
Banking & housing guidance
Get free consultation
Free initial consultation
About company
Machinify

Machinify is a leading healthcare intelligence company with expertise across the payment continuum, delivering value, transparency, and efficiency to health plan clients. Deployed by over 85 health plans representing more than 270 million lives, it brings together a fully configurable, AI-powered platform with best-in-class expertise.

Visit website
Job Details
Category data
Posted a month ago