Responsibilities
- Build and scale platform infrastructure for web scrapers that collect policy data from over 100 healthcare payers
- Improve scraper reliability and scalability while maintaining data accuracy
- Own the complete scraper pipeline from data ingestion to storage
- Resolve infrastructure bottlenecks that hinder new scraper development
- Use LLM systems to convert unstructured payer policy text into structured, searchable data
- Infer and tag relevant medical codes from policy language using AI models
- Determine coverage status for procedures based on LLM-driven analysis
- Build evaluation and testing harnesses to validate AI system output quality
- Build and maintain CI/CD pipelines for testing and deploying scraper jobs
- Improve scraper job infrastructure reliability through monitoring and fixes
- Diagnose and resolve system issues across the data pipeline
- Own architectural decisions and communicate them to the engineering team
- Design and maintain APIs that serve normalized policy data to customers
- Support the Snowflake-based data product used by client customers
- Ensure data outputs are structured for downstream usability
Requirements
- 3-5 years of professional backend engineering experience with demonstrated ownership of production data pipelines or platform-level systems
- Expert-level Python proficiency (3+ years in a professional backend engineering context)
- Strong SQL experience, including complex queries and schema design (Postgres or equivalent relational database)
- Hands-on cloud infrastructure experience deploying and managing services (AWS preferred; Azure or GCP with 2+ years accepted)
- Demonstrated experience working with messy, unstructured, or ambiguous data sources (data normalization, cleaning, or ingestion from inconsistent external sources)
- Experience owning CI/CD pipeline design and implementation for testing and deployment
- Client-facing English proficiency at B2+ (CEFR)
Nice to Have
- Experience building or productionizing LLM-based systems with evaluation or testing components (not limited to API calls)
- Web scraping framework experience (Scrapy, Playwright, or similar)
- Experience with RAG (Retrieval-Augmented Generation) pipelines
- Prior experience in healthcare, fintech, or another highly regulated data domain
- Infrastructure-as-Code experience (Terraform) or familiarity with FastAPI/Snowflake
Benefits
- Competitive Salary: Based on experience and skills
- Remote Work: Fully remote—work from anywhere
- Team Incentives: Recognition for maintaining 100% CRM hygiene and on-time reporting
- Generous PTO: In accordance with company policy
- Health Coverage for PH-based talents: HMO coverage after 3 months for full-time employees
- Direct Mentorship: Guidance from international industry experts
- Learning & Development: Ongoing access to resources for professional growth
- Global Networking: Connect with professionals worldwide
Compensation
Competitive salary based on experience and skills
Work Arrangement
Remote (Worldwide) — LATAM
Work Arrangement
Remote (Worldwide) — LATAM
Other
- Work Schedule: EST | Full overlap with US Eastern business hours (Monday–Friday)
- Work From Anywhere in LATAM
- Client-facing English proficiency at B2+ (CEFR)