Actively recruiting / 43 applicants
We’re here to help you
Jane Cervantes is in direct contact with the company and can answer any questions you may have. Email
Jane Cervantes, RecruiterRole Overview
We are looking for a skilled Freelance Developer to work on a high-impact data engineering project. This contract role focuses on transforming an early-stage data aggregation platform into a robust, production-grade infrastructure. Your efforts will be crucial in ensuring data pipeline reliability, system observability, and delivering accurate customer-facing data.
Responsibilities
- Conduct an independent audit of the live database to evaluate its structure, correctness, and potential risks, comparing findings against internal defect logs.
- Establish a control plane by implementing version control for schema and functions, setting up declarative migrations, a code review workflow, a staging environment, and a thoroughly tested point-in-time restore path.
- Migrate existing local scrapers to hosted runners to ensure unattended daily refresh of all registry feeds, complete with robust failure reporting.
- Repair broken enrichment fields, fix non-writing filters, and reconcile mismatched counts between connectors and dashboards.
- Build and maintain daily full-corpus exports to cloud storage buckets and develop a customer-facing REST API.
- Expand pipelines by creating ingestion workflows for additional international trial registries over time.
Required Skills
- Deep expertise in PostgreSQL, including advanced PL/pgSQL, query plan optimization, indexing strategies, and strict transaction boundary management.
- Proven experience with Supabase, including Supabase CLI, declarative migrations, branching workflows, Deno/TypeScript Edge Functions, RLS, and pg_cron.
- Strong background in data pipeline engineering, with a focus on idempotent incremental loads, Change Data Capture (CDC), automated reconciliation against source records, and safe backfill processes.
- Proficiency in Python or TypeScript for developing ingestion workers and scrapers.
- Strong focus on data quality, with the capability to identify silent failures, missing writes, bad denominators, and unverified data references.
- Excellent written communication skills in English for effective asynchronous, remote collaboration.
Nice to Have
- Experience with Google Cloud Storage and service account management.
- Web scraping experience, particularly with handling site defenses like proxy and residential IP management.
- Experience with translation pipelines or Chinese-language data processing.
- Familiarity with pharmaceutical, clinical trial, or life sciences data.