Platform Bioinformatician
AI Summary
Builds and operates cloud data infrastructure and bioinformatics pipelines to turn clinical, genomic, and imaging data into model-ready assets for AI and life sciences teams.
About this role
About the Role
Join a multidisciplinary AI and life sciences team building the cloud data platform that turns clinical, genomic, imaging, and other biomedical data into reliable, model-ready assets. This hands-on role combines platform engineering, data infrastructure, and bioinformatics workflows to support internal teams and external partners.
What You'll Do
Build and operate cloud data infrastructure, including object storage, lakehouse architecture, catalogs, lineage, versioning, and access layers.
Design and maintain governed data warehouses for clinical records, molecular features, metadata, and versioned dataset releases.
Develop reproducible omics pipelines for data such as whole-genome sequencing, RNA sequencing, and proteomics using containerized workflows and cloud batch compute.
Implement data quality checks, biological validation, and visualizations that help teams identify issues across processing steps.
Prepare training-ready dataset releases and build data deployment interfaces for models used in partner environments.
Build and operate production cloud infrastructure, Kubernetes clusters, CI/CD systems, and distributed workload orchestration, with attention to reliability, security, observability, and cost.
What We're Looking For
5+ years of experience in platform, infrastructure, or site reliability engineering, including operating cloud infrastructure in a high-growth startup or technically demanding environment.
Deep AWS or GCP experience, including production Kubernetes, and experience with Terraform or similar infrastructure-as-code tools.
Production programming experience in Python and/or Go, plus experience designing CI/CD pipelines and deployment systems.
Experience building job orchestration or scheduling systems for distributed workloads; GPU or machine learning infrastructure experience is a plus.
Familiarity with biological, clinical, or other sensitive data and production data pipelines. Experience with Nextflow or similar workflow managers, lakehouse and warehouse design, data catalogs, Docker, and cloud batch compute is relevant.
Python, R, SQL, and familiarity with standards such as OMOP or FHIR are useful.
Compensation & Benefits
Visa sponsorship is not available.
Location
On-site in San Francisco, California, United States.
Skills
Explore related jobs
More jobs at Clera
Similar AWS jobs
Browse these categories
Market data for this role
All reports →- SeriesRole reportsOne role family at a time: how many openings, what changed this week, who is hiring, what it pays.
- SeriesSalary reportsWhat employers publish in job postings, by level and workplace. Not self-reported pay.
- Market overviewState of tech hiring, September 2026: up 4.8%Tech hiring rose 4.8% month over month in September 2026, with 411,122 new listings. Customer support and account executive roles led the growth.
