Senior GCP Cloud / Platform Operations Engineer
Presupuesto: $20.0 - $45.0
HOURLY / FULL_TIME
β 4.93 (28)
United States
cloud-architecture, google-cloud-platform, python, devops, docker, terraform, cicd, infrastructure-as-code, hipaa, automated-deployment
Cualificaciones preferidas
- Tipo de talento: Independiente
- Experiencia: Experto
- InglΓ©s: Nativo
π οΈ Senior GCP Cloud / Platform Operations Engineer
We're a GCP-based healthcare AI company building production automations and multi-practice cloud infrastructure. Looking for a senior cloud / platform ops engineer to keep production reliable, harden our shared GCP platform with IaC, and take day-to-day operational load off our Cloud Lead.
Hands-on build-and-run role β not a product feature developer. You partner with the Cloud Lead; you own reliability, platform hygiene, and engineer enablement so app teams can ship safely.
ββββββββββββββββββββββββββββββββββββββββββββββ
π§° Tech stack
β’ Compute / serverless: Cloud Run (services + jobs), Cloud Functions, Compute Engine + Docker where needed
β’ Eventing / data: Cloud Scheduler, Pub/Sub, Firestore, Cloud SQL (MySQL)
β’ Security / access: IAM, Secret Manager, IAP, VPC / private networking / egress
β’ Delivery: Terraform (preferred), Cloud Build and/or GitHub Actions + Workload Identity, Artifact Registry
β’ Ops: Cloud Monitoring / Logging, dashboards, alerts, SLOs, billing / cost hygiene
We are all-in on GCP β real Google Cloud experience is important.
ββββββββββββββββββββββββββββββββββββββββββββββ
π§± What you'll do
Keep our GCP estate reliable β Cloud Run, Functions, Scheduler, Pub/Sub, Firestore, IAM, networking β with dashboards, alerts, SLOs, incident response, and post-mortems that become durable fixes
Own Infrastructure as Code and CI/CD β Terraform for new and existing projects; repeatable deploys via Cloud Build / GitHub Actions; no snowflake console as long-term source of truth
Help stand up and operate environments for engineering and practices β IAP-protected admin surfaces, Cloud SQL, private networking, secrets and service-account design (execution under Cloud Lead)
Enable engineers safely β group-based IAM, Artifact Registry readers, shared SQL access, deploy permissions, and documented runbooks instead of one-off personal grants
Maintain cost and compliance hygiene β billing visibility, quotas, org-policy awareness, and keeping regulated workloads on a compliant path
Grow into either a practice-hosting support lane or an automations / standards lane as we hire and split work β this role is the shared GCP core both need
ββββββββββββββββββββββββββββββββββββββββββββββ
β
Must-have
β’ Hands-on GCP: Cloud Run, Cloud Functions, Cloud Scheduler, Pub/Sub, IAM, VPC networking, Secret Manager, Cloud Monitoring / Logging
β’ Infrastructure as Code β Terraform strongly preferred (Pulumi OK)
β’ CI/CD and deployment automation on GCP
β’ Incident response β diagnose and own production failures under pressure
β’ GCP security basics β least-privilege IAM, service accounts, secrets, private networking
β’ HIPAA compliance experience and/or formal training
β’ Clear written runbooks; comfortable pairing with a Cloud Lead
β’ Senior-level experience (~5+ years cloud infrastructure / DevOps / SRE); can propose architecture and work independently
β Nice-to-have
β’ Cloud SQL (MySQL) ops β HA, backups, Auth Proxy / private IP
β’ Identity-Aware Proxy (IAP) for admin UIs and SSH
β’ Firestore (incl. multi-database / app patterns)
β’ Docker + digest-pinned container deploys
β’ Cost optimization / billing export / BigQuery
β’ GCP Professional Cloud Architect or Cloud DevOps Engineer cert
β’ Open-source EHR ops experience (e.g. OpenEMR) β a plus, not a filter; we train the product surface
β’ Healthcare automation / RCM exposure; familiarity with common practice systems a plus
β’ Playwright / headed-browser infra for login automation
β’ Vertex AI / model-serving ops (support, not research)
ββββββββββββββββββββββββββββββββββββββββββββββ
β±οΈ Engagement
Full-time Β· remote Β· long-term Β· partner daily with our Cloud Lead Priority is high β production automations, upcoming cloud migration work, and multi-practice infrastructure all need dedicated ops capacity. Most of the team works on PST and EST time zones.
ββββββββββββββββββββββββββββββββββββββββββββββ
β Screening questions
Describe a production incident you owned end-to-end on GCP (or similar). What failed, how you diagnosed it, and what durable fix you left behind?
Which GCP services have you operated in production day-to-day? Call out Cloud Run, IAM, networking, and monitoring specifically if applicable.
How do you keep infrastructure as code (Terraform or similar) as the source of truth when a team is moving fast?
Briefly describe your HIPAA (or equivalent regulated-data) experience in a cloud environment β training, BAAs, IAM / audit / boundary practices you've applied.
Abrir en Upwork
AI proposal draft
Generate a short cover letter for this job. Edit before sending.
Sign in to generate an AI proposal draft.
Entrar