Rezilient Health Logo

Rezilient Health

Senior Platform Engineer

Posted 5 Hours Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
Own the reliability, scalability, security, and performance of Rezilient’s healthcare platform. Design and operate cloud infrastructure, Infrastructure as Code, CI/CD pipelines, observability, container orchestration, incident response, automation, disaster recovery, and security controls for PHI. Establish SLOs, SLIs, and error budgets; reduce operational toil; lead postmortems; and partner with engineering, product, and security teams to deliver resilient services.
The summary above was generated by AI

At Rezilient, we’re redefining primary care by making access to healthcare more convenient, timely, and seamless. Our innovative CloudClinic model combines virtual provider visits with cutting-edge technology to create a personalized digital healthcare experience that puts patients at the center of their care. By streamlining care delivery and continuously expanding specialty services, we empower our care team to focus on patient well-being while providing the most comprehensive and accessible care possible.

We're seeking a Senior Platform Engineer to own the reliability, scalability, and performance of the systems that power Rezilient's care delivery. In a healthcare environment, uptime is not a convenience, it's patient care in progress, so this role is critical to ensuring that patients and clinical teams can depend on our platform every time they need it. You'll design and operate the infrastructure, observability, and automation that keep our services fast, resilient, and secure, while partnering closely with Full-Stack, Front-End, Product, and Security teams.

You'll define reliability standards, drive down operational toil through automation, and build the tooling and practices that let a fast-moving engineering team ship confidently without compromising availability or the security of protected health information (PHI). Your work will directly shape how dependably patients receive care and how our providers deliver it, advancing our mission to make humane healthcare truly abundant.

Key Responsibilities

  • Design, provision, and maintain cloud infrastructure using Infrastructure as Code (Terraform, CloudFormation, or equivalent), managing environments across development, staging, and production with repeatable, auditable configuration.
  • Own and evolve CI/CD pipelines to enable safe, frequent deployments, including automated testing gates, blue/green and canary rollouts, and fast, reliable rollbacks.
  • Build and operate observability across the stack, including metrics, logging, distributed tracing, dashboards, and alerting, and define SLOs, SLIs, and error budgets that align reliability work with product priorities.
  • Lead incident response: on-call rotation participation, triage, mitigation, and clear communication, followed by blameless postmortems that turn failures into systemic fixes.
  • Automate away operational toil through scripting and tooling, reducing manual intervention and improving mean time to detection and recovery.
  • Design for scalability and resilience, including capacity planning, load and failure testing, autoscaling, redundancy, and disaster recovery and backup strategies with defined RTO/RPO targets.
  • Manage containerized workloads and orchestration (Docker, Kubernetes), including networking, service discovery, and resource management.
  • Partner with Security and Engineering to harden infrastructure: secrets management, network segmentation, encryption, vulnerability scanning, patching, and audit logging consistent with HIPAA-aligned requirements for handling PHI.
  • Collaborate with development teams to embed reliability and operability into services from design through production, balancing "move fast" with "do no harm."

Requirements
  • Bachelor's degree in computer science, software engineering, or a related field, or equivalent hands-on experience.
  • 5+ years in Site Reliability, DevOps, or infrastructure engineering roles at startup or growth-stage organizations, ideally within healthcare, health tech, or another regulated, high-availability, data-sensitive industry.
  • Deep hands-on experience operating production systems on a major cloud platform (AWS, GCP, or Azure), including compute, networking, storage, and managed database services.
  • Strong proficiency with Infrastructure as Code (Terraform or equivalent) and configuration management.
  • Production experience with containerization and orchestration (Docker and Kubernetes), including deployment, scaling, and troubleshooting.
  • Expertise building CI/CD pipelines and release automation, with a track record of enabling safe, frequent deployments.
  • Hands-on experience with observability and monitoring tooling (e.g., Datadog, Prometheus, Grafana, CloudWatch, ELK/OpenSearch) and defining SLOs, SLIs, and error budgets.
  • Strong scripting and automation skills (Python, Go, or Bash) and fluency with command-line and cloud CLIs.
  • Experience leading incident response and on-call, including postmortem and root-cause analysis practices.
  • Solid understanding of infrastructure and network security, secrets management, and encryption, with familiarity handling sensitive or protected data (PHI/HIPAA experience strongly preferred).
  • Proficient with Git and version control workflows; familiarity with the Agile Development Framework and ideally the Atlassian toolset (Jira and Confluence).
  • Excellent verbal and written communication skills; calm, methodical, and detail-oriented under pressure, and comfortable owning ambiguity in an early-stage environment.

Benefits

At Rezilient Health, we're on a mission to make humane healthcare truly abundant, and we invest in the people who make that possible. This is a high-impact role where your ideas directly shape how care is delivered, within a supportive, collaborative, and diverse team that believes healthcare is fundamentally a human experience. We back that commitment with a robust total rewards package:

  • Meaningful ownership through stock options, allowing you to share directly in the value you help create as we grow
  • Competitive base compensation
  • Comprehensive medical, dental, and vision coverage options, with Rezilient contributing toward your premiums
  • Complimentary access to Rezilient's clinical programs for you and your household members, the same connected care we deliver to our patients
  • 401(k) retirement plan to support your long-term financial well-being
  • Flexible Paid Time Off, so you can recharge without counting days
  • Dedicated Paid Sick Leave, separate from your FTO, for when health needs come first
  • 11 paid company holidays each year
  • Paid family leave to support you through life's biggest moments
  • Optional ancillary benefits, including life insurance, disability coverage, and a Health Savings Account (HSA)

Similar Jobs

3 Hours Ago
Easy Apply
Remote
United States
Easy Apply
139K-235K Annually
Senior level
139K-235K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Build and operate production AI agents, agentic workflows, MCP servers, and CX applications. Develop Python and TypeScript services, AWS infrastructure, integrations, testing, observability, security controls, and CI/CD pipelines. Partner with Customer Success and other CX teams to prototype solutions, translate workflows into technical plans, and deliver reusable platforms. Mentor engineers and non-engineers, document systems, establish governance, and support reliable handoff and shared ownership.
Top Skills: Aws EcsAws LambdaAws RdsClaude CodeDockerFastapiFlaskGainsightGitGitlab Ci/CdGitlab Duo Agent PlatformGongGraphQLKantataLarge Language Models (Llms)Model Context Protocol (Mcp)OktaPythonReactRestSalesforceSnowflakeTerraformThought IndustriesTypescriptVue
23 Days Ago
Easy Apply
Remote
United States
Easy Apply
139K-235K Annually
Senior level
139K-235K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Design, build, operate, and scale GitLab Orbit backend services in Rust across distributed cloud-native environments. Improve Kubernetes deployments, infrastructure automation, observability, reliability, incident readiness, multi-tenant isolation, data workflows, graph query capabilities, indexing pipelines, cloud storage integrations, and API/MCP surfaces. Own technical designs through rollout while collaborating asynchronously with infrastructure, security, AI, data, delivery, and SRE teams.
Top Skills: Amazon S3AWSClickhouseGCPGoHelmKubernetesMcpNatsRubyRustSiphonTerraformTypescriptVue
Yesterday
In-Office or Remote
100K-137K Annually
Senior level
100K-137K Annually
Senior level
Fintech
Leads DevOps and platform engineering initiatives by designing CI/CD pipelines, automating cloud infrastructure, supporting Kubernetes deployments, integrating security and observability controls, and improving developer productivity. Partners across engineering, security, SRE, and architecture teams to establish standardized delivery practices, self-service platforms, engineering metrics, and scalable cloud-native solutions. Mentors engineers and drives technical standards across multiple teams.
Top Skills: AWSAzureAzure DevopsBashCi/CdDockerDora MetricsGitopsInfrastructure As CodeKubernetesObservabilityPowershellPythonTerraform

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account