Istari Digital Logo

Istari Digital

Sr. Cloud Infrastructure Engineer

Posted 5 Days Ago
Remote
Hiring Remotely in US
135K-220K Annually
Senior level
Remote
Hiring Remotely in US
135K-220K Annually
Senior level
Design, implement, and operate scalable AWS and Kubernetes infrastructure using Terraform and automation. Improve CI/CD, observability, security posture, and developer experience; troubleshoot production incidents, maintain runbooks, and support customer-managed deployments.
The summary above was generated by AI

We are seeking a Senior Cloud Infrastructure Engineer to join our Engineering team. This role is central to building, operating, and scaling the cloud infrastructure that powers our platform, and to making our software easy to deploy both in our cloud-hosted environments and in customer-managed environments.

This engineer will work closely with platform and product engineering to improve the reliability, automation, and operational maturity of our AWS and Kubernetes environments. Some of our environments serve regulated, security-sensitive customers; while deep security specialization is not required, this engineer should be comfortable operating within an established, security-conscious baseline.

The ideal candidate combines strong hands-on infrastructure expertise with a practical, execution-focused mindset and a bias toward automation.

Core Responsibilities

  • Design, implement, and maintain scalable, reliable infrastructure in AWS

  • Operate and improve Kubernetes-based environments, including production workloads

  • Build and maintain infrastructure as code using Terraform

  • Improve CI/CD pipelines, deployment workflows, and release automation in partnership with engineering teams

  • Build and maintain the packaging and reference architectures customers use to install our software in their own environments

  • Strengthen observability across the platform, including monitoring, logging, alerting, and actionable dashboards

  • Improve developer experience through tooling, environment automation, and self-service infrastructure

  • Operate within and preserve the established security and compliance posture of our environments

  • Monitor, troubleshoot, and resolve complex infrastructure issues with clear and timely communication

  • Participate in incident response and post-incident analysis

  • Develop and maintain documentation, runbooks, and technical standards

  • Identify opportunities to improve cost efficiency, performance, and resilience across environments

Required Qualifications

  • Minimum of 5 years of experience in DevOps, Infrastructure Engineering, Platform Engineering, or Site Reliability Engineering

  • Strong hands-on experience with AWS in production environments

  • Proven experience operating Kubernetes in production

  • Strong experience with Terraform and infrastructure-as-code practices

  • Proven experience building or improving CI/CD pipelines and deployment automation

  • Solid understanding of cloud networking, IAM, secrets management, and operational controls

  • Experience with monitoring, logging, and observability tooling

  • Scripting proficiency in Python, Go, Bash, or similar

  • Excellent troubleshooting and problem-solving skills in complex production environments

  • Strong communication skills with the ability to explain technical concepts to both technical and non-technical stakeholders

  • Must live/work in the U.S.

  •  Preferred Qualifications
  • Experience operating stateful workloads on Kubernetes, such as databases or message queues, including persistent storage and backup/recovery

  • Experience with GitOps-based deployment workflows

  • Experience packaging software for customer-managed or self-hosted deployment (e.g., Helm charts)

  • Familiarity with compliance or security frameworks such as FedRAMP, NIST, SOC 2, or similar

  • Experience with PostgreSQL, cloud storage platforms, and production networking patterns

  • Experience with configuration management tools such as Ansible

  • Experience with additional cloud platforms such as Azure or GCP

  • Experience with service mesh or advanced Kubernetes networking

  • Experience supporting customer-facing or mission-critical production infrastructure

  • Top Secret Security Clearance

Similar Jobs

7 Days Ago
Remote
United States
165K-165K Annually
Senior level
165K-165K Annually
Senior level
Security • Cybersecurity
Design, build, and maintain scalable AWS and Azure infrastructure using Infrastructure-as-Code. Develop automation, deployment pipelines, and production release processes. Contribute code to cloud services, manage networking and IAM, ensure availability and security, and respond to incidents while collaborating on observability and security practices.
Top Skills: AnsibleAWSAzureDockerGoIamKubernetesPackerPythonTerraform
3 Days Ago
Remote
USA
214K-267K Annually
Senior level
214K-267K Annually
Senior level
Artificial Intelligence • Robotics • Software
Lead and manage a cloud engineering team while architecting and implementing high-availability GCP systems and ML pipelines that ingest and process massive image volumes. Drive GPU resource optimization, IaC deployments, observability, and cross-functional collaboration with ML, data, and robotics teams to enable scalable model training, serving, and inference.
Top Skills: BigQueryCloud RunCloud StorageDataflowDockerGithub ActionsGkeGoGoogle Cloud PlatformGpuIso27001JenkinsKafkaKubeflowKubernetesMlopsPub/SubPulumiPythonSoc2Tensorflow ServingTerraformTypescriptVertex Ai
12 Days Ago
In-Office or Remote
4 Locations
168K-334K Annually
Senior level
168K-334K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Own design, automation, and lifecycle of Kubernetes platform for global network infrastructure. Build production-quality tooling for cluster provisioning, upgrades, GitOps delivery, observability, and recovery. Provide production support and incident response for network services on the platform, drive root-cause analysis, and establish platform standards and runbooks.
Top Skills: Ci/CdCluster Api (Capi)GitGitopsGoInfrastructure As CodeKubernetesKubernetes Controllers/OperatorsMetal3ObservabilityPythonTelemetry

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account