Satsuma.ai Logo

Satsuma.ai

Senior Site Reliability Engineer

Reposted 6 Days Ago
In-Office or Remote
Hiring Remotely in Austin, TX, USA
50K-80K Annually
Senior level
In-Office or Remote
Hiring Remotely in Austin, TX, USA
50K-80K Annually
Senior level
The Senior SRE will manage multi-cloud infrastructure, ensuring reliability and scalability. Responsibilities include building CI/CD pipelines, defining SLOs, and implementing automation.
The summary above was generated by AI

About Satsuma

Satsuma is a commerce iPaaS that builds merchant-specific APIs, MCP Servers, and MCP Apps, enabling retailers to connect their full commerce stack once and deploy branded shopping experiences across every AI channel. We work with enterprise retailers and move fast. Our infra has to match.

The role

We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-cloud infrastructure. You'll be the person who keeps things running, builds the systems that prevent fires, and makes on-call not terrible.

This is an infra-first role. But we're an AI-native company, and we expect you to use AI-assisted development (Claude Code) as a core part of your workflow — writing tooling, automating runbooks, building internal utilities.

What you'll do

  • Own infrastructure across AWS, GCP, and Azure environments
  • Build and maintain CI/CD pipelines, observability stacks, and incident response workflows
  • Define and enforce SLOs/SLIs; lead postmortems
  • Author and maintain IaC (Terraform preferred)
  • Write internal tooling and automation using AI-assisted development workflows
  • Partner closely with engineering on reliability reviews and architecture decisions

Requirements
  • 5-8 years in SRE, DevOps, or infrastructure engineering
  • Hands-on experience across at least two major cloud providers
  • Strong Kubernetes, Terraform, and observability tooling (Datadog, Grafana, or equivalent)
  • Comfortable reading and editing code; able to ship scripts and internal tools
  • Experience with AI-assisted development (Copilot, Cursor, Claude Code)
  • On-call maturity -- you've owned incidents end-to-end and made systems better afterward
  • Prior experience at a startup or high-growth SaaS company
  • Familiarity with API gateway infrastructure or commerce tech stacks
  • Hands-on experience with MCP or agentic AI infrastructure

Benefits
  • Unlimited PTO
  • 401(K)
  • Healthcare Stipend
  • Gym stipend

Similar Jobs

17 Days Ago
Remote
United States
180K-220K Annually
Senior level
180K-220K Annually
Senior level
Software • Defense
Work as an SRE embedded with product teams to improve reliability by fixing application code (primarily TypeScript), building observability (Prometheus, Loki, Grafana, Alloy), defining SLIs/SLOs, leading incident response and postmortems, automating toil, and supporting deployments across on‑prem DoD and AWS environments.
Top Skills: AlloyAWSBashContainersDockerGithub ActionsGitlab Ci/CdGoGrafanaJenkinsKubectlKubernetesLokiNode.jsPrometheusPythonTypescript
22 Days Ago
Remote or Hybrid
United States
Senior level
Senior level
Fintech • Software
Lead SRE efforts for DFIN SaaS: ensure availability, performance, scalability, and automation. Implement monitoring, CI/CD, IaC, container orchestration, AI-enhanced observability, incident response, RCA, and runbook automation while collaborating across engineering teams.
Top Skills: .NetAiopsAksAnsibleAppdynamicsAWSAzureAzure DevopsBashC#Ci/CdCloud Ai ServicesContainersCosmosDatadogDynatraceEksFirewallHarnessIdera Sql Diagnostic ManagerInfrastructure As Code (Iac)JavaJenkinsKubernetesLinuxLoad BalancingNew RelicPowershellPythonRedgate Sql MonitorSolarwinds Database Performance AnalyzerSQLTerraformWindows
2 Days Ago
In-Office or Remote
3 Locations
105K-192K Annually
Senior level
105K-192K Annually
Senior level
Information Technology • Legal Tech • Analytics
Partner with DBAs and SRE teams to ensure reliable, secure, and scalable database infrastructure. Lead incident management, vulnerability remediation, DR planning and resilience testing, infrastructure provisioning across VMs and Chainguard containers, cost optimization, and SRE capability building. Provide on-call support and coach junior staff.
Top Skills: Amazon Web Services (Aws)ChainguardGitlabGrafanaIaasJenkinsLinuxMySQLPmm3PostgresPrometheusTerraformVm

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account