Blitzy Logo

Blitzy

Site Reliability Engineer - Public Sector

Posted 19 Days Ago
Remote
Hiring Remotely in USA
140K-170K Annually
Mid level
Remote
Hiring Remotely in USA
140K-170K Annually
Mid level
Operate and harden Blitzy's self-hosted, Kubernetes-based AI platform inside customer-controlled secure cloud environments. Own deployments, upgrades, capacity planning, observability, incident response, and customer-facing technical coordination while championing security and feeding operational learnings back into the product roadmap.
The summary above was generated by AI
About Blitzy

Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software to unlock the next industrial revolution. We're transforming how enterprises build software, turning enterprise requirements into production-ready code with an agentic software development platform that can autonomously execute 80% of the quantum of software development work. We're backed by multiple tier 1 investors, and have proven success as founders of previous start-ups.


Location:
Remote (U.S.), with occasional travel for key customer workshops

Compensation: $140,000 - $170,000 base salary (bonus + equity commensurate with experience)

Eligibility: U.S. citizenship required (customer badging requirement); no security clearance required

The Role

As a Site Reliability Engineer on Blitzy's Public Sector team, you will be the backbone of our platform's reliability and operational excellence for a dedicated enterprise customer opportunity in a highly regulated industry. You'll deploy and operate Blitzy's self-hosted platform within the customer's secure cloud environment, serving as Blitzy's embedded engineer on the account. You'll work at the intersection of software engineering, infrastructure, and customer success, ensuring our AI-powered development platform remains highly available and performant in one of the most demanding security environments in enterprise software. This is a high-impact, hands-on role for an engineer who thrives in a fast-moving environment and takes deep ownership of the systems they operate.

What Success Looks Like
  • In 30 days: You have a deep understanding of Blitzy's self-hosted deployment architecture, have begun customer onboarding, and are actively supporting deployment planning alongside customer infrastructure teams.

  • In 90 days: You are operating inside the customer environment, have stood up and hardened the deployment, and have established monitoring, alerting, and incident response workflows that meet the customer's security requirements.

  • In 6 months: The platform is running reliably at scale with defined SLOs and validated capacity headroom, and you are the trusted technical voice for the customer's infrastructure, security, and platform teams.

Areas of Ownership
  • Deploy, operate, and maintain Blitzy's self-hosted platform within a customer-controlled, secure cloud environment.

  • Own the Kubernetes-based deployment: releases, upgrades, capacity planning, and performance benchmarking for compute-intensive AI workloads.

  • Design and maintain observability — logging, metrics, tracing, and alerting — that operates fully within the customer's security boundary.

  • Serve as Blitzy's on-account technical presence: partner with customer infrastructure, security, and governance teams on provisioning, reviews, documentation, and operational escalations.

  • Handle sensitive customer data in accordance with customer security requirements; champion security best practices across the deployment.

  • Feed lessons learned back into Blitzy's product and infrastructure roadmap to strengthen our self-hosted offering for future public sector customers.

Required Experience
  • U.S. citizenship (customer badging requirement) and ability to complete a customer background/badging process.

  • 3+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles.

  • Strong proficiency in Kubernetes and container orchestration; hands-on experience deploying software into customer-controlled or restricted environments.

  • Experience operating in isolated, restricted, or otherwise highly regulated network environments (defense, government, financial services, or similar).

  • Hands-on experience with infrastructure-as-code tools (Terraform, Pulumi, or equivalent) and at least one major cloud platform.

  • Deep expertise in observability tooling, incident management, and on-call practices.

  • Strong scripting and automation skills (Python, Go, Bash, or similar).

  • Excellent communication skills — you will work directly with customer engineering, security, and governance stakeholders and represent Blitzy on the account.

What Makes You Stand Out
  • Experience deploying or operating software in government-accredited or similarly certified cloud environments.

  • Familiarity with handling sensitive data and security frameworks common to regulated industries.

  • Experience supporting AI/ML workloads or their supporting infrastructure.

  • Prior forward-deployed, residency, or embedded-engineer experience at an enterprise customer site.

  • Prior experience in a high-growth startup environment where you wore multiple hats.

What Makes This Role Different

You won't be maintaining legacy systems or fighting fires in a sprawling monolith. You'll be a founding reliability engineer for Blitzy's public sector business — operating a greenfield deployment of an AI platform that is redefining how the world creates software, inside one of the most security-conscious enterprises in the country. Your work directly shapes the playbook for the government and defense customers that follow. You'll have direct influence over architectural decisions, work side-by-side with world-class engineers, and see the tangible impact of your work from day one.

Our Culture

Who we are:

Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.


How we work:

We move Blitzy Fast: Time is both our company's and our clients' most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.


Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.

Passion for Invention: We're pushing the frontier of what's possible, requiring constant innovation and iteration.


We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.


We believe in being 'everyday athletes'—taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for optimal mental
performance. It makes for a happier and more productive team.


Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger.

Similar Jobs

2 Days Ago
Easy Apply
Remote or Hybrid
2 Locations
Easy Apply
127K-249K Annually
Senior level
127K-249K Annually
Senior level
Big Data • Cloud • Software • Database
As a Senior Site Reliability Engineer, you'll design and build complex systems, support Atlas platform operations, automate processes, and ensure high availability of services.
Top Skills: AWSAzureDnsGCPGoHTTPLinuxPythonRubyTls
12 Days Ago
Remote
United States of America
185K-227K Annually
Senior level
185K-227K Annually
Senior level
Other
The Senior Site Reliability Engineer at Juul Labs ensures operational stability and performance of hybrid cloud infrastructure, leads automation, and handles critical incidents.
Top Skills: AWSBashCloudFormationGCPNutanixPowershellPythonTerraform
9 Minutes Ago
Remote or Hybrid
179K-322K Annually
Senior level
179K-322K Annually
Senior level
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Leads a specialized casualty claims team managing severe-exposure portfolios, complex litigation, coverage issues, reserves, settlements, vendors, and claim-handling protocols. Partners with legal, underwriting, actuarial, reinsurance, and other stakeholders to develop strategies, monitor litigation trends, address legal system abuse, manage financial outcomes, and improve operational performance. Coaches claims leaders, oversees staffing and portfolio capacity, and presents strategic recommendations to executives and business partners.
Top Skills: Guidewire Claims System

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account