Site Reliability Engineer

Blitzy

Completely RemoteFull TimeInformation Technology
Posted Today

Job description

About the Company

Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software. We're transforming how enterprises build software, turning enterprise requirements into production-ready code with an agentic software development platform.

Responsibilities

  • Deploy, operate, and maintain Blitzy's self-hosted platform within a customer-controlled, secure cloud environment
  • Own Kubernetes-based deployments, including releases, upgrades, capacity planning, and performance benchmarking
  • Design and maintain observability including logging, metrics, tracing, and alerting within security boundaries
  • Partner with customer infrastructure, security, and governance teams on provisioning and operational escalations
  • Handle sensitive customer data in accordance with security requirements and champion best practices
  • Feed operational lessons back into the product and infrastructure roadmap

Requirements

  • U.S. citizenship required for customer badging
  • 3+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering
  • Strong proficiency in Kubernetes and container orchestration
  • Experience deploying software into customer-controlled or restricted environments
  • Experience in highly regulated network environments (defense, government, or financial services)
  • Hands-on experience with Infrastructure-as-Code (Terraform, Pulumi, or equivalent)
  • Proficiency with at least one major cloud platform
  • Expertise in observability tooling and incident management
  • Strong scripting and automation skills in Python, Go, or Bash

Preferred Qualifications

  • Experience operating in government-accredited cloud environments
  • Familiarity with security frameworks for regulated industries
  • Experience supporting AI/ML workloads or infrastructure
  • Prior experience as a forward-deployed or embedded engineer at an enterprise site
  • Experience in a high-growth startup environment

Skills & tools

KubernetesTerraformPython

What the team is looking for

Use this list as a quick fit check before you apply.

  1. 01U.S. citizenship
  2. 023+ years SRE/DevOps experience
  3. 03Kubernetes proficiency
  4. 04Infrastructure-as-Code experience
  5. 05Scripting skills (Python, Go, or Bash)
NeverApplyAd

Wake up to a shortlist, not a search results page.

NeverApply scores every new listing against your CV, salary floor and visa. A handful of real matches by morning.

Get your daily matches