SRE role owning reliability for Blitzy's AI-powered software development platform at a highly regulated enterprise customer, with embedded on-account operations.
Responsibilities
- ▹Deploy, operate, and maintain Blitzy's self-hosted platform within a customer-controlled, secure cloud environment
- ▹Own the Kubernetes-based deployment: releases, upgrades, capacity planning, and performance benchmarking for compute-intensive AI workloads
- ▹Design and maintain observability (logging, metrics, tracing, alerting) that operates fully within the customer's security boundary
- ▹Serve as Blitzy's on-account technical presence, partnering with customer infrastructure, security, and governance teams on provisioning, reviews, documentation, and operational escalations
- ▹Handle sensitive customer data in accordance with security requirements and champion security best practices
- ▹Feed lessons learned back into Blitzy's product and infrastructure roadmap for future public sector customers
Requirements
- ▹U.S. citizenship (customer badging requirement) and ability to complete a customer background/badging process
- ▹3+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles
- ▹Strong proficiency in Kubernetes and container orchestration, with hands-on experience
Soft skills
Deep ownership of the systems operatedThriving in a fast-moving environmentBuilding trust with the customer's infrastructure, security, and platform teams
What we offer
- ▹$140,000-$170,000 base salary, plus bonus and equity commensurate with experience
- ▹Remote work (U.S.), with occasional travel for key customer workshops
- ▹No security clearance required
About the company
Blitzy is a Cambridge, MA based AI software development platform whose agentic development platform can autonomously execute up to 80% of enterprise software development work.
Similar jobs

Job
DevOps Engineer IV (Observability)
Jumio
AI/MLCloudformationDatadog
+6
💰 Salary: not specified
🌍 Remote
Anywhere in the World
🗣️ EN

Job
HPC Infrastructure Site Reliability Engineer
Radiant
AI/ML
+3
💰 Salary: not specified
🌍 Remote
Gloucestershire
🗣️ EN

Job
DevOps Engineer
Sophos
AI/MLCloudformation
+8
💰 Salary: not specified
🏢 On-site
🗣️ EN
Job
Cloud Security & DevOps Engineer
Ajaia
+5
$28,000–$35,000/yr
gross
🌍 Remote
🗣️ EN

Job
Site Reliability Engineer
Filevine
CloudformationPulumi
+2
$147,000–$168,000/yr
gross
🏢 On-site
🗣️ EN

Job
Streaming Infrastructure DevOps Engineer
Armis Security
Istio
+7
💰 Salary: not specified
🏢 On-site
Tel Aviv-Yafo
🗣️ EN
