Staff Site Reliability Engineer
StackBlitz · Remote
You apply on StackBlitz's own site — MulaJob never submits anything for you without your say-so.
h4 data-start= 233 data-end= 248 🚀 About Us /h4
p data-start= 250 data-end= 267 We’re Bolt.new by StackBlitz! /p
p data-start= 269 data-end= 558 We’re the team that brought you WebContainers, the first-of-its-kind technology that made it possible to run Node.js right inside your browser. That breakthrough kicked off our journey in 2019, and it’s what powers the blazing-fast online IDE used by over 1 million developers every month. /p
p data-start= 560 data-end= 1013 But we didn’t stop there. We doubled down on everything we learned and built a class= href= https://bolt.new target= _new data-start= 637 data-end= 669 strong data-start= 638 data-end= 650 Bolt.new /strong /a — the fastest way to go from idea to production without writing traditional code. It’s a next-gen, AI-powered app builder that helps you create, edit, and deploy full-stack web and mobile apps instantly, right in your browser. No installs. No setup. Just smart automation and instant dev environments that let you move at the speed of thought. /p
p data-start= 1015 data-end= 1161 We’re a fully remote team, globally distributed, deeply collaborative, and seriously passionate about building the future of software development. /p
p data-start= 1163 data-end= 1342 This is your chance to join a small team with a big vision. If you love shipping fast, solving real problems, and pushing the boundaries of what’s possible, we’d love to meet you. /p
h4 data-start= 1349 data-end= 1377 ✨ About This Opportunity /h4
p As a Staff Site Reliability Engineer, you'll be the reliability conscience of our engineering organization, embedding with product and platform teams from the earliest stages of a project, shaping designs, and making sure what we build is observable, scalable, and operable long before it reaches production. The heart of this role is making the pager ring less over time, but the pager is real. Every SRE here shares our on-call rotation, and sometimes the work genuinely is rolling up your sleeves and digging into a live incident. /p
p You'll set technical direction, define the standards other engineers build against, and drive initiatives that span multiple teams. This is a high-influence individual-contributor role: you won't manage people, but you will change how the whole organization thinks about reliability. You'll respond to incidents and share the on-call rotation alongside the rest of the team, but your lasting impact is the incidents that never happen because reliability was designed in from the start, at the scale of millions of developers building real products on Bolt.new every day. /p
h4 data-start= 1545 data-end= 1567 🛠️ How You'll Contribute /h4
ul
li strong Embed With Teams Early: /strong Partner with development teams throughout the project lifecycle, from design and architecture reviews through launch readiness. Bringing an SRE perspective before code is written, not after it breaks. Shepherd projects to completion with reliability designed in. /li
li strong Define Production-Readiness Standards: /strong Establish and evolve the design reviews, launch checklists, and operational acceptance criteria that projects pass through, and own how teams adopt them across the org. /li
li strong Make Reliability Measurable: /strong Define meaningful SLIs, SLOs, and error budgets in collaboration with product and engineering, and help teams use them to make real prioritization decisions. /li
li strong Build the Paved Roads: /strong Create the frameworks, tooling, and golden paths across AWS, GCP, and Azure, with Terraform as the common backbone, that make the reliable way the easy way for every engineer. /li
li strong Cross-Team Leadership: /strong Partner across engineering, product, and design to align reliability work with business objectives. Influence roadmaps, resolve technical disagreements, identify process and technical debt across the organization, and propose solutions that accelerate velocity for multiple teams. Mentor senior and mid-level engineers, raising the bar for operational excellence everywhere. /li
li strong Mature Our Incident Practice: /strong Lead by influence on incident management and blameless postmortems, turning failure modes and operational signals into systematic, durable improvements. /li
li strong Represent Us Externally: /strong Build relationships with our cloud and infrastructure provider teams to influence roadmaps and unlock early access to new capabilities, and represent StackBlitz in customer trust conversations and the broader reliability community. /li
li strong On-call rotation: /strong Every SRE shares our on-call rotation strong , currently one week per month /strong . /li
/ul
h4 data-start= 1663 data-end= 1706 💡 Qualifications /h4
ul
li strong Multi-Cloud Fluency: /strong General fluency across AWS, GCP, and Azure matters more to us than deep specialization in any one, we run across all three. Terraform is our common infrastructure-as-code layer everywhere. /li
li strong Our Stack: /strong Comfort supporting and contributing to TypeScript (frontend and backend) and Ruby on Rails (backend) services. We're opinionated about our stack, and you'll work alongside it daily. /li
li strong SRE / Production Engineering Experience: /strong Significant experience as an SRE, production/platform engineer, or software engineer with a deep reliability focus, including time operating at scale. /li
li strong Software Engineering Excellence: /strong Strong software engineering fundamentals; you write production-quality code and can go deep with the teams you partner with, balancing immediate needs against long-term maintainability. /li
li strong Technical Leadership Influence: /strong A track record of changing how teams work, not just how systems run, leading across team boundaries without formal authority. /li
li strong Strategic Execution: /strong Ability to take ambiguous, high-scope problems and drive them to completion with minimal oversight. /li
li strong Systems Thinking: /strong Ability to identify process, communication, and technical debt across the organization and propose solutions that accelerate velocity for multiple teams. /li
li strong Data-Driven Leadership: /strong Experience building measurement and evaluation frameworks, identifying patterns in operational data, and translating findings into organizational improvements. /li
li Strong verbal and written English communication skills are required, as this role involves frequent collaboration with team members, stakeholders, customers, and external audiences where English is the primary working language. /li
/ul
h4 data-start= 1850 data-end= 1873 🎯 Bonus Points /h4
ul
li Experience standing up or maturing an SRE practice at a growth-stage company. /li
li Background working as an embedded SRE or partnering closely with product teams. /li
li Experience designing chaos/resilience testing or progressive delivery practices. /li
/ul
h4 data-start= 1974 data-end= 1992 📌 A Few Notes /h4
ul data-start= 1994 data-end= 2199
li data-start= 1994 data-end= 2043
p data-start= 1996 data-end= 2043 You do strong data-start= 2003 data-end= 2010 not /strong need a college degree to apply /p
/li
li data-start= 2044 data-end= 2134
p data-start= 2046 data-end= 2134 You do strong data-start= 2053 data-end= 2060 not /strong need to be located in the U.S. — we’re remote-friendly /p
/li
li data-start= 2044 data-end= 2134
p data-start= 2046 data-end= 2134 You do strong data-start= 2144 data-end= 2151 not /strong need to meet every qualification listed above br br /p
/li
/ul