Staff Systems Engineer
Blacksmith — The fastest way to run your GitHub Actions
About the role
About Blacksmith
- We started by building infrastructure to run CI workloads really fast. Our first product helps companies run GitHub Actions substantially faster and cheaper by owning and operating our own global fleet of bare-metal machines rather than renting generic cloud VMs.
- Today, we orchestrate tens of millions of Firecracker VMs each month, running CI for 3,000+ companies and hit ~$10M in ARR in less than 2 years. We’ve more than tripled revenue since the start of 2026.
- We operate thousands of bare-metal machines across multiple regions, regularly schedule 100k+ vCPUs concurrently, and run a petabyte-scale Ceph cluster that we manage ourselves.
- We’ve raised $13.5M across Seed and Series A, led by Google Ventures (GV), and we’re intentionally building a small, but exceptional team.
- Blacksmith was founded by a team with deep systems and scaling experience, including building search/ads infrastructure at Faire, and operating large distributed systems at Cockroach Labs. Our GTM is led by Jon Boyer, formerly Head of Sales at Zapier.
- We’re now extending the same CI infrastructure into a broader platform: running agent sandboxes at scale and building our own background coding agent on top of it.
The Staff Systems Engineer at Blacksmith
We are looking for a Staff Engineer to help set the technical direction for our core infrastructure, critical to the Blacksmith product. This role operates at the organization level, owning architecture, solving complex systems problems, and driving initiatives that shape how Blacksmith grows, performs, and remains reliable.
You will be one of the key technical authorities and a force multiplier at Blacksmith. You’ll raise engineering standards across the team, mentor senior talent and own production outcomes.
This is a high impact, hands-on technical leader role.
What you’ll be doing
- Own and drive the architectural direction for critical infrastructure platforms.
- Translate business and product strategy into long-term technical roadmaps and execution plans.
- Accountable for production outcomes, including reliability, performance, and operational excellence.
- Mentor senior engineers and act as a force multiplier.
- Operate effectively in ambiguous problem spaces where both the problem and the solution need to be defined.
You’re a good fit if you have
- Deep familiarity with virtualization technologies such as Firecracker, QEMU, or Cloud Hypervisor, and an understanding of the tradeoffs between them.
- Experience with distributed storage systems (e.g., Ceph) and a strong grasp of low-level file system performance, I/O tuning, and storage reliability at scale.
- Comfort working at the kernel and filesystems level: experience with eBPF, cgroups, namespaces, or similar Linux primitives.
- Track record of leading through influence and shaping technical direction while operating as an individual contributor.
- Strong communication skills, with the ability to explain complex systems-level concepts to engineers, leadership, and non-technical stakeholders.
Compensation and benefits
- Medical, Vision, and Dental insurance.
- Competitive base + equity.
- 401K match.
- Unlimited PTO.
- Annual offsite.
- Early-exercise stock options
- 12 weeks fully paid parental leave (US)
What the index says about this role
- First seen by JobLarper — Aug 2, 2026, 6 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
- What DevOps & Site Reliability Engineer roles ask for — across 1,405 indexed openings: Cloud (50%), Python (41%), REST/APIs (33%), Go (31%), Backend (24%). This posting names Backend.
- Blacksmith is hiring actively — 9 open roles indexed.
Derived from the 27,000 roles JobLarper indexes daily from official company boards — not from the job description above.
More open roles at Blacksmith
- Enterprise Product ManagerNew York City · Manager · $200K – $240K • Offers Equity
- General Manager, SandboxesNew York City · Manager · $250K – $350K • Offers Equity
- Staff Product EngineerNew York City · Senior+ · $280K – $380K • Offers Equity
- Technical Support EngineerSF / NYC · Mid · $160K – $180K • Offers Equity
- Solutions EngineerNew York City · Mid · $200K – $240K • Offers Commission
- General ApplicationSF / NYC · Mid
- Senior Systems EngineerNew York City · Senior+ · $200K – $300K • Offers Equity
- Senior Product EngineerNew York City · Senior+ · $200K – $300K • Offers Equity
All 9 open roles at Blacksmith →
Similar DevOps & Site Reliability Engineer roles at other companies
- Finance Systems Engineer, TaxAnthropic · San Francisco, CA | Seattle, WA
- Senior SRE - VOIPRingCentral · Bangalore, India
- Senior OT / Edge DevOps Engineer (m/f/x)Reverion · Eresing · München
- Principal AI Platform EngineerSentinelOne · Brno, South Moravian, Czech Republic
- Lead Cloud DevOps Engineer (Oakland, CA Office)Fictiv · Oakland, CA Office
- Principal Site Reliability EngineerDell Technologies · Bengaluru, Karnataka, India
- Senior Application Support Engineer / Site Reliability Engineer (SRE)DTCC · Boston, MA, United States
- Senior Application Support Engineer (SRE)DTCC · Tampa, FL, United States
Browse all devops & site reliability engineer jobs in united states — 721 open roles across 306 companies.