JobLarper
JobLarperCompaniesPicogrid › Site Reliability Engineer

Site Reliability Engineer

Picogrid — Tracked from its ashby job board

El Segundo, CA Mid $170K – $195K • Offers Equity • Offers Bonus Posted Jul 30, 2026
React NativeCloud

About the role

Who we are

Picogrid is a leading venture-backed defense technology company founded to bridge the decades-long gap between modern technology and the critical demands of national security. Today, we're building the essential infrastructure to unify sensors, autonomy, and operators with our technology deployed in active operations around the world. Our mission is to deliver an operational advantage to secure the United States and its allies.

ABOUT THE ROLE

As Picogrid's first Site Reliability Engineer you will own production reliability across cloud and edge, from observability and incident response through node lifecycle, stateful workloads, and a fleet of hardware edge devices in the field. You will help build and define the systems, processes and best practices that ensure Picogrid's systems can be relied upon by our warfighters in even the toughest battlefield conditions. You will work with engineers to build a strong on-call culture where issues are root caused swiftly, and ensure our alerting and monitoring have exceptional coverage and signal-to-noise ratio.

Security is a shared responsibility across all our DevSecOps roles, and as part of a scrappy startup team you will be expected to help stand up new infrastructure and other related DevSecOps tasks as needed.

RESPONSIBILITIES

- Own, define and drive our reliability SLIs and SLOs for cloud deployments

- Own, define and drive our reliability SLIs and SLOs for our edge devices deployed in remote and sometimes contested areas

- Own the observability stack: Grafana, Prometheus, Loki, and OpenTelemetry, with dashboards versioned in git and alerting rules checked in alongside the code they watch

- Participate in on-call and incident response: log-first troubleshooting, blameless postmortems, and follow-up hardening

- Encode reliability into infrastructure as code

REQUIRED QUALIFICATIONS

- 3+ years of experience as an SRE or related roles

- Deep Kubernetes operations experience: node lifecycle, workload scheduling, StatefulSets, graceful drains, and live cluster debugging

- Experience designing comprehensive observability dashboards and high signal-to-noise ratio alerting rules

- You are a competent and experienced incident responder practicing methodical evidence-first triage, blameless postmortems, and turning incidents into durable guardrails

- Production Terraform or OpenTofu experience

- Fluent in AWS including IAM, networking, multi-account environments, and account and workload hardening

- Experience managing high availability database deployments

- IoT or edge fleet operation experience

- Comfortable operating in scrappy, fast-paced environments, and turning ambiguous requirements into concrete solutions

- You optimize for providing value early in projects and short iteration cycles

PREFERRED QUALIFICATIONS

- GovCloud, FIPS, or other regulated or air-gapped environment experience

- Constrained edge hardware such as NVIDIA Jetson platforms (AGX Thor, Orin Nano), including shared CPU and GPU memory and thermal constraints

- Overlay or mesh networking operations: Nebula, WireGuard, Tailscale, or similar

- Standing up SLO and error-budget tooling (sloth, Pyrra, or equivalent) from scratch

- Active security clearance

COMPENSATION & BENEFITS

- Base salary range: $170,000 - $195,000 per year. Base salary is just one part of your total compensation package at Picogrid.

- Significant stock options with a high potential upside as an early-stage company

- 401(k) with employer matching

- Full health coverage (medical, dental, and vision insurance)

- Relocation assistance provided (if applicable)

- Unlimited PTO (two-week minimum) and 11 paid holidays per year

- Paid parental leave for both parents

- Lunch provided when working in-office and a fully stocked kitchenette

- Free EV charging at the HQ

- Unique office in El Segundo, CA stocked with quality coffee, snacks, and craft beer

EXPORT CONTROL REQUIREMENTS

To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State.

EQUAL EMPLOYMENT OPPORTUNITY (EEO) POLICY

Picogrid is committed to providing a professional work environment free from discrimination, harassment, and retaliation. We are an equal opportunity employer and make all employment decisions based on merit, qualifications, and business needs.

#LI-DNP

Equal Employment Opportunity (EEO) Policy

Picogrid is committed to providing a professional work environment free from discrimination, harassment, and retaliation. We are an equal opportunity employer and make all employment decisions based on merit, qualifications, and business needs.

What the index says about this role

  • First seen by JobLarper — Aug 2, 2026, 6 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
  • What DevOps & Site Reliability Engineer roles ask for — across 1,405 indexed openings: Cloud (50%), Python (41%), REST/APIs (33%), Go (31%), Backend (24%). This posting names Cloud, React Native.
  • Picogrid is hiring actively — 11 open roles indexed.

Derived from the 27,000 roles JobLarper indexes daily from official company boards — not from the job description above.

⚡ JobLarper watched this role appear on Picogrid's official board on Jul 30, 2026. Sign up free to get alerted minutes after roles like this go live, and tailor your real résumé to the exact description — nothing invented.

More open roles at Picogrid

All 11 open roles at Picogrid →

Similar DevOps & Site Reliability Engineer roles at other companies