Cloud Infrastructure Engineer
Braintrust Data — Tracked from its ashby job board
About the role
ABOUT THE COMPANY
Braintrust is the agent observability platform. By actively applying intelligence to agent traces and automatically surfacing the most critical patterns, Braintrust gives teams the visibility to understand how agents behave in production and the tools to improve them.
Teams at Notion, Stripe, Box, OpenAI, and Cloudflare use Braintrust to trace their agents, find the issues in their observability data, and run evals that tell them how to improve.
ABOUT THE ROLE
We’re looking for a Cloud Infrastructure Engineer to help us build reliable, scalable infrastructure and give developers a world-class platform to ship code with speed and confidence. You’ll lead efforts across Terraform, Kubernetes, CI/CD, observability, and support, and play a key role in how we scale Braintrust both internally and for customers self-hosting our platform.
This is a high-impact role where you’ll contribute across our internal AWS environment and help customers deploy our stack in AWS, Azure, and GCP.
WHAT YOU’LL DO
- Build and maintain Terraform modules for both internal infrastructure and customer deployments
- Work directly with customers in Slack to support self-hosting and troubleshoot infrastructure issues. Build tools to make it easier for them to support themselves.
- Own and improve our CI/CD pipeline: reduce build times, improve failure visibility, and enable safer, faster releases
- Centralize and scale observability - including logs, metrics, dashboards, and alerts
- Partner with engineering teams to build and evolve a secure, developer-friendly infrastructure platform
- Support multi-cloud deployment patterns (AWS primarily, with Azure and GCP support for enterprise customers)
- Implement tools and automation to improve deployment, rollback, and infrastructure reliability
IDEAL CANDIDATE CREDENTIALS
- 5+ years of experience in DevOps, SRE, or Infrastructure Engineering roles
- Deep experience with Terraform and at least one major cloud provider (AWS strongly preferred)
- Strong Kubernetes skills: deploying, debugging, and scaling real workloads
- Proficient in scripting or programming (Python, Typescript, or Go)
- Experience supporting production systems and responding to incidents
- Comfortable working directly with customers in a support or deployment context
- Bonus: experience with multi-cloud environments or self-hosted enterprise software
BENEFITS INCLUDE
- Medical, dental, and vision insurance
- Daily lunch, snacks, and beverages
- Flexible time off
- Competitive salary and equity
- Wifi & cellphone stipend
EQUAL OPPORTUNITY
Braintrust is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.
What the index says about this role
- First seen by JobLarper — Aug 2, 2026, 6 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
- No pay range in our index for this listing. Across 177 indexed DevOps & Site Reliability Engineer roles in US that do publish one, the middle half sits between $165k and $250k, median $204k — JobLarper's read of the market, not a figure from Braintrust Data.
- What DevOps & Site Reliability Engineer roles ask for — across 1,405 indexed openings: Cloud (50%), Python (41%), REST/APIs (33%), Go (31%), Backend (24%). This posting names Cloud, Python, Go.
- Braintrust Data is hiring actively — 13 open roles indexed.
Derived from the 27,000 roles JobLarper indexes daily from official company boards — not from the job description above.
More open roles at Braintrust Data
- Developer Support Engineer (Singapore)Singapore · Mid · SGD 140K – SGD 190K • Offers Equity • Multiple Ranges
- Software Engineer, BackendSan Francisco · New York City · Seattle · Mid
- Data EngineerSan Francisco · New York City · Mid
- Developer Support EngineerSan Francisco · Mid · $157K – $205K • Offers Equity • Multiple Ranges
- Software Engineer, Developer ExperienceSan Francisco · New York City · Seattle · Mid
- Solutions Engineer (East Region)New York City · Mid
- Solutions Engineer (West Region)San Francisco · Seattle · Mid
- Open Source Engineer - GoRemote · New York City · Seattle · Mid
All 13 open roles at Braintrust Data →
Similar DevOps & Site Reliability Engineer roles at other companies
- Finance Systems Engineer, TaxAnthropic · San Francisco, CA | Seattle, WA
- Senior SRE - VOIPRingCentral · Bangalore, India
- Senior OT / Edge DevOps Engineer (m/f/x)Reverion · Eresing · München
- Principal AI Platform EngineerSentinelOne · Brno, South Moravian, Czech Republic
- Lead Cloud DevOps Engineer (Oakland, CA Office)Fictiv · Oakland, CA Office
- Principal Site Reliability EngineerDell Technologies · Bengaluru, Karnataka, India
- Senior Application Support Engineer / Site Reliability Engineer (SRE)DTCC · Boston, MA, United States
- Senior Application Support Engineer (SRE)DTCC · Tampa, FL, United States
Browse all devops & site reliability engineer jobs in united states — 721 open roles across 306 companies.