DevOps Team Lead
Fundamental — Tracked from its ashby job board
About the role
ABOUT FUNDAMENTAL
Fundamental is an AI company pioneering the future of enterprise decision-making. Founded by DeepMind alumni, Fundamental has developed NEXUS – the world's most powerful Large Tabular Model (LTM) – purpose-built for the structured records that actually drive enterprise decisions. Backed by world class investors and trusted by Fortune 100 companies, Fundamental unlocks trillions of dollars of value by giving businesses the Power to Predict.
At Fundamental, you'll work on unprecedented technical challenges in foundation model development and build technology that transforms how the world's largest companies make decisions. This is your opportunity to be part of a category-defining company from the ground-up. Join the team defining the future of enterprise AI.
KEY RESPONSIBILITIES
- Lead and mentor a team of DevOps engineers, fostering technical growth and collaboration
- Define and drive the infrastructure roadmap aligned with company objectives
- Architect and oversee cloud infrastructure design and implementation
- Establish best practices, standards, and processes for infrastructure development and operations
- Partner with Engineering, Research, and FDE to align infrastructure capabilities with business needs
- Drive the evolution of Kubernetes clusters optimized for GPU workloads, Production SaaS hosting and and varied enterprise deployment models
- Champion GitOps practices using ArgoCD for continuous deployment
- Establish infrastructure as code standards using Terraform
- Define monitoring and observability strategy for distributed systems
- Collaborate with ML engineers to optimize infrastructure for model training and serving
- Own infrastructure reliability, performance, and security posture
- Implement and maintain cost optimization strategies (FinOps) for cloud resources
MUST HAVE
- 7+ years of experience in cloud infrastructure and DevOps, with 3+ years in a technical leadership role
- Proven track record of building and leading high-performing infrastructure teams
- Strong experience with AWS, GCP and Azure
- Deep expertise in Kubernetes, including multi-cluster management, GPU workload optimization, resource scheduling and autoscaling, and network policies and security
- Extensive experience with cloud networking, including VPC design, load balancer configuration, network security and segmentation, and cross-cloud networking solutions
- Strong CI/CD expertise, preferably with GitHub Actions
- Proficiency in Terraform
- Proficiency with GitOps tools (ArgoCD preferred)
- 3+ years of experience with Python
- Experience with monitoring and observability tools
- Experience with FinOps practices and cloud cost optimization
- Excellent communication skills with ability to translate technical concepts for diverse audiences
NICE TO HAVE
- Experience with ML workflow tooling (MLflow, Kubeflow, or similar)
- Experience with FastAPI and backend applications
- Familiarity with data platforms like Databricks or Snowflake
- SRE practices experience or cloud security certifications
- Hands-on experience with Prometheus, Grafana, or Datadog
- Experience scaling infrastructure for AI/ML startups
BENEFITS
- Competitive compensation with salary and equity
- Comprehensive health coverage for you and your dependents
- Paid parental leave for all new parents, inclusive of adoptive and surrogate journeys
- Relocation support for employees moving to join the team in one of our office locations
- A mission-driven, low-ego culture that values diversity of thought, ownership, and bias toward action
What the index says about this role
- First seen by JobLarper — Aug 2, 2026, 6 days ago. Older postings collect hundreds of applicants — a tailored résumé matters more the longer a role has been live.
- What DevOps & Site Reliability Engineer roles ask for — across 1,405 indexed openings: Cloud (50%), Python (41%), REST/APIs (33%), Go (31%), Backend (24%). This posting names Cloud, Python, Backend, ML.
- Fundamental is hiring actively — 18 open roles indexed.
Derived from the 27,000 roles JobLarper indexes daily from official company boards — not from the job description above.
More open roles at Fundamental
- MLOps EngineerEurope · Israel (remote) · Mid
- MLOps Team LeadEurope · Israel (remote) · Senior+
- Data Scientist - ExtensionsEurope · Israel (remote) · Mid
- Backend Engineer - ExtensionsEurope · Israel (remote) · Mid
- Solutions ArchitectSan Francisco · Senior+
- Data Scientist (Forward Deployed)US (remote) · Mid
- Senior Technical Program ManagerSan Francisco · Israel (remote) · Manager
- Solutions ArchitectJapan · Senior+
All 18 open roles at Fundamental →
Similar DevOps & Site Reliability Engineer roles at other companies
- Senior Site Reliability Engineer (SRE & Platform Reliability)Affirm · Remote Poland · Remote
- Principal AI Platform EngineerSentinelOne · Brno, South Moravian, Czech Republic
- Senior Systems EngineerVoyager Space · Pittsburgh, PA; Remote · Remote
- Sr. Systems Engineer (Internal Support)Atlas Technica · Kyiv, Ukraine · Remote
- Staff AI Platform Engineer, Infrastructure ServicesSentinelOne · United States - Remote · Remote
- Senior AI Platform Engineer, Infrastructure ServicesSentinelOne · United States - Remote · Remote
- Staff AI Platform EngineerRelevance AI · Sydney, Australia · Remote
- Senior DevOps EngineerViz.ai · Tel Aviv, Israel · Remote
Browse all devops & site reliability engineer jobs in europe — 176 open roles across 114 companies.