About the role

What to ExpectYour role is to run and harden the infrastructure our own software and our AI systems run on: the servers, the network, the storage and the Kubernetes cluster behind them, all self-hosted on hardware we own. That is bare-metal Linux and a cluster on top of it, the network and identity by which people and machines reach the estate, backups proved by restores, monitoring and incident response, and security as a daily habit. You will take an estate built by hand and make it reproducible from version control, plan capacity and hardware ahead of demand, including the GPU compute our AI workloads run on, and extend the agents that carry routine operational work. We are looking for infrastructure engineers with experience and interest in running Linux and Kubernetes on hardware they own, networking and storage, backups and the restores that prove them, and security and automation. Every role works with our AI systems daily. What can be deterministic, must be. You need no AI background; we prefer people without one. We hire for your knowledge and experience in the field first, so you can steer the ship; the tooling is a learning curve we expect you to take on. We work on-site in El Segundo. This is hard work, but you will be rewarded with equity in a company we believe will become one of the world's most valuable.
What You'll Do:
  • Bare-metal Linux hosts: provisioning, systemd, disks and RAID, encryption at rest, out-of-band management, racking
  • The Kubernetes cluster: scheduling, ingress and TLS, persistent volumes, network policies, upgrades and rollbacks
  • Network, access and identity: routing and firewalling, VPN, DNS, certificates, mutual TLS, single sign-onStorage and backups: capacity layout, retention design, offsite copies, and restore drills actually run
  • Monitoring and incident response: metrics, structured logs, alert routing, on-call, postmortem follow-through
  • Security: hardening and patching, secrets management, key rotation, least privilege, external exposure review
  • Infrastructure as code: configuration in version control, self-hosted source control, registry and CI runners
  • Capacity and hardware planning: hosts, storage, network, and the GPU compute AI workloads need
  • What You'll Bring
  • A degree in computer engineering or computer science, a systems administration background, or equivalent experience
  • Hands-on Linux administration on machines you were responsible for: systemd, networking, storage and filesystems
  • Kubernetes in production: workloads and storage you sized, ingress you configured, upgrades you have run
  • Fluent with routing, firewalls, DNS, VPNs, TLS and certificates when something is unreachable
  • Backups proved by restores you have performed, and security discipline held to without being asked
  • Ready to put a hand-built estate into code, and to take an outage to its cause
Bonus: GPU and machine learning infrastructure, Ansible or Terraform on a real estate, or serious self-hosting Expected Compensation$100k to $250k base salary, plus equity Equity participation in a high-growth startup Comprehensive health, dental, and vision insurance 401(k) with company matching Bonus for living within 5 miles of our El Segundo facility On-site work with rare work-from-home exceptions Merit-based organization where contribution drives reward

Matching similar jobs

JOB OVERVIEW

Salary

$100,000 - $250,000 per year

Experience level

Lead

Location

El Segundo, CA

Occupation

Computer Systems Engineers/Architects

Industry

Computer Systems Design Services

Posted

2 days ago

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE