Senior SRE, AIOps Platform for GPU Data Centers
nvidiaSanta Clara, CA
Senior SRE, AIOps Platform for GPU Data Centers
nvidiaSanta Clara, CA
yesterday
Occupations
Computer Systems Engineers/ArchitectsSoftware DevelopersNetwork and Computer Systems AdministratorsIndustries
Computing Infrastructure Providers, Data Processing, Web Hosting, and Related ServicesSoftware PublishersComputer Systems Design ServicesAbout the role
NVIDIA Corporation in Santa Clara, CA is hiring a Dev Ops Engineer to operate our AI Data Center telemetry platform. You’ll own reliability, incident response, and postmortems for telemetry ingestion, processing, storage, and APIs/dashboards used by operators.
Expect to lead Kubernetes deployments end-to-end, build runbooks, and partner with Software and Systems Engineering to translate platform signals into actionable, trustworthy alerts and automation.
Matching similar jobs
JOB OVERVIEW
Experience level
Lead
Location
Santa Clara, CA
Occupation
Computer Systems Engineers/Architects
Industry
Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services
Posted
yesterday
Tired of running searches?
Rank the roles you'd take once, and matches like these arrive on their own.
CREATE PROFILE