About the role

Role: HPC Systems Administrator supporting Lilly's AI-driven drug discovery partnership with NVIDIA, focused on building and maintaining scalable high-performance computing platforms for advanced machine learning workloads. Responsibilities include deploying secure, high-availability AI and HPC systems; automating operations for efficiency; optimizing infrastructure for performance, reliability, and cost-effectiveness; and collaborating with scientists and engineers to facilitate model training and experimentation across GPU, cloud, and on-premises environments.
Requirements: Bachelor’s in Computer Science or related field; 5+ years supporting large-scale HPC or GPU compute environments; expertise in Linux administration, scripting (Python, Bash), automation tools (Ansible, Kubernetes), and HPC job schedulers (Slurm, Grid Engine). Preferred skills include experience with NVIDIA GPU infrastructure, cloud platforms (AWS, Azure, GCP), and working in regulated or critical environments. The role is based in Silicon Valley with a hybrid work model (3 days onsite, 2 remote). High-value specifics Focus on AI/ML workloads, GPU hardware management, distributed computing, and automation for scientific research in drug discovery.

Matching similar jobs

JOB OVERVIEW

Experience level

Senior

Location

South San Francisco, CA

Occupation

Network and Computer Systems Administrators

Industry

Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services

Posted

today

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE