LandMe

Senior HPC & GPU Infrastructure Engineer

Sciforium · San Francisco, California

$150,000–$220,000/yrOnsiteFull-timeDirect from the employer

Posted Oct 5, 2026 · Verified open Oct 9, 2026

Apply on the employer's siteFind more jobs like this on LandMe

About the job

Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary, high-efficiency serving platform. Backed by multi-million-dollar funding and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier AI models and real-time applications.

ABOUT THE ROLE

We are seeking a Senior HPC & GPU Infrastructure Engineer to take full ownership of the health, reliability, and performance of our GPU compute cluster. You will be the primary custodian of our high-density accelerator environment and the linchpin between hardware operations, distributed systems, and machine learning workflows. This role spans everything from hands-on Linux systems engineering and GPU driver bring-up to maintaining the ML software stack (CUDA/ROCm, PyTorch, JAX, vLLM). If you love squeezing every bit of performance out of hardware, enjoy debugging GPUs at scale, and want to build world-class AI infrastructure, this role is for you.

WHAT YOU'LL DO

1. System Health & Reliability (SRE)

2. Linux & Network Administration

3. GPU & ML Stack Engineering

IDEAL CANDIDATE PROFILE

NICE-TO-HAVE

BENEFITS INCLUDE

EQUAL OPPORTUNITY

Sciforium is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.