ML Systems Engineer — Accelerate GPU Inference & Training

Posted 2 days ago

aionia groupMenlo Park (CA)

SENIORITY

Lead

Apply

About the role

Aionia Group in Menlo Park, CA is hiring a Member of Technical Staff, ML Systems to accelerate model training and inference across image, video, and world-model workloads. You will work with a founding team on kernels, runtimes, and distributed engines that power production-scale ML stacks. You’ll optimize GPU performance, profile bottlenecks with Nsight, and implement low-level CUDA and Triton improvements.

Before you apply

Applying takes about a minute. These four things decide how fast it moves after that.

Your profile is current

It's what we read first. Occupations, seniority and locations matter more than a long history.

Two examples you can talk through

Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.

A number in mind

What you're on now and what would make you move. We negotiate better when we know both.

Your notice period

Employers plan around it, and it's the question that stalls offers most often.

Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.

More like this