ML Inference Engineer - Low-Latency GPU Systems
susquehanna international groupBala Cynwyd (PA)
ML Inference Engineer - Low-Latency GPU Systems
Posted yesterday
susquehanna international groupBala Cynwyd (PA)
SENIORITY
Mid
About the role
Susquehanna International Group, LLP is seeking a Machine Learning Engineer in Bala Cynwyd, PA. This role focuses on low-latency inference optimization for high-performance model serving systems.
You will collaborate with researchers to optimize performance, evaluate frameworks, and debug GPU memory issues while managing inference workloads effectively. A strong background in modern ML frameworks, programming experience, and understanding of production environments is essential.
Before you apply
Applying takes about a minute. These four things decide how fast it moves after that.
Your profile is current
It's what we read first. Occupations, seniority and locations matter more than a long history.
Two examples you can talk through
Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.
A number in mind
What you're on now and what would make you move. We negotiate better when we know both.
Your notice period
Employers plan around it, and it's the question that stalls offers most often.
Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.
More like this
