Machine Learning Engineer

Posted 2 days ago

cna searchDenver (CO)

SENIORITY

Lead

Apply

About the role

We’re looking for a Senior AI/ML Research Engineer to work on large-scale AI inference, synthetic data generation, and distributed training systems. This is a hands-on research and engineering role for someone who has built infrastructure supporting large-scale AI models and wants to work on difficult problems across inference performance, distributed systems, reinforcement learning, and synthetic data. No third parties
What You’ll Do: Design and build large-scale synthetic data generation pipelines and orchestration systems. Optimize AI inference workloads for performance, cost, memory, and compute utilization. Build and improve distributed inference infrastructure for large-scale AI models. Contribute to open-source libraries and frameworks for synthetic data generation and distributed reinforcement learning. Research and implement new approaches to model inference, training, and post-training. Collaborate with researchers and engineers working on large-scale AI systems. Contribute research suitable for publication at conferences such as ICML and NeurIPS.Translate highly technical work into clear technical content for developers and users. Stay current with advances in AI infrastructure, distributed inference, synthetic data, and model optimization.
What We’re Looking For: Strong AI/ML engineering background with experience building production systems for large-scale model inference or training. Experience designing and implementing end-to-end AI/ML pipelines. Deep understanding of distributed inference and optimization of large-model workloads. Hands-on experience with inference frameworks such as vLLM or SGLang. Strong understanding of GPU compute, memory optimization, throughput, latency, and resource utilization. Experience with one or more of the following is especially valuable: LLM inference infrastructurevLLMSGLang Distributed inference Distributed training Reinforcement learning Synthetic data generationGPU optimization Model post-training Large-scale ML systems

Before you apply

Applying takes about a minute. These four things decide how fast it moves after that.

Your profile is current

It's what we read first. Occupations, seniority and locations matter more than a long history.

Two examples you can talk through

Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.

A number in mind

What you're on now and what would make you move. We negotiate better when we know both.

Your notice period

Employers plan around it, and it's the question that stalls offers most often.

Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.

More like this