AI Accelerator Architect Ultra-Efficient LLM Inference HW
normal computingPalo Alto (CA)
AI Accelerator Architect Ultra-Efficient LLM Inference HW
Posted today
normal computingPalo Alto (CA)
SENIORITY
Lead
About the role
Normal Computing seeks a Hardware ASIC Architect to define the silicon and system microarchitecture for our unconventional compute platform. You will drive architectural trade-offs to enable a large leap in energy efficiency for LLM and diffusion model inference, translating transformer and diffusion flows into custom compute tiles and interconnects.
You will work with compiler, RTL, and analog teams to build performance models, write microarchitecture specs, and ensure maximum
Before you apply
Applying takes about a minute. These four things decide how fast it moves after that.
Your profile is current
It's what we read first. Occupations, seniority and locations matter more than a long history.
Two examples you can talk through
Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.
A number in mind
What you're on now and what would make you move. We negotiate better when we know both.
Your notice period
Employers plan around it, and it's the question that stalls offers most often.
Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.
More like this
