
Posted 4 months ago
Machine Learning Architect
AI Summary
A founding ML practitioner who deploys and trains models on prototype ML processors, collaborates with hardware teams to refine architecture, and helps prove out novel chip performance for AI workloads.
About this role
About Ayo
We are a venture-backed early-stage startup developing processors specialized for machine learning. The processor will provide orders or magnitude improvement in speed and power efficiency with a goal of unseating the GPU as the dominant computing platform for AI.
The Role
We are seeking a deep ML practitioner to join as a founding team member - someone with hands-on experience working on or alongside a foundation model at scale, who understands what happens under the hood when splitting jobs across thousands of GPUs, and who is excited to bring that depth to novel hardware. They have a full-stack understanding of machine learning architectures, love to optimize algorithms across disciplinary boundaries, and will deploy and train models directly on our prototype chips to help us prove out what our processor can do - no prior hardware experience required.
What You’ll Do:
Deploy and run trained models on prototype hardware and digital twins, producing working demonstrations on our chips.
Develop and adapt algorithms to train models on novel processing environments, including our prototype hardware.
Work with hardware engineers to define and refine processor architecture based on insights learned through model training and experimentation.
Maintain a deep curiosity about what makes machine learning systems work - and bring that curiosity to bear on how they run on new hardware.
Support go-to-market strategy development
What We’re Looking For:
PhD in machine learning, representation learning, theory of computation, or a related field - or equivalent industry experience working on foundation models at scale.
Experience training models at scale - distributed training across many GPUs, working with large datasets and compute.
Has built and trained neural networks from scratch
Deep knowledge of the structure and internal operation of neural networks - including how and why they behave the way they do (e.g. interpretability or explainability work is a plus).
Excitement about applying deep AI expertise to new and novel hardware environments - you don’t need prior experience with photonics or silicon, but you want to learn.
Fluent knowledge of Python
Fluency in PyTorch (preferred), TensorFlow, JAX, or other industry-standard ML software libraries
Why Ayo
You'll be a founding AI team member - shaping how we think about and deploy AI as a company.
We are working on a genuinely hard and interesting problem: unseating the GPU as the dominant AI compute platform.
Early-stage means real ownership, real impact, and meaningful equity.
Competitive salary. Equity commensurate with stage and seniority. Benefits package including health, dental, and vision.
Fully onsite in Boston - we are a collaborative, in-person team.
Skills
Explore related jobs
More jobs at Ayo Semiconductor
Similar Distributed Training jobs
Jobs in Boston
- CHospice RN - New England Area (MA / NH / RI)Carehospice · Boston, MA
- GSenior Geotechnical EngineerGannettfleming · Boston, MA
- SProject Engineer - Career Start - Boston, MA (July 2027 Start)Suffolkconstruction · Boston, MA
- SVDC Manager - VisualizationsSuffolkconstruction · Boston, MA
- ENurse Practitioner – House Calls (North Boston Region)Essenmed · Boston, MA
- RDirector of RehabilitationReliant Rehab · Boston, MA
Browse these categories
Market data for this role
All reports →- SeriesRole reportsOne role family at a time: how many openings, what changed this week, who is hiring, what it pays.
- SeriesSalary reportsWhat employers publish in job postings, by level and workplace. Not self-reported pay.
- Market overviewState of tech hiring, September 2026: up 4.8%Tech hiring rose 4.8% month over month in September 2026, with 411,122 new listings. Customer support and account executive roles led the growth.