R
Relace
ML Engineer

Machine Learning Scientist

On-siteMidML Engineerposted 11mo ago
✦Role summaryAI-generated

Mid-level ML Engineer role focused on scaling high-performance, low-latency language models for code generation, targeting OpenRouter’s fastest throughput (10k tok/s) while optimizing retrieval and application-specific inference for enterprise clients like Figma and Vercel.

Skills required

About this role

About Us

Relace is building the models and infrastructure that code agents reach for. We power the fastest model on OpenRouter (10,000 tok/s) and deliver optimized small language models designed for retrieval, application, and core code generation functions.

Our technology supports some of the world’s fastest-moving companies — including Lovable, Figma, and Vercel — as they deploy and scale code generation to hundreds of millions of users. We recently raised our Series A from a16z, and we’re growing quickly.

Our team is made up of mathematicians, physicists, and computer scientists who are deeply passionate about their craft. If you thrive on ambitious technical problems, care about elegant systems design, and want to build the foundation of how code gets written at scale, this is the place for you.

The Role

We’re looking for a Machine Learning Scientist to push the limits of small, high-performance language models. This is a deeply technical role focused on advancing the capabilities of our models for retrieval, application, and code generation.

The ideal candidate has a strong background in ML research and engineering, is comfortable working with both theory and production systems, and thrives in an environment where ideas turn into deployed infrastructure fast. This person should be excited to work on training methodology, optimization, evaluation, and model architecture at scale — and collaborate directly with infrastructure and product teams to get breakthroughs into production quickly.

This role is best suited for someone who loves both mathematical elegance and real-world impact.

Requirements

  • Strong background in machine learning, deep learning, or related fields.

  • 2+ years of experience working on ML research or production systems.

  • Fluency in Python and frameworks like PyTorch or JAX.

  • Experience with training and optimizing large or efficient models.

  • Strong understanding of applied optimization, distributed training, or model evaluation.

  • Familiarity with code models, retrieval systems, or language modeling a plus.

  • Advanced degree (MS or PhD) in a quantitative field, or equivalent industry experience.

  • Willingness to work in-person from our SF office in FiDi.

Apply on Relace →Opens in new tab
score your resume against this role

Similar open roles

UP
NEW

Machine Learning Engineer

United Parcel Service·55 GLENLAKE PARKWAY NE, ATLANTA, GA 30328, United States of America
On-siteMidML Engineer
yesterday
A
NEW

Engenheiro de Machine Learning Pleno

ArcelorMittal·Belo Horizonte, MG, Brazil
On-siteMidML Engineer
yesterday
S
NEW

Staff Machine Learning Engineer- Voice AI

Servicenow·Hyderabad, India
On-siteStaffML Engineer
yesterday
S
NEW

Senior Staff Machine Learning Engineer- Voice AI

Servicenow·Hyderabad, India
On-siteStaffML Engineer
yesterday
US
NEW

Data Scientist - AI and NLP

Uni Systems·Ispra, Province of Varese, Italy
On-siteSeniorML Engineer
yesterday
P
NEW

Applied Machine Learning Engineer

Permitflow·New York City, NY
On-siteML Engineer
$175k – $250k USD
yesterday
Apply on Relace →