Back to all jobs

Lead Engineer, Machine Learning

Ref: JO-2609-363171

  • Environment: Remote
  • Contract Type: Permanent
  • Starts: 2026-11-16
Apply Report issue

About the Opportunity

We are partnering with a fast-growing technology company developing a new generation of AI-native applications designed to make everyday tasks, communication, organization and workflows more intelligent and intuitive.

The team is building proactive AI experiences with a strong focus on persistent context, reliable long-running workflows and successful real-world task completion.

They are looking for a Lead Engineer – Machine Learning to own the execution layer that transforms advanced research and model capabilities into reliable, scalable production systems.

About the Role

As Lead Engineer – Machine Learning, you will work across the complete model lifecycle, including data, training, evaluation, inference and deployment.

This is a hands-on technical leadership position for someone who enjoys operating at the intersection of machine learning research, systems engineering and product development.

You will take ownership of production ML systems while helping establish the technical standards and infrastructure required to deploy, monitor and continuously improve large-scale AI models.

What You’ll Own

  • Own end-to-end ML systems, from data and training through to evaluation, inference and production deployment.
  • Build and evolve training and fine-tuning pipelines for large models.
  • Design evaluation systems that measure model capability, robustness, safety and real-world product performance.
  • Architect high-performance inference systems, optimizing latency, GPU utilization, memory, cost and reliability.
  • Build scalable data pipelines supporting high-quality real-world and synthetic training data.
  • Establish reliable production infrastructure for deploying, monitoring and continuously improving models.
  • Partner closely with research and application engineering teams to translate model capabilities into meaningful product improvements.
  • Identify, diagnose and resolve model and system issues within production environments.
  • Make pragmatic technical trade-offs and rapidly iterate based on measurable real-world performance.
  • Provide technical leadership and support to engineers working across machine learning systems.

What We’re Looking For

  • Proven experience building and shipping machine learning systems used in production, rather than solely developing research prototypes or demonstrations.
  • Strong understanding of modern large-model training, fine-tuning, evaluation and inference.
  • Strong software engineering and systems engineering fundamentals.
  • Experience operating ML workloads at meaningful scale, particularly within GPU-based environments.
  • Experience building reliable training pipelines, inference systems and production ML infrastructure.
  • Strong understanding of large models and their potential failure modes.
  • Strong technical judgement with the ability to navigate complex and ambiguous engineering problems independently.
  • A bias toward experimentation, measurement, iteration and shipping.
  • High standards for correctness, reliability and production quality.
  • Ability to write strong, maintainable, production-grade code.

Technology Environment

Experience across the following technologies and areas would be particularly relevant:

  • Python
  • PyTorch
  • JAX
  • GPU-based training and inference systems
  • Large Language Models (LLMs)
  • Model training and fine-tuning
  • Model evaluation
  • Distributed ML systems
  • Production inference infrastructure
  • Real-world and synthetic training data pipelines

What Success Looks Like

Success in this role means research and model capabilities consistently translate into production-ready solutions with clearly defined performance and quality targets.

ML pipelines, training loops and inference systems will be stable, efficient and maintainable, with production issues detected, diagnosed and resolved quickly.

You will continuously improve model and system performance through experimentation, evaluation and monitoring, ensuring improvements are measurable and ultimately enhance the end-user experience.

You will also help create an environment where engineers are aligned, technically supported and able to deliver high-impact ML work efficiently.

Salt is acting as an Employment Agency in relation to this vacancy.

Apply Report issue

Data, AI and Machine Learning jobs

Career and Job Insights

Apply for this job

Lead Engineer, Machine Learning

  • USA
  • Data, AI and Machine Learning, Technology
  • Remote
  • Permanent

Save jobs

Log in to save a job

Report job

Lead Engineer, Machine Learning

  • USA
  • Data, AI and Machine Learning, Technology
  • Remote
  • Permanent

"*" indicates required fields

Need talent? Request a callback

This form is for companies looking to hire talent.

I am looking for a job I have a general enquiry

"*" indicates required fields

E.g. “Senior Frontend Developer” or “Offshoring team for design.”
This field is hidden when viewing the form