Location
New York, New York, United States
Posted
July 26, 2026
Job Description
We are sharing a specialised full-time consulting opportunity for US-based MLOps and ML systems engineers with production experience in JAX, PyTorch, distributed training infrastructure, and custom GPU kernel development using Pallas or Triton.
This role supports a high-impact generative AI initiative focused on developing and evaluating advanced ML infrastructure tasks for frontier model training. Selected engineers will design technically challenging problems, produce rigorous solutions, assess model-generated outputs, and help establish evaluation standards across training pipelines, distributed systems, framework-level optimisation, and GPU kernel performance.
Key Responsibilities
ML Infrastructure & Training Systems
- Analyse and improve machine learning training infrastructure, deployment workflows, and model-development systems
- Guide research and engineering teams on MLOps, distribut...