
Senior Machine Learning Systems Engineer
Owner of end-to-end ML platform and MLOps patterns including data preparation, model management, experiment tracking, and model serving. Work includes building scalable graph data pipelines (Beam, Spark, Ray Data), optimizing distributed GPU training and costs, and operating on GCP (BigQuery, GCS) with infrastructure-as-code (Terraform). Requires 5+ years in ML infrastructure and hands-on experience with Python, PyTorch/TensorFlow, Ray, Kubernetes, and tools like MLflow or Wandb.






