xelys jobs xelys jobs

Machine Learning Engineer (MLOps/Eng)

SilverSearch, Inc.

full-remotemidcontractbackenddevops United States 131 days ago via LinkedIn

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

AWSAmazon SageMakerMLOpsMonitoringObservabilityDrift DetectionPyTorchTensorFlowGPU InferenceComputer Vision

About the role

Role overview

You’ll be an ML Engineering (MLOps) partner for production ML systems that power enterprise intelligence. The role centers on deploying, operating, scaling, and governing ML infrastructure and inference services for multimodal workloads (text, image, and video), with strong focus on runtime reliability, scalability, and observability on AWS.

What you’ll be doing

  • Design, deploy, and operate production ML infrastructure across Dev, QA, and Prod environments
  • Manage ML deployment pipelines and runtime operations using AWS SageMaker
  • Configure and optimize GPU/CPU infrastructure for large-scale inference
  • Implement monitoring, alerting, drift detection, and observability for ML systems
  • Build deployment governance (rollout, rollback, recovery)
  • Support high-throughput inference for text, image, and video pipelines
  • Optimize scalability, cost efficiency, and operational reliability
  • Collaborate with ML Engineers and Data Scientists to operationalize new models and workflows
  • Implement A/B testing and controlled rollout strategies for production systems

Required qualifications

  • Hands-on experience deploying and operating ML systems in production
  • Strong AWS SageMaker experience, including:
    • Pipelines
    • Endpoints
    • Monitoring
    • Multi-environment deployments
  • Experience with containerized ML deployment and orchestration
  • Experience operating inference systems in PyTorch and TensorFlow
  • Knowledge of autoscaling, infrastructure optimization, and runtime reliability
  • Experience implementing monitoring and observability for ML systems
  • Experience supporting distributed ML workloads in cloud environments

Strongly preferred

  • Experience with NLP and computer vision ML systems
  • Familiarity with semantic/vector search infrastructure
  • Experience with ranking/reranking systems
  • Knowledge of ANN/vector indexing approaches
  • Experience optimizing large-scale text, image, and video pipelines
  • Experience optimizing GPU-based infrastructure

Additional details

  • Fully remote
  • Preference for East Coast collaboration hours
  • 6-month contract-to-hire
  • Must be legally authorized to work in the United States and not require sponsorship

About SilverSearch, Inc.

SilverSearch, Inc. is recruiting for a globally recognized media and information organization that builds and operates large-scale production machine learning systems for enterprise intelligence products. The work focuses on deploying and governing ML infrastructure and inference services, with an emphasis on reliability, scalability, and observability on AWS-based platforms.

Scraped 5/17/2026