Skip to main content
Fundamental

MLOps Engineer

RemoteIsrael only
Published
Role
AI / ML
Experience
Senior
Employment
Full-time
Company size
Startup
Salary not disclosed
Check eligibility

Open to IL only. Set where you work from to check your eligibility.

No BS summary

MLOps Engineer with 5+ years in MLOps or DevOps, building MLOps infrastructure from scratch. Needs Kubernetes, cloud, IaC, Python/Bash/Go, model serving, observability, and AI/ML security/governance. Remote role with applicant location requirement: Israel.

Core skills

MLflowTritonKubernetes

Required skills

WandBPyTorchTensorFlowTorchServeTensorFlow ServingKServeAWS/GCP/AzureTerraformHelmGitOpsPythonBashGoPrometheusGrafanaDatadogOpenTelemetry

Optional skills

KubeflowFastAPIDatabricksSnowflakePrometheusGrafanaDatadog

What you'll do

  • Develop and manage scalable, automated machine learning pipelines, CI/CD workflows, and orchestration frameworks
  • Design and implement robust model serving infrastructure using platforms like TorchServe, TensorFlow, Triton etc.
  • Develop scalable inference architectures optimized with ultra-low latency and high throughput
  • Ensure seamless model deployment by implementing A/B testing, canary releases, and rollback capabilities
  • Develop logging, alerting, and monitoring solutions to track model development and reliability
  • Improve GPU usage, enable autoscaling, and streamline resource allocation to boost efficiency
  • Design, implement, and maintain feature stores, robust data pipelines, and scalable storage solutions to efficiently handle large volumes of data

What they require

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience
  • 5+ years of experience as MLOps engineer or DevOps roles, working with MLOps platforms and frameworks
  • Experience building and designing MLOps infrastructure from the ground up
  • Experience with model serving frameworks for high scalability and low latency inference
  • Experience in building and managing data pipelines to support both model training and inference
  • Experience with Kubernetes on a major cloud provider and with infrastructure as code
  • Strong software engineering skills with a focus on writing clean, maintainable, and scalable code
  • Experience in AI/ML systems security, compliance, and model governance
  • Proficient with observability and monitoring tools
  • Preferred: Experience with ML workflow tooling
  • Preferred: Experience with backend applications
  • Preferred: Familiarity with data platforms
  • Preferred: Exposure to SRE practices or cloud security certifications

Benefits

  • Competitive compensation with salary and equity
  • Comprehensive health coverage for you and your dependents
  • Paid parental leave for all new parents, inclusive of adoptive and surrogate journeys
  • Relocation support for employees moving to join the team in one of our office locations
  • A mission-driven, low-ego culture that values diversity of thought, ownership, and bias toward action

Fundamental is an AI company pioneering the future of enterprise decision-making. Founded by DeepMind alumni, Fundamental has developed NEXUS – the world's most powerful Large Tabular Model (LTM) – purpose-built for the structured records that actually drive enterprise decisions. Backed by world class investors and trusted by Fortune 100 companies, Fundamental unlocks trillions of dollars of value by giving businesses the Power to Predict.

AIStartup

Details

Apply routeDom
Salary not disclosed