SR Cloud Engineer

Summary

The Senior Cloud Engineer will design, deploy, and maintain GCP-based infrastructure to support AI and automation initiatives. The role involves using Terraform for infrastructure-as-code and ensuring the reliability of production AI/ML workloads.

How will you make an impact

You will be a core member of the Customer Recommender team within Metro Digital’s AI & Automation unit. Working at the intersection of cloud infrastructure and emerging AI technologies, you develop, deploy, and maintain the GCP-based solutions that powers Metro’s AI and automation initiatives. Your work directly enables your product team to bring ideas to production and keep them running reliably in a fast-moving field.

Your responsibilities

  • Design and implement cloud infrastructure on Google Cloud Platform (GCP) following architectural best practices for elastic, resilient, and lean systems
  • Automate infrastructure provisioning and configuration using Terraform, ensuring repeatable and auditable deployments
  • Deploy, operate, and maintain AI/ML workloads and services built on GCP
  • Take end-to-end ownership of bringing systems live, from initial setup through production rollout and ongoing operations
  • Monitor, alert, and respond to incidents; apply change and problem management practices to maintain service reliability
  • Identify and drive improvement opportunities proactively within the team and across the platform
  • Collaborate with internal stakeholders and external vendors on cloud implementations and governance topics
  • Contribute to cloud governance standards and ensure compliance within the team’s scope
  • Potentially act as a buddy or point of contact for less experienced team members in your discipline

Required key competencies and qualifications

  • Hands-on experience with Google Cloud Platform (GCP), designing, deploying, and operating cloud infrastructure
  • Proficiency with Terraform for infrastructure-as-code in real-world, production environments
  • Proven track record of taking systems from development to live production and maintaining them operationally
  • Ability to work independently on complex, ambiguous tasks and structure your own approach to solving them
  • Curiosity and adaptability, comfortable operating in a fast-moving field, learning continuously, and adjusting course quickly
  • Solid understanding of cloud architecture principles: availability, resilience, and lean system design
  • Familiarity with monitoring and observability tooling in cloud environments
  • Good communication skills to align with technical and non-technical stakeholders

Nice-to-have

  • Experience with AI/ML platforms or workflows, ideally on GCP
  • Background in software development, ability to write and read code, automate tooling, and build APIs
  • What We Offer at METRO

  • Hybrid and agile work: Thrive in a flexible, multicultural environment with self-organizing teams.
  • People development: Grow through individual and company-wide learning and development opportunities.
  • Ownership and impact: Contribute to a strategic Ultra Fresh platform improving availability, efficiency, and waste reduction.
  • International collaboration: Work with colleagues, business stakeholders, and partners across multiple countries and functions.
  • Support with individual solutions: We care about people and support you throughout your professional journey.

See also

DevOps jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available