Platform / DevOps Engineer (Junior – Early Intermediate)
About the role
Harmoney runs a mature cloud platform on GCP that powers our consumer lending platform. We're now growing our Platform Engineering practice, and you'll be one of the people who shapes it. This is a broad, high-leverage role: you'll operate and evolve a substantial platform while laying the foundations that let the business innovate safely, especially as we adopt AI.
You'll work alongside our existing platform team engineers and the Cloud Platform Lead in a small, high-trust, SOC2-conscious team, owning meaningful problems end-to-end and helping define how the practice operates as it matures.
Responsibilities:
Key responsibilities include, but are not limited to:
Support safe AI adoption: Assist with building foundations that allow responsible AI innovation, including vendor discovery, security reviews, AI observability, and cost visibility, while establishing guardrails.
Deliver platform capabilities: Support initiatives such as API improvements, governance of customer auth systems, platform operations, and technical debt reduction.
Maintain scalability and stability: Run platform and infrastructure upgrades, improve security, and contribute to reusable infrastructure patterns.
Drive cost efficiency: Improve cost visibility (compute and AI), optimize configurations, and apply cost-saving strategies.
Improve developer experience: Working closely with software engineers to enhance deployment pipelines, observability, documentation, and change management.
Strengthen incident and alert management: Help improve alerting standards, resolution processes, and metric retention.
Support security and compliance: Help maintain IAM, secrets management, and audit/change controls aligned with compliance requirements.
Help build the practice: Contribute to standards, ADRs, runbooks, and ways of working.
Experience ( nice to have ):
1–2+ years of exposure to cloud, DevOps, or software engineering—or a strong project portfolio demonstrating foundational platform knowledge.
Hands-on experience with a major public cloud (GCP preferred) and Kubernetes (GKE or equivalent).
Experience maintaining and upgrading live platforms (e.g., version upgrades, capacity tuning, incident response).
Exposure to regulated or security-sensitive environments (e.g., financial services, SOC2/ISO) is a strong advantage.
Comfortable working with ambiguity and eager to learn in emerging areas like AI enablement.
Experience embedding into software engineering teams, or a demonstrated understanding of what that implies.
Skills and Competencies:
Infrastructure as code: Working knowledge of Terraform with exposure to reusable patterns.
Kubernetes & GitOps: GKE, Helm, ArgoCD, and GCP services (IAM, networking, Cloud SQL, Secret Manager).
Observability & incident management: Tools like Grafana, Prometheus, Loki.
Cost & performance optimization: Ability to identify savings and optimize workloads.
CI/CD & automation: Git-based workflows, GitHub Actions, scripting (Bash/Python/Go).
Workflow/streaming platforms: Kafka exposure is a nice-to-have.
AI-readiness mindset: Interest or experience in AI/LLM observability, cost management, and AI-enabled operations. (No prior AI experience needed-we'll teach you the platform side of LLM tooling)
Typescript, Javascript, or Java language platform exposure (nice-to-have)
Skills
- AI
- Ai Enablement
- API
- Argo CD
- Automation
- Bash
- CI/CD
- Cloud
- Developer Experience
- DevOps
- GCP
- Git
- GitHub
- GitHub Actions
- GitOps
- GKE
- Grafana
- Helm
- IAM
- Infrastructure as Code
- Java
- JavaScript
- Kafka
- Kubernetes
- LLM
- Loki
- Networking
- Observability
- Prometheus
- Python
- Secrets Management
- SOC 2
- SQL
- Terraform
- TypeScript