Senior Platform Engineer
Summary
Senior Platform Engineer builds and scales real-time, event-driven infrastructure for AI-driven market intelligence at a fintech AI startup using AWS, Kubernetes, and distributed systems.
At Permutable, we're building real-time AI systems that transform large volumes of global news, market and proprietary data into market intelligence and systematic signals for financial institutions.
We're evolving our technology towards a more real-time, event-driven and agent-driven architecture and are looking for a Senior Platform Engineer to take ownership of the platform and infrastructure behind it.
As a small, ambitious team, you'll work directly with our founder and engineering team and have significant influence over how our next-generation architecture is designed and built.
What You'll Do
Design & Scale our Platform: Build and evolve highly available production infrastructure across AWS, Kubernetes/EKS, ECS, Docker, RDS/PostgreSQL and S3.
Build Event-Driven Architecture: Help move our architecture from scheduled pipelines towards real-time, event-driven services using messaging, queues, asynchronous processing and distributed systems.
Build Infrastructure for AI Agents: Develop the infrastructure required to run increasingly agent-driven AI workflows reliably in production, including state management, task execution, model routing, retries and observability.
Support AI Infrastructure: Work closely with our AI engineers supporting production NLP, LLM and quantitative model workloads, including our in-house GPU infrastructure.
Automate Infrastructure: Build and manage Infrastructure-as-Code using Pulumi and automate deployments through GitHub Actions.
Own Reliability & Observability: Build monitoring, logging, tracing and alerting across our data, model, service and agent infrastructure.
What We're Looking For
7+ years of professional experience: in Platform Engineering, Infrastructure, SRE, DevOps or distributed backend engineering.
Strong AWS experience, ideally including EKS/ECS, RDS/PostgreSQL, S3 and Lambda.
Distributed Systems Experience: Strong understanding of event-driven architectures, messaging, asynchronous processing and distributed workflows.
Infrastructure-as-Code: Experience with Pulumi, Terraform or equivalent.
CI/CD & DevOps: Experience with GitHub Actions or equivalent, automated testing, deployments, monitoring and production operations.
Strong Engineering Fundamentals: Python, Linux, networking, databases and cloud infrastructure.
Production Ownership: Comfortable diagnosing difficult production problems and taking responsibility for critical infrastructure.
Startup Mindset: Proactive, resourceful and comfortable taking ownership in a fast-moving engineering environment.
Bonus : with Apache Airflow, real-time data platforms, AI/ML infrastructure, LLM agents, GPU infrastructure or financial-market systems.
Why Permutable
Real Ownership: Take significant ownership of our production architecture and infrastructure.
Build the Next Generation: Help move our platform towards real-time, event-driven and agent-driven architecture.
AI at Scale : Work with LLMs, AI agents, quantitative models and large volumes of continuously arriving data.
Technical Influence: Have a meaningful say in architecture and technology decisions.
Real Production Impact: Your work will quickly reach production and support institutional financial clients.
Hybrid Flexibility: Spend 3+ days/week in our Vauxhall hub.