AI Platform Engineer

Wave Group · London Area, United Kingdom New
LinkedIn

Posted

Aug 28, 2026 (Yesterday)

Seniority

Not Specified

Work Model

Not Specified

Type

Contract

Category

DevOps & SRE

Salary

ยฃ115k+ โ‰ˆ $146,050 USD

Skills

CI/CD GCP Go LangChain Observability Python SRE Terraform

Description

๐Ÿ’ป Job Title: AI Platform Engineer ๐Ÿ’ฐ Salary: up to ยฃ115k (+ very generous early-stage equity, up to ~ยฃ90k) ๐Ÿ“ Location: Central London, EC1 (3 office day/week) ๐Ÿฆ Company: B2B FinTech / Fraud Prevention ๐Ÿ‘ฅ Employees : ~30 ๐Ÿ’ธ Funding : $15m+ (Series A) This London startup is building a new intelligence layer designed to bring more context and security to digital payments. Their technology analyses transactions in real time, gathering signals from multiple sources to determine whether a payment is legitimate or potentially fraudulent. The platform combines distributed data systems, real-time investigations and AI-driven decisioning to help financial institutions detect scams while allowing legitimate payments to flow without unnecessary friction. Within 2 years of being founded, they're working with most Tier 1 banks and payment providers in the UK - and are about to launch in the US! ๐Ÿš€ Hiring an AI Platform Engineer to help build the foundations that let a fast-moving engineering team ship safely to some of the country's largest financial institutions. Production is increasingly powered by non-deterministic AI agents, so this isn't a "keep the lights on" role - you'll be defining what good looks like for platform and infrastructure culture, not inheriting someone else's playbook. Key responsibilities: Owning the reliability and operability of production systems - monitoring, alerting, incident response and post-incident learning Maintaining and evolving the infrastructure-as-code estate, making it easy to ship safely and hard to ship dangerously Securing infrastructure defaults so the easy path is the safe path Designing observability across the stack - metrics, traces, logs, dashboards and alerts - including for AI agent behaviour, where "correct" isn't always the same twice Driving incident response maturity from detection through resolution to follow-up Building platform capabilities that unblock engineering teams - deployment pipelines, release tooling, developer experience Building the guardrails and automation (including AI-assisted triage and response) that let the wider team move fast without breaking things Supporting the platform's evolution from stability through to scalability as the company grows โœ… Must have requirements: At least 3-4 years in platform/DevOps/SRE, ideally with genuine ownership at an early-stage startup - comfortable with ambiguity Strong Terraform/IaC experience on real production infrastructure Deep cloud infrastructure experience (GCP a strong plus) Proven observability track record - monitoring, alerting and dashboards for distributed systems Software Engineering background / proficiency in Python or Go at a decent level Incident response experience - on-call, running incidents, and building the processes that make both better Security foundations - least-privilege access, secrets management, secure-by-default infrastructure CI/CD experience - deployment pipelines teams trust, with a focus on deploy velocity and rollback safety Genuine hands-on exposure to agentic AI frameworks (LangChain, LangGraph or similar) within the last ~12 months - not just conceptual awareness Fintech or other regulated-industry experience ๐Ÿ‘ Bonus points for: Experience building observability/reliability for non-deterministic or ML-powered systems specifically Exposure to compliance frameworks (ISO 27001, SOC 2) Experience with workflow orchestration engines (Temporal or similar) A track record of bootstrapping a platform/DevOps function from scratch