Site Reliability Engineer - Frontend
We’re Capital on Tap 👋💳 Capital on Tap started because small businesses were underserved. Big banks were slow, their products weren't fit for purpose, and small business owners often couldn't access what they needed. We set out to fix that. Today we're a financial platform - not just a credit card company. We offer a best-in-class business credit card, SME-focused spend management platform, a savings product that hit £1 billion in funds within its first year, and a growing suite of tools and financial products that make running a small business easier. 1,000+ employees, £20bn in annual card spend, 200,000+ customers, 17,000+ Trustpilot reviews averaging 4.7 stars, and we're profitable. We’ve done a pretty good job so far, but we’re just getting started! 📍London, Old Street | 🏢 2 Days in Office SRE at Capital On Tap 🌞 At Capital On Tap, we run a hybrid embedded SRE model - We aim to work closely with the teams to provide them the best support. As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You’ll design, build, and monitor systems, preventing issues before they happen. XP, the engineering group you will be embedded into, is the platform group that builds the foundations the rest of our engineers rely on: the shared UI frameworks for web and mobile, the backend-for-frontend layer, and the internal developer platform and shared backend libraries beneath them. Our job is to make it fast and safe for product teams to ship great features, by owning the architecture and paved paths they build on top of. What you’ll be doing 🗃️ Manage and automate resources in Azure, Datadog, NGINX & Cloudflare. Develop, deploy and monitor Kubernetes and Serverless resources. Build, manage, and evolve IAC using Terraform, Helm and Go CRDs. Improving systems, processes, and technologies; consulting stakeholders to enhance platform performance. Design and maintain comprehensive monitoring and alerting strategies using Datadog. Getting involved in new application architecture & design processes. Designing solutions to reduce toil, automate repetitive tasks and streamline workflows to reduce manual work and boost team productivity. Creating SLIs and SLOs; increasing application visibility Align with the Product team on SLAs and core service objectives Collaborating with core foundational teams to build and maintain reusable, automated solutions. Optimising CI/CD pipelines and developer workflows using Azure DevOps, Github, Octopus Deploy, MirrorD, and Flux. Leading incident response; taking an active part in communications, investigations, remediation and post-mortems to protect the customer experience. Our Values & Culture 🌞 Just Pilot: We never settle for “good enough”. We pilot new ideas fast, ask questions to figure it out, and scale quickly. Why Not Today? Fast is as slow as we go - speed and simplicity gives us a competitive advantage. Be a Buddy: We tap in from day one to help ...