In The Pocket

Senior Site Reliability Engineer

BelgiumFull timeSeniorPosted 20 days ago
Apply on In The Pocket →

Sign into see who you know at In The Pocket.

About us We're an independent European digital product studio with a strong track record of delivering market-defining strategies and products for enterprises and governments. With about 200 talented people across three countries, premium AI capabilities and a diverse portfolio of international clients, we define, design and ship high-quality digital products, strategies and software. From global cloud-based platforms and large-scale apps to medical devices, banking software and augmented reality experiences. In a nutshell: if it lives and breathes digital, it’s what we do best. Become a part of our team At In The Pocket, you join a multidisciplinary team of designers, engineers, strategists and data specialists who care about the work as much as the people they work with. We work in small, autonomous teams that stay close to our clients and their users, and we give everyone the space to take ownership, speak up and grow. You’ll learn from colleagues across our studios, ship products that reach thousands or even millions of people and help shape what digital can do for leading organisations in Europe. What you’ll do  We're building a brand-new RUN team from the ground up, a dedicated operational function that owns the after-care of the products we run, both our own and the ones we host or operate for customers. As one of the first people on the team, you won't inherit a fixed playbook: you'll write it, shaping how the team works, the practices and tooling it adopts, and how operational excellence becomes part of how we build and run products. It's a hands-on, senior role where you set the technical and operational direction for a small team of two to three people, leading by expertise and example rather than as a manager. AI is transforming how we work, and you'll bring that same fluency to how the RUN team operates from day one. If defining a new discipline energises you more than maintaining an established one, this is your challenge. Take operational ownership of internal and hosted products in production, keeping them reliable, secure, performant and cost-effective. Own the incident lifecycle end to end (detection, response, mitigation, resolution and blameless post-mortems) using incident.io, and take part in a fair, sustainable on-call rotation. Define and operationalise SLIs, SLOs and error budgets, and be a key figure in maintenance and SLA presales, shaping proposals, contracts and framework agreements. Build and maintain a modern observability stack with a strong focus on OpenTelemetry. Grow a platform engineering and inner-sourcing culture: reusable components (e.g. GitLab components, Terraform modules), self-service golden paths and production-readiness scorecards for product teams. Work hands-on across CI/CD, infrastructure as code, cloud-native architecture and networking, and design for resilience, performance and scale. Embed security into everything you run, secrets management, scanning (SAST/DAST), penetration tests and ...