Job title: SRE (Observability & AI)
Opportunity Type: Resource Augmentation
Area of Expertise: Front Office
Job published: 10-09-2026
Job ID: 219126

Job Description

We're currently working with an Investment Bank based in Warsaw for a Site Reliability Engineer position within a global Markets Technology environment, where you’ll have a genuine opportunity to influence the reliability, resilience and performance of critical production platforms rather than simply keeping the lights on.

Key Responsibilities
    • Provide support for critical production applications and services within the Markets Technology environment.
    • Own and manage production incidents through to successful resolution, minimising business impact.
    • Conduct root cause analysis and implement long-term solutions to prevent recurring issues.
    • Drive automation initiatives to reduce manual operational effort and improve service reliability.
    • Improve monitoring, observability and alerting capabilities across the technology estate.
    • Partner with development, infrastructure and platform engineering teams to enhance system stability and performance.
    • Support capacity planning, resilience testing and performance optimisation activities.
    • Contribute to continuous improvement initiatives and the adoption of SRE best practices.
    • Participate in on-call and out-of-hours support for high-priority production incidents when required.

Essential
    • Proven experience as a Site Reliability Engineer, Production Engineer, Platform Engineer or Senior Application Support Engineer.
    • Strong experience supporting large-scale, business-critical applications in complex enterprise environments.
    • Strong Linux/Unix knowledge.
    • Hands-on scripting and automation experience using Python, Shell or similar technologies.
    • Experience with monitoring (Grafana/Prometheus), alerting and observability platforms.
    • Cloud platform experience (AWS, Azure or GCP).
    • Strong troubleshooting, analytical and problem-solving capabilities.
    • Experience managing major incidents and working within highly available environments.
    • Excellent communication skills and the ability to collaborate across global teams.

Desirable
    • Experience within Financial Services, Capital Markets or Investment Banking.
    • Murex support experience.
    • Exposure to DevOps practices, CI/CD pipelines and infrastructure automation.
    • Experience supporting high-volume or low-latency applications.