XTB S.A. brand banner

Site Reliability Engineer

XTB S.A.Poznań, PolandPLN 214k–PLN 272k / yrPosted 7h ago
Jobs by Adzuna
XTB is a global company from the financial industry, focusing on online trading of financial instruments. We are the largest FinTech in Poland and a leader in Central and Eastern Europe, and the range of our operations covers several countries, including Asia and South America. At XTB, we focus on the development of our employees, giving them opportunities to gain knowledge and skills in various fields, as well as offering a number of training and development programs. If you are looking for challenges and want to gain valuable experience in an international business environment, XTB is the right place for you. We are a certified Great Place to Work company. We are looking for a Site Reliability Engineer to define and drive the reliability of XTB systems at the scale of millions of clients. In this role, you will strengthen SRE practices and shape the resilience of our entire technology stack through high-impact observability, ensuring our systems remain robust and scalable. Responsibilities Observability Platform Engineering: Develop a standardized observability ecosystem. Implement a conscious telemetry model focusing on structured events, distributed tracing, and intelligent sampling strategies - that provides deep, actionable insights into system behavior. Reliability Enablement: Act as a strategic partner to product engineering teams, providing the platform, standards, and data they need to own service reliability. Use error budgets and alerting as the primary language for balancing feature velocity with stability. Proactive Resilience & Protection: Enhance detection capabilities to identify issues before they impact the customer. Leverage early-warning systems and AI/ML for automated anomaly detection and intelligent data analysis to continuously verify and strengthen system resilience. Operations & Tooling: Build internal automation and tooling that streamlines SRE workflows, automates routine operational tasks, and enhances efficiency across the technology stack. Incident Management & On-Call Rotation: Participate in an on-call rotation to provide incident management, ensuring rapid incident resolution, effective communication, and post-incident analysis to drive continuous improvement. Requirements Professional Background: Professional experience in SRE, Infrastructure, or DevOps roles managing high-scale, distributed environments. Technical Engineering: Advanced programming skills in Python, with a strong focus on building scalable automation, internal tooling, and robust scripts. Cloud & Orchestration: Hands-on expertise in managing production-grade Kubernetes environments, configuration management tools like Ansible, and designing resilient infrastructure architectures within Azure Kubernetes Service and on-prem environments. Observability Engineering: Proficiency in building standardized telemetry ecosystems. You have mastered self-hosted opensource tools for observability data collection, storage and visualization, like Prometheus, Grafana, ELK Stack, Tempo, Thanos, Jaeger and similar. Operational & Soft Skills: Ability to drive incident management, conduct thorough post-incident analysis, and foster a culture of reliability and shared ownership. Nice to have Experience with commercial APM platforms (e.g., Datadog, Splunk, New Relic) and chaos engineering tooling. Experience with cloud cost management and FinOps principles. Experience defining and tracking SRE metrics (SLI/SLOs) and managing error budgets to drive reliability. Experience with AI/ML techniques for SRE tasks, such as AIOps, automated anomaly detection, log analysis, and optimizing reliability workflows. Experience in building and managing strategies to proactively manage technical debt and align team output with organizational goals. What we offer Real influence on the development of the company and the product. Work in an experienced team that is happy to share its knowledge. A clear vision of development thanks to regular feedback and clear career paths. Regular team-building meetings. Benefits A training budget for courses and conferences that interest you. An extra day off on your birthday. An extra day off for parents. Equipment tailored to your needs. Private medical care and group insurance. Access to an e-learning platform for learning English and a benefits platform. Access to a wellbeing platform and the opportunity to take advantage of workshops and private therapy sessions. Remote work, from the office in Warsaw or from a coworking space in your city.

Similar roles

  • Kaufman Rossin logo

    DevOps Engineer

    Kaufman Rossin·Bengaluru, India

    Why We Stand Out Seeking a new challenge where your professional and personal aspirations are not only possible but supported? Kaufman Rossin might be just the place for you! Kaufman Rossin Professional Services Private Limited’s (the “Company”) offices are located in the World Trade Center (WTC) in Bangalore, Karnataka, India, and at Udyog Vihar, in Gurgaon, Haryana, India. The Bangalore…

    • Hybrid
    • Full-time
  • Temporal Technologies logo

    Senior Platform Architect - UK

    Temporal Technologies·United Kingdom

    SUMMARY Join our Technical Services team as a Senior Platform Architect and drive transformational outcomes with Temporal. You’ll be the accountable technical owner from purchase to stable production, turning complex challenges into scalable, repeatable solutions. Shape delivery patterns, operational standards, and enablement strategies that accelerate customer success and create lasting competitive advantage across teams and environments. WHAT YOU'LL DO Own…

    • On-site
    • Full-time
  • JAC Recruitment logo

    Application Developer - JAC Recruitment

    JAC Recruitment·Singapore, Singapore

    Key Responsibilities 1. Execute equity trades. 2. Gather and analyze requirements from Portfolio Managers (PMs) and traders. 3. Develop tools using Bloomberg API and Enfusion API. 4. Automate order execution, market data retrieval, and position management processes. 5. Improve workflows and enhance operational efficiency. 6. Systems and Tools to be Developed Automation of trade and market data collection. 7. Automation…

    • Full-time
    • Tech.Passself-sponsored, 2-yr
  • Rubrik Job Board logo

    Software Engineer, Atlas Distributed Systems

    Rubrik Job Board·Palo Alto, United States

    ABOUT TEAM Rubrik Atlas is the core data path for all Rubrik products, whether in the data center, at the edge, or in the cloud. It is a distributed, scale-out, fault tolerant, performant, deduplicated user-space filesystem that uniquely combines block device (ext4/NFS/SMB) and S3 backend storage interfaces. It has been the cornerstone of Rubrik’s innovative tech stack since Day 1.…

    • Full-time
  • veritree logo

    QA Engineer

    veritree·Vancouver, Canada

    VERITREE AND JOB OVERVIEW veritree is an award-winning climate tech start-up based in Vancouver. Launched in 2021, our technology measures and verifies the impact of global restoration efforts from the ground up. We are on a mission to plant 1 billion verified trees by 2030, collaborating with businesses, planting organizations, and consumers who believe in the transformative power of verified…

    • Hybrid
    • Full-time
    • Express EntryPR day one, no employer
  • logen.ai logo

    WerkstudentIn AI Agent Developer

    logen.ai·Berlin, Germany

    Du hast Erfahrung mit KI und willst sie an echten Kundenprojekten einsetzen? Bei uns baust du AI Agents, die in Produktion gehen. Keine Proof-of-Concepts für die Schublade. Wir sind ein KI-Startup aus Berlin, spezialisiert auf Automatisierung im Kundenservice. Wir bauen AI Agents, Voicebots und Automatisierungslösungen für Unternehmen im deutschen Mittelstand. Gegründet 2023 von Oscar Schwarz und Ludwig Sickert, aktuell ein…

    • On-site
    • Full-time
    • Blue Cardtied; settle 21–33mo
  • Solirius Reply logo

    Test Engineer (Java/.Net)

    Solirius Reply·London, United Kingdom

    About Us: Solirius Reply, part of the Reply Group, is a technology consultancy and digital transformation partner that helps organisations solve complex challenges through strategy, design, engineering, and delivery. We work closely with our clients to deliver secure, accessible, user-focused services that evolve with their needs. By combining deep technical expertise with people-centred design, we create solutions that deliver meaningful,…

    • Hybrid
    • Full-time
  • Team17 Digital logo

    Platform Engineer (Mid-level)

    Team17 Digital·Wakefield, United Kingdom

    About the Role We are seeking a Platform Engineer (Mid-level) to join our Platform Engineering function. This hands-on technical role focuses on the platforms, infrastructure, automation, and tooling that support our development teams and business operations. The successful candidate will help maintain, improve, and modernise our platform estate across cloud services, infrastructure, monitoring, automation, CI/CD, and operational tooling. The role…

    • Full-time