Senior SRE

Link Group Warszawa 2026-10-07
  • Bachelor’s degree in Computer Science, or 10+ years of hands-on experience in infrastructure engineering, DevOps, or SRE, prominently featuring technical leadership.
  • Mastery of public cloud environments (specifically AWS) and IaC automation at an enterprise scale.
  • Profound understanding of Kubernetes architecture and the operational realities of running distributed, high-availability container platforms.
  • Demonstrable history of building and scaling automated delivery pipelines across decentralized engineering units.
  • Expert-level knowledge of modern observability stacks and enterprise telemetry standards.
  • Practical experience embedding AI or machine learning tools into operational workflows to boost system efficiency.
  • Strong ability to mentor senior talent and shape architectural direction across multiple teams through influence, rather than direct reporting lines.
  • Proven capability to design robust backends for varied application profiles, ranging from heavy data processing platforms to high-concurrency microservices.
  • Deep technical grasp of cloud networking, FinOps/cost-tuning, and large-scale distributed system security.
  • Excellent communication skills, adept at translating business objectives into technical realities alongside product and project managers.
  • A drive to innovate within a rapidly expanding, forward-thinking organization that highly values continuous learning, agility, and cross-functional teamwork.

We are seeking a highly experienced Principal Site Reliability Engineer to spearhead the performance, uptime, and scaling strategy for our core digital ecosystem. Acting as a technical compass within a fast-paced agile environment, you will collaborate with cross-functional leadership to define infrastructure standards, elevate engineering practices, and build fault-tolerant systems using modern technologies to support our rapidly growing enterprise.

,[Blueprint and enforce enterprise-wide SRE methodologies, ensuring peak performance and high availability, and setting a high technical bar for the entire engineering department., Drive the infrastructure-as-code (IaC) vision utilizing tools like Terraform or Pulumi to manage large-scale cloud (AWS) deployments, balancing reliability with aggressive cost optimization., Take full ownership of our containerization strategy, steering the long-term roadmap for expansive Kubernetes clusters to guarantee robust, resilient deployment environments., Collaborate with engineering and infrastructure leadership to revamp continuous integration and deployment (CI/CD) workflows, fostering a culture of pervasive automation., Design comprehensive monitoring and telemetry ecosystems (leveraging tools such as Prometheus, Datadog, Grafana, or Dynatrace) while commanding incident response, root cause analysis, and post-event reviews., Orchestrate disaster recovery protocols and capacity forecasting to ensure uninterrupted service delivery for mission-critical applications at scale., Elevate the engineering team through mentorship, guiding staff-level peers in SRE practices, and championing the integration of AI-assisted monitoring and auto-remediation tools., Strengthen cloud security postures and compliance frameworks, producing top-tier architectural documentation for globally distributed engineering hubs.] Requirements: Degree, DevOps, SRE, Public cloud, AWS, IaC, Kubernetes, AI, Machine learning, Boost, Cloud, Networking