Senior Site Reliability Engineer, KRAKÓW
- Bachelor’s degree in Computer Science, Engineering, or a related technical field.
- 5+ years of experience in Site Reliability Engineering, Platform Engineering or a related role.
- Strong experience with cloud-native architectures and distributed systems.
- Hands-on experience implementing observability solutions, including metrics, logging, tracing, and application performance monitoring.
- Experience designing scalable backend architecture using Node.js, TypeScript, .NET
- Strong knowledge of database administration, performance tuning, and optimization, including Azure Cosmos DB, MySQL, or similar platforms.
- Experience building automation frameworks, scripting solutions, and operational tooling.
- Familiarity with CI/CD pipelines and continuous integration practices using GitHub Actions or similar platforms.
- Experience with containerization technologies such as Docker.
- Strong understanding of event-driven architectures, messaging systems, and real-time data processing platforms.
- Experience with monitoring and observability tools such as Datadog, OpenTelemetry, ELK Stack, App Insights, Grafana, or Prometheus.
- Excellent troubleshooting, analytical, and problem-solving skills.
- Strong written and verbal communication skills with the ability to collaborate across technical and business teams.
- Advanced proficiency in English & Polish, both written and spoken (B2+).
This is a hands-on role for someone with deep expertise in cloud-native platforms, distributed systems, observability, and performance optimization. You’ll work closely with Engineering, Product, and Operations teams to improve platform reliability, scalability, and operational efficiency while reducing manual overhead through automation. Over time, you’ll have opportunities to influence platform architecture, lead reliability initiatives, and establish Site Reliability Engineering best practices across the organization. ,[Serve as the platform subject matter expert, mentoring engineering teams on reliability, scalability, security, and operational best practices., Design, implement, and maintain observability solutions covering logs, metrics, traces, APM, and alerting across all platform services., Benchmark, analyze, and optimize application performance, cloud services, integrations, databases, and distributed systems., Build and maintain performance, load, chaos, and resilience testing frameworks to proactively identify system weaknesses., Perform advanced analysis, tuning, and optimization of structured and unstructured databases to improve performance and scalability., Develop automation frameworks and operational tooling that eliminate manual processes and reduce human error., Lead incident response efforts, conduct root cause analysis, and drive continuous improvement through postmortem reviews and reliability initiatives., Optimize the performance of databases, caches, streaming platforms, message brokers, and backend services., Collaborate closely with Engineering, Product, and Operations teams to deliver highly available, production-grade platform solutions., Establish and maintain monitoring, alerting, and reliability standards across the platform., Create and maintain technical documentation, runbooks, architectural diagrams, and operational procedures., Evaluate emerging technologies and recommend solutions that improve platform reliability, scalability, observability, and operational efficiency.] Requirements: Docker, Datadog, Observability, Cosmos DB, MySQL, CI/CD, GitHub Actions, ELK Stack, Grafana, Prometheus, .NET Additionally: Sport subscription, Private healthcare, Flat structure, Small teams, Free coffee, Playroom, Free snacks, Free beverages, Free lunch, Bike parking, Free parking, In-house trainings, In-house hack days, Modern office, Startup atmosphere, No dress code.
Data publikacji: 2026-07-21
APLIKUJ
Podobne oferty
Senior Software Engineer Site Reliability, WARSZAWA, Asana
Senior Site Reliability Engineer SRE, REMOTE, RecruTec
Site Reliability Engineer II SRE Guidewire Cloud Platform Application, KRAKÓW, Guidewire Software
Senior Site Reliability Engineer, WARSZAWA, VISA
Site Reliability Engineer SRE [M/F], REMOTE, Stackmine
B2B Senior Site Reliability Engineer, REMOTE KATOWICE, Jamf
Service Reliability Engineer, WROCŁAW, Toyota Digital HUB & SSC (TME NV/SA)
Site Reliability Engineer SRE, REMOTE KRAKÓW, Connectis_
DevOps Site Reliability Engineer, WARSZAWA, Link Group
Site Reliability Engineer, WARSZAWA, AVENGA (Agencja Pracy, nr KRAZ: 8448)