DevOps

Mindbox Sp. z o.o. KRAKÓW 2026-09-28


  • 4+ years of experience in developing and supporting distributed systems (Java-based preferred).

  • Hands-on experience in platform support for large-scale Java and Python applications across multi-platform environments.

  • Practical experience with CI/CD tooling: Jenkins, Ansible, and configuration automation.

  • Knowledge of logging, monitoring, and alerting solutions (Grafana, Prometheus, InfluxDB, Splunk, Loki).

  • Familiarity with cloud platforms, preferably Google Cloud Platform (GCP).

  • Strong understanding of Unix/Linux systems, relational databases (Oracle, PostgreSQL), and disaster recovery processes.

  • Ability to conduct technical troubleshooting and lead RCA sessions across global teams.

  • Proven knowledge of Agile/Kanban methodologies in a software delivery context.

  • Excellent verbal and written communication skills (English fluency required).




Joining this project you’ll become part of Mindbox – a tech-driven company where consulting, engineering, and talent meet to build meaningful digital solutions. We’ll back you up every step of the way, accelerate your development, and ensure your skills make a difference. 



At Mindbox we connect top IT talents with technology projects for leading enterprises across Europe. 



 


We are seeking a DevOps / Site Reliability Engineer (SRE) to join a high-performing engineering team responsible for developing and maintaining microservices for Counterparty Credit Risk (CCR) across global environments. This is a unique opportunity to work on business-critical risk solutions as part of a multi-year cloud migration program, leveraging microservices, container-based deployments, and distributed computing on Google Cloud Platform (GCP).


This role combines elements of platform reliability, application support, incident management, and automation, making it ideal for engineers passionate about stability, performance, and observability in complex distributed systems.




Sounds like your kind of challenge? 



What you get in return


  • Flexible cooperation model 

  • Hybrid work setup – 2 days a week from the office

  • Collaborative team culture – work alongside experienced professionals eager to share knowledge 

  • Continuous development – access to training platforms and growth opportunities 

  • Comprehensive benefits – including Interpolska Health Care, Multisport card, Warta Insurance, and more 

  • High quality equipment – laptop and essential software provided 


,[Manage application support operations with a strong focus on platform resiliency, high availability, and performance monitoring., Investigate, triage, and resolve incidents through root cause analysis (RCA) and implement long-term fixes., Document recovery steps, create knowledge base content, and drive process improvements., Actively participate in Incident Management, Problem Management, and Service Delivery activities., Apply SRE principles to improve capacity, reliability, and observability metrics, reducing operational overhead., Develop and enhance monitoring and alerting frameworks, incident detection, and release safety tools using tools such as Grafana, Prometheus, InfluxDB, Splunk, or similar., Collaborate globally and coordinate with cross-functional engineering and support teams., Work in rotational shifts (Morning: 8 AM; Evening: 4 PM) and participate in on-call and weekend support rotation.] Requirements: Java, Python, Jenkins, Ansible, Grafana, Prometheus, InfluxDB, Splunk, Loki, Cloud platform, Oracle, PostgreSQL, Communication skills, Google Cloud Platform, Unix, Linux