Skip to content
Apply now

This job is listed on several sites

ITP IT PERFORMANCE  sp. z o.o.

Senior Site Reliability Engineer

ITP IT PERFORMANCE sp. z o.o. Show all offers
B2B
Hybrid
Senior
Kraków
1 month ago

In short

Senior Site Reliability Engineer needed for a multi-cloud platform. Responsibilities include automation, incident response, and improving reliability metrics. Requires 10+ years of experience in Software Engineering/DevOps/Infrastructure, 3+ years of AWS & Kubernetes experience. B2B contract, remote with occasional European travel.

AI-written summary based on the listing content.

Technologies we use

About the project

This is how we organize our work

This is how we work

Your responsibilities

  • Design and drive the reliability, scalability, and performance of the multi-cloud provisioning platform across production environments.
  • Design and implement end-to-end automation pipelines to eliminate manual processes and reduce technical toil.
  • Define, monitor, and continuously improve key reliability metrics and SLIs/SLOs, including latency, throughput, error rates, and capacity utilization.
  • Collaborate with product and engineering teams to integrate reliability and security considerations early in the software development lifecycle.
  • Own and lead incident response, including detection, triage, critical incident management, Root Cause Analysis (RCA), and implementation of long-term preventive solutions.
  • Identify platform bottlenecks, eliminate single points of failure, and continuously simplify complex systems.
  • Take ownership of architectural decisions and establish reliability standards across the platform.
  • Provide technical guidance and mentoring to other engineers within the team.

Our requirements

  • 10+ years of hands-on experience in Software Engineering, DevOps, or Systems Infrastructure.
  • 3+ years of advanced AWS experience, including designing, operating, and troubleshooting cloud environments.
  • 3+ years of production-level Kubernetes experience, including cluster management, scaling, and provisioning.
  • Strong understanding of site reliability, automation, cloud infrastructure, and production-grade systems.
  • Proven ability to work with a high degree of autonomy, making sound architectural decisions and delivering production-ready solutions without day-to-day technical supervision.
  • Strong experience in designing scalable and automated infrastructure.
  • Excellent technical communication skills and the ability to collaborate effectively with Technical Leads and cross-functional engineering teams.
  • Ability to remain effective and structured when dealing with high-pressure production incidents.

What we offer

  • B2B Contract
  • Remote work with occasional travel within Europe

This is how we work on a project

Published 2026-08-21
Expires 2026-09-20
Source