Ukraine Flag We stand with Ukraine! Learn More
  • Home
  • /
  • Vacancies
  • /
  • Middle/ Senior Site Reliability Engineer (SRE) – Azure / AKS

Middle/ Senior Site Reliability Engineer (SRE) – Azure / AKS

Serbia

Akvelon is looking for experienced Site Reliability Engineers to join a long-term project for a global technology leader. We are building a dedicated SRE team responsible for supporting business-critical cloud infrastructure with 24×7 DRI coverage. This is an opportunity to work on enterprise-scale Azure environments, modern cloud technologies, and highly available distributed systems that power services used worldwide.

We’re hiring 5 Engineers: Middle / Mid-Senior / Senior SRE. Location is strictly limited to Poland and Serbia.

Responsibilities:

  • Maintain and improve the reliability, availability, and performance of production services.
  • Support and operate large-scale Azure infrastructure.
  • Manage and troubleshoot Azure Kubernetes Service (AKS) environments.
  • Investigate and resolve production incidents across infrastructure and application layers.
  • Automate operational processes using scripting and infrastructure automation.
  • Work closely with software engineers to improve system reliability and deployment processes.
  • Manage configuration, templating, and infrastructure changes.
  • Troubleshoot networking issues involving HTTP, DNS, TCP/IP, BGP, and hybrid cloud networking.
  • Ensure compliance with security best practices, patching, dependency management, and authentication standards.
  • Participate in an on-call rotation as part of the 24×7 DRI support model.

Requirements:

  • Proven experience as a Site Reliability Engineer, Infrastructure Engineer, or in a similar production operations role.
  • Strong hands-on experience with Microsoft Azure.
  • Experience with Azure Kubernetes Service (AKS) and containerized workloads.
  • Strong understanding of automation, templating, and configuration management.
  • Solid networking knowledge, including HTTP, DNS, TCP/IP, and BGP.
  • Experience troubleshooting production environments across infrastructure and application layers.
  • Ability to perform basic code debugging and implement small fixes when needed.
  • Understanding of distributed systems and modern infrastructure architecture.
  • Experience with PowerShell or another scripting language.
  • Knowledge of security best practices, patch management, package management, and secure authentication.

Nice to have:

  • Experience managing Dell Edge servers or similar compute infrastructure.
  • Experience with Dell iDRAC or comparable remote management solutions.
  • Experience supporting large-scale, multi-region datacenter environments.
  • Experience using AI-assisted engineering tools such as GitHub Copilot or Azure AI.

Benefits:

  • Paid vacation and sick leave
  • Official state holidays — 9-11 public holidays per year
  • Professional growth through challenging projects and the opportunity to master new technologies
  • Flexible working schedule: 8 hours per day, 40 hours per week
  • Personal Career Development Plan (CDP)
  • Employee support programs (discounts, healthcare, legal assistance)
  • Paid external training, conferences, and professional certifications aligned with business goals
  • Internal workshops, seminars, and English classes
  • Corporate library with paper and e-books

If you’re passionate about reliability engineering and enjoy solving complex infrastructure challenges at scale, we’d love to hear from you.

Tetiana Petryshyna

IT Recruiter

Apply

or Refer a friend

Get up to $1,5K for
a referrence

Long-term partnership approach

Refer a friend