
Akvelon is looking for experienced Site Reliability Engineers to join a long-term project for a global technology leader. We are building a dedicated SRE team responsible for supporting business-critical cloud infrastructure with 24×7 DRI coverage. This is an opportunity to work on enterprise-scale Azure environments, modern cloud technologies, and highly available distributed systems that power services used worldwide.
We’re hiring 5 Engineers: Middle / Mid-Senior / Senior SRE. Location is strictly limited to Poland and Serbia.
Responsibilities:
- Maintain and improve the reliability, availability, and performance of production services.
- Support and operate large-scale Azure infrastructure.
- Manage and troubleshoot Azure Kubernetes Service (AKS) environments.
- Investigate and resolve production incidents across infrastructure and application layers.
- Automate operational processes using scripting and infrastructure automation.
- Work closely with software engineers to improve system reliability and deployment processes.
- Manage configuration, templating, and infrastructure changes.
- Troubleshoot networking issues involving HTTP, DNS, TCP/IP, BGP, and hybrid cloud networking.
- Ensure compliance with security best practices, patching, dependency management, and authentication standards.
- Participate in an on-call rotation as part of the 24×7 DRI support model.
Requirements:
- Proven experience as a Site Reliability Engineer, Infrastructure Engineer, or in a similar production operations role.
- Strong hands-on experience with Microsoft Azure.
- Experience with Azure Kubernetes Service (AKS) and containerized workloads.
- Strong understanding of automation, templating, and configuration management.
- Solid networking knowledge, including HTTP, DNS, TCP/IP, and BGP.
- Experience troubleshooting production environments across infrastructure and application layers.
- Ability to perform basic code debugging and implement small fixes when needed.
- Understanding of distributed systems and modern infrastructure architecture.
- Experience with PowerShell or another scripting language.
- Knowledge of security best practices, patch management, package management, and secure authentication.
Nice to have:
- Experience managing Dell Edge servers or similar compute infrastructure.
- Experience with Dell iDRAC or comparable remote management solutions.
- Experience supporting large-scale, multi-region datacenter environments.
- Experience using AI-assisted engineering tools such as GitHub Copilot or Azure AI.
Benefits:
- Paid vacation and sick leave
- Official state holidays — 9-11 public holidays per year
- Professional growth through challenging projects and the opportunity to master new technologies
- Flexible working schedule: 8 hours per day, 40 hours per week
- Personal Career Development Plan (CDP)
- Employee support programs (discounts, healthcare, legal assistance)
- Paid external training, conferences, and professional certifications aligned with business goals
- Internal workshops, seminars, and English classes
- Corporate library with paper and e-books
If you’re passionate about reliability engineering and enjoy solving complex infrastructure challenges at scale, we’d love to hear from you.