Job DescriptionAs a Site Reliability Engineer, you will:- Manage, monitor, and improve the reliability, availability, and performance of production systems.
- Take ownership of production environments and drive continuous operational improvements.
- Work with Azure cloud services, including compute, storage, and networking.
- Support distributed systems and microservices to ensure high system resilience.
- Collaborate with cross-functional teams to resolve incidents and optimize system performance.
- Implement automation using Infrastructure as Code (Terraform, ARM, Bicep) and CI/CD pipelines.
- Contribute to Agile ways of working while ensuring operational excellence.
What You Bring to the Table:- 6+ years of experience in Site Reliability Engineering (SRE), Reliability Engineering, or similar roles.
- Strong hands-on experience managing production systems and driving reliability improvements.
- Experience working with Microsoft Azure cloud services.
- Knowledge of distributed systems and microservices architecture.
- Experience with Infrastructure as Code tools such as Terraform, ARM, or Bicep.
- Familiarity with automation tools and CI/CD pipelines.
- Strong stakeholder management, communication, and collaboration skills.
- Experience working in Agile environments.
You should possess the ability to:- Ensure the reliability, scalability, and availability of production systems.
- Troubleshoot and resolve complex production issues efficiently.
- Automate infrastructure provisioning and deployment processes.
- Collaborate effectively with technical and business stakeholders across teams.
- Drive continuous improvement initiatives for system performance and operational efficiency.
- Work independently while taking ownership of critical production environments.
- Adapt to evolving technologies and operational requirements.
What We Bring to the Table:- An opportunity to work on large-scale, cloud-based production environments using modern SRE practices.
- Exposure to Azure cloud technologies, automation, and Infrastructure as Code.
- A collaborative Agile work environment with cross-functional teams.
- Opportunities to contribute to highly available, scalable, and resilient systems.
- A role that values ownership, innovation, and continuous improvement.
Let’s Connect
Want to discuss this opportunity in more detail? Feel free to reach out.
Recruiter: Aswin Dhanvandhar
Phone: +31 20 369 0609 ; Extn :141
Email: aswin.d@stafide.nlLinkedIn:
https://www.linkedin.com/in/aswin-dhanvandhar/RequirementsAs a Site Reliability Engineer, you will: Manage, monitor, and improve the reliability, availability, and performance of production systems. Take ownership of production environments and drive continuous operational improvements. Work with Azure cloud services, including compute, storage, and networking. Support distributed systems and microservices to ensure high system resilience. Collaborate with cross-functional teams to resolve incidents and optimize system performance. Implement automation using Infrastructure as Code (Terraform, ARM, Bicep) and CI/CD pipelines. Contribute to Agile ways of working while ensuring operational excellence. What You Bring to the Table: 6+ years of experience in Site Reliability Engineering (SRE), Reliability Engineering, or similar roles. Strong hands-on experience managing production systems and driving reliability improvements. Experience working with Microsoft Azure cloud services. Knowledge of distributed systems and microservices architecture. Experience with Infrastructure as Code tools such as Terraform, ARM, or Bicep. Familiarity with automation tools and CI/CD pipelines. Strong stakeholder management, communication, and collaboration skills. Experience working in Agile environments. You should possess the ability to: Ensure the reliability, scalability, and availability of production systems. Troubleshoot and resolve complex production issues efficiently. Automate infrastructure provisioning and deployment processes. Collaborate effectively with technical and business stakeholders across teams. Drive continuous improvement initiatives for system performance and operational efficiency. Work independently while taking ownership of critical production environments. Adapt to evolving technologies and operational requirements. What We Bring to the Table: An opportunity to work on large-scale, cloud-based production environments using modern SRE practices. Exposure to Azure cloud technologies, automation, and Infrastructure as Code. A collaborative Agile work environment with cross-functional teams. Opportunities to contribute to highly available, scalable, and resilient systems. A role that values ownership, innovation, and continuous improvement. Let’s Connect Want to discuss this opportunity in more detail? Feel free to reach out. Recruiter: Aswin Dhanvandhar Phone: +31 20 369 0609 ; Extn :141 Email: aswin.d@stafide.nl LinkedIn:https://www.linkedin.com/in/aswin-dhanvandhar/