Ensono Logo

Ensono

Site Reliability Engineer

Reposted One Month Ago
Easy Apply
Remote or Hybrid
Hiring Remotely in United Kingdom
Mid level
Easy Apply
Remote or Hybrid
Hiring Remotely in United Kingdom
Mid level
Design, deploy, and maintain reliable cloud-native infrastructure using Terraform, Azure, and Kubernetes. Troubleshoot incidents, lead incident resolution and post‑mortems, improve CI/CD and monitoring, reduce toil, manage security risks, and lead client and supplier engagements to expand SRE practices.
The summary above was generated by AI
Who are we?

Ensono is a global technology services provider dedicated to helping organizations navigate the complexity of digital transformation. Through Ensono Product, Consulting & Technology, our dedicated consulting arm, we partner with clients to design, build, and modernize digital capabilities across application development and modernization, data platforms and AI, and identity and access management. As we continue to expand our global consulting footprint, we are committed to delivering innovative, high‑quality outcomes that enable our clients to move faster, smarter, and more securely.

About the role and what you'll be doing: 

We are seeking an experienced Site Reliability Engineer (SRE) with expertise in Infrastructure as Code tools like Terraform, core CI/CD tools such as Azure DevOps, and monitoring tools including DataDog and AWS CloudWatch. The ideal candidate will have commercial experience in technologies like Dotnet or Java, and be skilled in troubleshooting, incident resolution, and improving service and change management processes. Strong leadership in client-facing discussions and engagement with third-party suppliers is essential. An SRE Foundation certificate and a cloud provider associate-level certification are highly beneficial. 

  • Commercial experience and proficiency with industry standard: 

  • IAC tooling (Terraform preferably, or ARM/bicep and CloudFront) 

  • Core CI/CD Tooling (Azure DevOps, GitHub Actions or Gitlab) 

  • Monitoring Tooling (DataDog, Splunk, NewRelic, Azure Monitor, AWS CloudWatch) 

  • Commercial experience in at least one core technology (Dotnet, Java, AI/Data Engineering, Golang) 

  • Troubleshooting issues and identifying systemic failings indicated by incidents/failures 

  • Implementing fixes 

  • Proposing solutions for reducing toil 

  • Providing leadership in the Incident resolution process, including creating and maintaining documentation, and providing key input to Post-mortem analysis 

  • Improving Service Requests and Change Management processes, both technically and through stakeholder management). 

  • Participate in the process for, and Proactively mitigate risks in a Security management process (Vulnerabilities in Code, Infrastructure, Dependencies) 

  • Lead discussion in client-facing meetings and discussions around the SRE process, and identifying areas for increasing SRE footprint. 

  • Engaging with suppliers and 3rd parties for support, requests and opportunities 


We want all new Associates to succeed in their roles at Ensono. That's why we've outlined the job requirements below. To be considered for this role, it's important that you meet all Required Qualifications. If you do not meet all of the Preferred Qualifications, we still encourage you to apply.  


Required Qualifications  

  • 3-9 Years experience  

  • Bachelor’s degree (or equivalent) in computer science or related discipline 

  • SRE Foundation certificate (DevOps Institute) and a Cloud provider (AWS, Azure, GCP) 'associate'-level certification, or completed during the probationary period. 

  • Proficiency in Azure and Kubernetes, with hands-on experience in managing and deploying applications. 

  • Expertise in Infrastructure as Code (IaC) using Terraform for efficient and scalable infrastructure management. 


Preferred Qualifications 

  • Certified Kubernetes Administrator / Application Developer 

  • Certified Azure DevOps Engineer   

  • Experience with monitoring tools such as NewRelic or Splunk for effective system monitoring and alerting. 

  • Familiarity with Harness for continuous delivery and deployment processes. 

  • Strong programming skills in .Net, Java, or JavaScript for developing robust and scalable applications 

Ensono Belfast, Northern Ireland Office

Belfast, United Kingdom

Similar Jobs

26 Days Ago
Remote or Hybrid
United Kingdom
Expert/Leader
Expert/Leader
Financial Services
Leads the design and operation of highly available, scalable, and observable production infrastructure. Defines SLOs, error budgets, and reliability targets; drives incident response, root-cause analysis, and postmortems; develops automation and deployment tooling; champions monitoring and observability; evaluates platform technologies; and mentors engineers. The role also applies secure AI-assisted engineering practices to improve incident triage, testing, and delivery workflows.
Top Skills: AWSBashGoJavaKubernetesPythonTerraform
7 Hours Ago
Remote
United Kingdom
Entry level
Entry level
Cloud • Software • Analytics
Supports 24/7 service reliability through incident response, monitoring, alerting, observability, automation, and production operations. The role manages major incidents, improves runbooks and SLO-based alerting, develops scripts and self-healing tools, and supports Linux, cloud, and Kubernetes environments. It partners with engineering, infrastructure, security, and product teams to reduce operational toil and improve system availability, MTTD, and MTTR.
Top Skills: AnsibleAWSAzureBashCloudwatchDatadogDnsDockerGCPGoGrafanaKubernetesLinuxPrometheusPythonSplunkTcp/IpTerraform
10 Days Ago
Remote
United Kingdom
Senior level
Senior level
Cloud • Information Technology • Consulting
The Site Reliability Engineer will operate and improve production cloud environments, build infrastructure automation, enhance observability, optimize distributed system performance, and support incident response. The role focuses on reliability for Voice/Unified Communications and AI-enabled services, including real-time media quality, graceful degradation, dependency isolation, and recovery procedures. Responsibilities include capacity planning, release validation, monitoring, troubleshooting, service-level management, and collaboration with development, Voice/UC, and AI engineering teams.
Top Skills: Amazon S3Ci/CdContainersDevOpsDistributed SystemsHdfsInfrastructure AutomationKubernetesLinuxNfsObservabilityRtpSession Border ControllersSipSrtpWebrtc

What you need to know about the Belfast Tech Scene

If asked to name the birthplace of the RMS Titanic, you might not say Belfast. Similarly, if asked to name Europe's leading destination for foreign direct investment in new software development, Belfast might not come to mind. Yet, both are true. The city has emerged as a tech powerhouse, recently ranked among the best in the U.K. for tech careers — especially for software developers. It also leads the U.K. with the highest percentage of software development jobs advertised.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account