Login Sign Up

Site Reliability Engineer

Insight Global

2 - 5 years

Hyderabad

Posted: 02/06/2026

Job Description

Required Skills & Experience


Bachelors degree in a related discipline, such as Computer Science, Electrical Engineering, or related; a masters degree is preferred. A college or advanced educational degree is not necessarily required for this position. We also value education achieved via work experience, coursework and training, and even self-guided learning. A minimum of 5+ years in DevOps, SRE, or Cloud Operations roles, with at least 3 years on Azure. Hands-on Terraform experience at the level of authoring and maintaining modules in production not just running someone elses. Production support experience: on-call rotation, incident response, and familiarity with ITSM ticket flows (ServiceNow or equivalent). Working knowledge of CI/CD pipelines Azure DevOps Pipelines or GitHub Actions including build, test, deploy stages and basic pipeline troubleshooting. Comfortable with at least one of Site24x7 or Datadog at the level of building dashboards and alerts, plus working knowledge of Azure Monitor (Log Analytics, KQL). Day-to-day use of an AI coding assistant (GitHub Copilot, Claude Code, ChatGPT, Cursor, or similar) in a professional setting. Solid PowerShell and Python scripting skills.


Nice to Have Skills & Experience


Experience transitioning a managed-service or outsourced function back in-house. Exposure to Microsoft Sentinel, Defender for Cloud, and Logic Apps SOAR playbooks. Familiarity with Kubernetes (AKS), Container Apps, or KEDA. Experience with infrastructure-as-code policy frameworks (Checkov, OPA, or similar). Healthcare-interface experience such as HL7 / DICOM / FHIR is a plus. Familiarity with Power BI / Microsoft Fabric observability for analytics workloads. Prior experience using LLM APIs (OpenAI, Anthropic, or Azure OpenAI) for ops automation. Comfortable working hours overlapping with the UK and Australia time zones for handovers.


Job Description


We are looking for a skilled SRE/Devops Engineer to join our team. This role sits within Core Operations and works closely with the Service & Support squad on escalations, but it is not a traditional reactive L2/L3 support position. The focus is on proactively maintaining and improving platform health through engineering-led work. You will contribute to Terraform modules that define our Azure infrastructure, operate and enhance CI/CD pipelines, and design observability solutions using Site24x7 and Datadog to surface issues before they impact users. The role also involves building lightweight automations and runbooks to eliminate recurring manual effort across the team. While you will step in to handle Sev1 and Sev2 incidents when escalations exceed documented runbooks, the day-to-day work is primarily engineering-driven rather than ticket-based. We expect regular use of AI-assisted development tools such as GitHub Copilot, Claude Code, or similar as part of your daily workflow.

Services you might be interested in

We Search & Apply Jobs for You!

Our team scans through 1000s of opportunities and applies to roles best suited to your profile

Save 100+ hours and focus on what matters - cracking interviews and landing offers.