Login Sign Up

Site Reliability Engineer

RIB Software

2 - 5 years

Nasik

Posted: 20/05/2026

Job Description

Job Description

Role: Site Reliability Engineer

Team: Cloud Operations & Site Reliability Engineering (SRE)

Experience: 3 to 5 Years

Location: Nashik

Employment Type: Full-Time


Role Overview

We are seeking a high-caliber SRE & Reliability Engineer with a strong background in .NET development to join our Cloud Operations team. This is a "Hybrid" engineering role: you will be responsible for ensuring the high availability and performance of our production environments while actively developing custom engineering solutions to enhance our monitoring and observability stack.

The ideal candidate thrives at the intersection of software development and systems engineering, possessing the "SRE mindset" to treat operational challenges as software problems.


Key Responsibilities

1. Reliability Engineering & Automation

Design, develop, and maintain internal tools and "sidecar" applications using .NET Core/C# to automate manual operational tasks (Toil).

Build and scale "auto-healing" systems to improve service uptime and reduce MTTR (Mean Time to Recovery).

Implement Infrastructure as Code (IaC) using Bicep, Terraform, or ARM Templates to ensure consistent and repeatable deployments.

2. Observability & Monitoring

Develop and refine sophisticated monitoring solutions using Azure Monitor, Application Insights, and Grafana.

Define, track, and report on SLIs, SLOs, and Error Budgets to balance feature velocity with system stability.

Implement distributed tracing and advanced logging (OpenTelemetry) to provide deep visibility into microservices architecture.

3. Command Center (COC) & Incident Management

Participate in system monitoring and high-priority alert handling.

Lead Root Cause Analysis (RCA) for production incidents, ensuring that "post-mortems" lead to permanent engineering fixes.

Create and maintain technical Runbooks and SOPs to empower the broader support and infra teams.

4. Cloud Infrastructure & DevOps

Manage and optimize cloud-native services including Azure App Services, Azure Functions, and Azure Kubernetes Service (AKS).

Maintain and enhance CI/CD pipelines (Azure DevOps / GitHub Actions) to ensure secure and seamless code promotion.


Technical Qualifications

Development: 3+ years of professional experience in .NET / .NET Core / ASP.NET. Expert-level C#, Web APIs, and a deep understanding of asynchronous programming and multithreading.

Database: Strong proficiency in SQL Server and experience with NoSQL environments (preferably Cosmos DB).

Cloud: Hands-on experience with the Microsoft Azure ecosystem and containerization (Docker & Kubernetes).

Scripting: Proficiency in PowerShell or Bash for system-level automation.

Security: Solid understanding of REST API security, including OAuth and JWT.


Soft Skills & Competencies

Problem-Solving: A natural curiosity to deconstruct complex system failures.

Communication: Ability to communicate clearly and calmly during high-pressure production incidents.

Accountability: A strong sense of ownership over the "health" of the services you manage.

Collaboration: Experience working closely with cross-functional Dev, Infra, and Support teams.

Services you might be interested in

We Search & Apply Jobs for You!

Our team scans through 1000s of opportunities and applies to roles best suited to your profile

Save 100+ hours and focus on what matters - cracking interviews and landing offers.