Site Reliability Engineer

Highlights:

 7.00 – 10.00 Years

 15.00 – 25.00 INR (Lacs)/Yearly

 Full-time

 Noida, Hyderabad, Gurugram

Skills

SRE – Site Reliability

SRE Production Support

SRE Implementation

dynatrace

Dynatrace Apm

Python

Bash Scripting

Shell Scripting

Kubernetes

cloud native

Ci/Cd

Devops

Docker

Linux

Roles & Responsibility

Site Reliability Engineer (SRE) – Dynatrace Platform

Administrator – Noida/GGN/Hyderabad

Role Summary

We are seeking an experienced Site Reliability Engineer (SRE) with strong hands-on Dynatrace Administration experience. This role requires an engineer who has designed, implemented, configured, administered, and optimized Dynatrace platforms in enterprise environments rather than simply using Dynatrace as an end user. The engineer will be responsible for ensuring platform reliability, observability, monitoring excellence, performance management, automation, and operational stability across critical business applications and infrastructure.

Key Responsibilities

  •  Administer, configure, and maintain Dynatrace SaaS/Managed environments.
  •  Install, upgrade, and manage Dynatrace OneAgents, ActiveGates, Extensions, and integrations.
  •  Design monitoring architecture and observability standards across cloud, hybrid, and on-premise environments.
  •  Create and maintain dashboards, alerts, tagging strategies, management zones, and monitoring policies.
  •  Perform root cause analysis using Dynatrace AI-driven insights and monitoring data.
  •  Support incident, problem, and change management processes for production systems.
  •  Optimize application performance, infrastructure health, and service reliability.
  •  Integrate Dynatrace with ticketing, DevOps, CI/CD, and ITSM platforms.
  •  Define SLI, SLO, and SLA metrics and support reliability engineering initiatives.
  •  Automate operational tasks using scripting and infrastructure automation tools.
  •  Support capacity planning, availability reviews, and continuous improvement initiatives.
  •  Collaborate with development, operations, cloud, and platform teams.

Requirements

Required Skills

  •  5+ years of experience in Site Reliability Engineering, Platform Operations, or Production Support.
  •  Strong hands-on Dynatrace Administration and Configuration experience.
  •  Experience with Dynatrace OneAgent deployment, ActiveGate administration, Synthetic Monitoring, Real User Monitoring, Infrastructure Monitoring, and Application
See also  Technology Lead - Java

Performance Monitoring.

  •  Expertise in dashboards, alerting, problem detection configuration, management zones, tagging, and automation.
  •  Strong Linux and/or Windows administration skills.
  •  Experience with AWS, Azure, or GCP environments.
  •  Knowledge of CI/CD pipelines, DevOps practices, and observability frameworks.
  •  Experience with incident management, RCA, and performance tuning.

Preferred Skills

  •  Python, Shell, or PowerShell scripting.
  •  Kubernetes, Docker, and cloud-native technologies.
  •  ITIL processes and enterprise monitoring practices.
  •  Dynatrace certifications are highly preferred.

Experience & Qualification

  •  7–10 years of overall IT experience.
  •  Minimum 3–5 years of hands-on Dynatrace administration, setup, and configuration experience.
  •  Bachelor’s degree in Computer Science, Information Technology, Engineering, or related discipline.

Key Traits

  •  Strong ownership and reliability mindset.
  •  Excellent troubleshooting and analytical skills.
  •  Ability to work in high-pressure production environments.
  •  Strong communication and stakeholder management skills.
  •  Open to supporting global teams and offshore delivery models.