Title: Dynatrace SRE
Location: Fort Mill SC
Duration: 12+ Months
Rate: $65/hr on c2c
Note: Don’t share any profile without passport number.
We are looking for a Site Reliability Engineer with deep expertise in Dynatrace and a strong background in observability, automation, and cloud operations. This role focuses on designing and implementing highly reliable, scalable solutions while driving proactive monitoring and operational excellence.
Key Responsibilities:
Lead the design and implementation of full-stack
observability solutions with Dynatrace as the primary platform.
Configure Dynatrace for application performance monitoring (APM),
infrastructure monitoring, and intelligent alerting.
Build advanced dashboards and integrate Dynatrace with event management systems
to enable proactive incident prevention and root cause analysis.
Collaborate with teams to optimize Dynatrace usage for AIOps-driven insights
and automated anomaly detection.
Provide oversight for production operations to maximize reliability and
automation.
Develop and evolve SRE best practices, runbooks, and tooling to ensure high
availability and resilience.
Implement data-driven operational strategies to improve decision-making and
reduce MTTR.
Hands-on experience with Dynatrace, Splunk, ELK, Grafana, Prometheus, and
(future) ThousandEyes.
Build and manage CI/CD pipelines and Infrastructure as Code (IaC) solutions
using Terraform, Jenkins, TeamCity, Octopus, Bamboo, and U-Deploy across
hybrid/multi-cloud environments.
Develop and manage DevOps pipelines in AWS, Azure, and GCP using Terraform and
cloud-native tooling.
Strong developer background with the ability to understand application layers
and infrastructure interactions.
Define and document standard operating procedures, architecture diagrams, and
system documentation using Jira, Confluence, and UML.
Identify areas for process and efficiency improvement within Platform Services
Operations; recommend and implement solutions.
Drive automation initiatives across all operational processes.
Proactively monitor system capacity and health indicators; provide analytics
and forecasts for scaling.
Preferred Qualifications
Expert-level experience with Dynatrace, including dashboard
creation, alert configuration, and integration with other observability tools.
Strong knowledge of AIOps, performance tuning, and proactive incident
management.
Familiarity with hybrid/multi-cloud environments and modern DevOps practices.
Excellent problem-solving skills and ability to work in a fast-paced,
collaborative environment."
Regards
Divyansh