Job Overview
- Design, implement, and maintain enterprise-wide observability solutions using Dynatrace across web, mobile, APIs, and cloud-native applications.
- Configure Real User Monitoring (RUM), Application Performance Monitoring (APM), infrastructure monitoring, synthetic monitoring, and distributed tracing.
- Develop executive and operational dashboards that provide real-time visibility into application health, user experience, service performance, and platform availability.
- Implement AI-driven monitoring using Dynatrace Davis AI to establish intelligent baselines, reduce alert noise, and proactively identify performance issues.
- Build monitoring solutions for cloud-native microservices and distributed systems running in Google Cloud Platform (GCP).
- Integrate Dynatrace with PagerDuty and other incident management platforms to automate alerting and streamline operational response.
- Collaborate with software engineers and platform teams to establish observability standards throughout the software development lifecycle.
- Monitor critical business transactions and user journeys to ensure optimal application performance and availability.
- Analyze application performance trends, identify bottlenecks, and recommend improvements to increase system reliability.
- Support production incident investigations through performance analysis, root cause identification, and post-incident reviews.
- Define monitoring standards, documentation, and best practices for enterprise observability.
- Partner with DevOps, SRE, and Infrastructure teams to continuously improve operational excellence and platform resilience.
Ready to Apply?
Take the next step in your career journey
Stand out with a professional resume tailored for this role