Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision!
Open 15d reposted 5× · 3 open copies
Platform Engineer – DevOps, Site Reliability Engineering (SRE) & Dynatrace
Posted Updated
Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision! Platform Engineer – DevOps, Site Reliability Engineering (SRE) & Dynatrace
Summary
The Platform Engineer ensures enterprise-scale infrastructure is highly available, scalable, and reliable by designing, automating, and monitoring systems using DevOps/SRE practices, with deep expertise in Dynatrace for observability and incident management.
Platform Engineer – DevOps, Site Reliability Engineering (SRE) & Dynatrace
Required Skills
- Strong experience as a Platform Engineer with expertise in DevOps and Site Reliability Engineering (SRE).
- Experience designing, implementing, automating, and supporting enterprise‑scale platform infrastructure.
- Strong knowledge of high availability, reliability, scalability, and performance engineering for mission‑critical applications.
- Hands‑on experience with Dynatrace, including monitoring, dashboard creation, synthetic monitoring, observability, automation, incident management, and a strong understanding of SRE best practices and cloud/platform engineering.
- Ability to drive operational excellence through automation, monitoring, reliability engineering, and continuous improvement initiatives.
Key Responsibilities
- Design, build, and maintain highly available, scalable, and resilient platform infrastructure.
- Implement modern Platform Engineering and Site Reliability Engineering (SRE) practices across enterprise applications.
- Define and maintain:
- Service Level Indicators (SLIs)
- Service Level Objectives (SLOs)
- Error Budgets
- Drive initiatives focused on:
- Reliability
- Availability
- Capacity planning
- Performance optimization
- Operational excellence
- Support production environments and participate in on‑call rotations when required.
Observability & Monitoring
- Lead the implementation and administration of enterprise monitoring and observability solutions.
- Develop and maintain Dynatrace monitoring strategies for complex distributed systems.
- Create and manage:
- Dynatrace dashboards
- Alerts
- Management Zones
- Reporting solutions
- Implement proactive monitoring for:
- Infrastructure
- Middleware
- Applications
- Databases
- APIs
- Cloud services
- Configure and optimize:
- Anomaly detection
- Problem management
- Root cause analysis
Dynatrace Expertise
- Hands‑on experience with:
- Dynatrace OneAgent deployment
- Dynatrace SaaS environments
- Dynatrace Managed environments
What they ask for
Required
- Strong experience as a Platform Engineer with DevOps and SRE expertise
- Experience designing, implementing, automating, and supporting enterprise-scale platform infrastructure
- Strong knowledge of high availability, reliability, scalability, and performance engineering for mission-critical applications
- Hands-on experience with Dynatrace (monitoring, dashboards, synthetic monitoring, observability, automation, incident management)
- Ability to drive operational excellence through automation, monitoring, reliability engineering, and continuous improvement