E-IT

Site Reliability Engineer

📍 Location
ontario, ontario
⏰ Job Type
Full-time
📅 Posted
June 01, 2026
Apply Now

Job Description

Job Description

Key Responsibilities:

  • Incident Management and Reliability: Lead the incident management process, ensuring high availability and performance of the applications. Develop and implement SRE practices to improve system reliability and resilience.
  • Monitoring and Observability: Utilize Dynatrace, Splunk, and Grafana to monitor system health, detect anomalies, and provide actionable insights for performance optimization.
  • Root Cause Analysis: Conduct thorough root cause analysis of incidents and outages, developing long-term solutions to prevent recurrence.
  • DevOps Practices: Collaborate with development and operations teams to streamline CI/CD pipelines, automate workflows, and implement infrastructure as code (IaC) for efficient service deployment and management.
  • Networking Expertise: Provide expertise in networking technologies (Cisco, Arista, AVI, etc.), ensuring...

Start Your Week Right!

Apply now and make every Monday exciting with E-IT

Apply for this Position