Job Title: Monitoring Tool Automation Engineer
Location: Toronto,Onsite
Job Type: Full-Time Permanent
Must-Have Skills:
8–10 years of overall IT/infrastructure experience, with 5+ years working with Windows and/or Linux server environments.
Strong hands-on expertise in Python and PowerShell scripting/automation.
Strong Windows administration and Linux/Unix Shell/Bash scripting skills.
Experience with infrastructure monitoring, log analysis, system health checks, and performance troubleshooting.
Knowledge of monitoring tools such as Nagios, ITRS, or Icinga.
Experience in incident management and root cause analysis.
Strong communication, documentation, and stakeholder-management skills.
Nice-to-Have Skills:
Ansible or Terraform for infrastructure automation.
Cloud platform exposure including Azure, AWS, or GCP.
Experience integrating monitoring platforms with enterprise ticketing and notification solutions.
Key Responsibilities:
Develop and maintain infrastructure automation solutions using Python and PowerShell.
Automate monitoring, health checks, operational activities, and server administration across Windows and Linux environments.
Configure and support infrastructure monitoring using Nagios, ITRS, and Icinga.
Develop monitoring dashboards, alerts, and performance reports.
Perform server performance tuning, troubleshooting, and proactive system-health monitoring.
Support incident management, root cause analysis, and problem resolution.
Integrate monitoring solutions with enterprise notification and ticketing systems.
Maintain automation scripts, operational documentation, and monitoring standards.
Support production operations in fast-paced enterprise environments.