Description

Job Title:<br/><br/>Senior Site Reliability/Observability Engineer
Duration:<br/><br/>3-Month Contract (Aug – October)
Location:<br/><br/>Toronto, ON (Hybrid – 2 to 3 days onsite per week)
Job Overview<br/><br/>We are seeking a Senior Site Reliability/Observability Engineer to support a critical operational excellence initiative focused on Application Performance Monitoring (APM), Operational Readiness, Compliance Automation, and Dashboarding.
The successful candidate will be responsible for designing, developing, and implementing monitoring solutions, audit workflows, compliance automation, and executive reporting dashboards that improve system visibility, operational governance, and business decision-making.
This is a highly hands-on role requiring strong full-stack development experience, integration expertise, and the ability to work closely with operations, compliance, and business stakeholders.
Responsibilities:<br/><br/>Application Performance Monitoring (APM)<br/><br/>Design and implement APM solutions to monitor application health and performance
Configure monitoring for latency, throughput, error rates, and user experience metrics
Establish performance baselines and alerting thresholds
Enable end-to-end transaction tracing and dependency mapping
Operational Readiness Framework<br/><br/>Develop pre-deployment validation processes and operational readiness checklists
Create operational runbooks, standards, and support documentation
Implement deployment validation and approval workflows
Audit & Compliance Automation<br/><br/>Design and implement automated audit trails for system changes and deployments
Build compliance workflows for evidence collection and reporting
Develop audit dashboards and compliance reporting capabilities
Dashboarding & Visualization<br/><br/>Build role-based dashboards for executives, operations, and technical teams
Provide real-time visibility into system health, incidents, and performance trends
Enable self-service reporting and operational insights
Qualifications<br/><br/>Must-Have Skills<br/><br/>8+ years of Full Stack Development experience
Strong experience building enterprise dashboards and reporting solutions
Experience implementing APM/Observability platforms (Dynatrace, AppDynamics, Datadog, New Relic, Splunk, etc.)
Strong experience with API development and system integrations
Experience building workflow automation and operational tooling
Front-end development experience (React, Angular, or similar)
Back-end development experience (.NET, Java, Node.js, Python, or similar)
Experience with SQL and data visualization technologies
Strong understanding of DevOps, monitoring, and operational best practices
Nice-to-Have Skills<br/><br/>Experience with compliance, audit, or governance automation
Experience building executive and operational dashboards
Knowledge of ITIL, operational readiness, or change management processes
Cloud experience (Azure, AWS, or GCP)
Experience with CI/CD pipelines and Infrastructure as Code

IndKyn

Please note this is for a contract position with one of our clients and not a fulltime employment role with Kyndryl Canada**<br/><br/>#J-18808-Ljbffr