{"id":1148417,"url":"https://alion.io/job/partsource-sr-specialist-mainframe-compute-platform-governance","title":"Sr. Specialist – Mainframe & Compute Platform Governance","company":{"id":1051221,"name":"PartSource","domain":"partsource.ca","url":"https://alion.io/company/partsource","size_band":"1001-5000","is_staffing_agency":false,"is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":{"grade":"B","score":80,"open_postings":10,"ghost_share":0,"stale_share":1,"repost_share":0,"time_to_fill_p50_days":14,"computed_at":"2026-09-24T05:45:00Z"}},"role":"Enterprise Apps","role_family":"Enterprise Apps","seniority":"senior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Toronto, Canada"],"countries":["CA"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":64000,"max":85000,"currency":"USD","period":"year","gross":null,"usd_annual":85000},"salary_estimate":null,"experience_years_min":10,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"New Relic","optional":false},{"name":"Dynatrace","optional":true},{"name":"ServiceNow","optional":true},{"name":"SLI/SLO/SLA","optional":true},{"name":"Splunk","optional":true}],"status":"live","first_seen_at":"2026-09-23T16:31:51Z","employer_posted_date":"2026-09-23","last_verified_at":"2026-09-24T15:00:40Z","board_verified":true,"closed_at":null,"days_open":1,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":1},"description":"What You’ll Do:\nThe Sr. Specialist - Mainframe & Compute Platform Governance serves as the enterprise technical owner for CTC's compute platforms, including IBM Mainframe (z/OS), AIX, IBM iSeries, IBM NOI, New Relic, and associated business-critical services and applications.\nThis role is accountable for platform governance, service ownership, lifecycle strategy, observability governance, vendor oversight, and reliability outcomes across the compute environment. Acting as the internal authority for compute platforms, the position ensures services remain reliable, resilient, secure, observable, supportable, and aligned with enterprise technology strategy, risk management, and modernization objectives.\nOperational execution, administration, monitoring response, maintenance activities, and day-to-day infrastructure support are performed by HCL/Harmony and other designated support providers. The Sr. Specialist provides governance, strategic direction, technical leadership, and vendor accountability to ensure services are delivered in accordance with enterprise standards, contractual commitments, and business requirements.\nThe role leads platform roadmap development, technology currency initiatives, service reliability improvements, observability strategy, operational governance, vendor management, and continuous improvement programs while promoting Site Reliability Engineering (SRE) principles, operational excellence, automation, and platform sustainability. The position drives platform maturity, reduces consultant and key-person dependency, and establishes clear ownership and accountability across the enterprise compute environment.\nMainframe Platform Ownership and Lifecycle Governance\nProvide governance and oversight of HCL/Harmony-delivered Mainframe services, including z/OS lifecycle planning, platform currency, capacity strategy, LPAR governance, and infrastructure roadmap planning.\nAssess current platform, operating system, middleware, and Mainframe-supported application versions; identify lifecycle, supportability, operational, security, and compliance risks; and develop renewal and modernization strategies.\nLead platform lifecycle governance activities, including hardware refresh planning, operating system upgrade strategy, maintenance governance, change oversight, and long-term roadmap alignment.\nGovern service ownership, support models, operational accountability, and vendor-delivered support services for Mainframe-hosted and Mainframe-dependent applications.\nDefine, govern, and periodically review monitoring and observability requirements for Mainframe infrastructure, middleware, and business-critical services.\nProvide governance and oversight of disaster recovery readiness, resiliency planning, recovery testing, and recovery capability validation delivered by managed service providers.\nIdentify and drive resolution of operational ownership, support model, documentation, and Statement of Work (SOW) gaps through the appropriate vendors, support teams, and stakeholders.\nProvide technical leadership and governance during major incidents, problem investigations, and corrective action planning, ensuring responsible teams execute required remediation activities.\nServe as the primary technical escalation and governance authority for Mainframe-related risks, service concerns, lifecycle issues, and vendor performance matters.\nMonitoring, Event Management and Observability Governance\nProvide governance and strategic oversight of enterprise monitoring, event management, and observability capabilities.\nEstablish governance standards for monitoring coverage, alerting requirements, event correlation, escalation models, service health dashboards, and operational reporting.\nDrive continuous improvement of observability maturity, service visibility, monitoring effectiveness, synthetic monitoring capabilities, and operational intelligence.\nGovern event management practices, including alert quality, escalation effectiveness, incident correlation, operational readiness, and service monitoring standards.\nDefine strategic direction and adoption roadmaps for observability platforms, monitoring technologies, event management tooling, and automation capabilities.\nAct as the Compute SRE representative for enterprise monitoring strategy, operational intelligence, and observability initiatives.\nManaged Service Governance and Platform Accountability\nProvide technical governance, vendor oversight, and escalation leadership for HCL/Harmony-managed services across Mainframe, AIX, IBM iSeries, storage, backup, monitoring, and supporting infrastructure platforms.\nValidate vendor-delivered services against contractual obligations, SOW commitments, service level expectations, operational controls, monitoring standards, and governance requirements.\nLead vendor performance reviews, operational scorecards, service reporting reviews, incident follow-ups, and continuous improvement initiatives.\nIdentify ownership gaps, operational risks, monitoring deficiencies, contractual concerns, and escalation requirements for leadership review.\nEnsure vendor-delivered services are properly documented, transitioned, supportable, and aligned with the enterprise operating model.\nDrive vendor accountability for service reliability, incident response effectiveness, operational reporting quality, remediation commitments, and lifecycle objectives.\nPlatform Governance, Controls and Continuous Improvement\nEnsure operational documentation, runbooks, support procedures, monitoring standards, and technical documentation are established, maintained, and periodically reviewed by responsible support teams and service providers.\nSupport audit, compliance, and control activities related to backup, recovery, disaster recovery, monitoring controls, operational governance, and infrastructure lifecycle management.\nIdentify opportunities to improve reliability, observability, automation, resiliency, operational efficiency, and platform sustainability.\nParticipate in change, incident, problem, risk, lifecycle, and service governance forums as the Compute platform owner.\nProvide reporting and recommendations on platform risks, technology currency gaps, vendor performance, monitoring effectiveness, remediation priorities, lifecycle initiatives, and service readiness.\nLead continuous improvement initiatives focused on service maturity, reliability, observability, sustainability, risk reduction, and operational excellence.\nWhat You Bring:\n10+ years of experience in platform ownership, infrastructure governance, Site Reliability Engineering (SRE), service management, enterprise technology leadership, or related disciplines.\nStrong background in Mainframe platform governance, including z/OS environments, lifecycle planning, capacity management, service ownership, and vendor-managed support models.\nExperience governing enterprise monitoring, observability, event management, or operational intelligence practices across large-scale technology environments.\nStrong understanding of service reliability principles, alert management, event correlation, service health monitoring, operational reporting, and risk management.\nKnowledge of AIX, IBM iSeries, enterprise middleware, and application hosting environments supporting business-critical workloads.\nExperience governing enterprise storage and backup services, including capacity planning, recovery validation, resiliency requirements, and restoration testing oversight.\nExperience supporting disaster recovery governance, DR exercises, recovery reporting, remediation tracking, and resiliency planning.\nExperience working within MSP-governed or vendor-managed service environments.\nStrong understanding of incident, problem, change, lifecycle, operational risk, vendor governance, and service management processes.\nDemonstrated ability to lead technical escalations, coordinate across multiple stakeholder groups, and hold vendors accountable to service expectations.\nStrong communication skills with the ability to translate technical risk into clear operational and business impact.\nProven ability to influence technical direction, challenge operational assumptions, and drive accountability across diverse stakeholder groups.\nPreferred Qualifications\nExperience supporting IBM Mainframe environments, including z/OS, middleware, enterprise schedulers, file transfer platforms, or related infrastructure services.\nExperience with observability and monitoring platforms such as IBM OMEGAMON, ITM/Tivoli, Netcool, New Relic, Dynatrace, Splunk, Elastic, ThousandEyes, or equivalent technologies.\nExperience defining monitoring standards, observability strategies, alert rationalization programs, service health reporting, event correlation approaches, and operational intelligence improvements.\nExperience supporting infrastructure audit controls, disaster recovery governance, compliance remediation activities, and evidence management processes.\nExperience supporting highly regulated, high-availability, or business-critical technology environments.\nFamiliarity with ServiceNow-based intake, incident, event, problem, change, risk, vendor, and lifecycle management workflows.\nExperience governing infrastructure refreshes, hardware lifecycle initiatives, maintenance planning, and data centre change activities.\nExperience working with outsourced infrastructure providers, including SOW governance, SLA management, contract oversight, and vendor performance management.\nExperience supporting SRE, service reliability, observability, platform governance, or operational excellence programs across hybrid infrastructure environments spanning Mainframe and distributed platforms.\nLocation: Toronto, Ontario OR Calgary, Alberta (Hybrid: In-office 4 days a week)\nWe’re always looking for great talent! In addition to competitive pay, we offer:\nComprehensive benefits and retirement programs\nPerformance incentives, Continuing Education Programs\nOther perks to support your well-being\nCareer growth opportunities and product discounts\nBroadband Salary Range: $64,000 - $106,000.\nOur typical hiring range is between $64,000 and $85,000. Salary decisions are also dependent on other factors such as your experience, industry benchmarks, internal equity and other role-specific requirements. For critical roles, the compensation offering will be reviewed to ensure alignment with market rate and conditions and the unique value you bring to the role.\nThis posting represents an existing vacancy within our organization.\nWe may use artificial intelligence tools as part of our recruitment process to assist in the initial screening of resumes. All hiring decisions, including candidate evaluation, selection, and disposition, are made by human recruiters.\nAbout Us\nCanadian Tire Corporation, Limited (“CTC”) is one of Canada’s most admired and trusted companies. With more than 90 Owned Brands, over 1,600 retail locations, financial services, exemplary e-commerce capabilities, and exciting market-leading merchandising strategies. We dream big and work as one to innovate with purpose for our customers at every level of our business, investing in new technologies and products, and doubling down on top talent to drive the company forward. We offer competitive salaries and wages to CTC employees, as well as store discounts, supported learning through our Triangle Learning Academy, Canadian Tire Profit Sharing, and retirement and savings programs for eligible employees. As part of our enhanced flex benefits program, we offer mental health benefits in the amount of $5,000 per year for benefits-eligible employees and their families, including total well-being, and mental health tools and resources for all employees. Join us in helping to make life in Canada better through living and working our Core Values: we are innovators and entrepreneurs at our core, outcomes drive us, inclusion is a must, we are stronger together and we take personal responsibility. It is an especially exciting...","description_format":"text","description_chars":13092,"description_truncated":true,"requirements":{"experience_years_min":10,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Equity","Growth opportunities"],"hiring_locations":[{"name":"Canada","iso":"CA","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Consumer Goods","Wood Products & Furniture","Automotive Components"],"lifecycle":[{"event":"open","at":"2026-09-23T16:31:51Z"}],"liveness":{"score":60,"band":"ok","label":"Likely open","p_open":1,"p_active":0.602,"p_room":1,"age_days":0,"expected_fill_days":14,"reasons":["conf:0","stale_co","win:early","comp:brand"],"computed_at":"2026-09-24T05:45:00Z"},"pay":{"stated_usd_annual":85000,"is_top_pay":false},"html_url":"https://alion.io/job/partsource-sr-specialist-mainframe-compute-platform-governance","json_url":"https://alion.io/job/partsource-sr-specialist-mainframe-compute-platform-governance.json","meta":{"generated_at":"2026-09-24T18:57:14Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers"}}