You will help modernize traditional infrastructure into reliable, automated platform services that enable application teams to deliver faster and more securely. They are an infrastructure engineer who thinks in terms of platforms, automation, reliability and scalable operating models, while supporting day-to-day operations, incident response, and continuous improvement initiatives.
Responsibilities for the Infrastructure Operations Engineer Role Operate and evolve hybrid infrastructure services across Azure, OpenShift and supporting on-premises technologies. Own and improve core infrastructure platform services such as Remote Desktop Services (RDS next gen.), automation workflows and operational tooling. Develop automation to reduce manual work, improve repeatability and support infrastructure lifecycle activities such as provisioning, patching, configuration, monitoring and remediation.
Contribute to the design and operationalisation of self-service infrastructure capabilities for internal technical teams. Improve reliability through monitoring, alerting, incident response patterns, runbooks, capacity management and resilience practices. Strong hands-on experience with Microsoft infrastructure technologies such as Windows Server, Active Directory, certificates and related enterprise services.
Minimum of 5 years of experience in infrastructure operations, cloud operations, site reliability engineering or a similar technical role. Practical automation experience using PowerShell, Python, YAML, Ansible or equivalent automation frameworks.
Experience using Git-based workflows to manage infrastructure as code and automation. Understanding of monitoring, observability, incident response and operational reliability practices. Ability to work collaboratively with cross-functional technical teams and communicate effectively with both technical and non-technical stakeholders. Bilingualism in French and English required.
Experience with Terraform or other infrastructure-as-code technologies.
Experience with OpenShift Virtualization, VMware migration, virtual machine lifecycle management or hybrid workload modernisation.
Experience with Ansible Automation Platform and reusable automation catalogue development.
Experience with Prometheus, Grafana, Azure Monitor or similar observability technologies. Knowledge of backup, disaster recovery, resilience engineering, performance and capacity management. Familiarity with ServiceNow integration, operational workflows, ITIL practices or incident management processes.
Fully adaptable so you can choose what matters most Retirement: Defined benefit pension plan and group Registered Retirement Savings Plan (RRSP) Employee stock purchase plan and numerous corporate discounts Personal and family programs: Flexible schedules and "California Fridays" year-round Social and community activities throughout the year! At CAE, our mission is clear: to help make the world a safer place. For nearly 80 years, we’ve driven innovation in simulation, training, and mission readiness to support critical operations worldwide.
By leveraging advanced technologies, we empower our customers to operate smarter, faster, and more sustainably. As a global leader in Civil Aviation and Defense & Security across 40 countries, we advance together as One CAE, fueled by thousands of passionate individuals operating in a culture where everyone can thrive. We hold ourselves accountable for building an inclusive community at every level, and our comprehensive benefits support you professionally and personally.
Join us and be part of something bigger—where your impact reaches beyond CAE and helps shape the world of tomorrow. CAE is committed to providing equal opportunities to all applicants, regardless of race, nationality, color, religion, sex, gender identity or expression, sexual orientation, disability, neurodiversity, veteran status, age, or other characteristics protected by law.
📌 Infrastructure and Operations Engineering - Telecommute (Montreal)
🏢 CAE
📍 Montreal