03 Oct
|
Artech
|
Vancouver
Title: Senior DevOps Engineer
Location: Vancouver, BC Canada
Duration: 12 Months
Hourly Rate: 59-65/hr
Introduction
Seeking a professional to provide a 1-year contingent engagement supporting our High Performance Computing (HPC) and Electronic Design Automation (EDA) infrastructure team.
This role will work directly within the IT Datacenter (ITDC) organization and is expected to operate independently at a senior level with minimal ramp-up time. The ideal candidate brings strong hands-on experience with Linux HPC environments, infrastructure automation, SLURM workload management, datacenter migration execution, and enterprise identity integration.
Required Skills & Qualifications
- 5+ years of experience in a DevOps, Platform Engineering, or Linux Systems Engineering role
- Hands-on HPC cluster administration experience, including SLURM or equivalent workload managers
- Demonstrated experience supporting EDA or scientific computing environments
- Strong Ansible automation skills with production-grade playbook and role development
- Experience with bare-metal provisioning tools (RackN, Cobbler, or equivalent)
- Proven ability to plan and execute datacenter or infrastructure migrations with minimal disruption
- Familiarity with enterprise Linux identity and authentication stacks (SSSD, LDAP, AD, NIS, Okta)
- Experience with NetApp or comparable enterprise storage platforms in HPC contexts
- Ability to author formal technical documentation (MOPs, runbooks, architecture diagrams)
- Robust written and verbal communication skills; capable of coordinating across multiple teams
- Prior work experience at client or in client's Industry
Preferred Skills & Qualifications
- Experience with SUSE Linux Enterprise Server (SLES) 12 and/or 15 in an enterprise environment
- Familiarity with RackN / Digital Rebar Provision for bare-metal OS deployment
- Hands-on experience with Kiwi NG or similar tools for custom OS image creation
- Knowledge of Client vSphere for HPC support VM provisioning
- Experience migrating configuration artifacts and binaries to Artifactory
- Background in semiconductor, storage, or high-tech manufacturing IT environments
Day-to-Day Responsibilities
- Support and administer SLURM-based HPC compute environments, including partition configuration and migration planning
- Plan and execute HPC/EDA compute and storage infrastructure migrations across datacenters
- Develop migration strategies and evaluate implementation options, risks, dependencies, and operational tradeoffs
- Author formal Method of Procedure (MOP) documents and runbooks for infrastructure changes and service cutovers
- Coordinate cross-functionally with EDA/SPE teams, storage teams, and IDAM to deliver coordinated platform changes
- Verify storage volumes, application access, and service continuity following migrations or infrastructure changes
- Define HPC storage service tiers and gather performance and capacity requirements for EDA workloads
For immediate consideration please click APPLY to begin the screening process with Alex
📌 Senior DevOps Engineer (Vancouver)
🏢 Artech
📍 Vancouver