06 Aug
|
Leadingtalent
|
Markham
06 Aug
Leadingtalent
Markham
Introduction A career in IBM Software means you’ll be part of a team that transforms our customer’s challenges into solutions. Seeking new possibilities and always staying curious, we are a team dedicated to creating the world’s leading AI-powered, cloud-native software solutions for our customers. Our renowned legacy creates endless global opportunities for our IBMers, so the door is always open for those who want to grow their career.
IBM’s product and technology landscape includes Research, Software, and Infrastructure. Entering this domain positions you at the heart of IBM, where growth and innovation thrive.
Your role and responsibilities As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In this role, you will lead the problem resolution process for our clients, from analysis and troubleshooting, to deploying the latest software updates & fixes.
Your primary responsibilities include
- 24×7 Observability: Be part of a worldwide team that monitors the health of production systems and services around the clock, ensuring continuous reliability and optimal customer experience.
- Cross-Functional Troubleshooting: Collaborate with engineering teams to provide initial assessments and possible workarounds for production issues. Troubleshoot and resolve production issues effectively.
- Deployment and Configuration: Leverage Continuous Delivery (CI/CD) tools to deploy services and configuration changes at enterprise scale.
- Maintenance and Support: Tasks related to applying security pa Required education High School Diploma/GED Required technical and professional expertise - System Monitoring and Troubleshooting: knowledge in monitoring/observability, issue response, and troubleshooting for optimal system performance.
- Automation:
knowledge in automation for production environment changes, streamlining processes for efficiency, and reducing toil.
- Linux: Knowledge of Linux operating systems.
- Operation and Support Experience: Understanding in handling day-to-day operations, alert management, incident support, migration tasks, and break-fix support.
- Scripting: knowledge or experience of Python, go or bash.
- Familiar with cloud providers like IBM Cloud, AWS, Azure or GCP. Preferred technical and professional experience - Kubernetes/OpenShift: knowledge or experience of Kubernetes/OpenShift environments.
- Automation/Scripting: knowledge or experience of Ansible, Python, Terraform, and CI/CD tools such as Jenkins, IBM Continuous Delivery, ArgoCD.
- Monitoring/Observability: knowledge or experience crafting alerts and dashboards using tools such as Instana, New Relic, Grafana/Prometheus - DBA: Interest or experience configuring and maintaining SQL, NoSQL, and data streaming technologies (e.g. PostgreSQL, CouchDB, Redis, Kafka, Spark, etc.).
Equal Chance IBM is proud to be an equal-opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender, gender identity or expression, sexual orientation, national origin, genetics, pregnancy, disability, neurodivergence, age, or other characteristics protected by the applicable law. IBM is also committed to compliance with all fair employment practices regarding citizenship and immigration status.
Other relevant job details Must have the ability to work in Canada without sponsorship. This role will involve working with technology that is covered by Export Regulations sanctions. If you are a Foreign National from any of the following US sanctioned countries (Cuba, Iran, North Korea, Syria, and the Crimea, Luhansk, Donetsk, Kherson, and Zaporizhia regions of Ukraine) on a work permit, you are not eligible for employment in this position.
📌 Site Reliability Engineering Professional 2026 Intership (16 Months) (Markham)
🏢 Leadingtalent
📍 Markham