Senior Cloud Site Reliability Engineer (Quebec City)

Senior Cloud Site Reliability Engineer (Quebec City)

07 Aug
|
Solace
|
Quebec City

07 Aug

Solace

Quebec City

Senior Cloud Site Reliability Engineer at Solace.
About the role As a Senior Cloud Site Reliability Engineer at Solace, you will play a crucial role in ensuring the reliability, availability, and performance of our cloud-based services. You will work closely with development and operations teams to implement best practices in site reliability engineering, focusing on automation and monitoring. Your expertise will help us maintain high service levels while continuously improving our infrastructure and processes. This position is ideal for someone who thrives in a fast-paced environment and is passionate about cloud technologies and operational excellence.
Key facts Location: Ottawa, Ontario
Engagement: Full-time
Salary: Competitive, commensurate with experience
Visa: Open to international candidates, sponsorship available
What you'll do Design, implement, and maintain scalable cloud infrastructure to support our applications and services.
Collaborate with cross-functional teams to define and implement service-level objectives (SLOs) and service-level indicators (SLIs).
Develop and maintain automated deployment pipelines using CI/CD tools to streamline application updates and infrastructure changes.
Monitor system performance and reliability, utilizing advanced monitoring tools to proactively identify and resolve issues.
Create and maintain comprehensive documentation for systems, processes, and procedures to enhance knowledge sharing within the team.
Lead incident response efforts, conducting post-mortem analyses to identify root causes and implement preventive measures.
Optimize cloud resource utilization and cost management through effective monitoring and analysis.




Mentor junior engineers and contribute to a culture of continuous learning and improvement within the team.
Participate in on-call rotations to provide support for production systems and ensure minimal downtime.
Stay updated on industry trends and emerging technologies to recommend improvements and innovations in our cloud architecture.
Collaborate with security teams to ensure compliance with best practices and regulatory requirements.
Engage in capacity planning and performance tuning to ensure our systems can handle growth and increased demand.
Requirements Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
A minimum of 5 years of experience in site reliability engineering, DevOps, or a related role, with a strong focus on cloud technologies.
Proficiency in cloud platforms such as AWS, Azure, or Google Cloud, with hands‑on experience in deploying and managing cloud services.
Strong understanding of containerization technologies, such as Docker and Kubernetes, and experience in orchestrating containerized applications.
Experience with infrastructure as code (IaC) tools like Terraform or CloudFormation to manage cloud resources.
Solid scripting skills in languages such as Python, Bash, or Go for automation and tooling purposes.




Familiarity with monitoring and logging tools such as Prometheus, Grafana, ELK Stack, or similar technologies.
Excellent problem‑solving skills and the ability to work effectively under pressure in a fast‑paced environment.
Strong communication skills, with the ability to collaborate effectively with technical and non-technical stakeholders.
Nice to have Experience with service mesh technologies such as Istio or Linkerd for managing microservices communications.
Knowledge of database management systems, including SQL and NoSQL databases, and their operational considerations.
Familiarity with Agile methodologies and experience working in an Agile development environment.
Contributions to open‑source projects or active participation in the tech community.
Relevant certifications in cloud architecture or site reliability engineering (e.g., AWS Certified Solutions Architect, Google Skilled Cloud Architect).
Skills & tools Cloud platforms: AWS, Azure, Google Cloud
Containerization: Docker, Kubernetes
Infrastructure as Code: Terraform, CloudFormation
Monitoring: Prometheus, Grafana, ELK Stack
Scripting: Python, Bash, Go
CI/CD: Jenkins, GitLab CI, CircleCI
Practical notes This position is based in Ottawa, Ontario, and offers a full-time engagement. The salary is competitive and will be determined based on your experience and qualifications. Solace is open to considering international candidates and can provide visa sponsorship for the right individual. The role requires participation in on-call rotations, so candidates should be prepared for occasional after-hours support.

#J-18808-Ljbffr

📌 Senior Cloud Site Reliability Engineer (Quebec City)
🏢 Solace
📍 Quebec City

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior cloud site reliability engineer (quebec city) / quebec city