17 Aug
|
United States Digital Space
|
Toronto
17 Aug
United States Digital Space
Toronto
We are seeking a highly motivated and experienced Senior Site Reliability Developer (SRE) to manage critical cloud infrastructure and site reliability operations for the the company Platform Services and Emerging Technologies organization. The team delivers high-value, exabyte-scale and cloud data platform components powering desktop, mobile, and web products. This enables our product teams to build cohesive in-product data experiences, our partners to integrate and expand our data, and our end-users to work with their data across all the company products.
This pivotal role focuses on ensuring the highest reliability, availability, and performance of our AWS-hosted cloud infrastructure. Reporting to the Engineering Manager, you will be leading design and development of resilient and scalable architecture and innovative solutions for the platform. Independently manage requirement analysis, solution design, implementation, and release planning Automate infrastructure deployment, scaling, and management using modern DevOps tools and practices Implement and maintain configuration management and infrastructure as code (IaC) using Terraform Promote and document security and best practices across all pillars of DevOps/SRE throughout system design Bachelor’s degree or higher in Computer Science, Engineering, or a related field ~5+ years of progressive experience in Site Reliability Engineering, DevOps, or a similar field ~ Proficiency with managing AWS resources and understanding of networking and security protocols ~ Expertise in defining and building CI/CD processes with tools like Jenkins, GitHub, and Artifactory ~ Experience with container-based technologies like Docker,
Kubernetes and AWS ECS ~ Experience with monitoring and logging tools such as Dynatrace, Grafana, DataDog, ELK Stack, and CloudWatch ~ Technology Stack: Java/SpringBoot, AWS (ECS Fargate, Elastic Cache, Lambda, Kinesis, DynamoDB, VPC, IAM policies, API Gateway, NLB/ALB, Route 53, CloudWatch, Kibana, Open Search), Kafka, Flink, Jenkins, GitHub, Jira, Google Apigee, ServiceNow, and Splunk Knowledge in applying AI and ML solutions for engineering processes and/or DevOps automation AWS Certified DevOps Engineer, AWS Site Reliability Engineer) Broad knowledge of AWS, Redis, server programming, databases, and cloud architectures Broad knowledge with data streaming pipelines like Kinesis, Firehose, and Kafka Knowledge on core Java and SpringBoot concepts in JVM optimization Solid interpersonal and communication skills to effectively collaborate in an Agile/Scrum-oriented environment Experience in Linux Systems Administration, scripting, and troubleshooting in a production environment Proficiency in programming languages such as UNIX, Python, Go, Bash, Groovy, and Node.js Amazing things are created every day with our software – from the greenest buildings and cleanest cars to the smartest factories and biggest hit movies.
We help innovators turn their ideas into reality, transforming not only how things are made, but what can be made. Our culture guides the way we work and treat each other, informs how we connect with customers and partners, and defines how we show up in the world. When you’re an Autodesker, you can do meaningful work that helps build a better world designed and made for all.
In addition to base salaries, our compensation package may include annual cash bonuses, commissions for sales roles, stock grants, and a comprehensive benefits package.
📌 Developer - medical (Toronto)
🏢 United States Digital Space
📍 Toronto