Senior Site Reliability Developer (Toronto)

Senior Site Reliability Developer (Toronto)

17 Aug
|
United States Digital Space
|
Toronto

17 Aug

United States Digital Space

Toronto

Position OverviewWe are seeking a highly motivated and experienced Senior Site Reliability Developer (SRE) to manage critical cloud infrastructure and site reliability operations for the the company Platform Services and Emerging Technologies organization. The team delivers high-value, exabyte-scale and cloud data platform components powering desktop, mobile, and web products. This enables our product teams to build cohesive in-product data experiences, our partners to integrate and expand our data, and our end-users to work with their data across all the company products. This pivotal role focuses on ensuring the highest reliability, availability, and performance of our AWS-hosted cloud infrastructure. Reporting to the Engineering Manager, you will be leading design and development of resilient and scalable architecture and innovative solutions for the platform. You will independently manage and deliver end-to-end solutions while engaging with key stakeholders and partners.ResponsibilitiesLead architecture, solution design, development and maintenance of cloud infrastructure for microservices architectureIndependently manage requirement analysis, solution design, implementation, and release planningEnsure high adherence to trust and security compliance, guidelines and standardsStreamline CI/CD processes, improve system reliability, and ensure infrastructure scalability and securityAutomate infrastructure deployment, scaling, and management using contemporary DevOps tools and practicesImplement and maintain configuration management and infrastructure as code (IaC) using TerraformLead Disaster Recovery (DR) strategies, failover exercises, gamedays, and period maintenance activitiesContribute to critical vulnerability (CVEs) remediation effortsPromote and document security and best practices across all pillars of DevOps/SRE throughout system designProvide real-time operational support and collaborate across functions to resolve system, infrastructure,



and CI/CD issuesParticipate in on-call rotations, providing critical 24x7 support for production systemsMinimum QualificationsBachelor's degree or higher in Computer Science, Engineering, or a related field5+ years of progressive experience in Site Reliability Engineering, DevOps, or a similar fieldProficiency with managing AWS resources and understanding of networking and security protocolsExpertise in infrastructure as code (IaC) and cloud automation tools such as Terraform, Serverless, and CloudFormationExpertise in defining and building CI/CD processes with tools like Jenkins, GitHub, and ArtifactoryExperience with container-based technologies like Docker, Kubernetes and AWS ECSExperience with monitoring and logging tools such as Dynatrace, Grafana, DataDog, ELK Stack, and CloudWatchTechnology Stack: Java/SpringBoot, AWS (ECS Fargate, Elastic Cache, Lambda, Kinesis, DynamoDB, VPC, IAM policies, API Gateway, NLB/ALB, Route 53, CloudWatch, Kibana, Open Search), Kafka, Flink, Jenkins, GitHub, Jira, Google Apigee, ServiceNow, and SplunkPreferred QualificationsKnowledge in applying AI and ML solutions for engineering processes and/or DevOps automationKnowledge of standardized observability frameworks such as OpenTelemetryRelevant certifications (e.G., AWS Certified DevOps Engineer, AWS Site Reliability Engineer)Broad knowledge of AWS, Redis, server programming, databases, and cloud architecturesBroad knowledge with data streaming pipelines like Kinesis, Firehose, and KafkaKnowledge on core Java and SpringBoot concepts in JVM optimizationKnowledge on build tools, e.G.



GradleStrong interpersonal and communication skills to effectively collaborate in an Agile/Scrum-oriented environmentSelf-directed team player and independent contributor, demonstrating accountability and end-to-end ownershipExperience in Linux Systems Administration, scripting, and troubleshooting in a production environmentProficiency in programming languages such as UNIX, Python, Go, Bash, Groovy, and Node.JsAbout the companyWelcome to the company! Amazing things are created every day with our software – from the greenest buildings and cleanest cars to the smartest factories and biggest hit movies. We help innovators turn their ideas into reality, transforming not only how things are made, but what can be made.We take great pride in our culture here at the company – it's at the core of everything we do. Our culture guides the way we work and treat each other, informs how we connect with customers and partners, and defines how we show up in the world.When you're an Autodesker, you can do meaningful work that helps build a better world designed and made for all. Ready to shape the world and your future? Join us!Salary transparencySalary is one part of the company's competitive compensation package. For Canada based roles, we expect a starting base salary between $107,000 and $157,300. Offers are based on the candidate's experience and geographic location, and may exceed this range. In addition to base salaries, our compensation package may include annual cash bonuses, commissions for sales roles, stock grants, and a comprehensive benefits package. BelongingBelongingWe take pride in cultivating a culture of belonging where everyone can thrive. Learn more here: https://www.The company.Com/company/global-belongingIn-Person Onboarding and Identity VerificationThis role may require in-person onboarding and/or in-person ID verification. #J-18808-Ljbffr

📌 Senior Site Reliability Developer (Toronto)
🏢 United States Digital Space
📍 Toronto

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability developer (toronto) / toronto

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability developer (toronto) / toronto