- As a Senior Software Engineering Manager for Managed Gateways SREs, you will lead the charge in defining the reliability, scalability, and operational excellence of Kong’s critical managed services
- Your team’s unwavering commitment to 99.99%+ uptime and seamless performance directly empowers developers globally, fueling the Agentic Era by providing the robust infrastructure that underpins modern API-driven applications
- You’ll build this team in Toronto from the ground up, and in the early stages, you’ll stay hands-on - directly in the work, not just directing it from a distance
- Build Kong’s Managed Gateways SRE team in Toronto from the ground up - hiring, setting the bar, and staying hands-on in enterprise implementations as the team ramps
- Act as a direct contributor to critical implementations and reliability work in the early stages, with outcomes that materially move the needle on Managed Gateways’ growth and business performance
- Lead, mentor, and grow a high-performing team of Site Reliability Engineers dedicated to Kong’s Managed Gateway offerings, in direct support of our critical enterprise customer base across the Americas and Europe
- Architect and implement robust, scalable, and fault-tolerant cloud-native systems using technologies like Kubernetes, Golang, and major cloud providers
- Own the end-to-end operational lifecycle, from proactive monitoring and alerting to incident response and blameless post-mortems, ensuring continuous service improvement
- Drive developer delight and operational efficiency through automation, self-service tooling,
and streamlined workflows for deploying and managing API gateways, while proactively preventing technical debt and reducing operational toil
- Define, track, and report on key SLOs and SLIs, and advocate for architectural best practices that keep Managed Gateways performant and resilient as it scales
- Collaborate cross-functionally with Product, engineering, and support teams to influence roadmap decisions and ensure operational readiness for new features
- Ownership: You take full responsibility for the reliability and performance of systems, driving initiatives from conception to completion
- Urgency: You thrive in dynamic environments, prioritizing critical issues and delivering impactful solutions with speed and precision
- Collaboration: You build strong relationships, foster open communication, and work effectively across global teams to achieve shared goals
Benefits
- Flexible time off: Take time to take care of yourself and the things that matter most
- Stock options: We want you to share in our success. That’s why stock options are offered to most Kongers
- U-First Fridays: Get 4 hours a month for continuous learning with a book, podcast, or course of your choice
- Virtual events:
Stay connected with Donut chats, trivia, fitness challenges, guided meditations, and more
- Home office stipend: Build a home office environment tailored to support your productivity
- Dedicated unplug days: Silence those notifications. Enjoy some well-deserved long weekend where the entire team unplugs
- Strong understanding of observability principles and experience with tools such as Prometheus, Grafana, OpenTelemetry, or similar
- Familiarity with API gateway technologies, service mesh, or network proxies is a significant plus
- Proven experience leading and managing Site Reliability Engineering or DevOps teams in a fast-paced, high-growth environment
- Deep expertise in designing, deploying, and operating highly available distributed systems on cloud platforms (AWS, Azure, or GCP)
- Proficiency in Golang or similar up-to-date programming languages for infrastructure automation and service development
- Demonstrated ability to manage critical incidents, perform root cause analysis, and implement effective preventative measures
- Extensive hands-on experience with Kubernetes (k8s) and container orchestration in production environments
- Experience with open-source contributions or active participation in SRE/cloud-native communities
- Certifications in cloud platforms (e.g., AWS Certified DevOps Engineer) or Kubernetes (e.g., CKA, CKAD)
- Background in companies focused on developer tools, infrastructure software, or API management
#J-18808-Ljbffr
📌 Senior Software Engineering Manager (Managed Gateways SREs) (Toronto)
🏢 Kong
📍 Toronto