Become a key player as a Cloud Reliability Engineer specializing in storage solutions. Work with a senior team to enhance cloud architecture remotely while ensuring operational excellence.
This role is pivotal in the development of distributed storage services, emphasizing operational durability and efficiency. Collaborate effectively to balance immediate engineering needs with long-term infrastructure goals. Bring your background in automation and cloud technologies to streamline processes and eliminate manual work.
Key Responsibilities:
• Develop reliable and fault-tolerant storage services
• Configure metrics for effective incident response
• Participate in continuous monitoring and on-call support
• Provide insights to optimize performance across systems
• Drive automation initiatives to minimize manual intervention
Requirements:
• 6+ years in software engineering for distributed systems
• Expertise in programming languages like Python, Go
• Experience managing scalable stateful storage systems
• Knowledge of containerization via Kubernetes
• Solid grasp of cloud infrastructure and Linux internals
Lead efforts in improving cloud storage reliability with a focus on enhanced performance and user-centric solutions.
#J-18808-Ljbffr