09 Aug
|
Atlantis IT Group
|
Montreal
09 Aug
Atlantis IT Group
Montreal
Overview
Site Reliability Engineer (Linux / Cloud Infrastructure) role with hands-on experience across Linux, distributed systems, scripting, databases, monitoring, containers, cloud SaaS integrations, messaging, load balancers, security, and incident management.
Responsibilities
- Provide hands-on administration of Linux 7.x and related infrastructure.
- Work with Service Oriented Architecture, distributed systems, and scripting (Python, shell).
- Manage relational databases (e.g., Sybase, DB2, SQL, Postgres) and application integration, configuration, and troubleshooting.
- Operate observability and monitoring tools: Open Telemetry, Prometheus, Grafana, Splunk, Ansible.
- Manage web servers (Apache, Nginx) and application servers (Tomcat, JBoss) for integration and troubleshooting.
- Work with Docker containers, Kubernetes, and SaaS platform integration.
- Understand messaging systems (e.g., Kafka) and their role in the architecture.
- Design and implement load balancing, web proxies, and storage platforms (NAS/SAN) from an implementation perspective.
- Apply basic security policies for secure hosting solutions, including Kerberos and encryption methods (SSL/TLS).
- Experience in managing large web-based, multi-tier (n-tier) applications in secure cloud environments.
- Apply SRE principles with appropriate tooling approach; robust Linux/Unix admin, storage, networking, and web technologies knowledge.
- Troubleshoot application issues and manage incidents effectively.
- Exhibit excellent verbal and written communication skills.
Qualifications
- Hands-on experience with Linux 7.x operating system (5+ years) at an advanced level.
- Hands-on experience with SOA, distributed systems, and scripting (Python, shell).
- Experience with relational databases (Sybase, DB2, SQL, Postgres).
- Exposure to tools: Open Telemetry, Prometheus, Grafana, Splunk, Ansible.
- Hands-on experience with web servers (Apache, Nginx) and application servers (Tomcat, JBoss).
- Experience with Docker, Kubernetes, and SaaS platform integration.
- Experience with Kafka and messaging technologies.
- Understanding of load balancers, web proxies, and NAS/SAN storage from an implementation perspective.
- Familiar with security policies for secure hosting, Kerberos, SSL/TLS.
- Experience managing large web-based n-tier applications in secure cloud environments.
- Strong knowledge of SRE principles and tooling.
- Strong infrastructure knowledge in Linux/Unix administration, storage, networking, and web technologies.
- Excellent troubleshooting and incident management capabilities.
Senioriry level
Mid-Senior level
Employment type
Contract
Job function
Information Technology
Industries
IT Services and IT Consulting
#J-18808-Ljbffr
📌 Site Reliability Engineer (Linux / Cloud Infrastructure) (Montreal)
🏢 Atlantis IT Group
📍 Montreal