<p style="font-family:"><b><strong style="font-size:18pt;white-space:pre-wrap;">About <a href="https://himalayas.app/companies/latitude-sh">Latitude.sh</a></strong></b></p><p style="font-family:"><a href="https://himalayas.app/companies/latitude-sh">Latitude.sh</a>'s global computing platform was launched in 2019, enabling businesses to programmatically deploy single-tenant Bare Metal instances in different parts of the world.<br><br>We are a team of passionate individuals about hardware, software, and network infrastructure looking to build the fastest, easiest-to-use, developer-centric single-tenant Cloud infrastructure. If you share this passion, join our growing team of talented people and help build the future of the Internet.</p><p style="font-family:"><b><strong style="font-size:18pt;white-space:pre-wrap;">Summary </strong></b></p><p style="font-family:">At <a href="https://himalayas.app/companies/latitude-sh">Latitude.sh</a>, the Reliability team is responsible for the health and resilience of the infrastructure that powers our global bare metal cloud. You’ll design and implement tools that automate operations, improve incident response, and enhance system observability—ensuring our platform is always ready for the workloads of our customers.</p><p style="font-family:">This might be a good opportunity if you’re passionate about reliability, automation, and creating cloud-like experiences for bare metal infrastructure.</p><p style="font-family:"><b><strong style="font-size:18pt;white-space:pre-wrap;">Key Responsabilities </strong></b></p><ul data-pattern="discCircleSquare" data-depth="1" style="font-family:"><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Continuously improve <a href="https://himalayas.app/companies/latitude-sh">Latitude.sh</a>’s platform reliability and performance</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Design, build, and maintain tools to automate operational tasks and incident response</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Implement and improve observability solutions, including monitoring, alerting, and tracing</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Collaborate with engineering and platform teams to design scalable and resilient systems</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Participate in on-call rotations and lead post-incident reviews with a focus on learning</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Develop and document processes and runbooks that ensure operational excellence</li><li style="color:rgb(0,0,0);margin:0pt 0px 7pt;font-size:12pt;line-height:1.38;">Contribute to SLOs/SLIs definition and reliability metrics adoption across teams</li></ul><p style="font-family:"><b><strong style="font-size:18pt;white-space:pre-wrap;">Skills and Qualifications</strong></b></p><ul data-pattern="discCircleSquare" data-depth="1" style="font-family:"><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Strong verbal and written English communication skills</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Advanced knowledge of Linux/Unix systems in production environments</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Experience with Kubernetes and container orchestration</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Proficiency with infrastructure automation tools (e.g., Terraform, Ansible)</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Experience with observability stacks (e.g., Prometheus, Grafana, Loki, ELK)</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Familiarity with scripting and programming languages such as Bash, Python, Go, or Ruby</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Working knowledge of Git and CI/CD pipelines</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:12pt;line-height:1.38;">Solid understanding of incident management and root cause analysis processes</li><li style="color:rgb(0,0,0);margin:0pt 0px 7pt;font-size:12pt;line-height:1.38;">Knowledge of cloud-native reliability and security best practices</li></ul><p style="font-family:"><b><strong style="color:rgb(0,0,0);font-size:18pt;white-space:pre-wrap;">What do we offer?</strong></b></p><ul data-pattern="discCircleSquare" data-depth="1" style="font-family:"><li style="color:rgb(0,0,0);margin:12pt 0px 0pt;font-size:11pt;line-height:1.38;">Contractor (PJ)</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:11pt;line-height:1.38;">Paid Time Off</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:11pt;line-height:1.38;">Competitive Compensation</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:11pt;line-height:1.38;">Wellhub (former Gympass)</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:11pt;line-height:1.38;">Annual Bonus based on company and team performance</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:11pt;line-height:1.38;">Flexible work hours</li><li style="color:rgb(0,0,0);margin:0pt 0px;font-size:11pt;line-height:1.38;">Opportunities for professional growth and development</li></ul><p style="font-family:"><b><strong style="color:rgb(0,0,0);font-size:18pt;white-space:pre-wrap;">Why <a href="https://himalayas.app/companies/latitude-sh">Latitude.sh</a>?</strong></b></p><p style="font-family:">We're a lean, agile team of passionate professionals who believe in the power of innovation and creative problem-solving. As part of our team, you won't be lost in the crowd – you'll be an essential contributor, making a real impact from day one.</p><p style="font-family:">Our values at <a href="https://himalayas.app/companies/latitude-sh">Latitude.sh</a> guide us in all our work and partnerships. We're proud to be an inclusive company, and we welcome all applicants for our open positions, regardless of their background, religion, sexual orientation, gender identity, age, nationality, or disability.
About Latitude.sh
Latitude.sh's global computing platform was launched in 2019, enabling businesses to programmatically deploy single-tenant Bare Metal instances in different parts of the world.
We are a team of passionate individuals about hardware, software, and network infrastructure looking to build the fastest, easiest-to-use, developer-centric single-tenant Cloud infrastructure. If you share this passion, join our growing team of talented people and help build the future of the Internet.
Summary
At Latitude.sh, the Reliability team is responsible for the health and resilience of the infrastructure that powers our global bare metal cloud. As a Senior Site Reliability Engineer (SRE), you’ll focus on building reliable, observable, and self-healing systems at scale.
SREs at Latitude.sh work at the intersection of software engineering and infrastructure. You’ll design and implement tools that automate operations, improve incident response, and enhance system observability—ensuring our platform is always ready for the workloads of our customers.
This might be a good opportunity if you’re passionate about reliability, automation, and creating cloud-like experiences for bare metal infrastructure.
Key Responsabilities
Skills and Qualifications
What do we offer?
Why Latitude.sh?
We're a lean, agile team of passionate professionals who believe in the power of innovation and creative problem-solving. As part of our team, you won't be lost in the crowd – you'll be an essential contributor, making a real impact from day one.
Our values at Latitude.sh guide us in all our work and partnerships. We're proud to be an inclusive company, and we welcome all applicants for our open positions, regardless of their background, religion, sexual orientation, gender identity, age, nationality, or disability. If these values speak to you, we'd love for you to become a part of our team.
Originally posted on Himalayas