Skip to content
Jobsearch.ing

Senior Site Reliability Engineer

MegaportSao Paulo, BR

EngineeringSeniorOther
Source-verified: read directly from this employer's own lever job board, not a repost.HybridPosted (24 days ago)Last verified (today)

At a glance

Location
Sao Paulo, BR
Workplace
Hybrid
Pay
Not published by the employer
Employment type
Other
Experience
Not stated
Education
No degree requirement stated
Job family
Engineering
Seniority
Senior
Posted by employer
13 August 2026
Last verified open
7 September 2026
Region and country
BR
Team
Latitude
Listed via
Lever

What the employer wrote

The Role At Latitude.sh, the Reliability team is responsible for the health and resilience of the infrastructure that powers our global bare metal cloud. As a Senior Site Reliability Engineer (SRE), you’ll focus on building reliable, observable, and self-healing systems at scale.

SREs at Latitude.sh work at the intersection of software engineering and infrastructure. You’ll design and implement tools that automate operations, improve incident response, and enhance system observability—ensuring our platform is always ready for the workloads of our customers.

This might be a good opportunity if you’re passionate about reliability, automation, and creating cloud-like experiences for bare metal infrastructure.

What You'll Be Doing • Continuously improve Latitude.sh’s platform reliability and performance

• Design, build, and maintain tools to automate operational tasks and incident response

• Implement and improve observability solutions, including monitoring, alerting, and tracing

• Collaborate with engineering and platform teams to design scalable and resilient systems

• Participate in on-call rotations and lead post-incident reviews with a focus on learning

• Develop and document processes and runbooks that ensure operational excellence

• Contribute to SLOs/SLIs definition and reliability metrics adoption across teams

What We're Looking For • Strong verbal and written English communication skills

• Advanced knowledge of Linux/Unix systems in production environments

• Experience with Kubernetes and container orchestration

• Proficiency with infrastructure automation tools (e.g., Terraform, Ansible)

• Experience with observability stacks (e.g., Prometheus, Grafana, Loki, ELK)

• Familiarity with scripting and programming languages such as Bash, Python, Go, or Ruby

• Working knowledge of Git and CI/CD pipelines

• Solid understanding of incident management and root cause analysis processes

• Knowledge of cloud-native reliability and security best practices

What We Offer • Contractor (PJ)

• Paid Time Off

• Competitive Compensation

• Wellhub (former Gympass)

• Annual Bonus based on company and team performance

• Flexible work hours

• Opportunities for professional growth and development

Where this record came from

Read from Megaport's own Lever job board on , and last confirmed still open on . The employer published it on 13 August 2026. Jobsearch.ing did not write, edit or rank this posting, and does not vet the employer. View the original posting.

More roles like this one