Senior Site Reliability Engineer
Role Summary
Join Megaport as a Senior Site Reliability Engineer, enhancing production reliability and championing SRE practices in a collaborative tech environment.
About the Organisation
We’re not your typical tech company – and we don’t want to be. Megaport is the global leader in Network as a Service (NaaS), and has transformed the way businesses connect to the cloud, data centers, and each other. We’re publicly listed on the Australian Stock Exchange and partnered with the biggest names in tech like Amazon, Microsoft, Google, Oracle, IBM, and more. Headquartered in Brisbane with a crew of over 600 people spread across Asia-Pacific, Europe, and the Americas, our employees enjoy an environment that is collaborative, supportive, and (actually) fun.
Our Team Culture
We’re a team of problem solvers, pixel pushers, code slingers, and cloud fanatics. Culture is more than a poster on the wall here – collaboration beats hierarchy, curiosity fuels our growth, and everyone’s voice matters. We take our work seriously, but not ourselves. We work across time zones to execute on our global vision, trust each other to get things done, and never compromise our values for commercial gain. Most importantly, we place our customers at the center of everything we do.
We’re committed to increasing representation in the tech industry and welcome applicants from all backgrounds. Don’t meet every requirement? That’s okay. If you’re excited about this role, we encourage you to apply.
Minimum Requirements
- 5+ years administering Linux systems and related infrastructure in production environments
- A collaborative SRE mindset, with familiarity around SLIs/SLOs/SLAs, error budgets, blast radius, and blameless postmortems
- A focus on automation, reducing toil, and preventing problem recurrence
- A track record of writing runbooks that work for the broader team, not just yourself
- Strong Kubernetes and broader ecosystem fundamentals
- Cloud infrastructure experience; AWS strongly preferred and bare-metal is a bonus
- Strong tool development - Bash, plus either Python or Go preferred, or similar
- Infrastructure-as-code tooling experience - Terraform preferred
- CI/CD and version control, GitHub preferred
- Database experience - one of Postgres, Cassandra, or ClickHouse preferred
- Experience operating a production observability stack (metrics, logs, traces), with an eye for signal over noise
- Comfortable working on live production infrastructure, with strong troubleshooting instincts and ownership of incident response
- A history of continual professional development
- A self-directed style suited to an async, globally distributed team, and comfortable picking up adjacent work when the situation calls for it
Working Conditions
- Flexible working environment – a remote-first culture with coworking options available.
- Generous leave plans – including 4 weeks of paid annual leave, parental leave, birthday leave, and a purchased annual leave program.
- Health and wellness support – through a wellness allowance and employee wellbeing initiatives.
- Comprehensive learning support – generous study and training allowance plus 5 days of paid study leave
- Creative, modern workspaces – designed to inspire when you're not working remotely
- Motivated, inclusive team – work alongside industry experts and fresh talent
- Recognition programs – celebrate achievements with our Legend and Kudos awards
Eligibility Criteria
The job does not specifically mention eligibility requirements for applicants, including those from Africa.
⚠️ Disclaimer: PathwayAI Africa does not guarantee employment, scholarships, visas, or admission. Always verify all opportunities through official sources before submitting personal information.
Check if You Qualify
Upload your CV to compare it against this specific opportunity. You can still apply even if improvements are recommended.
