Powered by pgvector · cosine kNN
Salient Group
Observability & Reliability Engineer | Melbourne 💰 $150k–$170k + Super + Bonus📍 Melbourne | Hybrid When you’re building a platform that sits behind high-stakes decisions, knowing what to work on and when is importan…
Your match
See how you fit
Scored against this job in seconds
Your account
Sign in to apply
Your profile and your match for this job appear right here.
sign in above to apply · via LinkedIn
About the role
Observability & Reliability Engineer | Melbourne
💰 $150k–$170k + Super + Bonus📍 Melbourne | Hybrid
When you’re building a platform that sits behind high-stakes decisions, knowing what to work on and when is important.
As the platform grows, so has the complexity underneath it - more services, more dependencies, more signals and, inevitably, more noise.
It’s easy to spend too much time working out which alerts actually matter.
That's the problem you're coming in to solve.
This isn’t just configuring another alert. It’s designing what should be monitored in the first place, understanding what happens as systems scale, dealing with alert fatigue, tracing problems across distributed services and learning which signals genuinely predict something going wrong.
If you've built or significantly evolved an observability system before, you'll probably have some scars from doing it.
That's exactly what they're looking for.
So what could a week actually look like?
Monday, you're looking at the existing environment and working out why engineers are getting so much noise - what can disappear, what needs changing and what's currently missing altogether.
Tuesday, you're tracing a request across a distributed AWS environment, finding a degradation that previously would have been difficult to spot until it became an incident.
Wednesday, you're working directly with software engineers on instrumentation, getting deeper visibility into what's happening inside the application rather than simply watching the infrastructure around it.
Thursday, you're looking at where static thresholds stop being useful and where anomaly detection or automation could identify and resolve problems earlier.
Friday, something you've changed means an engineer doesn't get pulled away from their work by another meaningless alert.
Over time, that compounds. More engineering time spent actually building the product.And that's what I think makes the role interesting from a career perspective.
You're not joining a huge reliability organisation and inheriting one small component of somebody else's system.
You get to shape the capability.
You'll take ownership of questions like:
What are the signals that actually tell us something is wrong?How do we identify degradation before it becomes an incident?How do we trace what's happening across distributed services?How do we stop one underlying problem generating a wall of noise?What should wake an engineer up, and what shouldn't?What needs to change as the platform and traffic scale?What can we automate away entirely?
You'll work across AWS, distributed and event-driven systems, OpenTelemetry, metrics, logging, tracing, anomaly detection and automation, with enough ownership to make decisions about how it all fits together.
And if you do it well, you'll be able to point to a production platform and say: “I built how we know this thing is healthy.”
You might currently be called an SRE, Reliability Engineer, Observability Engineer, Platform Engineer or DevOps Engineer.
I'm much more interested in what you've built than what you're called.
If you've tackled this problem before and want more ownership of it next time around, apply here or drop me Jon Holland a message on Linkedin.
sign in above to apply · via LinkedIn
IMPROVE - 3Rs concepts to improve the quality of biomedical science (CA21139)
🚀 WE'RE HIRING💼 SITE RELIABILITY ENGINEER (SRE)📍 AUSTRALIA🕒 FULL-TIME✨ ENSURE UPTIME. AUTOMATE OPERATIONS. BUILD SCALABLE SYSTEMS.We are looking for a talented and proactive Site Reliability Engineer (SRE) to join…
HUB24
About HUB24 At HUB24, we’re rethinking the way wealth management works, combining platform, technology and data to create better outcomes for financial professionals and their clients. Our purpose is simple: Empower b…
Tyro Payments
Why Tyro? At Tyro, we’re into business big time. Through our integrated payments, banking and lending solutions, we’re here to ensure nothing stands in the way of Australian business success. With over 21 years' exper…
Block
Block is one company built from many blocks, all united by the same purpose of economic empowerment. The blocks that form our foundational teams — People, Finance, Counsel, Hardware, Information Security, Platform Inf…
iterate
We're hiring a Site Reliability Engineer with a strong observability focus to help keep our clients stores, warehouses, and digital channels running when it matters most. You'll own our Prometheus and Grafana stack, t…
ASX
ASX: Powering Australia's financial markets Why join the ASX? When you join ASX, you’re joining a company with a strong purpose – to power a stronger economic future by enabling a fair and dynamic marketplace for all.…
Your job hunt, handled
Ask about any role and get a straight answer on your fit. Then stop searching: new matches land in your WhatsApp the moment they’re listed.
Free for jobseekers