HUB24 Limited Logo

HUB24 Limited

Site Reliability Engineer

Posted 24 Days Ago
Be an Early Applicant
In-Office
Sydney, New South Wales, AUS
Senior level
In-Office
Sydney, New South Wales, AUS
Senior level
Ensure reliability, performance and scalability of platform services by designing and operating highly available systems, embedding SRE practices (observability, SLOs, error budgets), monitoring and resolving incidents, performing RCA and post-incident reviews, partnering with engineering to build resilient applications, automating deployments and participating in on-call rotation.
The summary above was generated by AI

About HUB24


At HUB24, we’re rethinking the way wealth management works, combining platform, technology and data to create better outcomes for financial professionals and their clients.


Our purpose is simple:  Empower better financial futures, together.


What sets us apart is how we work. We back bold thinking, move with pace, and turn ideas into action. You’ll have the opportunity to make a real impact across your team, the business, and for the clients we support every day.


HUB24 Limited is an ASX-listed company (ASX: HUB) and part of the ASX100. We have over 1,100 employees across Australia, with offices in Sydney, Melbourne, Brisbane, Perth and the Gold Coast. 


Why you’ll enjoy working here


We create an environment where you can do your best work and see the impact of it.

  • Work with smart, collaborative people who get things done.
  • Your ideas won’t sit in a backlog, they’ll be heard, tested and actioned.
  • Grow your career your way, with support to learn, stretch and explore new opportunities.

We also offer benefits to support you inside and outside of work:

  • Genuinely flexible and hybrid ways of working.
  • Employee Share Scheme.
  • Additional leave and wellbeing support.
  • Enhanced parental leave and support through different life stages.
  • Everyday benefits, including discounts and financial wellbeing support.

Why this is an exciting opportunity


HUB24 is expanding its Site Reliability Engineering function and investing in Dynatrace as our core monitoring and observability platform. This is an opportunity to be part of that team which will play a critical role in ensuring the reliability, performance and scalability across our technology platforms.
Working closely with engineering, infrastructure and operations teams, you will embed SRE best practices, proactively manage system health and help design resilient, high-availability services that support our customers and business growth.

What you’ll be doing


  • Demonstrated hands-on experience in Site Reliability Engineering (SRE), DevOps, or Infrastructure Engineering, supporting reliable, high-performing, and scalable production environments.
  • Provide technical support for complex production incidents, ensuring timely resolution and minimising customer impact.
  • Lead incident triage, troubleshooting, root cause analysis, and Post Incident Reviews (PIRs), driving permanent fixes and continuous service improvements.
  • Monitor system performance, availability, and security across HUB platforms, proactively identifying and escalating risks.
  • Apply SRE best practices, including SLIs, SLOs, error budgets, and reliability engineering principles, to improve service resilience and operational excellence.
  • Support and troubleshoot AWS and GCP cloud environments, ensuring stability, performance, and operational efficiency.
  • Administer and maintain Linux and Windows servers, including performance tuning, configuration management, and operational support.
  • Support containerised environments using Docker and Kubernetes, helping maintain scalable and resilient workloads.
  • Manage server patching and vulnerability remediation activities in line with compliance, security, and risk management requirements.
  • Contribute to automation initiatives, developing operational scripts and runbooks to reduce manual effort and improve consistency.
  • Participate in an on-call roster, providing support for critical systems, platforms, and services as required.

What you bring


You don’t need to tick every box, but experience in the below will set you up for success.


Experience & Skills


  • 3–5 years of hands-on experience in Site Reliability Engineering, Observability Engineering, DevOps, or Infrastructure Engineering, including support for production systems and incident response.
  • Practical experience in incident management, including triage, troubleshooting, root cause analysis, and participation in on-call support for high-availability environments.
  • Working knowledge of cloud platforms such as AWS and/or GCP, including infrastructure troubleshooting, operational support, and optimisation activities.
  • Experience using observability tools, preferably Dynatrace, for monitoring, alerting, dashboarding, and performance analysis.
  • Solid system administration skills across Linux and Windows environments, including troubleshooting and basic performance tuning.
  • Exposure to containerisation and orchestration technologies such as Docker and Kubernetes is preferred.
  • Experience supporting server patching and vulnerability remediation, including prioritisation based on risk and compliance requirements.
  • Experience with automation and Infrastructure as Code tools such as Terraform, Ansible, or similar is desirable.
  • Strong communication skills, with the ability to document findings, explain technical issues clearly, and collaborate effectively across teams.

Our process


We aim to keep the process simple and respectful of your time:

  • You’ll receive an acknowledgement after applying.
  • Our Talent team will review your application and keep you updated.
  • If shortlisted, we’ll connect to learn more about you.
  • Interviews may be virtual or in person.
  • You’ll receive an outcome and feedback.

If you need any adjustments, please let us know - we’re here to support you.



Our commitment

We’re committed to building an inclusive environment where everyone feels valued and supported to do their best work. We welcome applications from people of all backgrounds, identities and experiences.


Agencies, we work with a panel of preferred suppliers and are not accepting any unsolicited CVs.

HQ

HUB24 Limited Sydney, New South Wales, AUS Office

Level 2, 7 Macquarie Place, Sydney, NSW, Australia, 2000

Similar Jobs

14 Days Ago
In-Office
Sydney, New South Wales, AUS
Senior level
Senior level
Software
Lead transformation of APAC LiveOps into a modern SRE and CloudOps capability. Define operating model, improve reliability, observability, incident response, automation and infrastructure as code. Partner with Product, Platform, Security and regional teams to reduce downtime, increase resilience and embed SRE practices across the region while building and coaching high-performing teams.
Top Skills: AlertingCi/CdCloudCloudopsContainersDatadogGrafanaInfrastructure As CodeKubernetesMonitoringNew RelicObservabilitySaaSSre
10 Days Ago
Remote or Hybrid
Sydney, New South Wales, AUS
Senior level
Senior level
Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Big Data Analytics • Automation
Join the SRE team to automate operational tasks, optimize capacity, manage product releases, implement monitoring/alerting, resolve incidents, and maintain scalable cloud infrastructure. Participate in on-call rotations to ensure production stability and continuous improvement.
Top Skills: AWSAzureCGCPGoJavaKubernetesPythonShell
23 Days Ago
In-Office
Sydney, New South Wales, AUS
Senior level
Senior level
Fintech • Financial Services
Lead production support and SRE for Citi's digital asset platforms, owning incident management, observability, troubleshooting, and automation. Partner with product and engineering teams to diagnose root causes, implement scalable fixes, manage risks, and mentor junior engineers to improve platform reliability and performance across blockchain-connected systems.
Top Skills: BlockchainDigital WalletsDistributed Ledger TechnologiesElkGcoGrafanaJavaLinuxOtelPrometheusPythonSplunkSQL

What you need to know about the Sydney Tech Scene

From opera to comedy shows, the Sydney Opera House hosts more than 1,600 performances a year, yet its entertainment sector isn't the only one taking center stage. The city's tech sector has earned a reputation as one of the fastest-growing in the region. More specifically, its IT sector stands out as the country's third-largest, growing at twice the rate of overall employment in the past decade as businesses continue to digitize their operations to stay competitive.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account