Hiring.Camp

Production Reliability Engineer

Asx

·

Today

Location
Sydney Office, Australia
Workplace
Hybrid
Type
Full-time
Department
Engineering
Source
Workday

Description

ASX: Powering Australia's financial markets

Why join the ASX?

When you join ASX, you’re joining a company with a strong purpose – to power a stronger economic future by enabling a fair and dynamic marketplace for all.

In your new role, you’ll be part of a leading global securities exchange with a strong brand. We are known for being a trusted market operator and an exciting data hub. 

Want to know why we are a great place to work, click on the link to learn more.

www.asx.com.au/about/careers/a-great-place-to-work

We are more than a securities exchange!

The ASX team brings together talented people from a diverse range of disciplines. 

We run critical market infrastructure, with 1 in 3 people employed within technology.  Yet we have a unique complexity of roles across a range of disciplines such as operations, program delivery, financial products, investor engagement, risk and compliance.

We’re proud to foster a workplace where diversity is celebrated and inclusion is part of our everyday culture. Our employee-led networks champion LGBTIQ+ inclusion, promote gender equality, accessibility and wellbeing, inspire giving and volunteering, and celebrate cultural and religious events, creating a sense of belonging for all. As an AWEI Bronze employer and member of the Champions of Change Coalition for gender equality, we’re committed to a fair and inclusive workplace where everyone can thrive.

Key purpose of the role

The role is accountable for ensuring the reliability, availability, performance, observability, and operational resilience of critical business platforms through proactive monitoring, automation, incident management, capacity planning, and continuous service improvement, enabling secure and stable technology services that meet business and customer expectations.

Your responsibilities

  • Ensure the availability, reliability, scalability, and resilience of critical technology platforms and services.

  • Design, implement, and maintain comprehensive monitoring, logging, alerting, and observability solutions.

  • Develop meaningful dashboards, operational metrics, and health indicators to provide real-time visibility of platform performance.

  • Drive continuous improvement initiatives to reduce service disruptions and improve platform stability.

  • Proactively identify, assess, and mitigate reliability risks across production environments.

  • Conduct root cause analyses (RCA) and post-incident reviews, driving permanent resolutions and preventative actions.

  • Maintain operational runbooks, playbooks, and recovery procedures to support rapid incident resolution.

  • Support 24x7 production environments and participate in on-call rotations where required.

  • Develop and maintain infrastructure, deployment, and operational automation using Infrastructure as Code (IaC) principles.

  • Build automated recovery mechanisms to minimise manual intervention.

  • Champion engineering best practices across operational and development teams.

  • Conduct capacity planning, performance analysis, and workload forecasting to ensure services can meet current and future demand.

  • Identify and resolve performance bottlenecks across applications, infrastructure, databases, and integrations.

  • Partner with development teams to improve deployment reliability, release automation, and operational supportability.

  • Work closely with Engineering, Architecture, Security, Infrastructure, Product, and Operations teams to improve service outcomes.

Your experience and qualifications

Must have

  • 5+ years’ experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Engineering, Platform Engineering, or Production Support environments.

  • Strong scripting experience in Python, Java or PowerShell.

  • Strong knowledge of Linux, Windows, networking, and distributed systems.

  • Experience with CI/CD tooling and automated deployment pipelines.

  • Strong understanding of containerisation and orchestration technologies, including Docker and Kubernetes.

  • Experience with observability platforms such as Grafana, Splunk, ITRS Geneos, OpenTelemetry, AWS CloudWatch or similar.

  • Knowledge of database technologies, including Oracle, SQL Server, PostgreSQL or NoSQL platforms.

  • Familiarity with integration technologies, including APIs, real-time messaging and streaming architectures such as Kafka.

  • Non-functional test planning and execution experience.

Nice to have

  • Experience working in the Capital Markets industry, with a solid understanding of product and transaction lifecycles.

  • Passionate about solving problems, troubleshooting software issues and triaging environment issues.

  • Excellent verbal and written communication skills, with the ability to work with internal and external stakeholders.

  • Lateral thinker who brings forward ideas that will automate repeatable tasks to reduce work effort and ensure quality deliverables.

  • Learns quickly and enjoys the challenge of learning new systems.

  • Compassionate, empathetic and self-motivated.

  • Able to take ownership and not be afraid to acknowledge failure.

We make hiring decisions based on your skills, capabilities and experience, and how you’ll help us to live our values. We encourage you to apply even if you don’t meet all the criteria of this role.

If you need any adjustments during the application or interview process to help you present your best self, please let us know at [email protected].

At ASX Group, our diverse workforce is essential to build and maintain a fair and dynamic marketplace. We support flexible working and offer hybrid working options. Even if our roles are advertised as full-time, we encourage you to apply if you are interested in part-time or other flexible working arrangements.

We will arrange for successful candidates to have background checks, including reference and police checks, completed as part of the on-boarding process.

To be considered for this position, candidates must be legally authorised to work in Australia on a permanent basis without any restrictions.

Skills

PythonJavaAWSDockerKubernetesCI/CDLinuxSQLPostgreSQLOracleSQL ServerSplunkDevOpsCompliance

Similar Jobs

30

Production Reliability Engineer

Asx · Sydney Office, Australia · Hybrid

4 weeks ago

Production Reliability Engineer / SRE (Hong Kong)

Io Tech Solutions Limited · HongKong

1 week ago

Associate, Software Production Management & Reliability Engineer

Morgan Stanley · Tokyo,JP, JP

3 weeks ago

Associate, Software Production Management & Reliability Engineer

Ms · Otemachi Financial City, Japan

3 weeks ago

Associate, Software Production Management & Reliability Engineer

Ms · Otemachi Financial City, Japan

3 weeks ago

Senior Site Reliability Engineer, Production Engineer - ThousandEyes

Cisco · USA-SAN FRANCISCO, United States of America · Hybrid

3 weeks ago

Production Engineer, Site Reliability (Application Software)

Spacex · Hawthorne, CA +1

1 month ago

Site Reliability / Production Engineer

Storyteller · Remote

1 month ago

Site Reliability / Production Engineer

Storyteller · Remote

1 month ago

Site Reliability / Production Engineer

Storyteller · Remote

1 month ago

Site Reliability / Production Engineer

Storyteller · Remote

1 month ago

Site Reliability Engineer III- Production Management

JPMorgan Chase · NY, United States, US

1 month ago

Site Reliability Engineer III- Production Management

JP Morgan Chase · NY, United States, US

1 month ago

Lead Infrastructure Production Management & Reliability Engineer - Director

Morgan Stanley · HK

1 month ago

Lead Infrastructure Production Management & Reliability Engineer - Director

Ms · International Commerce Centre, Hong Kong

1 month ago

Lead Infrastructure Production Management & Reliability Engineer - Director

Ms · International Commerce Centre, Hong Kong

1 month ago

Production Reliability Engineer (Starship Electronics)

Spacex · Hawthorne, CA +1

1 month ago

Production Reliability Engineer - Calca Solutions

Newmarket · Westlake, LA, US

2 months ago

Production Reliability Engineer II

earlywarningservices · Scottsdale, United States of America · Hybrid

3 months ago

Electrical Engineering Student - Bitumen Production Reliability

CNRL Professional · Fort McMurray, AB, Canada

1 week ago

Production/Maintenance/Repair/Facilities Supervisor/Lead - Daventry, UK, Reliability, Maintenance and Engineering (RME)

Amazon · Onsite

1 week ago

Manager - Production Operations & Site Reliability Engineering

Alcon · Lake Forest - Main Campus, United States of America

2 weeks ago

Senior Software Production Management & Reliability Engineering

Morgan Stanley · Tokyo,JP, JP

3 weeks ago

Senior Software Production Management & Reliability Engineering

Ms · Otemachi Financial City, Japan

3 weeks ago

Senior Software Production Management & Reliability Engineering

Ms · Otemachi Financial City, Japan

3 weeks ago

Vice President, Lead Software Production Management & Reliability Engineering

Morgan Stanley · HK

1 month ago

Vice President, Lead Software Production Management & Reliability Engineering

Ms · International Commerce Centre, Hong Kong

1 month ago

Vice President, Lead Software Production Management & Reliability Engineering

Ms · International Commerce Centre, Hong Kong

1 month ago

Associate, Software Production Management & Reliability Engineering

Morgan Stanley · Tokyo,JP, JP

1 month ago

Associate, Software Production Management & Reliability Engineering

Morgan Stanley · HK

1 month ago