Powered by pgvector · cosine kNN
Tribus
About the role
Site Reliability Engineer (Trading Infrastructure) Multiple organisationsSponsorship AvailableSydney or Hong Kong | Onsite | Global Trading Environment
Join a high-performance engineering team responsible for the reliability, scalability and operational excellence of the infrastructure powering a global electronic trading platform.This is a hands-on Site Reliability Engineering role sitting close to the trading stack, where you'll help build resilient production systems, improve automation, and work alongside software engineers to support latency-sensitive applications operating across global financial markets.
Unlike traditional SRE environments focused primarily on SLI/SLO metrics, this team takes an event-driven approach to reliability engineering. You'll design intelligent monitoring and automated operational workflows that identify abnormal system behaviour, infrastructure anomalies and production events before they impact trading. The focus is on actionable signals, rapid diagnosis and engineering-led remediation rather than simply measuring service health.
What You'll Be Doing
Design, build and maintain highly reliable Linux-based production infrastructure.Develop Infrastructure as Code using Terraform and Ansible.Build observability platforms using Prometheus, Grafana and Splunk.Create event-driven monitoring, intelligent alerting and automated remediation workflows.Improve operational tooling and incident response across business-critical trading systems.Work with PostgreSQL and InfluxDB supporting production data platforms.Support workflow orchestration using Prefect.Administer virtualisation and storage platforms including CEPH, VMware, KVM and Proxmox.Partner closely with software engineers developing C++, Java, C# and Python applications.Improve CI/CD, deployment tooling and overall platform reliability through automation.
What We're Looking For
Strong Linux systems administration and troubleshooting experience.Python software development skills, specifically building tools for internal teamsExperience with Infrastructure as Code using Terraform, Ansible or similar tools.Hands-on experience with observability platforms such as Prometheus, Grafana or Splunk.Experience designing monitoring strategies that focus on operational events, anomaly detection and actionable alerts rather than purely SLI/SLO-driven metrics.Familiarity with PostgreSQL, InfluxDB or other production databases.Strong networking fundamentals including TCP/IP, routing and firewalls.Experience with Docker and modern development tooling.Exposure to virtualisation platforms such as VMware, KVM, Proxmox or CEPH.Experience supporting low-latency, distributed or mission-critical production environments is highly regarded.
Why This Role?
Work on infrastructure that directly supports real-time trading.Solve complex reliability challenges where milliseconds matter.Influence how monitoring, automation and operational engineering are built from the ground up.Collaborate with experienced infrastructure and software engineers in a highly technical environment.Work in a culture that values engineering ownership, continuous improvement and pragmatic problem solving.
If you're passionate about Linux, automation, observability and building resilient production platforms, we'd love to hear from you.
Your match
See how you fit
Scored against this job in seconds
Your account
Sign in to apply
Your profile and your match for this job appear right here.
By continuing you agree to our Terms and Privacy Policy.
Next step
Apply for this role
via LinkedIn — opens their site
Applications via LinkedIn
Your job hunt, handled
Ask about any role and get a straight answer on your fit. Then stop searching: new matches land in your WhatsApp the moment they’re listed.
Free for jobseekers
ASX
ASX: Powering Australia's financial markets Why join the ASX? When you join ASX, you’re joining a company with a strong purpose – to power a stronger economic future by enabling a fair and dynamic marketplace for all.…
IMPROVE - 3Rs concepts to improve the quality of biomedical science (CA21139)
🚀 WE'RE HIRING💼 SITE RELIABILITY ENGINEER (SRE)📍 AUSTRALIA🕒 FULL-TIME✨ ENSURE UPTIME. AUTOMATE OPERATIONS. BUILD SCALABLE SYSTEMS.We are looking for a talented and proactive Site Reliability Engineer (SRE) to join…
HUB24
About HUB24 At HUB24, we’re rethinking the way wealth management works, combining platform, technology and data to create better outcomes for financial professionals and their clients. Our purpose is simple: Empower b…
CMC Markets
If you also believe that everyone should be able to achieve their financial potential, then you’ll love contributing to CMC Markets’ company vision of providing the ultimate trading experience. Seize the opportunity t…
CMC Markets
If you also believe that everyone should be able to achieve their financial potential, then you’ll love contributing to CMC Markets’ company vision of providing the ultimate trading experience. Seize the opportunity t…
Arista Networks
Company Description Arista Networks is an industry leader in data-driven, client-to-cloud networking for large data center, campus and routing environments. What sets us apart is our relentless pursuit of innovation. …