Powered by pgvector · cosine kNN
Australian Payments Plus
About the role
Life @ AP+:We are one connected team in pursuit of one inspiring purpose - to unite people and technology to power better experiences. Each of us has a part to play in making that happen. You'll be encouraged to bring your big ideas forward and make a difference through your work. Taking steps forward in your career whilst still having room for fun, friendships, and flexibility in your daily life.We're driven by our core values: lead with heart, learn for tomorrow and live our legacy. A purpose like ours takes the inspired impact of an incredible team. Ready to change the game? We're ready to help you do it.The Purpose:Reliability matters when you're supporting technology that Australians depend on every day.As a Cloud Site Reliability Engineer, you'll help keep AP+'s cloud and platform services reliable, available, secure and resilient. This is a hands-on engineering role where you'll combine cloud infrastructure, automation, observability and SRE practices to improve the performance and reliability of critical technology services.Working primarily across AWS and Kubernetes environments, you'll engineer for resilience, respond to complex incidents, reduce operational toil through automation and continually improve how our platforms perform at scale. Key Responsibilities the Role Owns:Operate, monitor and continually improve cloud and platform infrastructure, engineering for high availability, resilience, performance and security. Act as a first responder to cloud and platform incidents, troubleshooting complex reliability, capacity and performance issues while improving MTTD and MTTR through lessons learned. Engineer and optimise AWS and Kubernetes environments, including EKS, EC2, EFS, VPC and IAM, supporting scalable and resilient workloads. Build and improve CI/CD pipelines and automate operational tasks to reduce manual effort, minimise risk and enable safe, frequent and reliable delivery. Strengthen observability across infrastructure, applications and cloud services, using meaningful metrics to identify issues and drive improvements in availability and performance. Build resilience into our platforms through disaster recovery, controlled change, configuration management and close collaboration with Engineering, Security and Service Management teams. You'll likely be a strong fit for this role if:You bring hands-on experience in Site Reliability, Cloud or Platform Engineering, supporting complex and highly available cloud environments. You have strong AWS expertise, ideally across EKS, EC2, EFS, VPC and IAM, underpinned by a solid understanding of cloud networking. You're experienced with Kubernetes and containerised workloads, including cluster management, scaling, upgrades and optimisation. Infrastructure-as-code and automation are central to how you work, with experience using tools such as Terraform, CloudFormation, CDK or Ansible, along with Python, TypeScript or similar scripting languages. You've designed and operated CI/CD pipelines using platforms such as GitHub, GitLab or Bitbucket and understand the importance of strong observability and operational monitoring. You bring an engineering mindset focused on reliability and resilience, with experience across cloud security, incident response, disaster recovery and controlled change, ideally within a regulated or high-availability environment.What happens next:At AP+, we believe in the power of passion, pride and purpose. Our team is driven by a shared mission to make a difference in the world of payments, and we're proud to work together towards this common goal.If you're an SRE or Cloud Engineer who gets excited about automation, observability and engineering platforms that need to be there when it matters, we'd love to hear from you.If you're ready to be a game changer, please submit your application. The Talent Acquisition team will endeavour to review your application and notify you of the outcome within the next two weeks.We want to remove all barriers to inclusion, so if you need advice or support with your application, we're here to help. Please reach out to recruitment@auspayplus.com.au. We also encourage you to let us know your pronouns at any point during the recruitment process.AP+ are not partnering with Recruitment agencies for this role
Your match
See how you fit
Scored against this job in seconds
Your account
Sign in to apply
Your profile and your match for this job appear right here.
By continuing you agree to our Terms and Privacy Policy.
Next step
Apply for this role
via apply.workable.com — opens their site
Applications via apply.workable.com
Your job hunt, handled
Ask about any role and get a straight answer on your fit. Then stop searching: new matches land in your WhatsApp the moment they’re listed.
Free for jobseekers
Hatch
This is a Cloud Site Reliability Engineer role with Australian Payments Plus based in Sydney, NSW, AU == Australian Payments Plus == Role Seniority - mid level, senior More about the Cloud Site Reliability Engineer ro…
Tyro Payments
Why Tyro? At Tyro, we’re into business big time. Through our integrated payments, banking and lending solutions, we’re here to ensure nothing stands in the way of Australian business success. With over 21 years' exper…
IMPROVE - 3Rs concepts to improve the quality of biomedical science (CA21139)
🚀 WE'RE HIRING💼 SITE RELIABILITY ENGINEER (SRE)📍 AUSTRALIA🕒 FULL-TIME✨ ENSURE UPTIME. AUTOMATE OPERATIONS. BUILD SCALABLE SYSTEMS.We are looking for a talented and proactive Site Reliability Engineer (SRE) to join…
Future Secure AI
About the Company At Future Secure AI, we're building something genuinely new — and we're looking for people bold enough to build it with us. We work at the frontier of AI, tackling big, real-world problems for global…
Universal Music Australia Pty Limited
Job Summary We are UMG, the Universal Music Group. We are the world’s leading music company. In everything we do, we are committed to artistry, innovation and entrepreneurship. We own and operate a broad array of busi…
Cover Genius
About the Company Cover Genius is a Series E Insurtech that protects the global customers of the world’s largest digital companies including Booking Holdings, owner of Priceline, Kayak and Booking.com, Intuit, Hopper,…