RoleHunter

Job description

  • 🗓️ Full-time | Permanent | September/October 2026 start
  • 📍 Sydney based (CBD)
  • 🏡 In office 2 days/week, WFH 3 days/week

Hands-on mentoring from our Australian SRE team | cutting edge tooling & technology | five wellness days per year | flexible benefits package of $1,000 per year | fantastic team culture | work with purpose. 

👋 Meet AlayaCare! We’re a fast-growing SaaS scaleup on a mission to transform aged and disability care across Australia, Canada, the US and beyond. Our platform helps care providers deliver exceptional service in homes, communities, and residential settings.

We're big on tech with purpose, passionate about improving care outcomes, and just as passionate about building a workplace where our people can grow and do their best work.

The Role:

We’re on the lookout for a Graduate Site Reliability Engineer who’s curious, keen to learn, and excited about how modern cloud platforms are built and run. No previous production experience with cloud infrastructure, Kubernetes or SRE practices is required, just a solid computer science foundation and some hands-on exposure (through coursework, a personal project, or an internship) to build on. Reporting to the SRE Engineering Manager, you’ll join our global SRE team of 9 (including 3 existing SREs in Australia) working across AWS and Azure cloud environments, and gradually take on ownership of shared platform services across products and regions.

Your days will involve:

Development, Automation, and Tooling

  • Develop scripts, configuration and operational tools, learning the ropes as you go
  • Contribute to reviewed infrastructure-as-code changes
  • Get comfortable using AI coding tools to help write and troubleshoot code

Reliability and Operations

  • Monitor, Design and improve multitenant architecture that serves millions of requests daily
  • Debug, investigate, fix and implement RCA for incidents
  • Learn how to architect reliable infrastructure on cloud
  • Join the help desk and on call rotation once you’re up to speed. Note that on call is compensated & shared across the global SRE team

Learning and Collaboration

  • Work closely with the SRE, Product and Engineering teams on operational requirements.
  • Improve monitoring, alerting and technical documentation as you build confidence
  • Learn how we choose which architecture to use and why

Growing Into the Role

  • Take on more ownership of our shared platform services as you build confidence, including databases, messaging, logging, search, and tenant provisioning
  • Get involved in projects that go beyond firefighting. Our team has shifted toward proactive infrastructure improvement, not just keeping the lights on

You won’t be expected to know or own everything on day one. You’ll start with guided tasks, pairing and structured onboarding before gradually taking on more responsibility

You’ll thrive in this role if you:

  • Have completed a degree in Computer Science within the last 2 years
  • Can explain basic networking, operating systems and debugging concepts, and can troubleshoot straightforward issues with some guidance
  • Can write and modify simple scripts (Python, Go or Bash) to automate a task
  • Are comfortable navigating Linux and the command line, running common commands and reading logs
  • Have used or deployed to a public cloud platform (AWS or Azure) through coursework, a personal project or an internship
  • Are comfortable using AI coding tools like Cursor or Claude Code to help generate code and troubleshoot
  • Are curious and calm under pressure, and genuinely interested in how things break and get fixed
  • Communicate well and aren’t afraid to ask questions or ask for help

You don’t need professional experience in all of these areas. We’re looking for foundational knowledge, some hands-on exposure and the ability to learn quickly.

We believe great work should be rewarded. Here’s how we show our appreciation:

  • 🏡 Hybrid work (2 days in office, 3 WFH per week)
  • 🧘 5 Wellness days per year
  • 💳 $1,000/year flexible benefits package
  • 🧡 2 days company-paid volunteer leave to support causes you care about
  • 🍕 Team lunches, events & wellness activities
  • 🤝 A genuinely open, inclusive, and collaborative culture
  • 💡 A chance to do purposeful work in the fast-paced tech sector, whilst making real impact in the care space.

How to Apply:

Send us your resume, plus a short cover letter which tells us:

  • What technical concept or tool have you taught yourself?
  • Describe a project where something broke and how you investigated it.
  • Why are you interested in SRE, reliability and automation?

Please take the time to answer the questions thoughtfully. We’re looking for specific examples that show how you learn, investigate problems and approach unfamiliar technical challenges.

In your resume we want to see:

  • Projects: university coursework, capstone or personal projects, hackathons. Show us what you built, not just what you studied.
  • Motivation: some sign of genuine interest in reliability, infrastructure or automation, not just software development.
  • Any relevant experience: internships/placements, or involvement in a CS/tech student society, if you’ve got it.

Belonging matters.

We’re committed to building an organisation that reflects the communities we serve. Diversity, equity, inclusion, and accessibility aren’t just buzzwords here, they’re woven into everything we do.

Need adjustments to participate in the recruitment process? We’ve got you. Just reach out to our HR team: people-anz@alayacare.com. We do not accept unsolicited CVs from Recruitment Agencies.