Reliability Engineer jobs
- OmiliaAustralia
- Champion a culture of reliability, performance, and continuous improvement across teams.
- Collaborate with development and cloud engineering teams to embed…
- Rio TintoPerth WA 6000
- Parental leave
- Health insurance
- Employee discount
- Relocation assistance
- Salary packaging
- Delivering reliability studies and improvement work across fixed assets, including bad actor analysis, repeat failure reviews and investigation of key…
- View all Rio Tinto jobs - Perth jobs - Reliability Engineer jobs in Perth WA
- Salary Search: Engineer Reliability - Fixed Assets salaries in Perth WA
- See popular questions & answers about Rio Tinto
- Hudson ManpowerSouth Australia
- Monitor equipment performance and analyze reliability data to minimize downtime and operational risks.
- The role involves monitoring equipment condition,…
- Northern Star Resources LimitedGoldfields WA
- Parental leave
- Health insurance
- Annual leave
- Employee assistance program
- Salary packaging
- Employee rewards program
- Experience: The successful applicant will have relevant similar experience in the reliability sector - mineral processing environment or similar.
- VisyNorth Melbourne VIC
- The Reliability Engiener is responsible for maintaining, troubleshooting, and repairing mechanical systems and equipment in a paper manufacturing plant.
- View all Visy jobs - North Melbourne jobs
- Salary Search: Reliability Engineer salaries
- See popular questions & answers about Visy
- BHPPerth WA
- Parental leave
- Monitor performance and analyse reliability trends to identify emerging risks and improvement opportunities.
- Experience analysing asset performance, identifying…
- View all BHP jobs - Perth jobs - Reliability Engineer jobs in Perth WA
- Salary Search: Specialist Comms Engineer Reliability salaries in Perth WA
- See popular questions & answers about BHP
- VisyTumut NSW
- Employee discount
- Visa sponsorship
- Partner with operations and maintenance to improve reliability and performance.
- Reporting to the Technical Manager, the Process Engineer will work across…
- View all Visy jobs - Tumut jobs
- Salary Search: Process Engineer salaries
- See popular questions & answers about Visy
- Tata Consultancy ServicesSydney NSW 2000
- Monitor application health and performance using tools such as Splunk, Dynatrace, Grafana, or Datadog.
- Perform end-to-end application troubleshooting, root…
- View all Tata Consultancy Services jobs - Sydney jobs
- Salary Search: Site Reliability Engineer salaries in Sydney NSW
- See popular questions & answers about Tata Consultancy Services
- Firmus TechnologiesLaunceston TAS
- For our customers, this approach allows us to make every watt count and deliver low-cost AI tokens globally.
- Perform hardware diagnostics, systems functionality…
- View all Firmus Technologies jobs - Launceston jobs
- Salary Search: Site Reliability Engineer salaries
- Firmus TechnologiesLaunceston TAS
- For our customers, this approach allows us to make every watt count and deliver low-cost AI tokens globally.
- Perform hardware diagnostics, systems functionality…
- View all Firmus Technologies jobs - Launceston jobs
- Salary Search: Site Reliability Engineer salaries
- Coca-Cola Europacific PartnersNorthmead NSW
- Employee stock purchase plan
- Employee discount
- Strong experience in reliability engineering, maintenance planning or continuous improvement.
- Develop and implement action plans to improve equipment…
- View all Coca-Cola Europacific Partners jobs - Northmead jobs
- Salary Search: Reliability Engineer salaries
- See popular questions & answers about Coca-Cola Europacific Partners
View similar jobs with this employerAmazon Commercial Services Pty LtdRavenhall VIC- From the introduction of Amazon Prime, to the use of advanced technology for package delivery; Amazon consistently drives change from the front of the pack.
- SourceoLaunceston TAS
- For our customers, this approach allows us to make every watt count and deliver low-cost AI tokens globally.
- Perform hardware diagnostics, systems functionality…
- View all Sourceo jobs - Launceston jobs
- Salary Search: Site Reliability Engineer salaries
- Coronado Global ResourcesBlackwater QLD 4717
- Housing allowance
- Free food
- Annual leave
- Gym membership
- Hands-on experience with condition monitoring and reliability analysis.
- Demonstrated experience in reliability engineering within mining or fixed plant…
- View all Coronado Global Resources jobs - Blackwater jobs
- Salary Search: Reliability Engineer salaries in Blackwater QLD
- Virgin AustraliaBrisbane QLD
- Insurance services
- Onsite gym
- Work from home
- You enjoy solving complex problems and building reusable capabilities that other engineers choose to adopt.
- Experience in platform engineering, site reliability…
- BHPPerth WA
- Parental leave
- Monitor signalling asset performance and analyse reliability trends to identify emerging risks and improvement opportunities.
- View all BHP jobs - Perth jobs - Signalling Engineer jobs in Perth WA
- Salary Search: Specialist Signals Engineer Reliability salaries in Perth WA
- See popular questions & answers about BHP
Job Post Details
Senior Site Reliability Engineer - job post
Job details
Job type
- Full-time
Shift and schedule
- On call
Location
Full job description
Description
We are looking for a Senior Site Reliability Engineer with Cloud platform experience. This individual will be part of a team responsible for operating and maintaining production clusters and developing our observability solutions; they will collaborate with team members to develop automation strategies, monitoring & alerting, and ensuring overall platform reliability. Your goal will be to become an integral part of the team, making every challenge of the platform – your own challenge, and solving them accordingly.
Responsibilities
- Ensure platform reliability and availability across production and pre-production environments through proactive monitoring, alerting, and automation.
- First response for incidents, contribute to problem management and root cause analysis.
- Supporting the development team's effort towards reliability, creating a solid reliability culture within the development lifecycle.
- Develop troubleshooting documentation for production support resources.
- Collaborate with Engineering teams to develop optimised and productive runbooks, operational documentation and automation of operational tasks.
- Collaborate with development and cloud engineering teams to embed reliability and performance into the software delivery lifecycle.
- Design, implement, and evolve observability solutions (metrics, logs, traces, dashboards) using tools such as Prometheus, Grafana, and ELK.
- Participate in on-call rotations and continuously improve alert quality and response processes.
- Champion a culture of reliability, performance, and continuous improvement across teams.
Requirements
- Bachelor's Degree or MS in Engineering or equivalent.
- Experience in operating at least one container orchestration cluster (Kubernetes, Docker Swarm).
- Experience developing or maintaining software for production services at scale.
- Experience with ELK.
- Experience with AWS.
- Experience with Grafana/Prometheus stack.
- Strong scripting skills (Bash, Python or Go).
- Excellent communication skills.
- Thinking out of the box and anticipating challenges. It is imperative we are not simply reactive; we must expect challenges and question technologies, procedures and thinking already in place. You will be expected to constantly review and challenge at all levels.
- Versatility. We work with agile/lean methods. We'd much rather iterate and learn than assume we know all the answers.
- Being a team player. You don't (always) work in isolation and are excited by the thought of using your team whilst involving product, experience design, engineering, and more in the process.
Will be considered as a plus:
- Telephony knowledge (SIP, VoIP);
- Experience in Linux Administration (RedHat, CentOS, AL);
- Working knowledge in Configuration Management tools (Terraform, Ansible);
- Experience with TCP/IP and general networking concepts;
- RDBMS knowledge (MySQL, Postgres);
- NoSQL knowledge (Redis).
Benefits
- Fixed compensation;
- Long-term employment with the working days vacation;
- Development in professional growth (courses, training, etc);
- Being part of successful cutting-edge technology products that are making a global impact in the service industry;
- Proficient and fun-to-work-with colleagues;
- Apple gear.
Omilia is proud to be an equal opportunity employer and is dedicated to fostering a diverse and inclusive workplace. We believe that embracing diversity in all its forms enriches our workplace and drives our collective success. We are committed to creating an environment where everyone feels welcomed, valued, and empowered to contribute their unique perspectives without regard to factors such as race, color, religion, gender, gender identity or expression, sexual orientation, national origin, heredity, disability, age, or veteran status, all eligible candidates will be given consideration for employment.