Jobs Career Advice Post Job
X

Send this job to a friend

X

Did you notice an error or suspect this job is scam? Tell us.

  • Posted: Aug 27, 2026
    Deadline: Not specified
    • @gmail.com
    • @yahoo.com
    • @outlook.com
  • The Sun International brand has a proud legacy in the gaming, hospitality and entertainment sector. Its superior hotels and resorts portfolio makes it a recognized premium brand. The Sun International Group has a diverse portfolio of assets including world class five star hotels, modern and well located casinos, and some of the world’s premier resorts. Our...

     

    Site Reliability Engineer

    Job Description

    • The Site Reliability Engineer (SRE) is a technical contributor responsible for improving the reliability, resilience and operational performance of Sun International’s technology services, platforms and infrastructure.
    • The role focuses on observability, monitoring, alerting, event management, reliability analytics, automation and auto-remediation, proactively identifying potential issues before they impact the business.
    • Working across software engineering, infrastructure, enterprise applications and service management teams, the SRE implements monitoring and automation solutions that strengthen service reliability and reduce operational.

    Core behavioural & Technical / proficiency competencies:

    • Cloud platform management across Azure, AWS and/or GCP.
    • Containerisation technologies, including Docker and Kubernetes.
    • Monitoring, observability and alerting tools such as Prometheus, Grafana and ELK.
    • Scripting and operational automation using Python, Bash and/or PowerShell.
    • Development of automation, auto-remediation capabilities and operational runbooks.
    • Incident management and proactive identification of service reliability risks.
    • Root Cause Analysis (RCA) and analysis of recurring operational failures.
    • Linux and Windows system administration.
    • Networking fundamentals.
    • Git version control.
    • Troubleshooting and problem-solving across technology services and infrastructure.
    • Analysis of operational telemetry, event data and service performance trends.
    • Application of reliability standards, security requirements, operational controls and governance practices.
    • Cross-functional collaboration with software engineering, infrastructure, enterprise applications and service management teams.
    • Analytical thinking and evidence-based decision-making.
    • Continuous improvement and operational excellence.
    • Collaboration and knowledge sharing.
    • Operational excellence and accountability.

    Job Requirements

    Qualifications:

    • Degree in Computer Science, Engineering, Information Technology or a related discipline (required)
    • Cloud platform certification, e.g. Azure Administrator Associate or AWS Certified SysOps Administrator (preferred)
    • ITIL Foundation Certification (preferred)

    Experience:

    • 2–5 years’ experience in infrastructure operations, technology operations, monitoring platforms, cloud operations, Site Reliability Engineering, observability tooling or a related technology discipline.
    • Practical experience with monitoring and observability, cloud platforms, automation/scripting, incident management and troubleshooting aligned to an SRE or technology operations environment

    Check how your CV aligns with this job

    Method of Application

    Interested and qualified? Go to Sun International on suninternationaljobs.mcidirecthire.com to apply

    Build your CV for free. Download in different templates.

  • Get new ICT / Computer jobs like this on Telegram.Subscribe on Telegram
  • Send your application

    Back To Home

Career Advice

View All Career Advice
 

Subscribe to Job Alert

 

Join our happy subscribers

 
 
Send your application through

GmailGmail YahoomailYahoomail