Site Reliability Engineer – AWS, Python (Contract)
A prestigious investment bank in Hong Kong seeks a Site Reliability Engineer – AWS, Python (Contract) to improve the reliability and performance of mission-critical systems while driving automation in a collaborative technology environment.
Key responsibilities:
As a Site Reliability Engineer, you will ensure the reliability, scalability, and security of critical banking systems through automation, monitoring, CI/CD, and cross-team collaboration.
- Drive system reliability, availability, and performance through proactive monitoring, incident management, and SRE practices
- Build and enhance observability platforms, including monitoring, alerting, dashboards, and telemetry pipelines
- Manage Linux server environments (RHEL 7/8/9), ensuring security, stability, patch compliance, and operational efficiency
- Analyse logs, metrics, and system behaviour to troubleshoot incidents, resolve performance issues, and improve service resilience
- Automate infrastructure, deployment, and operational workflows using AWS services and Python
- Partner with engineering teams to improve platform scalability, security, automation, and operational excellence
- Administer Kubernetes environments to support reliable workloads, platform health, and infrastructure observability
- Support business continuity through on-call participation, disaster recovery exercises, continuous improvement, and adoption of emerging SRE technologies
Candidate profile:
To excel in this role, you bring experience supporting large-scale systems, strong DevOps expertise in CI/CD, cloud and containers, and the ability to collaborate effectively under pressure.
- Bachelor's degree in Computer Science, Engineering, or a related discipline, with 3–7+ years' experience in SRE, platform engineering, or production support
- Strong expertise in monitoring and observability, including Prometheus, Elasticsearch, Grafana, Kibana, and enterprise monitoring tools
- Hands-on experience with Linux (RHEL 7/8/9), Kubernetes, and AWS in large-scale, high-availability production environments
- Strong understanding of SRE principles, incident management, disaster recovery, automation, and CI/CD
- Experience with networking and distributed systems troubleshooting, with strong problem-solving and communication skills
- Proficiency in Python, Bash, and Ansible for scripting and automation; AI/ML infrastructure experience is an advantage
About this company:
A well-established global financial institution provides innovative capital markets and investment solutions to clients across international markets. The organisation fosters a collaborative, technology-driven environment with strong emphasis on professional growth, innovation, and building the next generation of financial services capabilities.
Keywords: Site Reliability Engineering, AWS, Python, DevOps, Kubernetes
What’s next:
Build resilient systems, strengthen critical banking platforms, and advance your career in financial technology. Apply now!
About the job
Contract Type: Temp
Specialism: Tech & Transformation
Focus: DevOps, SRE Engineer & Application support
Industry: Banking
Salary: Up to HKD50,000 per month
Workplace Type: On-site
Experience Level: Associate
Location: Central
TEMPORARYJob Reference: 2IIUZ3-AFCE50CC
Date posted: 22 September 2026
Consultant: Melanie Wu
hong-kong tech-transformation/devops 2026-09-22 2026-10-22 banking Central Central and Western District HK HKD 50000 50000 50000 MONTH Robert Walters https://www.robertwalters.com.hk https://www.robertwalters.com.hk/content/dam/robert-walters/global/images/logos/web-logos/square-logo.png true