Site Reliability Engineer (SRE)

TEKsystems Hong Kong Hong Kong
Full time Permanent On-site Competitive

About the job

  • APAC wide role at Global FS firm
  • Variety of different project exposure
  • Design and hands on implementation responsabilities
We are looking for proactive and hands-on Site Reliability Engineer (SRE) / DevOps Engineer to join a growing technology team. This is an L3-focused role where you'll work on ad-hoc technical requests, platform improvements, and operational excellence initiatives.

This position offers significant autonomy and is best suited to someone who enjoys taking ownership, solving complex problems independently, and continuously improving platform reliability and performance.

Unlike a purely operational support role, you will be expected to actively build, maintain, and enhance the underlying platform rather than simply administer existing systems.

Key Responsibilities
  • Provide L3 support for complex infrastructure and platform-related issues.
  • Investigate and resolve production incidents and service disruptions.
  • Build, maintain, and optimize Kubernetes environments.
  • Develop Python scripts and automation solutions to improve operational efficiency.
  • Implement platform enhancements and reliability improvements.
  • Collaborate with engineering teams to improve scalability, performance, and system resilience.
  • Support monitoring, alerting, and observability initiatives.
  • Drive continuous improvement and automation across the technology stack.
  • Act as an individual contributor with a high degree of ownership and accountability.