Hong Kong, Hong Kong SAR, Hong Kong

Mid-Level SRE | Trading Technology | Performance & Observability

 Job Description:

We're looking for a Mid-Level SRE to join a fast-paced technology environment, helping to improve system performance, reliability and observability across production systems.

What you'll be doing

  • Monitor and improve system performance, reliability and availability
  • Build and maintain observability and telemetry solutions
  • Develop meaningful monitoring, metrics, dashboards and alerts
  • Troubleshoot production issues and identify root causes
  • Build and maintain CI/CD pipelines to support reliable and efficient releases
  • Write automation and engineering tools using Python or other programming languages
  • Work closely with development and infrastructure teams to improve system performance and operational efficiency
  • Support continuous improvements in a fast-moving, results-driven environment

What we're looking for

  • MinimumĀ 5 years of experience in SRE, DevOps, Platform Engineering, Infrastructure Engineering or a related field
  • Hands-on experience with observability, monitoring or telemetry
  • Solid experience with CI/CD
  • Strong programming skills in Python or another programming language such as Java, Go, Scala or Rust
  • Experience troubleshooting and improving production system performance
  • Strong problem-solving skills with a practical, hands-on approach
  • Comfortable working in a fast-paced and results-driven environment

Nice to have

  • Prometheus / Grafana / OpenTelemetry / Datadog / ELK or similar observability tools
  • Kubernetes / AWS or other cloud platforms
  • Terraform / Infrastructure as Code
  • Experience with distributed systems, data platforms or streaming technologies
  • Experience in fintech, trading or other high-performance technology environments
  Required Skills:

Environment Distributed Systems Performance CI/CD pipelines Go FinTech Data Cloud Prometheus Support Scala ROOT Development Grafana Pipelines Trading CI/CD Programming Languages Metrics Operational Efficiency Reliability Availability DevOps Infrastructure Automation Programming AWS Kubernetes Troubleshooting Engineering Java Python