Site Reliability Engineer
Indexed description
Site Reliability Engineer – Trading Technology
Hong Kong
I’m working with a technology-driven trading firm in Hong Kong looking to add a Site Reliability Engineer to its engineering team.
This is a hands-on role sitting across SRE, platform engineering, observability and system performance. You’ll work with production systems where reliability and performance matter, using telemetry, data and automation to understand system behaviour and build better engineering solutions.
What you’ll work on:
- Build and improve observability, monitoring and telemetry across production systems
- Troubleshoot Linux performance, networking, CPU, memory and application issues
- Develop internal tooling and automation using Python / Go
- Work across Kubernetes / EKS and containerised infrastructure
- Improve CI/CD and deployment pipelines
- Automate infrastructure using Terraform / IaC
- Use production data to identify bottlenecks and drive long-term reliability improvements
Technology environment:
- Linux | Python | Go | AWS | EKS | Kubernetes | Terraform | Prometheus | Grafana | ClickHouse | Kafka | Airflow | ArgoCD | CI/CD
- What we’re looking for:
- Ideally 5–7+ years across SRE, Production Engineering, DevOps or Platform Engineering, with strong Linux fundamentals, coding/automation skills and experience operating production systems.
Experience within trading, financial markets, digital assets or another high-performance technology environment would be advantageous but isn't essential.
If you’re based in Hong Kong or interested in relocating, feel free to message me directly.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search