All vacancies
Site Reliability Engineer
CXM
Remote · worldwide, UTC+4 acceptedSalary not disclosedfull-timeVerified 4 days agoHimalayas
Join our Platform & Production Reliability team and help ensure the reliability, performance, and availability of our mission-critical trading systems. As an Application Site Reliability Engineer (SRE), you will own the day-to-day reliability of our .NET/C# services running on Windows, starting with our in-house liquidity bridge that connects MetaTrader trading servers to external liquidity providers.
Responsibilities
- Participate in the on-call rotation for production trading systems and lead incident response during service disruptions.
- Investigate production incidents, perform root cause analysis, and implement preventive actions to eliminate recurring issues.
- Build and maintain Grafana dashboards, Prometheus alerts, and operational health views across applications, infrastructure, and databases.
- Instrument .NET services to improve telemetry, metrics, logging, and visibility into service health and customer impact.
- Define, implement, and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets.
- Troubleshoot issues across
- .NET/C# applications
- Windows Server
- Aurora PostgreSQL databases
- AWS infrastructure
Languages
- Work format
- Remote
- Seniority
- Mid
- Posted
- 3 Sept 2026 (1w ago)
- Last verified
- 13 Sept 2026
- Apply by
- 2 Nov 2026
