
Flow Traders is hiring Reliability Engineers to safeguard the production performance of its global trading platform. You prevent undetected degradation from costing P&L by the minute. Most failures become smaller, shorter and easier to contain before anything breaks. You automate repetitive triage, keep alerts focused on what needs human attention, and close monitoring gaps. When something does break, you take command as Incident Commander under the Global Incident Management framework, sizing up the issue, deciding who joins the call, and coordinating the response. You sit alongside traders, developers, infrastructure and specialist teams with the standing to push back when operational standards slip.
What you'll do
- Automate repeatable triage work so first-line responds faster and more consistently, including alert enrichment, routing, correlation and operational workflows
- Push for monitoring, alerting and visibility where gaps exist
- Track reliability and availability across critical trading applications and the platforms they depend on, working with users, development teams and IT to find where service levels are degrading
- Call out weak ownership, missed SLAs, poor alerts and ineffective runbooks, and drive the owning teams to fix them
- Triage incoming alerts, issues and escalations, and assess impact, urgency and ownership
- Declare incidents when criteria are met and act as Incident Commander
- Coordinate responders and stakeholders, keeping incident calls focused on facts, mitigation and recovery
- Maintain clear timelines, actions and status updates throughout an incident
- Recover or stabilise systems using approved runbooks, and escalate cleanly through the defined support and development path when the issue goes beyond documented recovery steps
- Perform common operational tasks across adjacent teams where needed
- Support PIR follow-up and recurring issue review
- Hand over cleanly between EMEA, AMER and APAC under one global model, one incident standard, one handover process
What you bring
- Experience in production operations, SRE, NOC/command centre, trading operations or a similar first-line technical role, ideally in a trading, financial services or other latency-sensitive environment
- Track record of running or coordinating major incidents, and comfort taking command of a call with senior people on it
- Strong triage and prioritisation. You can separate facts from assumptions under time pressure and keep the response moving
- Clear verbal and written communication. Your status updates are readable by a trader and an engineer at the same time
- Strong judgment and escalation discipline. You know when to keep going and when to pull in a specialist
- Willingness to hold the line on process and to push back when poor operational behaviour creates risk for trading
- Technically broad rather than deep. You understand how most teams operate to be useful across domains, not to be the specialist resolver
- Solid Linux and networking fundamentals, and the ability to read alerts, logs, dashboards and symptoms quickly
- Working knowledge of common operational tasks across adjacent teams (application support, infrastructure, connectivity, data)
- Familiarity with incident and observability tooling: PagerDuty or equivalent, Jira Service Management or equivalent, Grafana, Prometheus, log search
- Scripting and automation ability (Python preferred; Bash, Go a plus) applied to triage, enrichment, routing and correlation
Nice to have
- Exposure to containerised and cloud-hosted production systems (Kubernetes, Docker, GCP)
What we offer
- Competitive salary and annual discretionary bonus
- Flow Academy for continuous learning and opportunities to attend domain-related conferences
- Comprehensive health insurance coverage
- In-house lounge with a bar, pool table and console games
- Daily catered breakfast and lunch with healthy snacks and drinks available throughout the day
- In-house hairdresser and massage therapist
- Personal trainers, weekly boot camps and subsidized gym membership
- Annual company trip and a variety of events throughout the year
- Global rotations between offices worldwide
About Flow Traders
Flow Traders is a global trading firm with a non-hierarchical approach that stimulates innovation and collaboration. The company provides extensive onboarding, access to continuous learning opportunities, the latest technology, and a strong emphasis on retaining talent and maintaining a small business culture across its worldwide offices.
What engineering roles in crypto pay
411 salaries · our own dataMost engineering roles in crypto pay between $153k and $250k, with a median of $203k.