Senior DevOps Engineer
Our client is an exceptional proprietary trading firm seeking a Senior DevOps Engineer to join a high-impact platform engineering team at the heart of a rapidly growing global trading business. Renowned for its technology-first culture, entrepreneurial environment, and continued investment in infrastructure, the firm is scaling its next-generation platform that underpins mission-critical trading systems. This is an outstanding opportunity for a Senior DevOps Engineer to work on large-scale Kubernetes infrastructure, advanced observability, automation, and platform engineering challenges within a high-performance trading environment.
They are looking for a Senior DevOps Engineer who thrives on ownership, enjoys solving complex engineering problems, and wants to build platforms that directly impact trading performance and developer productivity. The ideal DevOps Engineer will combine strong software engineering capability with deep infrastructure expertise, bringing a proactive, can-do mindset and a passion for automation. This role would suit a DevOps Engineer who wants to move beyond operational support and play a key role in shaping the future of a modern, business-critical platform.
Key Responsibilities
- Own and evolve a large-scale Kubernetes platform supporting critical trading and engineering workloads across the firm.
- Drive cluster lifecycle management, upgrades, networking, storage, autoscaling, and disaster recovery initiatives.
- Scale and enhance a firm-wide observability platform, improving visibility, reliability, and operational performance.
- Build, optimise, and maintain CI/CD and GitOps workflows to accelerate software delivery.
- Develop automation and self-service tooling that empowers engineering teams and reduces operational friction.
- Partner directly with trading desks and software engineers to onboard new workloads and improve platform adoption.
- Troubleshoot complex production issues across Kubernetes, Linux, networking, and distributed systems environments.
- Contribute to platform strategy, architecture decisions, documentation, and engineering best practices.
- Deliver innovative solutions across areas such as workflow orchestration, AI-enabled tooling, and infrastructure automation.
Key Skills & Experience
- 5+ years of experience as a DevOps Engineer, Platform Engineer, or Site Reliability Engineer within complex production environments.
- Strong hands-on experience operating Kubernetes at scale, including cluster management, troubleshooting, networking, and security.
- Deep Linux systems knowledge covering networking, storage, system performance, and operational troubleshooting.
- Experience with observability technologies such as Prometheus, Grafana, OpenTelemetry, and large-scale metrics, logging, and tracing platforms.
- Strong experience building and maintaining CI/CD pipelines and GitOps environments using tools such as ArgoCD, GitLab CI, GitHub Actions, or similar.
- Expertise in Infrastructure as Code using Terraform, Ansible, or equivalent technologies.
- Strong programming skills in Python or Go, with a focus on automation and platform development.
- Experience supporting on-premise or bare-metal infrastructure environments.
- Excellent communication skills with the ability to work closely with developers, platform users, and business stakeholders.
- Experience within trading, financial markets, hedge funds, or other performance-sensitive environments would be highly advantageous.
- Exposure to Kafka, Airflow, Kubeflow, advanced Kubernetes networking, service mesh technologies, or AI-assisted operational tooling is beneficial.
- A collaborative, ownership-driven mindset with a desire to build scalable platforms and influence technical direction.
This is a unique opportunity to join a highly regarded trading firm where engineering excellence is a genuine competitive advantage. You'll work on greenfield initiatives, own critical infrastructure from end to end, and help shape a platform used across the business, all while operating in a fast-paced environment where your work has immediate and measurable impact.
