Summary
An experienced DevOps engineer is sought to own the stability, reliability, and efficiency of cloud infrastructure supporting demanding research and trading workflows. You will manage cloud resources and IAM, monitor costs and budgets, automate operational tasks with Python, and act as a bridge between engineering, research, and systems teams.
Highlights
High-ownership operations role responsible for the reliability and efficiency of cloud-based trading and research infrastructure, with broad exposure to cloud operations, automation, cost controls, and cross-team collaboration.
Description
Our technology team is looking for an experienced DevOps Engineer to own the stability, reliability, and efficiency of our trading and research infrastructure.
This is fundamentally an operations role: you will be responsible for keeping our systems running, our GCP environment well-managed, and our research and trading workflows uninterrupted.
You will serve as the connective tissue between engineering, research, and systems teams — someone who understands the nuance of working across business lines, exercises strong judgment and discretion, and takes ownership of the operational health of our platform.
Python proficiency is required for scripting and automation, but this role is defined by operational excellence, not software development.
In This Role, You Will Be Responsible For
Owning the day-to-day operational health of our Google Cloud Platform (GCP) environment, including resource provisioning, IAM management, cost controls, and budget monitoring Serving as the operational bridge between engineering, research, and systems teams — coordinating cross-functionally to ensure production systems run smoothly and issues are resolved quickly Monitoring and maintaining production trading and research systems, with a focus on reliability, uptime, and proactive incident response Managing the full research pipeline infrastructure on GCP, including job scheduling, resource allocation, and platform stability Administering and maintaining CI/CD pipelines and Kubernetes clusters to support reliable, repeatable deployments Managing fleets of VMs, including metrics collection, alerting, and log ingestion policies Administering high-performance database instances and supporting data pipeline reliability Benchmarking and tuning critical services for optimal performance across internal systems and external trading venues Building and maintaining operational tooling for monitoring, backup, deployment automation, and testing Integrating existing solutions and operational best practices rather than developing new systems from scratch
If you possess the following, we would love to explore what is available for you with our team:
3+ years of experience in a DevOps, platform operations, or site reliability engineering role — this is not an entry-level position Demonstrated ability to work cross-functionally across engineering, research, and business teams with professionalism and discretion Hands-on GCP experience, including IAM, compute, networking, and cost management Experience using Terraform (or similar Infrastructure as Code solutions) to manage, automate, and tag cloud infrastructure Strong working knowledge of Kubernetes cluster and application management Experience managing fleets of VMs, including metrics collection, alert configuration, and log ingestion Proven track record of implementing budget controls and monitoring cloud spend Strong Python/Bash scripting skills in a Linux environment, with an emphasis on automation and operational tooling Experience with CI/CD systems, including design, support, and ongoing management Experience with relational and columnar databases Knowledge of streaming processes or Kafka Previous experience in financial or exchange environments is a plus Strong analytical and problem-solving skills with a bias toward operational excellence Excellent communication skills and sound judgment in a fast-paced, high-stakes environment Bachelor’s degree in Computer Science, Computer Engineering, or equivalent experience
While we are serious about our work at Vatic, we also promote a fun environment! You can expect:
Ping-pong and poker gamesFun team outings Unlimited office snacksFree breakfast, lunch, and dinnerGym membershipFull health insurance coverage for employees and dependents
The base salary range for this role is between $150,000 and $250,000.
The base salary range does not include any other form of compensation, such as any bonus amounts, or any benefits.
Factors that may impact the agreed upon base salary within the range for a particular candidate include years of experience, level of education obtained, skill set, and other factors.