Senior DevOps / Site Reliability Engineer

Omniphi — Canada · Posted ~21 hours ago

Senior

Skills

DevOps Site reliability engineering GitHub CI/CD Cloud infrastructure Linux Release management Production monitoring Incident response Cloud

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior DevOps and Site Reliability role with end-to-end ownership of infrastructure, software delivery, and production reliability. You will build and operate GitHub-based CI/CD and safe-release processes, manage cloud and Linux environments, monitor service health, and lead production incident response.

Highlights

Senior ownership role at an early-stage company with broad responsibility for infrastructure, software delivery, reliability, and production operations. The position offers significant autonomy in building CI/CD, cloud infrastructure, monitoring, and incident-response systems.

Description

ABOUT THIS JOB OmniPhi is an early-stage company building an agentic trading environment where users can research, write, test, deploy and run a strategy live, all in one place, with an AI agent working alongside them rather than a collection of disconnected tools. The platform is live and running real strategies for real customers. We are a small team that ships quickly, and we are hiring a senior engineer to own the infrastructure, software delivery and reliability systems underneath it. You will manage the operational path from an approved GitHub change to reliable production service designing, building, operating and monitoring software delivery, cloud infrastructure, service health and production incident response. RESPONSIBILITIES Build and operate GitHub-based CI/CD, release and safe-deployment processes, and own release readiness, promotion and post-deployment monitoring.Operate cloud services and private Linux machines, including provisioning, configuration, capacity, recovery and cost control.Build monitoring, health checks, logging, alerting and incident-response procedures.Own database and cache operations across PostgreSQL, TimescaleDB and Redis/Valkey, including access, monitoring, migrations, backups and recovery.Maintain operational security across networking, DNS, TLS, IAM, secrets, system access and workload isolation.Maintain infrastructure automation, deployment tooling and actionable runbooks.Participate in scheduled on-call coverage and help recruit and mentor future infrastructure engineers.REQUIRED QUALIFICATIONS Bachelor's degree in Computer Science, Software Engineering, Computer Engineering or another relevant technical field, or equivalent practical experience.Six or more years of senior-level experience in DevOps, SRE, platform engineering, cloud engineering or production operations.Hands-on experience with at least one major cloud provider, such as AWS, Microsoft Azure, Google Cloud Platform or Oracle Cloud Infrastructure.Strong Linux, Docker, systemd, networking and production-troubleshooting skills.Strong Python and Bash experience, with the ability to read, diagnose and safely modify application and infrastructure code.Experience designing and operating modern CI/CD pipelines and controlled production deployment systems.Operational knowledge of PostgreSQL or comparable SQL databases, time-series systems and Redis or comparable cache services.Experience with production monitoring, incident response, recovery and cloud security.Ability to work on-site and participate in scheduled on-call coverage.STRONGLY PREFERRED Kubernetes or comparable container-orchestration experience.Infrastructure as Code using Terraform, OpenTofu, Pulumi or comparable tooling.Distributed Linux fleet management and controlled software delivery.Financial technology, trading or another reliability-sensitive production environment.Experience establishing operational practices or mentoring engineers.Responsible use of AI-assisted engineering tools such as Claude Code or Codex.You do not need to have used every technology listed. We are interested in senior engineers who have personally owned production systems, exercise sound operational judgment and can learn unfamiliar tools quickly. Ability to commute/relocate: Vancouver, BC: reliably commute or plan to relocate before starting work (required)Work Location: In person