Description
This is an exciting opportunity to be part of a psychological science-based tech startup.
This role sits within our product team and will be responsible for building, maintaining, and scaling the infrastructure that powers Hive’s scalable products.
This DevOps Engineer will roll up their sleeves and help us continually evolve the foundation of our platform, enabling faster delivery speed and greater scale, as well as supporting new generative AI and machine learning technologies.
We’re looking for someone who can bring strong infrastructure and automation expertise, and who can work closely with engineering and data science teams to support rapid product development & launch.
What You’ll Be Doing
As a DevOps Engineer (Product), you’ll be responsible for the reliability, scalability, and security of our entire infrastructure stack; from CI/CD pipelines to production deployments, from infrastructure orchestration to security governance.
You’ll be part of a high-powered team located in London.
You will build robust systems, move quickly, but be ready to scale and ensure production-grade reliability as our product grows.
You will constantly need to be at the cutting edge as we deploy and scale the latest AI capabilities within our core platform.
We can’t define everything you will be doing because some of it is unknown based on the disruptive world we live in and you need to be the kind of person ready to pivot at speed and stay at the bleeding edge of this new world.
Infrastructure & Cloud Engineering
Design, provision, and manage scalable cloud infrastructure using Infrastructure-as-Code (Terraform, CloudFormation) across AWS (must have deep experience), GCP, or Azure.Architect and maintain highly available, fault-tolerant systems that support our AI/ML workloads, web applications, and data pipelines.Manage containerization and orchestration platforms (Docker, Kubernetes, ECS) to support microservices and ML model deployments.
CI/CD & Automation
Build and maintain robust CI/CD pipelines (GitHub Actions, CircleCI, Jenkins) to automate testing, builds, and deployments across dev/staging/production environments.Implement MLOps workflows to streamline model deployment, versioning, and monitoring for our AI/ML products.Automate infrastructure provisioning (Terraform), configuration management, and deployment processes using scripting (Bash, Python) and automation tools.
Monitoring, Observability & Reliability
Implement comprehensive monitoring, logging, and alerting systems (Prometheus, Grafana, CloudWatch, Datadog, Sentry) to ensure system reliability and rapid incident response.Establish SLOs/SLIs and implement observability best practices to maintain high availability and performance.Lead incident response, root cause analysis, and implement preventive measures to improve system resilience.
Security & Governance
Implement and maintain security best practices including network security, firewalls, role-based access control (IAM), encryption at rest and in transit, and secrets management (AWS Secrets Manager, HashiCorp Vault).Develop and enforce governance frameworks for working with LLM APIs and AI services, including data protection, PII safeguards, and compliance requirements.Conduct security audits, vulnerability assessments, and implement remediation strategies to maintain a secure infrastructure.
Collaboration & Technical Support
Work closely with full-stack engineers and data scientists to support application deployments, optimize performance, and troubleshoot infrastructure issues.Support ETL/ELT workflows and data pipeline infrastructure for training and inference workloads across databases (SQL, NoSQL, Vector DBs, Graph DBs).Provide technical guidance and mentorship on DevOps best practices, infrastructure design, and deployment strategies.
Strong experience provisioning and managing secure cloud infrastructure (AWS preferred, also GCP or Azure)Expertise with Infrastructure-as-Code tools (Terraform, CloudFormation, Pulumi)Strong experience with containerization and orchestration (Docker, Kubernetes, ECS, Fargate)Proven track record building and maintaining CI/CD pipelines (GitHub Actions, CircleCI, Jenkins, GitLab CI)Experience with MLOps and supporting ML model deployment workflows (AWS Sagemaker, Lambda, containerized deployments)Proficiency in scripting and automation (Python, Bash, Go)Strong experience with monitoring and observability tools (CloudWatch, Prometheus, Grafana, Datadog, Sentry, New Relic)Experience with database administration and optimization across SQL, NoSQL, vector databases (Pinecone, FAISS), and graph databases (Neo4j)Knowledge of networking, security best practices, IAM configuration, and secrets managementExperience supporting data pipelines, ETL workflows, and cloud data platforms (Databricks, Snowflake)Strong experience with the set up / design / governance and security of Clean Rooms and clean room integrationsPrevious experience in early-stage product teams or high-growth startupsAbility to balance rapid prototyping with building scalable, production-grade infrastructureStrong problem-solving skills and ability to work independently in a fast-paced environment
Overall Work Experience
You may have come from a platform engineering team at a tech company or from a startup where you wore every hat.
You are fluent in both infrastructure theory and hands-on implementation, and you get a thrill out of building reliable, scalable systems that enable rapid product innovation and support cutting-edge AI/ML workloads.
As a Fast-paced Startup, Each Day Is Different From The One Before.
We’re Nimble And Creative, And Value Intellectual Humility.
We Work Really Hard Because We’re All 100% Dedicated To The Future We’re Building.
Our Work Is Stimulating, Challenging, And Exciting.
And Our Team Is Awesome.
At Hive We Only Hire Exceptional People, So You’ll Be In Good Company; Surrounded By Passionate, Insanely Smart People Who Want To Build The Future Of Customer Intelligence.
Specifically We’re Looking For Someone Who Will Thrive In This Type Of Environment
Fast-paced startup with competing demands and multiple priorities ongoingOwn critical infrastructure decisions that directly shape the products we buildA ‘solve the problem’ mentalityScrappy and creativeStrong passion for the Hive Science mission and a love of the scientific method
This is an in-person role in London, UK - we cannot consider candidates who do not currently live within commuting distance.