Senior DevOps Engineer

Gethookdai — Bulgaria · Posted ~22 hours ago

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Description

If you want to own infrastructure end to end, not operate a corner of someone else’s, keep reading. About GetHookd GetHookd.ai is an ad intelligence platform used by 5,500+ eCommerce and marketing teams to find and create winning ad creatives. We maintain a searchable library of 100+ million ads across Meta and TikTok, adding tens of thousands daily through a scraper fleet that runs 24/7 and has to stay ahead of TLS fingerprinting and platform changes. Semantic search over the whole thing. A data moat that grows every hour we stay online. In July we launched our API and MCP server. AI agents now query our platform programmatically, and it’s our fastest-growing usage channel. The infrastructure you’d run is increasingly what other people’s AI systems depend on. GetHookd is built and run by Meliora, a product studio in Sofia with two Shopify app exits behind us. GetHookd is our current focus and our biggest bet. A startup team with a studio’s track record, profitable, no outside funding, allergic to bureaucracy. The Role Straight version: our production knowledge currently sits with one senior engineer who built most of the platform. We’re hiring you to change that. You’ll own infrastructure end to end, make the systems legible, and have real influence over architecture, tooling, and long-term direction. You’re not joining an infra team. You’re becoming the platform function, working alongside the engineer who knows where the bodies are buried. The stack: GCP-primary (GKE) with Cloudflare at the edge, containerized services, TypeSense-powered semantic search on highmem nodes, GitHub Actions, and a product team shipping daily. What You’ll Own GCP and Cloudflare infrastructure, end to endMonitoring and alerting built on leading indicators. Disk trends, queue depth, error rates. Our worst outages were preceded by warnings nobody was watching. That ends with you.CI/CD pipelines across backend, frontend, and scraper servicesUptime and resilience of the scraper fleet and data pipelineIncident response and disaster recovery, including runbooks that someone other than their author has actually testedLoad testing and capacity planning ahead of Black Friday, our biggest traffic seasonEnvironment provisioning, secrets management, and access controlCost optimization at real scale. Cloud spend is a meaningful line item and you’ll see the actual numbers.Security posture as we move upmarket Who We’re Looking For 5+ years hands-on DevOps or platform work in production SaaS. You’ve personally carried a pager.Strong GCP and Kubernetes experience. AWS familiarity is a plus.Cloudflare-native thinking (Workers, R2, edge patterns) is a real bonusProduction-grade CI/CD (GitHub Actions or equivalent) and Infrastructure as Code (Terraform or Pulumi)Scripting in Python, Bash, Go, or TypeScriptMonitoring and observability in your bones. Grafana, Sentry, Datadog, pick your poison.Experience running search infrastructure (TypeSense, Elasticsearch, OpenSearch) at scale is a strong plusYou make trade-offs quickly and explain them clearlyYou run toward production incidents, not away from themYou treat AI-generated remediation plans as hypotheses to validate, not commands to execute. We will ask about this in the interview. There’s a story behind it. This Role Is Not For You If You prefer tickets and playbooks over ownership and initiativeDocumentation is a chore you do later. Later never comes and we both know it.You need a large team and long feedback cycles to functionYou want to replatform in month one. The architecture has scars but it works and it’s profitable.You’re more precious about stack preferences than about what actually works What We Offer Hybrid: week 1 in-office, then Mon–Wed office / Thu–Fri remote20+ days paid leave, growing with tenurePrivate health insurance and an optional sports cardReal product ownership. No red tape. No fluff.A small, honest team that ships things that matter How We Work No Black Boxes. Share context early, often, and publicly.Feed Forward. Feedback is fast, clear, and focused on what’s next.Own the Outcome. We finish the job, not the task.Hack a Better Way. Prototype, test, adapt.Flex & Flow. Plans change, we move. The Process Four steps, two weeks total: a short written screen (3 questions, 20 minutes), a technical conversation including a code review of a real PR from our codebase, a paid 3-hour practical, and a founder conversation. We pay for the practical because your time is worth money and unpaid take-homes tell you how a company treats people. Apply with a CV or LinkedIn. We read every application ourselves.