DevOps / Site Reliability Engineer

Xfarm Technologies — Italy · Posted ~21 hours ago

Senior Full-time

Skills

DevOps Site Reliability Engineering Kubernetes Cloud infrastructure Infrastructure operations CI/CD Cloud architecture Cloud Infrastructure as Code

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a fast-growing technology team as a hands-on SRE focused on keeping a large backend platform reliable across multiple regions. You’ll own cloud operations, work deeply with Kubernetes, collaborate with architects and engineers, and help establish robust operational practices for a rapidly scaling digital ecosystem.

Highlights

Own critical cloud infrastructure supporting a large-scale digital platform across multiple regions. The role provides strong ownership, close collaboration with architecture and engineering teams, and an opportunity to establish SRE practices.

Description

Our mission at xFarm Technologies is to drive the digital transformation in Agriculture by improving the lives of millions of farmers, acting for the increase of environmental, economic and social sustainability of the sector. We do this by developing digital tools and consulting projects, providing a Smart Farming digital ecosystem for all the actors of the agri-food supply chain, consisting of apps, sensors, software integrations and advanced analytics tools. Today, xFarm products are already used by more than 420 thousand farms on more than five million hectares, in Europe and South America. Now it's time to accelerate and we need you! We are looking for a hands-on DevOps / Site Reliability Engineer to join our CloudOps team and own the operational infrastructure that keeps our backend platform running across Europe and Latin America. You will work closely with our Cloud Architect, and with the Tech organization's engineering teams, as the first dedicated SRE / Kubernetes hire in a newly established, high-impact function. Your Mission Kubernetes Operations: Deploy, scale, secure and troubleshoot containerized backend services on Kubernetes, managing workloads, resource limits, networking and rollouts; CI/CD Pipelines: Own and improve GitLab CI/CD pipelines end to end, setting up new projects and pipeline scaffolding for new services while keeping build/test/deploy fast and reliable; Observability: Govern the Elastic Cloud stack (logs and metrics) and operate the Prometheus/Grafana stack, building and maintaining dashboards, alerts and log/metrics pipelines, acting as the reference point for teams debugging production issues; Reliability & Incident Response: Define and track SLIs/SLOs, respond to incidents, drive root-cause analysis and reduce toil through automation; Infrastructure as Code: Manage infrastructure declaratively (Terraform / Helm / GitOps or equivalent) to keep environments reproducible and auditable; Developer Enablement: Support engineering teams with tooling, environments and fast feedback loops, documenting runbooks and self-service paths; Security & Hygiene: Apply secrets management, access control and dependency/image hygiene across the pipeline and cluster. Opportunities Be the first dedicated SRE / Kubernetes hire in a newly established CloudOps function; Work directly alongside the Cloud Architect and the Tech organization's engineering teams across EU and LATAM; Play a high-leverage, high-visibility role: your work at the infrastructure layer helps the entire engineering organization ship faster and more safely; Help shape how CloudOps operates and grows from a formative moment onward. Qualifications Proven hands-on experience as a DevOps / SRE / Platform Engineer in production; Solid experience with Kubernetes in production — deployment, operations and troubleshooting, not just kubectl apply; Experience designing and maintaining CI/CD pipelines, ideally GitLab CI/CD, including bootstrapping new repositories and pipelines from scratch; Observability experience with both the Elastic stack (Elasticsearch / Kibana / Elastic Cloud) and the Prometheus/Grafana stack (metrics collection, PromQL, dashboards, alerting); Infrastructure as Code experience (any tool) and solid cloud fundamentals (containers, networking, secrets, IAM basics); Strong scripting ability (Bash / Python or similar); Full European-hours availability (09:00–18:00 CEST) to support engineering teams in real time; Solid English for working across EU and LATAM chapters; Java/JVM knowledge, AWS and/or GCP experience are a plus. Complete the Profile Collaboration and clear communication with non-infra colleagues Analytical, root-cause-driven problem-solving Composure under pressure during incidents Self-motivation and ownership Reliability and accountability Continuous learning Compensation RAL: €45.000 - € 50.000 (commensurate with experience) Other Information xFarm Technologies is a company that is committed to creating a work environment where diversity and inclusiveness are critical to the well-being and growth of all employees. This role operates in a hybrid work model, combining on-site presence and remote work.