Python Developer

Clustervision — Netherlands · Posted ~8 hours ago

Full-time

Skills

Python HPC AI infrastructure Cloud Machine learning Storage systems AI Machine Learning Storage

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Python development opportunity focused on advanced HPC and AI infrastructure. You will contribute to complex computing, storage, machine learning, and cloud environments, with opportunities to work on scalable cluster-management technology and open-source software.

Highlights

Join a technically focused environment working on high-performance computing, AI, storage, machine learning, and cloud technologies. The role offers exposure to complex infrastructure, open-source development, and scalable cluster-management solutions.

Description

About us- ClusterVision’s mission is to lead the market in full-service HPC/AI, Storage, Machine Learning and Cloud enablement. As the enabler of our customers’ complex IT requirements, our customers' success is our success. Over almost 25 years, we have developed, built and serviced some of the fastest and most complex supercomputers in the world, and have won awards for innovation and technology leadership. With headquarters in Hoofddorp- close to Amsterdam in the Netherlands, and extensive coverage across Europe and the Middle-East, our team is made up of knowledgeable and enthusiastic professionals. In addition to our core services, ClusterVision is actively developing TrinityX, our in-house open-source cluster management system. Released as an open-source platform, TrinityX is designed to address the evolving needs of HPC/AI infrastructure management. By offering a flexible and scalable solution, we aim for TrinityX to become the new standard in HPC/AI environments. Alongside it, we are designing, building and exploring a range of AI projects that bring LLMs and machine learning into the Linux and HPC realm. What we’re looking for - A senior Python developer who has worked with real systems — servers and networks — and writes code knowing where it will runSomeone who has carried what they built into production: diagnosed it under load, and provided fixes on production serversSomeone who wants to shape TrinityX, and to help us design and build the AI projects growing alongside itAn engineer who takes a problem statement rather than a ticket: designs it, builds it, ships it, and owns it in production Location: Hoofddorp/Netherlands What you will do and what you can expect- Become a key contributor to TrinityX. Think of Luna and its API, the CLI, node provisioning, image management and packaging — the parts our customers depend on every dayDesign as well as build: API endpoints, data models, and the structure of a codebase that has to stay maintainable for yearsWork in the open: TrinityX is open source, so the work you do here is visible to the whole HPC/AI communitySolve bugs and assist our engineering department with the problems they hit on real clusters in the field: root cause analysis, patches and permanent solutionsWork on custom solutions for and with customers, where what they need sits outside what the standard product covers todayContribute to the AI work ClusterVision is designing, building and exploring — a next-generation RAG system bringing LLM agents into the Linux realm, Python pipelines that extract and analyse metrics, logs and traces, and AI-assisted troubleshooting that makes our engineers measurably fasterHelp decide which of those ideas are worth pursuing, and turn key ideas and insights from State of The Art publications into things that actually runOwn a design end to end: from a vague problem to a defensible architecture, a shipped result, and the operational reality that follows itMake the calls that come with our environment: on-premise and airgapped deployment, clusters of thousands of nodes, mixed Linux distributions, and what customer data may never leave their siteSet the engineering standard in a young codebase — tests, packaging, CI, observability — and raise the level of the engineers around youPresent and demonstrate your work: it is not uncommon here to give a presentation or a demo of TrinityX, or of the part you built yourself, to colleagues, to customers or at an eventTake initiative on work nobody has asked for yet, planting seeds for upcoming ClusterVision projectsOur development team works in Scrum, so you will take part in the sprint cycle — planning, refinement and review. Within that cadence the position requires a self-motivated and independent professional who is comfortable owning a piece of work from start to finish. Required skills. You bring at least 8 years of professional software engineering experienceYou are fluent in Python 3 — an absolute requirement for this role. Other languages are a plusYou design and build software, not only automate it: APIs, data models, and code that lives in a product for yearsYou know Linux well — you have run Linux systems, not just developed on them: internals, systemd, permissions and namespaces, packaging (RPM/DEB), and the real differences between the RHEL- and Debian-family distributionsYou have a good understanding of networking in general — routing, subnetting, DNS and DHCP — and are familiar with IPv6You place a high value on the quality of your work and on producing clean code that the next person can maintainYou understand databases and query languages, and have a sense of what a query costsYou have built backend services with FastAPI, Flask, Django or similar, and worked with containersYou are comfortable using AI in your own workflow: you know how to offload the parts of the work that should be offloaded, and how to review what comes back — it makes you faster without making you carelessYou work autonomously: from a vague problem to a defensible design and a shipped result, without being managed through itYou hold a Bachelor Degree or Higher (preferably in Computer Science or related fields), or a track record that makes the question irrelevant Nice to have ( you don’t need to check all the boxes, any combination of the below is appreciated ) : HPC/AI or cluster management exposure: Slurm, MPI, InfiniBand, provisioning, parallel filesystemsVue.js and Node.jsMonitoring systems: Prometheus, Logstash, Elasticsearch, Grafana, InfluxDB, Jaeger, OpenTelemetry or similarFamiliarity with distributed systems, microservices, Docker and/or KubernetesML and scientific libraries: scikit-learn, pandas, numpy, matplotlib, PyTorch, TensorFlowFrontend and visualisation technologies like HTML, JavaScript, TypeScript, d3js, or also Gradio and StreamlitWorking knowledge of Scrum, or a certification such as PSM IAny exposure to AI topics applied to Linux systems, or Open source contributionsHands-on LLM work — a plus, not a requirement: retrieval pipelines, embeddings and vector search, context design, and evaluating output quality — Qdrant, Haystack, LlamaIndex, Ollama, vLLM ( or also LangChain and similar – if you like that stuff )AI topics such as: Signal Processing, Anomaly detection, NLP, Entity Recognition and Extraction, Information retrieval and query systems Other skills and characteristics you will need in this job: A high degree of self-motivation and a genuine “can-do” attitude: you go and find the answer rather than waiting for itThe judgement to know what to build and what to leave aloneTeam player who works productively with a wide range of people, and can explain a technical decision to someone who does not share your backgroundA passion for technology.The desire to make a real difference to a successful and rapidly growing organisation. The Offer ClusterVision offers an informal working atmosphere with energetic people who enjoy being part of a rapidly growing and successful organization. We have an open management culture in which we encourage all colleagues to contribute to the process of improving our products, services and processes. We offer competitive pay packages, but more importantly, a exciting place to work where you can develop your skills and build a career. You will be eligible for a Full OTE/Benefits package reflecting the senior nature of this role, your skills, qualifications and experience. If this sounds like a good fit, please send your CV to: surbi@tauruseu.com