Python Developer
Key Skills
Job Description
About us- ClusterVision’s mission is to lead the market in full-service HPC/AI, Storage, Machine Learning and Cloud enablement. As the enabler of our customers’ complex IT requirements, our customers' success is our success. Over almost 25 years, we have developed, built and serviced some of the fastest and most complex supercomputers in the world, and have won awards for innovation and technology leadership. With headquarters in Hoofddorp- close to Amsterdam in the Netherlands, and extensive coverage across Europe and the Middle-East, our team is made up of knowledgeable and enthusiastic professionals. In addition to our core services, ClusterVision is actively developing TrinityX, our in-house open-source cluster management system. Released as an open-source platform, TrinityX is designed to address the evolving needs of HPC/AI infrastructure management. By offering a flexible and scalable solution, we aim for TrinityX to become the new standard in HPC/AI environments. Alongside it, we are designing, building and exploring a range of AI projects that bring LLMs and machine learning into the Linux and HPC realm. What we’re looking for - A senior Python developer who has worked with real systems — servers and networks — and writes code knowing where it will run Someone who has carried what they built into production: diagnosed it under load, and provided fixes on production servers Someone who wants to shape TrinityX, and to help us design and build the AI projects growing alongside it An engineer who takes a problem statement rather than a ticket: designs it, builds it, ships it, and owns it in production Location: Hoofddorp/Netherlands What you will do and what you can expect- Become a key contributor to TrinityX. Think of Luna and its API, the CLI, node provisioning, image management and packaging — the parts our customers depend on every day Design as well as build: API endpoints, data models, and the structure of a codebase that has to stay maintainable for years Work in the open: TrinityX is open source, so the work you do here is visible to the whole HPC/AI community Solve bugs and assist our engineering department with the problems they hit on real clusters in the field: root cause analysis, patches and permanent solutions Work on custom solutions for and with customers, where what they need sits outside what the standard product covers today Contribute to the AI work ClusterVision is designing, building and exploring — a next-generation RAG system bringing LLM agents into the Linux realm, Python pipelines that extract and analyse metrics, logs and traces, and AI-assisted troubleshooting that makes our engineers measurably faster Help decide which of those ideas are worth pursuing, and turn key ideas and insights from State of The Art publications into things that actually run Own a design end to end: from a vague problem to a defensible architecture, a shipped result, and the operational reality that follows it Make the calls that come with our environment: on-premise and airgapped deployment, clusters of thousands of nodes, mixed Linux distributions, and what customer data may never leave their site Set the engineering standard in a young codebase — tests, packaging, CI, observability — and raise the level of the engineers around you Present and demonstrate your work: it is not uncommon here to give a presentation or a demo of TrinityX, or of the part you built yourself, to colleagues, to customers or at an event Take initiative on work nobody has asked for yet, planting seeds for upcoming ClusterVision projects Our development team works in Scrum, so you will take part in the sprint cycle — planning, refinement and review. Within that cadence the position requires a self-motivated and independent professional who is comfortable owning a piece of work from start to finish. Required skills. You bring at least 8 years of professional software engineering experience You are fluent in Python 3 — an absolute requirement for this role. Other languages are a plus You design and build software, not only automate it: APIs, data models, and code that lives in a product for years You know Linux well — you have run Linux systems, not just developed on them: internals, systemd, permissions and namespaces, packaging (RPM/DEB), and the real differences between the RHEL- and Debian-family distributions You have a good understanding of networking in general — routing, subnetting, DNS and DHCP — and are familiar with IPv6 You place a high value on the quality of your work and on producing clean code that the next person can maintain You understand databases and query languages, and have a sense of what a query costs You have built backend services with FastAPI, Flask, Django or similar, and worked with containers You are comfortable using AI in your own workflow: you know how to offload the parts of the work that should be offloaded, and how to review what comes back — it makes you faster without making you careless You work autonomously: from a vague problem to a defensible design and a shipped result, without being managed through it You hold a Bachelor Degree or Higher (preferably in Computer Science or related fields), or a track record that makes the question irrelevant Nice to have ( you don’t need to check all the boxes, any combination of the below is appreciated ) : HPC/AI or cluster management exposure: Slurm, MPI, InfiniBand, provisioning, parallel filesystems Vue.js and Node.js Monitoring systems: Prometheus, Logstash, Elasticsearch, Grafana, InfluxDB, Jaeger, OpenTelemetry or similar Familiarity with distributed systems, microservices, Docker and/or Kubernetes ML and scientific libraries: scikit-learn, pandas, numpy, matplotlib, PyTorch, TensorFlow Frontend and visualisation technologies like HTML, JavaScript, TypeScript, d3js, or also Gradio and Streamlit Working knowledge of Scrum, or a certification such as PSM I Any exposure to AI topics applied to Linux systems, or Open source contributions Hands-on LLM work — a plus, not a requirement: retrieval pipelines, embeddings and vector search, context design, and evaluating output quality — Qdrant, Haystack, LlamaIndex, Ollama, vLLM ( or also LangChain and similar – if you like that stuff ) AI topics such as: Signal Processing, Anomaly detection, NLP, Entity Recognition and Extraction, Information retrieval and query systems Other skills and characteristics you will need in this job: A high degree of self-motivation and a genuine “can-do” attitude: you go and find the answer rather than waiting for it The judgement to know what to build and what to leave alone Team player who works productively with a wide range of people, and can explain a technical decision to someone who does not share your background A passion for technology. The desire to make a real difference to a successful and rapidly growing organisation. The Offer ClusterVision offers an informal working atmosphere with energetic people who enjoy being part of a rapidly growing and successful organization. We have an open management culture in which we encourage all colleagues to contribute to the process of improving our products, services and processes. We offer competitive pay packages, but more importantly, a exciting place to work where you can develop your skills and build a career. You will be eligible for a Full OTE/Benefits package reflecting the senior nature of this role, your skills, qualifications and experience. If this sounds like a good fit, please send your CV to: [email protected]
Core Responsibilities
Design, build, and maintain TrinityX features and customer-specific solutions, including APIs, data models, provisioning, and packaging, while diagnosing and resolving issues on production clusters. Contribute to AI initiatives involving LLMs, RAG, and analysis pipelines, and own designs from initial problem definition through deployment and operational support.
Requirements
Requires at least 8 years of professional software engineering experience, strong Python 3 skills, hands-on Linux and networking knowledge, and experience building backend services and working with databases and containers. Candidates should be autonomous, produce maintainable software, and have a bachelor’s degree or higher, preferably in computer science or a related field, or equivalent experience.
Benefits
- Competitive Pay
- Full OTE/Benefits Package
About ClusterVision
Industry: Computer Hardware Manufacturing
Company size: 11-50 employees
ClusterVision specialises in high-performance computing (HPC) software. We develop, deliver, and support TrinityX, our open-source cluster management platform, empowering researchers and innovators to harness the full potential of HPC with ease. HPC accelerates scientific discovery. With this in mind, ClusterVision provides advanced software solutions that simplify the management of compute and GPU clusters. Our in-house development team continuously enhances TrinityX, ensuring it integrates seamlessly with leading HPC tools like Slurm, Kubernetes, and OpenHPC. By offering a powerful, adaptable software stack alongside expert support and training, we enable organisations across Europe to optimise their HPC environments, reduce complexity, and focus on groundbreaking research and innovation. HPC accelerates scientific discovery. With this in mind, ClusterVision provides advanced software solutions that simplify the management of compute and GPU clusters. Our in-house development team continuously enhances TrinityX, ensuring it integrates seamlessly with leading HPC tools like Slurm, Kubernetes, and OpenHPC. By offering a powerful, adaptable software stack alongside expert support and training, we enable organisations across Europe to optimise their HPC environments, reduce complexity, and focus on groundbreaking research and innovation.