Python Developer
RemoteHoofddorp, North Holland, NetherlandsseniorFull-time
- Posted
- today
- Source
- LinkedIn (remote, Europe)
- Field
- Engineering
Skills
Machine LearningElasticsearchJavaScriptTypeScriptKubernetesTensorFlowLeadershipNode.jsPyTorchVuePythonDockerPandasScrumNumPyLinuxNext.jsHTMLLLMsNLPGoAI
Description
About us-
ClusterVision’s mission is to lead the market in full-service HPC/AI, Storage, Machine Learning and Cloud enablement.
As the enabler of our customers’ complex IT requirements, our customers' success is our success. Over almost 25 years, we have developed, built and serviced some of the fastest and most complex supercomputers in the world, and have won awards for innovation and technology leadership. With headquarters in Hoofddorp- close to Amsterdam in the Netherlands, and extensive coverage across Europe and the Middle-East, our team is made up of knowledgeable and enthusiastic professionals.
In addition to our core services, ClusterVision is actively developing TrinityX, our in-house open-source cluster management system. Released as an open-source platform, TrinityX is designed to address the evolving needs of HPC/AI infrastructure management. By offering a flexible and scalable solution, we aim for TrinityX to become the new standard in HPC/AI environments. Alongside it, we are designing, building and exploring a range of AI projects that bring LLMs and machine learning into the Linux and HPC realm.
What we’re looking for -
- A senior Python developer who has worked with real systems — servers and networks — and writes code knowing where it will run
- Someone who has carried what they built into production: diagnosed it under load, and provided fixes on production servers
- Someone who wants to shape TrinityX, and to help us design and build the AI projects growing alongside it
- An engineer who takes a problem statement rather than a ticket: designs it, builds it, ships it, and owns it in production
Location: Hoofddorp/Netherlands
What you will do and what you can expect-
- Become a key contributor to TrinityX. Think of Luna and its API, the CLI, node provisioning, image management and packaging — the parts our customers depend on every day
- Design as well as build: API endpoints, data models, and the structure of a codebase that has to stay maintainable for years
- Work in the open: TrinityX is open source, so the work you do here is visible to the whole HPC/AI community
- Solve bugs and assist our engineering department with the problems they hit on real clusters in the field: root cause analysis, patches and permanent solutions
- Work on custom solutions for and with customers, where what they need sits outside what the standard product covers today
- Contribute to the AI work ClusterVision is designing, building and exploring — a next-generation RAG system bringing LLM agents into the Linux realm, Python pipelines that extract and analyse metrics, logs and traces, and AI-assisted troubleshooting that makes our engineers measurably faster
- Help decide which of those ideas are worth pursuing, and turn key ideas and insights from State of The Art publications into things that actually run
- Own a design end to end: from a vague problem to a defensible architecture, a shipped result, and the operational reality that follows it
- Make the calls that come with our environment: on-premise and airgapped deployment, clusters of thousands of nodes, mixed Linux distributions, and what customer data may never leave their site
- Set the engineering standard in a young codebase — tests, packaging, CI, observability — and raise the level of the engineers around you
- Present and demonstrate your work: it is not uncommon here to give a presentation or a demo of TrinityX, or of the part you built yourself, to colleagues, to customers or at an event
- Take initiative on work nobody has asked for yet, planting seeds for upcoming ClusterVision projects
Our development team works in Scrum, so you will take part in the sprint cycle — planning, refinement and review. Within that cadence the position requires a self-motivated and independent professional who is comfortable owning a piece of work from start to finish.
Required skills.
- You bring at least 8 years of professional software engineering experience
- You are fluent in Python 3 — an absolute requirement for this role. Other languages are a plus
- You design and build software, not only automate it: APIs, data models, and code that lives in a product for years
- You know Linux well — you have run Linux systems, not just developed on them: internals, systemd, permissions and namespaces, packaging (RPM/DEB), and the real differences between the RHEL- and Debian-family distributions
- You have a good understanding of networking in general — routing, subnetting, DNS and DHCP — and are familiar with IPv6
- You place a high value on the quality of your work and on producing clean code that the next person can maintain
- You understand databases and query languages, and have a sense of what a query costs
- You have built backend services with FastAPI, Flask, Django or similar, and worked with containers
- You are comfortable using AI in your own workflow: you know how to offload the parts of the work that should be offloaded, and how to review what comes back — it makes you faster without making you careless
- You work autonomously: from a vague problem to a defensible design and a shipped result, without being managed through it
- You hold a Bachelor Degree or Higher (preferably in Computer Science or related fields), or a track record that makes the question irrelevant
Nice to have ( you don’t need to check all the boxes, any combination of the below is appreciated ) :
- HPC/AI or cluster management exposure: Slurm, MPI, InfiniBand, provisioning, parallel filesystems
- Vue.js and Node.js
- Monitoring systems: Prometheus, Logstash, Elasticsearch, Grafana, InfluxDB, Jaeger, OpenTelemetry or similar
- Familiarity with distributed systems, microservices, Docker and/or Kubernetes
- ML and scientific libraries: scikit-learn, pandas, numpy, matplotlib, PyTorch, TensorFlow
- Frontend and visualisation technologies like HTML, JavaScript, TypeScript, d3js, or also Gradio and Streamlit
- Working knowledge of Scrum, or a certification such as PSM I
- Any exposure to AI topics applied to Linux systems, or Open source contributions
- Hands-on LLM work — a plus, not a requirement: retrieval pipelines, embeddings and vector search, context design, and evaluating output quality — Qdrant, Haystack, LlamaIndex, Ollama, vLLM ( or also LangChain and similar – if you like that stuff )
- AI topics such as: Signal Processing, Anomaly detection, NLP, Entity Recognition and Extraction, Information retrieval and query systems
Other skills and characteristics you will need in this job:
- A high degree of self-motivation and a genuine “can-do” attitude: you go and find the answer rather than waiting for it
- The judgement to know what to build and what to leave alone
- Team player who works productively with a wide range of people, and can explain a technical decision to someone who does not share your background
- A passion for technology.
- The desire to make a real difference to a successful and rapidly growing organisation.
The Offer
ClusterVision offers an informal working atmosphere with energetic people who enjoy being part of a rapidly growing and successful organization. We have an open management culture in which we encourage all colleagues to contribute to the process of improving our products, services and processes. We offer competitive pay packages, but more importantly, a exciting place to work where you can develop your skills and build a career. You will be eligible for a Full OTE/Benefits package reflecting the senior nature of this role, your skills, qualifications and experience.
If this sounds like a good fit, please send your CV to: surbi@tauruseu.com
JobMatch aggregates public listings. Always apply through the original posting.