Senior Platform Engineer
Salva questo lavoro e mantieni la tua ricerca organizzata
Crea un account gratuito per salvare lavori, creare avvisi e tornare a questa inserzione dalla tua dashboard.
Continuando accetti i nostri Termini & Informativa sulla privacy.
Questa posizione è in Domyn
Il processo di selezione sarà interamente gestito Domyn.
--
Senior Platform Engineer We are looking for an experienced Platform Engineer to join our growing team in Milan and help shape the future of our flagship project, Colosseum, one of Europe’s most powerful AI supercomputers, currently in development. Designed to run our proprietary AI models at scale, it forms the compute backbone behind the intelligence we deliver to the world’s most demanding industries. In this role, you will play a key part in ensuring optimal resource utilization across our AI infrastructure. You will build and operate the foundational platform capabilities required to orchestrate workloads and resources with precision and efficiency. Furthermore, you will be responsible for scaling and dynamically redistributing resources across tenants while integrating critical infrastructure metrics into the resource management framework to enable intelligent scheduling and optimization. You will also work closely with Site Reliability Engineers and AI Cloud Engineers to design, implement, and continuously improve the platform. You will embed security controls directly into the platform foundation by implementing firewalls, enforcing least-privilege access policies, and applying network hardening practices, ensuring that security is built in from the ground up—before code is even deployed.
What You Have
- Bachelor's or Master's degree in Computer Science, Software Engineering, Data Science, or a related field.
- At least 6 years of experience as a Platform Engineer or in similar roles.
- Knowledge of NVIDIA infrastructure technologies, including NicO, DSX, and DCGM.
- Expertise in leak detection technologies and monitoring systems.
- Strong experience with container orchestration and virtualization technologies, including Kubernetes implementations for strong multitenancy, bare metal provisioning and system network configuration.
- Knowledge of digital twin technologies and platforms such as NVIDIA Omniverse.
- Solid understanding of MCP/A2A or equivalent
- Proficiency in Python and experience with software development best practices, including version control, testing, and CI/CD pipelines.
- Experience with Run:ai or similar AI workload orchestration and job scheduling frameworks.
- Expertise in parallel file systems, including Weka, NetApp, and Hammerspace.
- Strong knowledge of HPC networking technologies, including Ethernet, InfiniBand, NVLink, and NVIDIA BlueField.
- Experience with workload optimization and resource management frameworks.
Who you are
- A versatile engineer, comfortable operating in complex and fast-paced environments.
- Driven and fearless, you proactively tackle challenges and overcome obstacles with determination.
- A systems thinker, capable of understanding the broader architecture and identifying dependencies across platforms and technologies.
- A collaborative team player who is enthusiastic, curious, and passionate about problem-solving, thriving both independently and within cross-functional teams.
- An effective communicator with strong interpersonal skills, able to engage with diverse stakeholders and foster collaboration.
- Fluent in English and eager to contribute in a multicultural and international environment.
What We Offer
- Compensation. We offer a competitive base salary, as well as the opportunity to receive a variable bonus and company equity. The typical base salary for this role ranges between €50.000 and €70.000, based on experience. As you gain experience and make more significant contributions to the business, your compensation will be reviewed to match your impact. Employment terms are governed by the CCNL Commercio, the Italian National Collective Bargaining Agreement for the Commerce, Distribution and Services sector.
- Learning. When our people grow, we grow with them. That’s why we give every team member a dedicated learning budget to invest in books, courses, conferences, or anything else that fuels their curiosity and supports their role.
- Flexible working. We offer a flexible work policy that lets employees choose to work from home or the office, whatever suits them best.
- Wellbeing. We provide resources to support the mental wellbeing of our team members.
Why work at Domyn?
Ambition, talent and teamwork-this is how we shape your future, combining a start-up mindset with the impact of a corporation. Domyn is an equal opportunity employer. We welcome applicants from all backgrounds and are committed to building a diverse and i